mi research lab
Yasin Abbasi-Yadkori

Yasin Abbasi-Yadkori

founder | senior researcher

Yasin has contributed to the advancement of machine intelligence through the foundational work in online reinforcement learning, uncertainty quantification and exploration in Richard Sutton's lab from 2006. His NeurIPS 2011 paper, "Improved Algorithms for Linear Stochastic Bandits," informs much of modern RL exploration. At Google DeepMind he applied UQ research to hallucination detection in large language models. His recent work on hierarchical and recursive architectures led to a Nature-invited paper and related generation advancements. He founded the machine intelligence research lab to advance collaboration between human and AI through the application of online RL.

selected work

Google Scholar
Improved algorithms for linear stochastic bandits
Y Abbasi-Yadkori, C Szepesvári, D Pál · NeurIPS, 2011
2781 cited
Regret Bounds for the Adaptive Control of Linear Quadratic Systems
Y Abbasi-Yadkori, C Szepesvári · COLT, 2011
537 cited
To Believe or Not to Believe Your LLM: Iterative Prompting for Estimating Epistemic Uncertainty
Y Abbasi-Yadkori, I Kuzborskij, A György, C Szepesvári · NeurIPS, 2024
227 cited