Yasin Abbasi-Yadkori
Yasin Abbasi-Yadkori
DeepMind
Adresse e-mail validée de google.com - Page d'accueil
Titre
Citée par
Citée par
Année
Improved algorithms for linear stochastic bandits
Y Abbasi-Yadkori, C Szepesvári, D Pal
Advances in Neural Information Processing Systems, 2312-2320, 2011
6572011
Regret Bounds for the Adaptive Control of Linear Quadratic Systems.
Y Abbasi-Yadkori, C Szepesvári
COLT, 1-26, 2011
1672011
Fast approximate nearest-neighbor search with k-nearest neighbor graph
K Hajebi, Y Abbasi-Yadkori, H Shahbazi, H Zhang
Twenty-Second International Joint Conference on Artificial Intelligence, 2011
1642011
Online-to-Confidence-Set Conversions and Application to Sparse Stochastic Bandits.
Y Abbasi-Yadkori, D Pal, C Szepesvari
AISTATS 22, 1-9, 2012
912012
Online Learning in Markov Decision Processes with Adversarially Chosen Transition Probability Distributions
Y Abbasi-Yadkori, P Bartlett, V Kanade, Y Seldin, C Szepesvari
Neural Information Processing Systems, 2013
62*2013
Sharp Convergence Rates for Langevin Dynamics in the Nonconvex Setting
X Cheng, NS Chatterji, Y Abbasi-Yadkori, PL Bartlett, MI Jordan
arXiv preprint arXiv:1805.01648, 2018
612018
Model-Free Linear Quadratic Control via Reduction to Expert Prediction
Y Abbasi-Yadkori, N Lazic, C Szepesvari
The 22nd International Conference on Artificial Intelligence and Statistics, 2019
58*2019
Bayesian Optimal Control of Smoothly Parameterized Systems
Y Abbasi-Yadkori, C Szepesvári
Proceedings of the Conference on Uncertainty in Artificial Intelligence, 2015
40*2015
Prediction with limited advice and multiarmed bandits with paid observations
Y Seldin, P Bartlett, K Crammer, Y Abbasi-Yadkori
International Conference on Machine Learning, 280-287, 2014
372014
Conservative contextual linear bandits
A Kazerouni, M Ghavamzadeh, YA Yadkori, B Van Roy
Advances in Neural Information Processing Systems, 3910-3919, 2017
352017
Online Learning for Linearly Parametrized Control Problems
Y Abbasi-Yadkori
University of Alberta, 2012
332012
Linear Programming for Large-Scale Markov Decision Problems
Y Abbasi-Yadkori, P Bartlett, A Malek
Proceedings of the 31st International Conference on Machine Learning (ICML …, 2014
31*2014
Online least squares estimation with self-normalized processes: An application to bandit problems
Y Abbasi-Yadkori, D Pál, C Szepesvári
arXiv preprint arXiv:1102.2670, 2011
312011
POLITEX: Regret bounds for policy iteration using expert prediction
Y Abbasi-Yadkori, P Bartlett, K Bhatia, N Lazic, C Szepesvári, G Weisz
Proceedings of the 36th International Conference on Machine Learning 97 …, 2019
212019
POLITEX: Regret Bounds for Policy Iteration Using Expert Prediction
Y Abbasi-Yadkori, PL Bartlett, K Bhatia, N Lazic, C Szepesvári, G Weisz
212019
Learning when to stop thinking and do something!
B Póczos, Y Abbasi-Yadkori, C Szepesvári, R Greiner, N Sturtevant
Proceedings of the 26th Annual International Conference on Machine Learning …, 2009
212009
Forced-exploration based algorithms for playing in stochastic linear bandits
Y Abbasi-Yadkori, A Antos, C Szepesvári
COLT Workshop on On-line Learning with Limited Feedback, 2009
192009
Minimax Time Series Prediction
WM Koolen, A Malek, PL Bartlett, Y Abbasi-Yadkori
Advances in Neural Information Processing Systems, 2548-2556, 2015
182015
Tracking Adversarial Targets
Y Abbasi-Yadkori, QUT EDU, P Bartlett, B EDU, V Kanade
182014
Extending rapidly-exploring random trees for asymptotically optimal anytime motion planning
Y Abbasi-Yadkori, J Modayil, C Szepesvari
2010 IEEE/RSJ International Conference on Intelligent Robots and Systems …, 2010
172010
Le système ne peut pas réaliser cette opération maintenant. Veuillez réessayer plus tard.
Articles 1–20