Skip to content

Alessandro Lazaric

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

69

Venues

13

Active years

2007–2025

Best venue rank

A*

Where they publish

Papers

69 indexed papers, newest first.

YearVenueTitleAuthors
2025ICLRZero-Shot Whole-Body Humanoid Control via Behavioral Foundation Models.Andrea Tirinzoni, Ahmed Touati, Jesse Farebrother, Mateusz Guzek, Anssi Kanervisto, Yingchen Xu, Alessandro Lazaric, Matteo Pirotta
2025ICMLTemporal Difference Flows.Jesse Farebrother, Matteo Pirotta, Andrea Tirinzoni, Rmi Munos, Alessandro Lazaric, Ahmed Touati
2024ICLRFast Imitation via Behavior Foundation Models.Matteo Pirotta, Andrea Tirinzoni, Ahmed Touati, Alessandro Lazaric, Yann Ollivier
2024ICMLSimple Ingredients for Offline Reinforcement Learning.Edoardo Cetin, Andrea Tirinzoni, Matteo Pirotta, Alessandro Lazaric, Yann Ollivier, Ahmed Touati
2023AISTATSOn the Complexity of Representation Learning in Contextual Linear Bandits.Andrea Tirinzoni, Matteo Pirotta, Alessandro Lazaric
2023ALTReaching Goals is Hard: Settling the Sample Complexity of the Stochastic Shortest Path.Liyu Chen, Andrea Tirinzoni, Matteo Pirotta, Alessandro Lazaric
2023ICLRContextual bandits with concave rewards, and an application to fair ranking.Virginie Do, Elvis Dohmatob, Matteo Pirotta, Alessandro Lazaric, Nicolas Usunier
2023ICLRLinear Convergence of Natural Policy Gradient Methods with Log-Linear Policies.Rui Yuan, Simon Shaolei Du, Robert M. Gower, Alessandro Lazaric, Lin Xiao
2023ICMLLayered State Discovery for Incremental Autonomous Exploration.Liyu Chen, Andrea Tirinzoni, Alessandro Lazaric, Matteo Pirotta
2022AISTATSTop K Ranking for Multi-Armed Bandit with Noisy Evaluations.Evrard Garcelon, Vashist Avadhanula, Alessandro Lazaric, Matteo Pirotta
2022AISTATSAdaptive Multi-Goal Exploration.Jean Tarbouriech, Omar Darwiche Domingues, Pierre Mnard, Matteo Pirotta, Michal Valko, Alessandro Lazaric
2022AISTATSA general sample complexity analysis of vanilla policy gradient.Rui Yuan, Robert M. Gower, Alessandro Lazaric
2022CoRLLearning Goal-Conditioned Policies Offline with Self-Supervised Reward Shaping.Lina Mezghani, Sainbayar Sukhbaatar, Piotr Bojanowski, Alessandro Lazaric, Karteek Alahari
2022ICLRDirect then Diffuse: Incremental Unsupervised Skill Discovery for State Covering and Goal Reaching.Pierre-Alexandre Kamienny, Jean Tarbouriech, Sylvain Lamprier, Alessandro Lazaric, Ludovic Denoyer
2022ICLRA Reduction-Based Framework for Conservative Bandits and Reinforcement Learning.Yunchang Yang, Tianhao Wu, Han Zhong, Evrard Garcelon, Matteo Pirotta, Alessandro Lazaric, Liwei Wang, Simon Shaolei Du
2022ICLRMastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning.Denis Yarats, Rob Fergus, Alessandro Lazaric, Lerrel Pinto
2022ICMLScaling Gaussian Process Optimization by Evaluating a Few Unique Candidates Multiple Times.Daniele Calandriello, Luigi Carratino, Alessandro Lazaric, Michal Valko, Lorenzo Rosasco
2022UAITemporal abstractions-augmented temporally contrastive learning: An alternative to the Laplacian in RL.Akram Erraqabi, Marlos C. Machado, Mingde Zhao, Sainbayar Sukhbaatar, Alessandro Lazaric, Ludovic Denoyer, Yoshua Bengio
2021ALTSample Complexity Bounds for Stochastic Shortest Path with a Generative Model.Jean Tarbouriech, Matteo Pirotta, Michal Valko, Alessandro Lazaric
2021ICMLLeveraging Good Representations in Linear Contextual Bandits.Matteo Papini, Andrea Tirinzoni, Marcello Restelli, Alessandro Lazaric, Matteo Pirotta
2021ICMLReinforcement Learning with Prototypical Representations.Denis Yarats, Rob Fergus, Alessandro Lazaric, Lerrel Pinto
2020AAAIImproved Algorithms for Conservative Exploration in Bandits.Evrard Garcelon, Mohammad Ghavamzadeh, Alessandro Lazaric, Matteo Pirotta
2020AISTATSConservative Exploration in Reinforcement Learning.Evrard Garcelon, Mohammad Ghavamzadeh, Alessandro Lazaric, Matteo Pirotta
2020AISTATSA single algorithm for both restless and rested rotting bandits.Julien Seznec, Pierre Mnard, Alessandro Lazaric, Michal Valko
2020AISTATSA Novel Confidence-Based Algorithm for Structured Bandits.Andrea Tirinzoni, Alessandro Lazaric, Marcello Restelli
2020AISTATSFrequentist Regret Bounds for Randomized Least-Squares Value Iteration.Andrea Zanette, David Brandfonbrener, Emma Brunskill, Matteo Pirotta, Alessandro Lazaric
2020ICMLEfficient Optimistic Exploration in Linear-Quadratic Regulators via Lagrangian Relaxation.Marc Abeille, Alessandro Lazaric
2020ICMLNear-linear time Gaussian process optimization with adaptive batching and resparsification.Daniele Calandriello, Luigi Carratino, Alessandro Lazaric, Michal Valko, Lorenzo Rosasco
2020ICMLMeta-learning with Stochastic Linear Bandits.Leonardo Cella, Alessandro Lazaric, Massimiliano Pontil
2020ICMLNo-Regret Exploration in Goal-Oriented Reinforcement Learning.Jean Tarbouriech, Evrard Garcelon, Michal Valko, Matteo Pirotta, Alessandro Lazaric
2020ICMLLearning Near Optimal Policies with Low Inherent Bellman Error.Andrea Zanette, Alessandro Lazaric, Mykel J. Kochenderfer, Emma Brunskill
2020UAIActive Model Estimation in Markov Decision Processes.Jean Tarbouriech, Shubhanshu Shekhar, Matteo Pirotta, Mohammad Ghavamzadeh, Alessandro Lazaric
2019ACLWord-order Biases in Deep-agent Emergent Communication.Rahma Chaabouni, Eugene Kharitonov, Alessandro Lazaric, Emmanuel Dupoux, Marco Baroni
2019AISTATSRotting bandits are no harder than stochastic ones.Julien Seznec, Andrea Locatelli, Alexandra Carpentier, Alessandro Lazaric, Michal Valko
2019AISTATSActive Exploration in Markov Decision Processes.Jean Tarbouriech, Alessandro Lazaric
2019COLTGaussian Process Optimization with Adaptive Sketching: Scalable and No Regret.Daniele Calandriello, Luigi Carratino, Alessandro Lazaric, Michal Valko, Lorenzo Rosasco
2018ICMLImproved Regret Bounds for Thompson Sampling in Linear Quadratic Control Problems.Marc Abeille, Alessandro Lazaric
2018ICMLImproved Large-Scale Graph Learning through Ridge Spectral Sparsification.Daniele Calandriello, Ioannis Koutis, Alessandro Lazaric, Michal Valko
2018ICMLEfficient Bias-Span-Constrained Exploration-Exploitation in Reinforcement Learning.Ronan Fruit, Matteo Pirotta, Alessandro Lazaric, Ronald Ortner
2017AAAIParallel Higher Order Alternating Least Square for Tensor Recommender System.Romain Warlop, Alessandro Lazaric, Jrmie Mary
2017AISTATSLinear Thompson Sampling Revisited.Marc Abeille, Alessandro Lazaric
2017AISTATSThompson Sampling for Linear-Quadratic Control Problems.Marc Abeille, Alessandro Lazaric
2017AISTATSDistributed Adaptive Sampling for Kernel Matrix Approximation.Daniele Calandriello, Alessandro Lazaric, Michal Valko
2017AISTATSTrading off Rewards and Errors in Multi-Armed Bandits.Akram Erraqabi, Alessandro Lazaric, Michal Valko, Emma Brunskill, Yun-En Liu
2017AISTATSExploration-Exploitation in MDPs with Options.Ronan Fruit, Alessandro Lazaric
2017ICMLSecond-Order Kernel Online Convex Optimization with Adaptive Sketching.Daniele Calandriello, Alessandro Lazaric, Michal Valko
2017ICMLActive Learning for Accurate Estimation of Linear Models.Carlos Riquelme, Mohammad Ghavamzadeh, Alessandro Lazaric
2016AISTATSImproved Learning Complexity in Combinatorial Pure Exploration Bandits.Victor Gabillon, Alessandro Lazaric, Mohammad Ghavamzadeh, Ronald Ortner, Peter L. Bartlett
2016COLTReinforcement Learning of POMDPs using Spectral Methods.Kamyar Azizzadenesheli, Alessandro Lazaric, Animashree Anandkumar
2016COLTOpen Problem: Approximate Planning of POMDPs in the class of Memoryless Policies.Kamyar Azizzadenesheli, Alessandro Lazaric, Animashree Anandkumar
2016UAIAnalysis of Nystrm method with sequential ridge leverage scores.Daniele Calandriello, Alessandro Lazaric, Michal Valko
2015IJCAIMaximum Entropy Semi-Supervised Inverse Reinforcement Learning.Julien Audiffren, Michal Valko, Alessandro Lazaric, Mohammad Ghavamzadeh
2015IJCAIDirect Policy Iteration with Demonstrations.Jessica Chemali, Alessandro Lazaric
2015ISITThe replacement bootstrap for dependent data.Amir Sani, Alessandro Lazaric, Daniil Ryabko
2014ICMLOnline Stochastic Optimization under Correlated Bandit Feedback.Mohammad Gheshlaghi Azar, Alessandro Lazaric, Emma Brunskill
2012AAAIConservative and Greedy Approaches to Classification-Based Policy Iteration.Mohammad Ghavamzadeh, Alessandro Lazaric
2012AAMASA truthful learning mechanism for multi-slot sponsored search auctions with externalities.Nicola Gatti, Alessandro Lazaric, Francesco Trov
2012ICMLA Dantzig Selector Approach to Temporal Difference Learning.Matthieu Geist, Bruno Scherrer, Alessandro Lazaric, Mohammad Ghavamzadeh
2011ALTUpper-Confidence-Bound Algorithms for Active Learning in Multi-armed Bandits.Alexandra Carpentier, Alessandro Lazaric, Mohammad Ghavamzadeh, Rmi Munos, Peter Auer
2011ICMLClassification-based Policy Iteration with a Critic.Victor Gabillon, Alessandro Lazaric, Mohammad Ghavamzadeh, Bruno Scherrer
2011ICMLFinite-Sample Analysis of Lasso-TD.Mohammad Ghavamzadeh, Alessandro Lazaric, Rmi Munos, Matthew W. Hoffman
2010ICMLBayesian Multi-Task Reinforcement Learning.Alessandro Lazaric, Mohammad Ghavamzadeh
2010ICMLAnalysis of a Classification-based Policy Iteration Algorithm.Alessandro Lazaric, Mohammad Ghavamzadeh, Rmi Munos
2010ICMLFinite-Sample Analysis of LSTD.Alessandro Lazaric, Mohammad Ghavamzadeh, Rmi Munos
2009COLTHybrid Stochastic-Adversarial On-line Learning.Alessandro Lazaric, Rmi Munos
2009ICMLWorkshop summary: On-line learning with limited feedback.Jean-Yves Audibert, Peter Auer, Alessandro Lazaric, Rmi Munos, Daniil Ryabko, Csaba Szepesvri
2008ICMLTransfer of samples in batch reinforcement learning.Alessandro Lazaric, Marcello Restelli, Andrea Bonarini
2007AAMASBifurcation Analysis of Reinforcement Learning Agents in the Selten's Horse Game.Alessandro Lazaric, Enrique Munoz de Cote, Fabio Dercole, Marcello Restelli
2007ICINCOPiecewise constant reinforcement learning for robotic applications.Andrea Bonarini, Alessandro Lazaric, Marcello Restelli