| 2025 | COLT | Thompson Sampling for Bandit Convex Optimisation. | Alireza Bakhtiari, Tor Lattimore, Csaba Szepesvri |
| 2024 | COLT | Online Newton Method for Bandit Convex Optimisation Extended Abstract. | Hidde Fokkema, Dirk van der Hoeven, Tor Lattimore, Jack J. Mayo |
| 2023 | COLT | A Second-Order Method for Stochastic Bandit Convex Optimisation. | Tor Lattimore, Andrs Gyrgy |
| 2023 | COLT | A Lower Bound for Linear and Kernel Regression with Adaptive Covariates. | Tor Lattimore |
| 2023 | ICML | Distributed Contextual Linear Bandits with Minimax Optimal Communication Cost. | Sanae Amani, Tor Lattimore, Andrs Gyrgy, Lin Yang |
| 2023 | ICML | Leveraging Demonstrations to Improve Online Learning: Quality Matters. | Botao Hao, Rahul Jain, Tor Lattimore, Benjamin Van Roy, Zheng Wen |
| 2022 | COLT | Minimax Regret for Partial Monitoring: Infinite Outcomes and Rustichini's Regret. | Tor Lattimore |
| 2022 | COLT | Return of the bias: Almost minimax optimal high probability bounds for adversarial linear bandits. | Julian Zimmert, Tor Lattimore |
| 2022 | ICML | Contextual Information-Directed Sampling. | Botao Hao, Tor Lattimore, Chao Qin |
| 2021 | AAAI | Gated Linear Networks. | Joel Veness, Tor Lattimore, David Budden, Avishkar Bhoopchand, Christopher Mattern, Agnieszka Grabska-Barwinska, Eren Sezener, Jianan Wang, Peter Toth, Simon Schmitt, Marcus Hutter |
| 2021 | AISTATS | Online Sparse Reinforcement Learning. | Botao Hao, Tor Lattimore, Csaba Szepesvri, Mengdi Wang |
| 2021 | COLT | Asymptotically Optimal Information-Directed Sampling. | Johannes Kirschner, Tor Lattimore, Claire Vernade, Csaba Szepesvri |
| 2021 | COLT | Improved Regret for Zeroth-Order Stochastic Convex Bandits. | Tor Lattimore, Andrs Gyrgy |
| 2021 | COLT | Mirror Descent and the Information Ratio. | Tor Lattimore, Andrs Gyrgy |
| 2021 | ICML | Sparse Feature Selection Makes Batch Reinforcement Learning More Sample Efficient. | Botao Hao, Yaqi Duan, Tor Lattimore, Csaba Szepesvri, Mengdi Wang |
| 2021 | ICML | On the Optimality of Batch Policy Optimization Algorithms. | Chenjun Xiao, Yifan Wu, Jincheng Mei, Bo Dai, Tor Lattimore, Lihong Li, Csaba Szepesvri, Dale Schuurmans |
| 2021 | UAI | Matrix games with bandit feedback. | Brendan O'Donoghue, Tor Lattimore, Ian Osband |
| 2020 | AISTATS | Adaptive Exploration in Linear Contextual Bandit. | Botao Hao, Tor Lattimore, Csaba Szepesvri |
| 2020 | COLT | Information Directed Sampling for Linear Partial Monitoring. | Johannes Kirschner, Tor Lattimore, Andreas Krause |
| 2020 | COLT | Exploration by Optimisation in Partial Monitoring. | Tor Lattimore, Csaba Szepesvri |
| 2020 | ICLR | Behaviour Suite for Reinforcement Learning. | Ian Osband, Yotam Doron, Matteo Hessel, John Aslanides, Eren Sezener, Andre Saraiva, Katrina McKinney, Tor Lattimore, Csaba Szepesvri, Satinder Singh, Benjamin Van Roy, Richard S. Sutton, David Silver, Hado van Hasselt |
| 2020 | ICML | Learning with Good Feature Representations in Bandits and in RL with a Generative Model. | Tor Lattimore, Csaba Szepesvri, Gellrt Weisz |
| 2020 | ICML | Linear bandits with Stochastic Delayed Feedback. | Claire Vernade, Alexandra Carpentier, Tor Lattimore, Giovanni Zappella, Beyza Ermis, Michael Brckner |
| 2019 | AIES | Degenerate Feedback Loops in Recommender Systems. | Ray Jiang, Silvia Chiappa, Tor Lattimore, Andrs Gyrgy, Pushmeet Kohli |
| 2019 | ALT | Cleaning up the neighborhood: A full classification for adversarial partial monitoring. | Tor Lattimore, Csaba Szepesvri |
| 2019 | COLT | An Information-Theoretic Approach to Minimax Regret in Partial Monitoring. | Tor Lattimore, Csaba Szepesvri |
| 2019 | ICML | Garbage In, Reward Out: Bootstrapping Exploration in Multi-Armed Bandits. | Branislav Kveton, Csaba Szepesvri, Sharan Vaswani, Zheng Wen, Tor Lattimore, Mohammad Ghavamzadeh |
| 2019 | ICML | Online Learning to Rank with Features. | Shuai Li, Tor Lattimore, Csaba Szepesvri |
| 2019 | IJCAI | Iterative Budgeted Exponential Search. | Malte Helmert, Tor Lattimore, Levi H. S. Lelis, Laurent Orseau, Nathan R. Sturtevant |
| 2019 | UAI | BubbleRank: Safe Online Learning to Re-Rank via Implicit Click Feedback. | Chang Li, Branislav Kveton, Tor Lattimore, Ilya Markov, Maarten de Rijke, Csaba Szepesvri, Masrour Zoghi |
| 2019 | UAI | On First-Order Bounds, Variance and Gap-Dependent Bounds for Adversarial Bandits. | Roman Pogodin, Tor Lattimore |
| 2017 | AISTATS | The End of Optimism? An Asymptotic Analysis of Finite-Armed Linear Bandits. | Tor Lattimore, Csaba Szepesvri |
| 2017 | ALT | Soft-Bayes: Prod for Mixtures of Experts with Log-Loss. | Laurent Orseau, Tor Lattimore, Shane Legg |
| 2017 | IJCAI | On Thompson Sampling and Asymptotic Optimality. | Jan Leike, Tor Lattimore, Laurent Orseau, Marcus Hutter |
| 2016 | COLT | Regret Analysis of the Finite-Horizon Gittins Index Strategy for Multi-Armed Bandits. | Tor Lattimore |
| 2016 | ICML | Conservative Bandits. | Yifan Wu, Roshan Shariff, Tor Lattimore, Csaba Szepesvri |
| 2016 | UAI | Thompson Sampling is Asymptotically Optimal in General Environments. | Jan Leike, Tor Lattimore, Laurent Orseau, Marcus Hutter |
| 2014 | ALT | On Learning the Optimal Waiting Time. | Tor Lattimore, Andrs Gyrgy, Csaba Szepesvri |
| 2014 | ALT | Bayesian Reinforcement Learning with Exploration. | Tor Lattimore, Marcus Hutter |
| 2014 | CEC | Free Lunch for optimisation under the universal distribution. | Tom Everitt, Tor Lattimore, Marcus Hutter |
| 2014 | UAI | Optimal Resource Allocation with Semi-Bandit Feedback. | Tor Lattimore, Koby Crammer, Csaba Szepesvri |
| 2013 | ALT | Concentration and Confidence for Discrete Bayesian Sequence Predictors. | Tor Lattimore, Marcus Hutter, Peter Sunehag |
| 2013 | ALT | Universal Knowledge-Seeking Agents for Stochastic Environments. | Laurent Orseau, Tor Lattimore, Marcus Hutter |
| 2013 | ICML | The Sample-Complexity of General Reinforcement Learning. | Tor Lattimore, Marcus Hutter, Peter Sunehag |
| 2013 | TAMC | On Martin-Lf Convergence of Solomonoff's Mixture. | Tor Lattimore, Marcus Hutter |
| 2012 | ALT | PAC Bounds for Discounted MDPs. | Tor Lattimore, Marcus Hutter |
| 2011 | ALT | Asymptotically Optimal Agents. | Tor Lattimore, Marcus Hutter |
| 2011 | ALT | Time Consistent Discounting. | Tor Lattimore, Marcus Hutter |
| 2011 | ALT | Universal Prediction of Selected Bits. | Tor Lattimore, Marcus Hutter, Vaibhav Gavane |