Skip to content

Thomas Mesnard

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

6

Venues

2

Active years

2018–2024

Best venue rank

A*

Where they publish

Papers

6 indexed papers, newest first.

YearVenueTitleAuthors
2024ICMLRLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback.Harrison Lee, Samrat Phatale, Hassan Mansoor, Thomas Mesnard, Johan Ferret, Kellie Lu, Colton Bishop, Ethan Hall, Victor Carbune, Abhinav Rastogi, Sushant Prakash
2024ICMLNash Learning from Human Feedback.Rmi Munos, Michal Valko, Daniele Calandriello, Mohammad Gheshlaghi Azar, Mark Rowland, Daniel Guo, Yunhao Tang, Matthieu Geist, Thomas Mesnard, Cme Fiegel, Andrea Michi, Marco Selvi, Sertan Girgin, Nikola Momchev, Olivier Bachem, Daniel J. Mankowitz, Doina Precup, Bilal Piot
2023ICMLCuriosity in Hindsight: Intrinsic Exploration in Stochastic Environments.Daniel Jarrett, Corentin Tallec, Florent Altch, Thomas Mesnard, Rmi Munos, Michal Valko
2023ICMLQuantile Credit Assignment.Thomas Mesnard, Wenqi Chen, Alaa Saade, Yunhao Tang, Mark Rowland, Theophane Weber, Clare Lyle, Audrunas Gruslys, Michal Valko, Will Dabney, Georg Ostrovski, Eric Moulines, Rmi Munos
2021ICMLCounterfactual Credit Assignment in Model-Free Reinforcement Learning.Thomas Mesnard, Theophane Weber, Fabio Viola, Shantanu Thakoor, Alaa Saade, Anna Harutyunyan, Will Dabney, Thomas S. Stepleton, Nicolas Heess, Arthur Guez, Eric Moulines, Marcus Hutter, Lars Buesing, Rmi Munos
2018ICLRExtending the Framework of Equilibrium Propagation to General Dynamics.Benjamin Scellier, Anirudh Goyal, Jonathan Binas, Thomas Mesnard, Yoshua Bengio