Thomas Mesnard
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
6
Venues
2
Active years
2018–2024
Best venue rank
A*
Where they publish
Papers
6 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2024 | ICML | RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback. | Harrison Lee, Samrat Phatale, Hassan Mansoor, Thomas Mesnard, Johan Ferret, Kellie Lu, Colton Bishop, Ethan Hall, Victor Carbune, Abhinav Rastogi, Sushant Prakash |
| 2024 | ICML | Nash Learning from Human Feedback. | Rmi Munos, Michal Valko, Daniele Calandriello, Mohammad Gheshlaghi Azar, Mark Rowland, Daniel Guo, Yunhao Tang, Matthieu Geist, Thomas Mesnard, Cme Fiegel, Andrea Michi, Marco Selvi, Sertan Girgin, Nikola Momchev, Olivier Bachem, Daniel J. Mankowitz, Doina Precup, Bilal Piot |
| 2023 | ICML | Curiosity in Hindsight: Intrinsic Exploration in Stochastic Environments. | Daniel Jarrett, Corentin Tallec, Florent Altch, Thomas Mesnard, Rmi Munos, Michal Valko |
| 2023 | ICML | Quantile Credit Assignment. | Thomas Mesnard, Wenqi Chen, Alaa Saade, Yunhao Tang, Mark Rowland, Theophane Weber, Clare Lyle, Audrunas Gruslys, Michal Valko, Will Dabney, Georg Ostrovski, Eric Moulines, Rmi Munos |
| 2021 | ICML | Counterfactual Credit Assignment in Model-Free Reinforcement Learning. | Thomas Mesnard, Theophane Weber, Fabio Viola, Shantanu Thakoor, Alaa Saade, Anna Harutyunyan, Will Dabney, Thomas S. Stepleton, Nicolas Heess, Arthur Guez, Eric Moulines, Marcus Hutter, Lars Buesing, Rmi Munos |
| 2018 | ICLR | Extending the Framework of Equilibrium Propagation to General Dynamics. | Benjamin Scellier, Anirudh Goyal, Jonathan Binas, Thomas Mesnard, Yoshua Bengio |