Johan Ferret
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
9
Venues
5
Active years
2020–2025
Best venue rank
A*
Where they publish
Papers
9 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICLR | Diversity-Rewarded CFG Distillation. | Geoffrey Cideron, Andrea Agostinelli, Johan Ferret, Sertan Girgin, Romuald Elie, Olivier Bachem, Sarah Perrin, Alexandre Ram |
| 2025 | ICLR | BOND: Aligning LLMs with Best-of-N Distillation. | Pier Giuseppe Sessa, Robert Dadashi-Tazehozi, Lonard Hussenot, Johan Ferret, Nino Vieillard, Alexandre Ram, Bobak Shahriari, Sarah Perrin, Abram L. Friesen, Geoffrey Cideron, Sertan Girgin, Piotr Stanczyk, Andrea Michi, Danila Sinopalnikov, Sabela Ramos Garea, Amlie Hliou, Aliaksei Severyn, Matthew Hoffman, Nikola Momchev, Olivier Bachem |
| 2025 | ICML | On Teacher Hacking in Language Model Distillation. | Daniil Tiapkin, Daniele Calandriello, Johan Ferret, Sarah Perrin, Nino Vieillard, Alexandre Ram, Mathieu Blondel |
| 2024 | EMNLP | Conditional Language Policy: A General Framework For Steerable Multi-Objective Finetuning. | Kaiwen Wang, Rahul Kidambi, Ryan Sullivan, Alekh Agarwal, Christoph Dann, Andrea Michi, Marco Gelmi, Yunxuan Li, Raghav Gupta, Avinava Dubey, Alexandre Ram, Johan Ferret, Geoffrey Cideron, Le Hou, Hongkun Yu, Amr Ahmed, Aranyak Mehta, Lonard Hussenot, Olivier Bachem, Edouard Leurent |
| 2024 | ICML | RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback. | Harrison Lee, Samrat Phatale, Hassan Mansoor, Thomas Mesnard, Johan Ferret, Kellie Lu, Colton Bishop, Ethan Hall, Victor Carbune, Abhinav Rastogi, Sushant Prakash |
| 2024 | ICML | WARM: On the Benefits of Weight Averaged Reward Models. | Alexandre Ram, Nino Vieillard, Lonard Hussenot, Robert Dadashi, Geoffrey Cideron, Olivier Bachem, Johan Ferret |
| 2023 | ACL | Factually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback. | Paul Roit, Johan Ferret, Lior Shani, Roee Aharoni, Geoffrey Cideron, Robert Dadashi, Matthieu Geist, Sertan Girgin, Lonard Hussenot, Orgad Keller, Nikola Momchev, Sabela Ramos Garea, Piotr Stanczyk, Nino Vieillard, Olivier Bachem, Gal Elidan, Avinatan Hassidim, Olivier Pietquin, Idan Szpektor |
| 2021 | ICLR | Adversarially Guided Actor-Critic. | Yannis Flet-Berliac, Johan Ferret, Olivier Pietquin, Philippe Preux, Matthieu Geist |
| 2020 | IJCAI | Self-Attentional Credit Assignment for Transfer in Reinforcement Learning. | Johan Ferret, Raphal Marinier, Matthieu Geist, Olivier Pietquin |