Lonard Hussenot
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
11
Venues
5
Active years
2021–2025
Best venue rank
A*
Where they publish
Papers
11 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICLR | BOND: Aligning LLMs with Best-of-N Distillation. | Pier Giuseppe Sessa, Robert Dadashi-Tazehozi, Lonard Hussenot, Johan Ferret, Nino Vieillard, Alexandre Ram, Bobak Shahriari, Sarah Perrin, Abram L. Friesen, Geoffrey Cideron, Sertan Girgin, Piotr Stanczyk, Andrea Michi, Danila Sinopalnikov, Sabela Ramos Garea, Amlie Hliou, Aliaksei Severyn, Matthew Hoffman, Nikola Momchev, Olivier Bachem |
| 2024 | EMNLP | Conditional Language Policy: A General Framework For Steerable Multi-Objective Finetuning. | Kaiwen Wang, Rahul Kidambi, Ryan Sullivan, Alekh Agarwal, Christoph Dann, Andrea Michi, Marco Gelmi, Yunxuan Li, Raghav Gupta, Avinava Dubey, Alexandre Ram, Johan Ferret, Geoffrey Cideron, Le Hou, Hongkun Yu, Amr Ahmed, Aranyak Mehta, Lonard Hussenot, Olivier Bachem, Edouard Leurent |
| 2024 | ICML | MusicRL: Aligning Music Generation to Human Preferences. | Geoffrey Cideron, Sertan Girgin, Mauro Verzetti, Damien Vincent, Matej Kastelic, Zaln Borsos, Brian McWilliams, Victor Ungureanu, Olivier Bachem, Olivier Pietquin, Matthieu Geist, Lonard Hussenot, Neil Zeghidour, Andrea Agostinelli |
| 2024 | ICML | WARM: On the Benefits of Weight Averaged Reward Models. | Alexandre Ram, Nino Vieillard, Lonard Hussenot, Robert Dadashi, Geoffrey Cideron, Olivier Bachem, Johan Ferret |
| 2023 | ACL | Factually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback. | Paul Roit, Johan Ferret, Lior Shani, Roee Aharoni, Geoffrey Cideron, Robert Dadashi, Matthieu Geist, Sertan Girgin, Lonard Hussenot, Orgad Keller, Nikola Momchev, Sabela Ramos Garea, Piotr Stanczyk, Nino Vieillard, Olivier Bachem, Gal Elidan, Avinatan Hassidim, Olivier Pietquin, Idan Szpektor |
| 2022 | AAAI | Offline Reinforcement Learning as Anti-exploration. | Shideh Rezaeifar, Robert Dadashi, Nino Vieillard, Lonard Hussenot, Olivier Bachem, Olivier Pietquin, Matthieu Geist |
| 2022 | ICML | Continuous Control with Action Quantization from Demonstrations. | Robert Dadashi, Lonard Hussenot, Damien Vincent, Sertan Girgin, Anton Raichuk, Matthieu Geist, Olivier Pietquin |
| 2021 | ICLR | What Matters for On-Policy Deep Actor-Critic Methods? A Large-Scale Study. | Marcin Andrychowicz, Anton Raichuk, Piotr Stanczyk, Manu Orsini, Sertan Girgin, Raphal Marinier, Lonard Hussenot, Matthieu Geist, Olivier Pietquin, Marcin Michalski, Sylvain Gelly, Olivier Bachem |
| 2021 | ICLR | Primal Wasserstein Imitation Learning. | Robert Dadashi, Lonard Hussenot, Matthieu Geist, Olivier Pietquin |
| 2021 | ICML | Offline Reinforcement Learning with Pseudometric Learning. | Robert Dadashi, Shideh Rezaeifar, Nino Vieillard, Lonard Hussenot, Olivier Pietquin, Matthieu Geist |
| 2021 | ICML | Hyperparameter Selection for Imitation Learning. | Lonard Hussenot, Marcin Andrychowicz, Damien Vincent, Robert Dadashi, Anton Raichuk, Sabela Ramos, Nikola Momchev, Sertan Girgin, Raphal Marinier, Lukasz Stafiniak, Manu Orsini, Olivier Bachem, Matthieu Geist, Olivier Pietquin |