Sertan Girgin
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
19
Venues
13
Active years
2006–2025
Best venue rank
A*
Where they publish
Papers
19 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICLR | Diversity-Rewarded CFG Distillation. | Geoffrey Cideron, Andrea Agostinelli, Johan Ferret, Sertan Girgin, Romuald Elie, Olivier Bachem, Sarah Perrin, Alexandre Ram |
| 2025 | ICLR | BOND: Aligning LLMs with Best-of-N Distillation. | Pier Giuseppe Sessa, Robert Dadashi-Tazehozi, Lonard Hussenot, Johan Ferret, Nino Vieillard, Alexandre Ram, Bobak Shahriari, Sarah Perrin, Abram L. Friesen, Geoffrey Cideron, Sertan Girgin, Piotr Stanczyk, Andrea Michi, Danila Sinopalnikov, Sabela Ramos Garea, Amlie Hliou, Aliaksei Severyn, Matthew Hoffman, Nikola Momchev, Olivier Bachem |
| 2024 | ICML | MusicRL: Aligning Music Generation to Human Preferences. | Geoffrey Cideron, Sertan Girgin, Mauro Verzetti, Damien Vincent, Matej Kastelic, Zaln Borsos, Brian McWilliams, Victor Ungureanu, Olivier Bachem, Olivier Pietquin, Matthieu Geist, Lonard Hussenot, Neil Zeghidour, Andrea Agostinelli |
| 2024 | ICML | Nash Learning from Human Feedback. | Rmi Munos, Michal Valko, Daniele Calandriello, Mohammad Gheshlaghi Azar, Mark Rowland, Daniel Guo, Yunhao Tang, Matthieu Geist, Thomas Mesnard, Cme Fiegel, Andrea Michi, Marco Selvi, Sertan Girgin, Nikola Momchev, Olivier Bachem, Daniel J. Mankowitz, Doina Precup, Bilal Piot |
| 2023 | ACL | Factually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback. | Paul Roit, Johan Ferret, Lior Shani, Roee Aharoni, Geoffrey Cideron, Robert Dadashi, Matthieu Geist, Sertan Girgin, Lonard Hussenot, Orgad Keller, Nikola Momchev, Sabela Ramos Garea, Piotr Stanczyk, Nino Vieillard, Olivier Bachem, Gal Elidan, Avinatan Hassidim, Olivier Pietquin, Idan Szpektor |
| 2022 | EMNLP | Decoding a Neural Retriever's Latent Space for Query Suggestion. | Leonard Adolphs, Michelle Chen Huebscher, Christian Buck, Sertan Girgin, Olivier Bachem, Massimiliano Ciaramita, Thomas Hofmann |
| 2022 | ICML | Continuous Control with Action Quantization from Demonstrations. | Robert Dadashi, Lonard Hussenot, Damien Vincent, Sertan Girgin, Anton Raichuk, Matthieu Geist, Olivier Pietquin |
| 2022 | ICML | Scalable Deep Reinforcement Learning Algorithms for Mean Field Games. | Mathieu Laurire, Sarah Perrin, Sertan Girgin, Paul Muller, Ayush Jain, Theophile Cabannes, Georgios Piliouras, Julien Prolat, Romuald Elie, Olivier Pietquin, Matthieu Geist |
| 2021 | ICLR | What Matters for On-Policy Deep Actor-Critic Methods? A Large-Scale Study. | Marcin Andrychowicz, Anton Raichuk, Piotr Stanczyk, Manu Orsini, Sertan Girgin, Raphal Marinier, Lonard Hussenot, Matthieu Geist, Olivier Pietquin, Marcin Michalski, Sylvain Gelly, Olivier Bachem |
| 2021 | ICML | Hyperparameter Selection for Imitation Learning. | Lonard Hussenot, Marcin Andrychowicz, Damien Vincent, Robert Dadashi, Anton Raichuk, Sabela Ramos, Nikola Momchev, Sertan Girgin, Raphal Marinier, Lukasz Stafiniak, Manu Orsini, Olivier Bachem, Matthieu Geist, Olivier Pietquin |
| 2017 | ICMI | Text based user comments as a signal for automatic language identification of online videos. | A. Seza Dogruz, Natalia Ponomareva, Sertan Girgin, Reshu Jain, Christoph Oehler |
| 2010 | ICDM | Advertising Campaigns Management: Should We Be Greedy? | Sertan Girgin, Jrmie Mary, Philippe Preux, Olivier Nicol |
| 2009 | AIME | A Novel Multilingual Report Generation System for Medical Applications. | Kaya Kuru, Sertan Girgin, Kemal Arda |
| 2009 | KSEM | Developing Diagnostic DSSs Based on a Novel Data Collection Methodology. | Kaya Kuru, Sertan Girgin, Kemal Arda, Ugur Bozlar, Veysel Akgn |
| 2008 | EUROGP | Feature Discovery in Reinforcement Learning Using Genetic Programming. | Sertan Girgin, Philippe Preux |
| 2008 | ICMLA | Basis Function Construction in Reinforcement Learning Using Cascade-Correlation Learning Architecture. | Sertan Girgin, Philippe Preux |
| 2007 | IJCAI | State Similarity Based Approach for Improving Performance in RL. | Sertan Girgin, Faruk Polat, Reda Alhajj |
| 2006 | ECAI | Learning by Automatic Option Discovery from Conditionally Terminating Sequences. | Sertan Girgin, Faruk Polat, Reda Alhajj |
| 2006 | IDEAL | Effectiveness of Considering State Similarity for Reinforcement Learning. | Sertan Girgin, Faruk Polat, Reda Alhajj |