Olivier Pietquin
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
93
Venues
25
Active years
2002–2025
Best venue rank
A*
Where they publish
- MulticonferenceICASSP14 papers
- A*ICML14 papers
- AInterspeech10 papers
- A*ICLR7 papers
- A*IJCAI7 papers
- A*AAAI6 papers
- AAISTATS5 papers
- A*ACL3 papers
- BLREC3 papers
- BICONIP3 papers
- BSIGdial3 papers
- BESANN3 papers
- A*EMNLP2 papers
- ANAACL2 papers
- CACML1 paper
- BIJCNN1 paper
- A*ECCV1 paper
- A*CVPR1 paper
- CASRU1 paper
- AECAI1 paper
- NationalHCI1 paper
- CICMLA1 paper
- BIJCNLP1 paper
- BMDAI1 paper
- NationalAIMSA1 paper
Papers
93 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICASSP | Biodenoising: Animal Vocalization Denoising without Access to Clean Data. | Marius Miron, Sara Keen, Jen-Yu Liu, Benjamin Hoffman, Masato Hagiwara, Olivier Pietquin, Felix Effenberger, Maddie Cusimano |
| 2025 | ICLR | Self-Improving Robust Preference Optimization. | Eugene Choi, Arash Ahmadian, Matthieu Geist, Olivier Pietquin, Mohammad Gheshlaghi Azar |
| 2025 | ICLR | NatureLM-audio: an Audio-Language Foundation Model for Bioacoustics. | David Robinson, Marius Miron, Masato Hagiwara, Olivier Pietquin |
| 2024 | AAAI | Learning Discrete-Time Major-Minor Mean Field Games. | Kai Cui, Gke Dayanikli, Mathieu Laurire, Matthieu Geist, Olivier Pietquin, Heinz Koeppl |
| 2024 | ACL | Back to Basics: Revisiting REINFORCE-Style Optimization for Learning from Human Feedback in LLMs. | Arash Ahmadian, Chris Cremer, Matthias Gall, Marzieh Fadaee, Julia Kreutzer, Olivier Pietquin, Ahmet stn, Sara Hooker |
| 2024 | ACL | Countering Reward Over-Optimization in LLM with Demonstration-Guided Reinforcement Learning. | Mathieu Rita, Florian Strub, Rahma Chaabouni, Paul Michel, Emmanuel Dupoux, Olivier Pietquin |
| 2024 | EMNLP | Contrastive Policy Gradient: Aligning LLMs on sequence-level scores in a supervised-friendly fashion. | Yannis Flet-Berliac, Nathan Grinsztajn, Florian Strub, Eugene Choi, Bill Wu, Chris Cremer, Arash Ahmadian, Yash Chandak, Mohammad Gheshlaghi Azar, Olivier Pietquin, Matthieu Geist |
| 2024 | ICML | MusicRL: Aligning Music Generation to Human Preferences. | Geoffrey Cideron, Sertan Girgin, Mauro Verzetti, Damien Vincent, Matej Kastelic, Zaln Borsos, Brian McWilliams, Victor Ungureanu, Olivier Bachem, Olivier Pietquin, Matthieu Geist, Lonard Hussenot, Neil Zeghidour, Andrea Agostinelli |
| 2023 | ACL | Factually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback. | Paul Roit, Johan Ferret, Lior Shani, Roee Aharoni, Geoffrey Cideron, Robert Dadashi, Matthieu Geist, Sertan Girgin, Lonard Hussenot, Orgad Keller, Nikola Momchev, Sabela Ramos Garea, Piotr Stanczyk, Nino Vieillard, Olivier Bachem, Gal Elidan, Avinatan Hassidim, Olivier Pietquin, Idan Szpektor |
| 2023 | ICML | Regularization and Variance-Weighted Regression Achieves Minimax Optimality in Linear MDPs: Theory and Practice. | Toshinori Kitamura, Tadashi Kozuno, Yunhao Tang, Nino Vieillard, Michal Valko, Wenhao Yang, Jincheng Mei, Pierre Mnard, Mohammad Gheshlaghi Azar, Rmi Munos, Olivier Pietquin, Matthieu Geist, Csaba Szepesvri, Wataru Kumagai, Yutaka Matsuo |
| 2022 | AAAI | Generalization in Mean Field Games by Learning Master Policies. | Sarah Perrin, Mathieu Laurire, Julien Prolat, Romuald lie, Matthieu Geist, Olivier Pietquin |
| 2022 | AAAI | Offline Reinforcement Learning as Anti-exploration. | Shideh Rezaeifar, Robert Dadashi, Nino Vieillard, Lonard Hussenot, Olivier Bachem, Olivier Pietquin, Matthieu Geist |
| 2022 | AISTATS | Implicitly Regularized RL with Implicit Q-values. | Nino Vieillard, Marcin Andrychowicz, Anton Raichuk, Olivier Pietquin, Matthieu Geist |
| 2022 | ICLR | On the role of population heterogeneity in emergent communication. | Mathieu Rita, Florian Strub, Jean-Bastien Grill, Olivier Pietquin, Emmanuel Dupoux |
| 2022 | ICML | Continuous Control with Action Quantization from Demonstrations. | Robert Dadashi, Lonard Hussenot, Damien Vincent, Sertan Girgin, Anton Raichuk, Matthieu Geist, Olivier Pietquin |
| 2022 | ICML | Scalable Deep Reinforcement Learning Algorithms for Mean Field Games. | Mathieu Laurire, Sarah Perrin, Sertan Girgin, Paul Muller, Ayush Jain, Theophile Cabannes, Georgios Piliouras, Julien Prolat, Romuald Elie, Olivier Pietquin, Matthieu Geist |
| 2022 | NAACL | Learning Natural Language Generation with Truncated Reinforcement Learning. | Alice Martin, Guillaume Quispe, Charles Ollion, Sylvain Le Corff, Florian Strub, Olivier Pietquin |
| 2021 | ICASSP | Learning From Heterogeneous Eeg Signals with Differentiable Channel Reordering. | Aaqib Saeed, David Grangier, Olivier Pietquin, Neil Zeghidour |
| 2021 | ICLR | What Matters for On-Policy Deep Actor-Critic Methods? A Large-Scale Study. | Marcin Andrychowicz, Anton Raichuk, Piotr Stanczyk, Manu Orsini, Sertan Girgin, Raphal Marinier, Lonard Hussenot, Matthieu Geist, Olivier Pietquin, Marcin Michalski, Sylvain Gelly, Olivier Bachem |
| 2021 | ICLR | Primal Wasserstein Imitation Learning. | Robert Dadashi, Lonard Hussenot, Matthieu Geist, Olivier Pietquin |
| 2021 | ICLR | Adversarially Guided Actor-Critic. | Yannis Flet-Berliac, Johan Ferret, Olivier Pietquin, Philippe Preux, Matthieu Geist |
| 2021 | ICML | Offline Reinforcement Learning with Pseudometric Learning. | Robert Dadashi, Shideh Rezaeifar, Nino Vieillard, Lonard Hussenot, Olivier Pietquin, Matthieu Geist |
| 2021 | ICML | Hyperparameter Selection for Imitation Learning. | Lonard Hussenot, Marcin Andrychowicz, Damien Vincent, Robert Dadashi, Anton Raichuk, Sabela Ramos, Nikola Momchev, Sertan Girgin, Raphal Marinier, Lukasz Stafiniak, Manu Orsini, Olivier Bachem, Matthieu Geist, Olivier Pietquin |
| 2021 | IJCAI | Mean Field Games Flock! The Reinforcement Learning Way. | Sarah Perrin, Mathieu Laurire, Julien Prolat, Matthieu Geist, Romuald lie, Olivier Pietquin |
| 2021 | IJCAI | Don't Do What Doesn't Matter: Intrinsic Motivation with Action Usefulness. | Mathieu Seurin, Florian Strub, Philippe Preux, Olivier Pietquin |
| 2020 | AAAI | On the Convergence of Model Free Learning in Mean Field Games. | Romuald Elie, Julien Prolat, Mathieu Laurire, Matthieu Geist, Olivier Pietquin |
| 2020 | AAAI | Deep Conservative Policy Iteration. | Nino Vieillard, Olivier Pietquin, Matthieu Geist |
| 2020 | ACML | Foolproof Cooperative Learning. | Alexis Jacq, Julien Prolat, Matthieu Geist, Olivier Pietquin |
| 2020 | AISTATS | Momentum in Reinforcement Learning. | Nino Vieillard, Bruno Scherrer, Olivier Pietquin, Matthieu Geist |
| 2020 | EMNLP | Supervised Seeded Iterated Learning for Interactive Language Learning. | Yuchen Lu, Soumye Singhal, Florian Strub, Olivier Pietquin, Aaron C. Courville |
| 2020 | ICML | Countering Language Drift with Seeded Iterated Learning. | Yuchen Lu, Soumye Singhal, Florian Strub, Aaron C. Courville, Olivier Pietquin |
| 2020 | IJCAI | Self-Attentional Credit Assignment for Transfer in Reinforcement Learning. | Johan Ferret, Raphal Marinier, Matthieu Geist, Olivier Pietquin |
| 2020 | IJCNN | "I'm Sorry Dave, I'm Afraid I Can't Do That" Deep Q-Learning from Forbidden Actions. | Mathieu Seurin, Philippe Preux, Olivier Pietquin |
| 2020 | Interspeech | A Machine of Few Words: Interactive Speaker Recognition with Reinforcement Learning. | Mathieu Seurin, Florian Strub, Philippe Preux, Olivier Pietquin |
| 2019 | ICML | A Theory of Regularized Markov Decision Processes. | Matthieu Geist, Bruno Scherrer, Olivier Pietquin |
| 2019 | ICML | Learning from a Learner. | Alexis Jacq, Matthieu Geist, Ana Paiva, Olivier Pietquin |
| 2018 | AAAI | Deep Q-learning From Demonstrations. | Todd Hester, Matej Vecerk, Olivier Pietquin, Marc Lanctot, Tom Schaul, Bilal Piot, Dan Horgan, John Quan, Andrew Sendonaris, Ian Osband, Gabriel Dulac-Arnold, John P. Agapiou, Joel Z. Leibo, Audrunas Gruslys |
| 2018 | AISTATS | Actor-Critic Fictitious Play in Simultaneous Move Multistage Games. | Julien Prolat, Bilal Piot, Olivier Pietquin |
| 2018 | ECCV | Visual Reasoning with Multi-hop Feature Modulation. | Florian Strub, Mathieu Seurin, Ethan Perez, Harm de Vries, Jrmie Mary, Philippe Preux, Aaron C. Courville, Olivier Pietquin |
| 2018 | ICASSP | End-to-End Automatic Speech Translation of Audiobooks. | Alexandre Berard, Laurent Besacier, Ali Can Kocabiyikoglu, Olivier Pietquin |
| 2018 | ICLR | Noisy Networks For Exploration. | Meire Fortunato, Mohammad Gheshlaghi Azar, Bilal Piot, Jacob Menick, Matteo Hessel, Ian Osband, Alex Graves, Volodymyr Mnih, Rmi Munos, Demis Hassabis, Olivier Pietquin, Charles Blundell, Shane Legg |
| 2017 | AISTATS | Learning Nash Equilibrium for General-Sum Markov Games from Batch Data. | Julien Prolat, Florian Strub, Bilal Piot, Olivier Pietquin |
| 2017 | CVPR | GuessWhat?! Visual Object Discovery through Multi-modal Dialogue. | Harm de Vries, Florian Strub, Sarath Chandar, Olivier Pietquin, Hugo Larochelle, Aaron C. Courville |
| 2017 | IJCAI | End-to-end optimization of goal-driven and visually grounded dialogue systems. | Florian Strub, Harm de Vries, Jrmie Mary, Bilal Piot, Aaron C. Courville, Olivier Pietquin |
| 2016 | AISTATS | On the Use of Non-Stationary Strategies for Solving Two-Player Zero-Sum Markov Games. | Julien Prolat, Bilal Piot, Bruno Scherrer, Olivier Pietquin |
| 2016 | ICML | PAC learning of Probabilistic Automaton based on the Method of Moments. | Hadrien Glaude, Olivier Pietquin |
| 2016 | ICML | Softened Approximate Policy Iteration for Markov Games. | Julien Prolat, Bilal Piot, Matthieu Geist, Bruno Scherrer, Olivier Pietquin |
| 2016 | Interspeech | A Stochastic Model for Computer-Aided Human-Human Dialogue. | Merwan Barlier, Romain Laroche, Olivier Pietquin |
| 2016 | LREC | MultiVec: a Multilingual and Multilevel Representation Learning Toolkit for NLP. | Alexandre Berard, Christophe Servan, Olivier Pietquin, Laurent Besacier |
| 2015 | ASRU | Spectral learning with non negative probabilities for finite state automaton. | Hadrien Glaude, Cyrille Enderli, Olivier Pietquin |
| 2015 | ICML | Proceedings of the 4th Workshop on Machine Learning for Interactive Systems (MLIS-2015). | Heriberto Cuayhuitl, Nina Dethlefs, Lutz Frommberger, Martijn van Otterlo, Olivier Pietquin |
| 2015 | ICML | Approximate Dynamic Programming for Two-Player Zero-Sum Markov Games. | Julien Prolat, Bruno Scherrer, Bilal Piot, Olivier Pietquin |
| 2015 | ICML | Imitation Learning Applied to Embodied Conversational Agents. | Bilal Piot, Olivier Pietquin, Matthieu Geist |
| 2015 | ICONIP | Optimism in Active Learning with Gaussian Processes. | Timoth Collet, Olivier Pietquin |
| 2015 | ICONIP | Non-negative Spectral Learning for Linear Sequential Systems. | Hadrien Glaude, Cyrille Enderli, Olivier Pietquin |
| 2015 | IJCAI | Inverse Reinforcement Learning in Relational Domains. | Thibaut Munzer, Bilal Piot, Matthieu Geist, Olivier Pietquin, Manuel Lopes |
| 2015 | SIGdial | Human-Machine Dialogue as a Stochastic Game. | Merwan Barlier, Julien Prolat, Romain Laroche, Olivier Pietquin |
| 2014 | ICASSP | Ordinal regression for interaction quality prediction. | Layla El Asri, Hatim Khouzaimi, Romain Laroche, Olivier Pietquin |
| 2014 | Interspeech | Predicting when to laugh with structured classification. | Bilal Piot, Olivier Pietquin, Matthieu Geist |
| 2014 | LREC | NASTIA: Negotiating Appointment Setting Interface. | Layla El Asri, Rmi Lemonnier, Romain Laroche, Olivier Pietquin, Hatim Khouzaimi |
| 2014 | LREC | DINASTI: Dialogues with a Negotiating Appointment Setting Interface. | Layla El Asri, Romain Laroche, Olivier Pietquin |
| 2013 | ICASSP | Random projections: A remedy for overfitting issues in time series prediction with echo state networks. | Lucie Daubigney, Matthieu Geist, Olivier Pietquin |
| 2013 | IJCAI | Inverse reinforcement learning for interactive systems. | Olivier Pietquin |
| 2013 | Interspeech | Particle swarm optimisation of spoken dialogue system strategies. | Lucie Daubigney, Matthieu Geist, Olivier Pietquin |
| 2013 | SIGdial | Model-free POMDP optimisation of tutoring systems with echo-state networks. | Lucie Daubigney, Matthieu Geist, Olivier Pietquin |
| 2012 | ECAI | A Reinforcement Learning Approach to Optimize the longitudinal Behavior of a Partial Autonomous Driving Assistance System. | Olivier Pietquin, Fabio Tango |
| 2012 | ICASSP | Clustering behaviors of Spoken Dialogue Systems users. | Senthilkumar Chandramohan, Matthieu Geist, Fabrice Lefvre, Olivier Pietquin |
| 2012 | ICASSP | Off-policy learning in large-scale POMDP-based dialogue systems. | Lucie Daubigney, Matthieu Geist, Olivier Pietquin |
| 2012 | NAACL | Statistical User Simulation for Spoken Dialogue Systems: What for, Which Data, Which Future? | Olivier Pietquin |
| 2011 | ESANN | Single-trial P300 detection with Kalman filtering and SVMs. | Lucie Daubigney, Olivier Pietquin |
| 2011 | HCI | Automation Effects on Driver's Behaviour When Integrating a PADAS and a Distraction Classifier. | Fabio Tango, Luca Minin, Raghav Aras, Olivier Pietquin |
| 2011 | ICMLA | A Non-parametric Approach to Approximate Dynamic Programming. | Hadrien Glaude, Fadi Akrimi, Matthieu Geist, Olivier Pietquin |
| 2011 | IJCAI | Sample Efficient On-Line Learning of Optimal Dialogue Policies with Kalman Temporal Differences. | Olivier Pietquin, Matthieu Geist, Senthilkumar Chandramohan |
| 2011 | IJCNLP | Training a BN-based user model for dialogue simulation with missing data. | Stphane Rossignol, Olivier Pietquin, Michel Ianotto |
| 2011 | Interspeech | User Simulation in Dialogue Systems Using Inverse Reinforcement Learning. | Senthilkumar Chandramohan, Matthieu Geist, Fabrice Lefvre, Olivier Pietquin |
| 2011 | Interspeech | Uncertainty Management for On-Line Optimisation of a POMDP-Based Large-Scale Spoken Dialogue System. | Lucie Daubigney, Milica Gasic, Senthilkumar Chandramohan, Matthieu Geist, Olivier Pietquin, Steve J. Young |
| 2010 | ESANN | Online speaker diarization with a size-monitored growing neural gas algorithm. | Jean-Louis Gutzwiller, Herv Frezza-Buet, Olivier Pietquin |
| 2010 | ICASSP | Bayesian framework for artifact reduction on ECG IN MRI. | Julien Oster, Olivier Pietquin, Michel Kraemer, Jacques Felblinger |
| 2010 | Interspeech | Optimizing spoken dialogue management with fitted value iteration. | Senthilkumar Chandramohan, Matthieu Geist, Olivier Pietquin |
| 2010 | Interspeech | Single-speaker/multi-speaker co-channel speech classification. | Stphane Rossignol, Olivier Pietquin |
| 2010 | MDAI | Revisiting Natural Actor-Critics with Value Function Approximation. | Matthieu Geist, Olivier Pietquin |
| 2010 | SIGdial | Sparse Approximate Dynamic Programming for Dialog Management. | Senthilkumar Chandramohan, Matthieu Geist, Olivier Pietquin |
| 2009 | ESANN | Kernelizing Vector Quantization Algorithms. | Matthieu Geist, Olivier Pietquin, Gabriel Fricout |
| 2009 | ICASSP | A specific QRS detector for electrocardiography during MRI: Using wavelets and local regularity characterization. | Julien Oster, Olivier Pietquin, Roger Abcherli, Michel Kraemer, Jacques Felblinger |
| 2009 | ICONIP | Tracking in Reinforcement Learning. | Matthieu Geist, Olivier Pietquin, Gabriel Fricout |
| 2008 | ICASSP | Functional semi-automated segmentation of renal DCE-MRI sequences. | Beatrice Chevaillier, Yannick Ponvianne, Jean-Luc Collette, Damien Mandry, Michel Claudon, Olivier Pietquin |
| 2008 | ICASSP | Adaptive RR prediction for cardiac MRI. | Julien Oster, Olivier Pietquin, Gilles Bosser |
| 2007 | ICASSP | Learning to Ground in Spoken Dialogue Systems. | Olivier Pietquin |
| 2007 | Interspeech | Machine learning for spoken dialogue systems. | Oliver Lemon, Olivier Pietquin |
| 2006 | AIMSA | Machine Learning for Spoken Dialogue Management: An Experiment with Speech-Based Database Querying. | Olivier Pietquin |
| 2006 | ICASSP | Dynamic Bayesian Networks for NLU Simulation with Applications to Dialog Optimal Strategy Learning. | Olivier Pietquin, Thierry Dutoit |
| 2005 | Interspeech | Comparing ASR modeling methods for spoken dialogue simulation and optimal strategy learning. | Olivier Pietquin, Richard Beaufort |
| 2002 | ICASSP | ASR system modeling for automatic evaluation and optimization of dialogue systems. | Olivier Pietquin, Steve Renals |