Skip to content

Olivier Pietquin

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

93

Venues

25

Active years

2002–2025

Best venue rank

A*

Where they publish

Papers

93 indexed papers, newest first.

YearVenueTitleAuthors
2025ICASSPBiodenoising: Animal Vocalization Denoising without Access to Clean Data.Marius Miron, Sara Keen, Jen-Yu Liu, Benjamin Hoffman, Masato Hagiwara, Olivier Pietquin, Felix Effenberger, Maddie Cusimano
2025ICLRSelf-Improving Robust Preference Optimization.Eugene Choi, Arash Ahmadian, Matthieu Geist, Olivier Pietquin, Mohammad Gheshlaghi Azar
2025ICLRNatureLM-audio: an Audio-Language Foundation Model for Bioacoustics.David Robinson, Marius Miron, Masato Hagiwara, Olivier Pietquin
2024AAAILearning Discrete-Time Major-Minor Mean Field Games.Kai Cui, Gke Dayanikli, Mathieu Laurire, Matthieu Geist, Olivier Pietquin, Heinz Koeppl
2024ACLBack to Basics: Revisiting REINFORCE-Style Optimization for Learning from Human Feedback in LLMs.Arash Ahmadian, Chris Cremer, Matthias Gall, Marzieh Fadaee, Julia Kreutzer, Olivier Pietquin, Ahmet stn, Sara Hooker
2024ACLCountering Reward Over-Optimization in LLM with Demonstration-Guided Reinforcement Learning.Mathieu Rita, Florian Strub, Rahma Chaabouni, Paul Michel, Emmanuel Dupoux, Olivier Pietquin
2024EMNLPContrastive Policy Gradient: Aligning LLMs on sequence-level scores in a supervised-friendly fashion.Yannis Flet-Berliac, Nathan Grinsztajn, Florian Strub, Eugene Choi, Bill Wu, Chris Cremer, Arash Ahmadian, Yash Chandak, Mohammad Gheshlaghi Azar, Olivier Pietquin, Matthieu Geist
2024ICMLMusicRL: Aligning Music Generation to Human Preferences.Geoffrey Cideron, Sertan Girgin, Mauro Verzetti, Damien Vincent, Matej Kastelic, Zaln Borsos, Brian McWilliams, Victor Ungureanu, Olivier Bachem, Olivier Pietquin, Matthieu Geist, Lonard Hussenot, Neil Zeghidour, Andrea Agostinelli
2023ACLFactually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback.Paul Roit, Johan Ferret, Lior Shani, Roee Aharoni, Geoffrey Cideron, Robert Dadashi, Matthieu Geist, Sertan Girgin, Lonard Hussenot, Orgad Keller, Nikola Momchev, Sabela Ramos Garea, Piotr Stanczyk, Nino Vieillard, Olivier Bachem, Gal Elidan, Avinatan Hassidim, Olivier Pietquin, Idan Szpektor
2023ICMLRegularization and Variance-Weighted Regression Achieves Minimax Optimality in Linear MDPs: Theory and Practice.Toshinori Kitamura, Tadashi Kozuno, Yunhao Tang, Nino Vieillard, Michal Valko, Wenhao Yang, Jincheng Mei, Pierre Mnard, Mohammad Gheshlaghi Azar, Rmi Munos, Olivier Pietquin, Matthieu Geist, Csaba Szepesvri, Wataru Kumagai, Yutaka Matsuo
2022AAAIGeneralization in Mean Field Games by Learning Master Policies.Sarah Perrin, Mathieu Laurire, Julien Prolat, Romuald lie, Matthieu Geist, Olivier Pietquin
2022AAAIOffline Reinforcement Learning as Anti-exploration.Shideh Rezaeifar, Robert Dadashi, Nino Vieillard, Lonard Hussenot, Olivier Bachem, Olivier Pietquin, Matthieu Geist
2022AISTATSImplicitly Regularized RL with Implicit Q-values.Nino Vieillard, Marcin Andrychowicz, Anton Raichuk, Olivier Pietquin, Matthieu Geist
2022ICLROn the role of population heterogeneity in emergent communication.Mathieu Rita, Florian Strub, Jean-Bastien Grill, Olivier Pietquin, Emmanuel Dupoux
2022ICMLContinuous Control with Action Quantization from Demonstrations.Robert Dadashi, Lonard Hussenot, Damien Vincent, Sertan Girgin, Anton Raichuk, Matthieu Geist, Olivier Pietquin
2022ICMLScalable Deep Reinforcement Learning Algorithms for Mean Field Games.Mathieu Laurire, Sarah Perrin, Sertan Girgin, Paul Muller, Ayush Jain, Theophile Cabannes, Georgios Piliouras, Julien Prolat, Romuald Elie, Olivier Pietquin, Matthieu Geist
2022NAACLLearning Natural Language Generation with Truncated Reinforcement Learning.Alice Martin, Guillaume Quispe, Charles Ollion, Sylvain Le Corff, Florian Strub, Olivier Pietquin
2021ICASSPLearning From Heterogeneous Eeg Signals with Differentiable Channel Reordering.Aaqib Saeed, David Grangier, Olivier Pietquin, Neil Zeghidour
2021ICLRWhat Matters for On-Policy Deep Actor-Critic Methods? A Large-Scale Study.Marcin Andrychowicz, Anton Raichuk, Piotr Stanczyk, Manu Orsini, Sertan Girgin, Raphal Marinier, Lonard Hussenot, Matthieu Geist, Olivier Pietquin, Marcin Michalski, Sylvain Gelly, Olivier Bachem
2021ICLRPrimal Wasserstein Imitation Learning.Robert Dadashi, Lonard Hussenot, Matthieu Geist, Olivier Pietquin
2021ICLRAdversarially Guided Actor-Critic.Yannis Flet-Berliac, Johan Ferret, Olivier Pietquin, Philippe Preux, Matthieu Geist
2021ICMLOffline Reinforcement Learning with Pseudometric Learning.Robert Dadashi, Shideh Rezaeifar, Nino Vieillard, Lonard Hussenot, Olivier Pietquin, Matthieu Geist
2021ICMLHyperparameter Selection for Imitation Learning.Lonard Hussenot, Marcin Andrychowicz, Damien Vincent, Robert Dadashi, Anton Raichuk, Sabela Ramos, Nikola Momchev, Sertan Girgin, Raphal Marinier, Lukasz Stafiniak, Manu Orsini, Olivier Bachem, Matthieu Geist, Olivier Pietquin
2021IJCAIMean Field Games Flock! The Reinforcement Learning Way.Sarah Perrin, Mathieu Laurire, Julien Prolat, Matthieu Geist, Romuald lie, Olivier Pietquin
2021IJCAIDon't Do What Doesn't Matter: Intrinsic Motivation with Action Usefulness.Mathieu Seurin, Florian Strub, Philippe Preux, Olivier Pietquin
2020AAAIOn the Convergence of Model Free Learning in Mean Field Games.Romuald Elie, Julien Prolat, Mathieu Laurire, Matthieu Geist, Olivier Pietquin
2020AAAIDeep Conservative Policy Iteration.Nino Vieillard, Olivier Pietquin, Matthieu Geist
2020ACMLFoolproof Cooperative Learning.Alexis Jacq, Julien Prolat, Matthieu Geist, Olivier Pietquin
2020AISTATSMomentum in Reinforcement Learning.Nino Vieillard, Bruno Scherrer, Olivier Pietquin, Matthieu Geist
2020EMNLPSupervised Seeded Iterated Learning for Interactive Language Learning.Yuchen Lu, Soumye Singhal, Florian Strub, Olivier Pietquin, Aaron C. Courville
2020ICMLCountering Language Drift with Seeded Iterated Learning.Yuchen Lu, Soumye Singhal, Florian Strub, Aaron C. Courville, Olivier Pietquin
2020IJCAISelf-Attentional Credit Assignment for Transfer in Reinforcement Learning.Johan Ferret, Raphal Marinier, Matthieu Geist, Olivier Pietquin
2020IJCNN"I'm Sorry Dave, I'm Afraid I Can't Do That" Deep Q-Learning from Forbidden Actions.Mathieu Seurin, Philippe Preux, Olivier Pietquin
2020InterspeechA Machine of Few Words: Interactive Speaker Recognition with Reinforcement Learning.Mathieu Seurin, Florian Strub, Philippe Preux, Olivier Pietquin
2019ICMLA Theory of Regularized Markov Decision Processes.Matthieu Geist, Bruno Scherrer, Olivier Pietquin
2019ICMLLearning from a Learner.Alexis Jacq, Matthieu Geist, Ana Paiva, Olivier Pietquin
2018AAAIDeep Q-learning From Demonstrations.Todd Hester, Matej Vecerk, Olivier Pietquin, Marc Lanctot, Tom Schaul, Bilal Piot, Dan Horgan, John Quan, Andrew Sendonaris, Ian Osband, Gabriel Dulac-Arnold, John P. Agapiou, Joel Z. Leibo, Audrunas Gruslys
2018AISTATSActor-Critic Fictitious Play in Simultaneous Move Multistage Games.Julien Prolat, Bilal Piot, Olivier Pietquin
2018ECCVVisual Reasoning with Multi-hop Feature Modulation.Florian Strub, Mathieu Seurin, Ethan Perez, Harm de Vries, Jrmie Mary, Philippe Preux, Aaron C. Courville, Olivier Pietquin
2018ICASSPEnd-to-End Automatic Speech Translation of Audiobooks.Alexandre Berard, Laurent Besacier, Ali Can Kocabiyikoglu, Olivier Pietquin
2018ICLRNoisy Networks For Exploration.Meire Fortunato, Mohammad Gheshlaghi Azar, Bilal Piot, Jacob Menick, Matteo Hessel, Ian Osband, Alex Graves, Volodymyr Mnih, Rmi Munos, Demis Hassabis, Olivier Pietquin, Charles Blundell, Shane Legg
2017AISTATSLearning Nash Equilibrium for General-Sum Markov Games from Batch Data.Julien Prolat, Florian Strub, Bilal Piot, Olivier Pietquin
2017CVPRGuessWhat?! Visual Object Discovery through Multi-modal Dialogue.Harm de Vries, Florian Strub, Sarath Chandar, Olivier Pietquin, Hugo Larochelle, Aaron C. Courville
2017IJCAIEnd-to-end optimization of goal-driven and visually grounded dialogue systems.Florian Strub, Harm de Vries, Jrmie Mary, Bilal Piot, Aaron C. Courville, Olivier Pietquin
2016AISTATSOn the Use of Non-Stationary Strategies for Solving Two-Player Zero-Sum Markov Games.Julien Prolat, Bilal Piot, Bruno Scherrer, Olivier Pietquin
2016ICMLPAC learning of Probabilistic Automaton based on the Method of Moments.Hadrien Glaude, Olivier Pietquin
2016ICMLSoftened Approximate Policy Iteration for Markov Games.Julien Prolat, Bilal Piot, Matthieu Geist, Bruno Scherrer, Olivier Pietquin
2016InterspeechA Stochastic Model for Computer-Aided Human-Human Dialogue.Merwan Barlier, Romain Laroche, Olivier Pietquin
2016LRECMultiVec: a Multilingual and Multilevel Representation Learning Toolkit for NLP.Alexandre Berard, Christophe Servan, Olivier Pietquin, Laurent Besacier
2015ASRUSpectral learning with non negative probabilities for finite state automaton.Hadrien Glaude, Cyrille Enderli, Olivier Pietquin
2015ICMLProceedings of the 4th Workshop on Machine Learning for Interactive Systems (MLIS-2015).Heriberto Cuayhuitl, Nina Dethlefs, Lutz Frommberger, Martijn van Otterlo, Olivier Pietquin
2015ICMLApproximate Dynamic Programming for Two-Player Zero-Sum Markov Games.Julien Prolat, Bruno Scherrer, Bilal Piot, Olivier Pietquin
2015ICMLImitation Learning Applied to Embodied Conversational Agents.Bilal Piot, Olivier Pietquin, Matthieu Geist
2015ICONIPOptimism in Active Learning with Gaussian Processes.Timoth Collet, Olivier Pietquin
2015ICONIPNon-negative Spectral Learning for Linear Sequential Systems.Hadrien Glaude, Cyrille Enderli, Olivier Pietquin
2015IJCAIInverse Reinforcement Learning in Relational Domains.Thibaut Munzer, Bilal Piot, Matthieu Geist, Olivier Pietquin, Manuel Lopes
2015SIGdialHuman-Machine Dialogue as a Stochastic Game.Merwan Barlier, Julien Prolat, Romain Laroche, Olivier Pietquin
2014ICASSPOrdinal regression for interaction quality prediction.Layla El Asri, Hatim Khouzaimi, Romain Laroche, Olivier Pietquin
2014InterspeechPredicting when to laugh with structured classification.Bilal Piot, Olivier Pietquin, Matthieu Geist
2014LRECNASTIA: Negotiating Appointment Setting Interface.Layla El Asri, Rmi Lemonnier, Romain Laroche, Olivier Pietquin, Hatim Khouzaimi
2014LRECDINASTI: Dialogues with a Negotiating Appointment Setting Interface.Layla El Asri, Romain Laroche, Olivier Pietquin
2013ICASSPRandom projections: A remedy for overfitting issues in time series prediction with echo state networks.Lucie Daubigney, Matthieu Geist, Olivier Pietquin
2013IJCAIInverse reinforcement learning for interactive systems.Olivier Pietquin
2013InterspeechParticle swarm optimisation of spoken dialogue system strategies.Lucie Daubigney, Matthieu Geist, Olivier Pietquin
2013SIGdialModel-free POMDP optimisation of tutoring systems with echo-state networks.Lucie Daubigney, Matthieu Geist, Olivier Pietquin
2012ECAIA Reinforcement Learning Approach to Optimize the longitudinal Behavior of a Partial Autonomous Driving Assistance System.Olivier Pietquin, Fabio Tango
2012ICASSPClustering behaviors of Spoken Dialogue Systems users.Senthilkumar Chandramohan, Matthieu Geist, Fabrice Lefvre, Olivier Pietquin
2012ICASSPOff-policy learning in large-scale POMDP-based dialogue systems.Lucie Daubigney, Matthieu Geist, Olivier Pietquin
2012NAACLStatistical User Simulation for Spoken Dialogue Systems: What for, Which Data, Which Future?Olivier Pietquin
2011ESANNSingle-trial P300 detection with Kalman filtering and SVMs.Lucie Daubigney, Olivier Pietquin
2011HCIAutomation Effects on Driver's Behaviour When Integrating a PADAS and a Distraction Classifier.Fabio Tango, Luca Minin, Raghav Aras, Olivier Pietquin
2011ICMLAA Non-parametric Approach to Approximate Dynamic Programming.Hadrien Glaude, Fadi Akrimi, Matthieu Geist, Olivier Pietquin
2011IJCAISample Efficient On-Line Learning of Optimal Dialogue Policies with Kalman Temporal Differences.Olivier Pietquin, Matthieu Geist, Senthilkumar Chandramohan
2011IJCNLPTraining a BN-based user model for dialogue simulation with missing data.Stphane Rossignol, Olivier Pietquin, Michel Ianotto
2011InterspeechUser Simulation in Dialogue Systems Using Inverse Reinforcement Learning.Senthilkumar Chandramohan, Matthieu Geist, Fabrice Lefvre, Olivier Pietquin
2011InterspeechUncertainty Management for On-Line Optimisation of a POMDP-Based Large-Scale Spoken Dialogue System.Lucie Daubigney, Milica Gasic, Senthilkumar Chandramohan, Matthieu Geist, Olivier Pietquin, Steve J. Young
2010ESANNOnline speaker diarization with a size-monitored growing neural gas algorithm.Jean-Louis Gutzwiller, Herv Frezza-Buet, Olivier Pietquin
2010ICASSPBayesian framework for artifact reduction on ECG IN MRI.Julien Oster, Olivier Pietquin, Michel Kraemer, Jacques Felblinger
2010InterspeechOptimizing spoken dialogue management with fitted value iteration.Senthilkumar Chandramohan, Matthieu Geist, Olivier Pietquin
2010InterspeechSingle-speaker/multi-speaker co-channel speech classification.Stphane Rossignol, Olivier Pietquin
2010MDAIRevisiting Natural Actor-Critics with Value Function Approximation.Matthieu Geist, Olivier Pietquin
2010SIGdialSparse Approximate Dynamic Programming for Dialog Management.Senthilkumar Chandramohan, Matthieu Geist, Olivier Pietquin
2009ESANNKernelizing Vector Quantization Algorithms.Matthieu Geist, Olivier Pietquin, Gabriel Fricout
2009ICASSPA specific QRS detector for electrocardiography during MRI: Using wavelets and local regularity characterization.Julien Oster, Olivier Pietquin, Roger Abcherli, Michel Kraemer, Jacques Felblinger
2009ICONIPTracking in Reinforcement Learning.Matthieu Geist, Olivier Pietquin, Gabriel Fricout
2008ICASSPFunctional semi-automated segmentation of renal DCE-MRI sequences.Beatrice Chevaillier, Yannick Ponvianne, Jean-Luc Collette, Damien Mandry, Michel Claudon, Olivier Pietquin
2008ICASSPAdaptive RR prediction for cardiac MRI.Julien Oster, Olivier Pietquin, Gilles Bosser
2007ICASSPLearning to Ground in Spoken Dialogue Systems.Olivier Pietquin
2007InterspeechMachine learning for spoken dialogue systems.Oliver Lemon, Olivier Pietquin
2006AIMSAMachine Learning for Spoken Dialogue Management: An Experiment with Speech-Based Database Querying.Olivier Pietquin
2006ICASSPDynamic Bayesian Networks for NLU Simulation with Applications to Dialog Optimal Strategy Learning.Olivier Pietquin, Thierry Dutoit
2005InterspeechComparing ASR modeling methods for spoken dialogue simulation and optimal strategy learning.Olivier Pietquin, Richard Beaufort
2002ICASSPASR system modeling for automatic evaluation and optimization of dialogue systems.Olivier Pietquin, Steve Renals