Matthieu Geist
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
66
Venues
26
Active years
2009–2025
Best venue rank
A*
Where they publish
- A*ICML16 papers
- A*ICLR7 papers
- A*AAAI5 papers
- AInterspeech5 papers
- A*IJCAI4 papers
- A*ICRA3 papers
- AAISTATS3 papers
- MulticonferenceICASSP3 papers
- BICONIP2 papers
- BSIGdial2 papers
- A*EMNLP1 paper
- AIROS1 paper
- AUAI1 paper
- A*ACL1 paper
- UnrankedCoRL1 paper
- CACML1 paper
- CCoDIT1 paper
- A*ICCV1 paper
- CISCC1 paper
- CUIC1 paper
- A*PERCOM1 paper
- NationalICAISC1 paper
- CFUSION1 paper
- CICMLA1 paper
- BMDAI1 paper
- BESANN1 paper
Papers
66 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICLR | Self-Improving Robust Preference Optimization. | Eugene Choi, Arash Ahmadian, Matthieu Geist, Olivier Pietquin, Mohammad Gheshlaghi Azar |
| 2024 | AAAI | Learning Discrete-Time Major-Minor Mean Field Games. | Kai Cui, Gke Dayanikli, Mathieu Laurire, Matthieu Geist, Olivier Pietquin, Heinz Koeppl |
| 2024 | EMNLP | Contrastive Policy Gradient: Aligning LLMs on sequence-level scores in a supervised-friendly fashion. | Yannis Flet-Berliac, Nathan Grinsztajn, Florian Strub, Eugene Choi, Bill Wu, Chris Cremer, Arash Ahmadian, Yash Chandak, Mohammad Gheshlaghi Azar, Olivier Pietquin, Matthieu Geist |
| 2024 | ICLR | On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes. | Rishabh Agarwal, Nino Vieillard, Yongchao Zhou, Piotr Stanczyk, Sabela Ramos Garea, Matthieu Geist, Olivier Bachem |
| 2024 | ICLR | Closing the Gap between TD Learning and Supervised Learning - A Generalisation Point of View. | Raj Ghugare, Matthieu Geist, Glen Berseth, Benjamin Eysenbach |
| 2024 | ICML | MusicRL: Aligning Music Generation to Human Preferences. | Geoffrey Cideron, Sertan Girgin, Mauro Verzetti, Damien Vincent, Matej Kastelic, Zaln Borsos, Brian McWilliams, Victor Ungureanu, Olivier Bachem, Olivier Pietquin, Matthieu Geist, Lonard Hussenot, Neil Zeghidour, Andrea Agostinelli |
| 2024 | ICML | Nash Learning from Human Feedback. | Rmi Munos, Michal Valko, Daniele Calandriello, Mohammad Gheshlaghi Azar, Mark Rowland, Daniel Guo, Yunhao Tang, Matthieu Geist, Thomas Mesnard, Cme Fiegel, Andrea Michi, Marco Selvi, Sertan Girgin, Nikola Momchev, Olivier Bachem, Daniel J. Mankowitz, Doina Precup, Bilal Piot |
| 2024 | IROS | DRIFT: Deep Reinforcement Learning for Intelligent Floating Platforms Trajectories. | Matteo El Hariry, Antoine Richard, Vivek Muralidharan, Matthieu Geist, Miguel A. Olivares-Mndez |
| 2024 | UAI | Towards Minimax Optimality of Model-based Robust Reinforcement Learning. | Pierre Clavier, Erwan Le Pennec, Matthieu Geist |
| 2023 | ACL | Factually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback. | Paul Roit, Johan Ferret, Lior Shani, Roee Aharoni, Geoffrey Cideron, Robert Dadashi, Matthieu Geist, Sertan Girgin, Lonard Hussenot, Orgad Keller, Nikola Momchev, Sabela Ramos Garea, Piotr Stanczyk, Nino Vieillard, Olivier Bachem, Gal Elidan, Avinatan Hassidim, Olivier Pietquin, Idan Szpektor |
| 2023 | ICLR | Extreme Q-Learning: MaxEnt RL without Entropy. | Divyansh Garg, Joey Hejna, Matthieu Geist, Stefano Ermon |
| 2023 | ICML | A Connection between One-Step RL and Critic Regularization in Reinforcement Learning. | Benjamin Eysenbach, Matthieu Geist, Sergey Levine, Ruslan Salakhutdinov |
| 2023 | ICML | Regularization and Variance-Weighted Regression Achieves Minimax Optimality in Linear MDPs: Theory and Practice. | Toshinori Kitamura, Tadashi Kozuno, Yunhao Tang, Nino Vieillard, Michal Valko, Wenhao Yang, Jincheng Mei, Pierre Mnard, Mohammad Gheshlaghi Azar, Rmi Munos, Olivier Pietquin, Matthieu Geist, Csaba Szepesvri, Wataru Kumagai, Yutaka Matsuo |
| 2023 | ICML | Policy Mirror Ascent for Efficient and Independent Learning in Mean Field Games. | Batuhan Yardim, Semih Cayci, Matthieu Geist, Niao He |
| 2023 | ICRA | Pose-graph SLAM Using Multi-order Ultrasonic Echoes and Beamforming for Long-range Inspection Robots. | Othmane-Latif Ouabi, Neil Zeghidour, Nico F. Declercq, Matthieu Geist, Cdric Pradalier |
| 2022 | AAAI | Generalization in Mean Field Games by Learning Master Policies. | Sarah Perrin, Mathieu Laurire, Julien Prolat, Romuald lie, Matthieu Geist, Olivier Pietquin |
| 2022 | AAAI | Offline Reinforcement Learning as Anti-exploration. | Shideh Rezaeifar, Robert Dadashi, Nino Vieillard, Lonard Hussenot, Olivier Bachem, Olivier Pietquin, Matthieu Geist |
| 2022 | AISTATS | A general class of surrogate functions for stable and efficient reinforcement learning. | Sharan Vaswani, Olivier Bachem, Simone Totaro, Robert Mller, Shivam Garg, Matthieu Geist, Marlos C. Machado, Pablo Samuel Castro, Nicolas Le Roux |
| 2022 | AISTATS | Implicitly Regularized RL with Implicit Q-values. | Nino Vieillard, Marcin Andrychowicz, Anton Raichuk, Olivier Pietquin, Matthieu Geist |
| 2022 | ICML | Continuous Control with Action Quantization from Demonstrations. | Robert Dadashi, Lonard Hussenot, Damien Vincent, Sertan Girgin, Anton Raichuk, Matthieu Geist, Olivier Pietquin |
| 2022 | ICML | Large Batch Experience Replay. | Thibault Lahire, Matthieu Geist, Emmanuel Rachelson |
| 2022 | ICML | Scalable Deep Reinforcement Learning Algorithms for Mean Field Games. | Mathieu Laurire, Sarah Perrin, Sertan Girgin, Paul Muller, Ayush Jain, Theophile Cabannes, Georgios Piliouras, Julien Prolat, Romuald Elie, Olivier Pietquin, Matthieu Geist |
| 2022 | ICRA | Combined Grid and Feature-based Mapping of Metal Structures with Ultrasonic Guided Waves. | Othmane-Latif Ouabi, Ayoub Ridani, Pascal Pomarede, Neil Zeghidour, Nico F. Declercq, Matthieu Geist, Cdric Pradalier |
| 2021 | CoRL | Learning Behaviors through Physics-driven Latent Imagination. | Antoine Richard, Stphanie Aravecchia, Matthieu Geist, Cdric Pradalier |
| 2021 | ICLR | What Matters for On-Policy Deep Actor-Critic Methods? A Large-Scale Study. | Marcin Andrychowicz, Anton Raichuk, Piotr Stanczyk, Manu Orsini, Sertan Girgin, Raphal Marinier, Lonard Hussenot, Matthieu Geist, Olivier Pietquin, Marcin Michalski, Sylvain Gelly, Olivier Bachem |
| 2021 | ICLR | Primal Wasserstein Imitation Learning. | Robert Dadashi, Lonard Hussenot, Matthieu Geist, Olivier Pietquin |
| 2021 | ICLR | Adversarially Guided Actor-Critic. | Yannis Flet-Berliac, Johan Ferret, Olivier Pietquin, Philippe Preux, Matthieu Geist |
| 2021 | ICML | Offline Reinforcement Learning with Pseudometric Learning. | Robert Dadashi, Shideh Rezaeifar, Nino Vieillard, Lonard Hussenot, Olivier Pietquin, Matthieu Geist |
| 2021 | ICML | Hyperparameter Selection for Imitation Learning. | Lonard Hussenot, Marcin Andrychowicz, Damien Vincent, Robert Dadashi, Anton Raichuk, Sabela Ramos, Nikola Momchev, Sertan Girgin, Raphal Marinier, Lukasz Stafiniak, Manu Orsini, Olivier Bachem, Matthieu Geist, Olivier Pietquin |
| 2021 | IJCAI | Mean Field Games Flock! The Reinforcement Learning Way. | Sarah Perrin, Mathieu Laurire, Julien Prolat, Matthieu Geist, Romuald lie, Olivier Pietquin |
| 2020 | AAAI | On the Convergence of Model Free Learning in Mean Field Games. | Romuald Elie, Julien Prolat, Mathieu Laurire, Matthieu Geist, Olivier Pietquin |
| 2020 | AAAI | Deep Conservative Policy Iteration. | Nino Vieillard, Olivier Pietquin, Matthieu Geist |
| 2020 | ACML | Foolproof Cooperative Learning. | Alexis Jacq, Julien Prolat, Matthieu Geist, Olivier Pietquin |
| 2020 | AISTATS | Momentum in Reinforcement Learning. | Nino Vieillard, Bruno Scherrer, Olivier Pietquin, Matthieu Geist |
| 2020 | IJCAI | Self-Attentional Credit Assignment for Transfer in Reinforcement Learning. | Johan Ferret, Raphal Marinier, Matthieu Geist, Olivier Pietquin |
| 2020 | ICRA | Image-Based Place Recognition on Bucolic Environment Across Seasons From Semantic Edge Description. | Assia Benbihi, Stphanie Arravechia, Matthieu Geist, Cdric Pradalier |
| 2019 | CoDIT | Deep Reinforcement Learning-based Continuous Control for Multicopter Systems. | Anush Manukyan, Miguel A. Olivares-Mndez, Matthieu Geist, Holger Voos |
| 2019 | ICCV | ELF: Embedded Localisation of Features in Pre-Trained CNN. | Assia Benbihi, Matthieu Geist, Cdric Pradalier |
| 2019 | ICML | A Theory of Regularized Markov Decision Processes. | Matthieu Geist, Bruno Scherrer, Olivier Pietquin |
| 2019 | ICML | Learning from a Learner. | Alexis Jacq, Matthieu Geist, Ana Paiva, Olivier Pietquin |
| 2019 | ICONIP | Semi-supervised Domain Adaptation with Representation Learning for Semantic Segmentation Across Time. | Assia Benbihi, Matthieu Geist, Cdric Pradalier |
| 2019 | ISCC | Learning Sensor Placement from Demonstration for UAV networks. | Assia Benbihi, Matthieu Geist, Cdric Pradalier |
| 2019 | UIC | Image-Based Text Classification using 2D Convolutional Neural Networks. | Erinc Merdivan, Anastasios Vafeiadis, Dimitrios Kalatzis, Sten Hanke, Joahannes Kroph, Konstantinos Votis, Dimitrios Giakoumis, Dimitrios Tzovaras, Liming Chen, Raouf Hamzaoui, Matthieu Geist |
| 2018 | PERCOM | A Deep Learning Approach for Privacy Preservation in Assisted Living. | Ismini Psychoula, Erinc Merdivan, Deepika Singh, Liming Chen, Feng Chen, Sten Hanke, Johannes Kropf, Andreas Holzinger, Matthieu Geist |
| 2016 | ICML | Softened Approximate Policy Iteration for Markov Games. | Julien Prolat, Bilal Piot, Matthieu Geist, Bruno Scherrer, Olivier Pietquin |
| 2015 | ICML | Imitation Learning Applied to Embodied Conversational Agents. | Bilal Piot, Olivier Pietquin, Matthieu Geist |
| 2015 | IJCAI | Inverse Reinforcement Learning in Relational Domains. | Thibaut Munzer, Bilal Piot, Matthieu Geist, Olivier Pietquin, Manuel Lopes |
| 2014 | Interspeech | Predicting when to laugh with structured classification. | Bilal Piot, Olivier Pietquin, Matthieu Geist |
| 2013 | ICASSP | Random projections: A remedy for overfitting issues in time series prediction with echo state networks. | Lucie Daubigney, Matthieu Geist, Olivier Pietquin |
| 2013 | Interspeech | Particle swarm optimisation of spoken dialogue system strategies. | Lucie Daubigney, Matthieu Geist, Olivier Pietquin |
| 2013 | SIGdial | Model-free POMDP optimisation of tutoring systems with echo-state networks. | Lucie Daubigney, Matthieu Geist, Olivier Pietquin |
| 2012 | ICAISC | Monte-Carlo Swarm Policy Search. | Jrmy Fix, Matthieu Geist |
| 2012 | ICASSP | Clustering behaviors of Spoken Dialogue Systems users. | Senthilkumar Chandramohan, Matthieu Geist, Fabrice Lefvre, Olivier Pietquin |
| 2012 | ICASSP | Off-policy learning in large-scale POMDP-based dialogue systems. | Lucie Daubigney, Matthieu Geist, Olivier Pietquin |
| 2012 | ICML | A Dantzig Selector Approach to Temporal Difference Learning. | Matthieu Geist, Bruno Scherrer, Alessandro Lazaric, Mohammad Ghavamzadeh |
| 2012 | ICML | Approximate Modified Policy Iteration. | Bruno Scherrer, Victor Gabillon, Mohammad Ghavamzadeh, Matthieu Geist |
| 2011 | FUSION | Performance evaluation for particle filters. | Remi Chou, Yvo Boers, Martin Podt, Matthieu Geist |
| 2011 | ICMLA | A Non-parametric Approach to Approximate Dynamic Programming. | Hadrien Glaude, Fadi Akrimi, Matthieu Geist, Olivier Pietquin |
| 2011 | IJCAI | Sample Efficient On-Line Learning of Optimal Dialogue Policies with Kalman Temporal Differences. | Olivier Pietquin, Matthieu Geist, Senthilkumar Chandramohan |
| 2011 | Interspeech | User Simulation in Dialogue Systems Using Inverse Reinforcement Learning. | Senthilkumar Chandramohan, Matthieu Geist, Fabrice Lefvre, Olivier Pietquin |
| 2011 | Interspeech | Uncertainty Management for On-Line Optimisation of a POMDP-Based Large-Scale Spoken Dialogue System. | Lucie Daubigney, Milica Gasic, Senthilkumar Chandramohan, Matthieu Geist, Olivier Pietquin, Steve J. Young |
| 2010 | Interspeech | Optimizing spoken dialogue management with fitted value iteration. | Senthilkumar Chandramohan, Matthieu Geist, Olivier Pietquin |
| 2010 | MDAI | Revisiting Natural Actor-Critics with Value Function Approximation. | Matthieu Geist, Olivier Pietquin |
| 2010 | SIGdial | Sparse Approximate Dynamic Programming for Dialog Management. | Senthilkumar Chandramohan, Matthieu Geist, Olivier Pietquin |
| 2009 | ESANN | Kernelizing Vector Quantization Algorithms. | Matthieu Geist, Olivier Pietquin, Gabriel Fricout |
| 2009 | ICONIP | Tracking in Reinforcement Learning. | Matthieu Geist, Olivier Pietquin, Gabriel Fricout |