Skip to content

Matthieu Geist

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

66

Venues

26

Active years

2009–2025

Best venue rank

A*

Where they publish

Papers

66 indexed papers, newest first.

YearVenueTitleAuthors
2025ICLRSelf-Improving Robust Preference Optimization.Eugene Choi, Arash Ahmadian, Matthieu Geist, Olivier Pietquin, Mohammad Gheshlaghi Azar
2024AAAILearning Discrete-Time Major-Minor Mean Field Games.Kai Cui, Gke Dayanikli, Mathieu Laurire, Matthieu Geist, Olivier Pietquin, Heinz Koeppl
2024EMNLPContrastive Policy Gradient: Aligning LLMs on sequence-level scores in a supervised-friendly fashion.Yannis Flet-Berliac, Nathan Grinsztajn, Florian Strub, Eugene Choi, Bill Wu, Chris Cremer, Arash Ahmadian, Yash Chandak, Mohammad Gheshlaghi Azar, Olivier Pietquin, Matthieu Geist
2024ICLROn-Policy Distillation of Language Models: Learning from Self-Generated Mistakes.Rishabh Agarwal, Nino Vieillard, Yongchao Zhou, Piotr Stanczyk, Sabela Ramos Garea, Matthieu Geist, Olivier Bachem
2024ICLRClosing the Gap between TD Learning and Supervised Learning - A Generalisation Point of View.Raj Ghugare, Matthieu Geist, Glen Berseth, Benjamin Eysenbach
2024ICMLMusicRL: Aligning Music Generation to Human Preferences.Geoffrey Cideron, Sertan Girgin, Mauro Verzetti, Damien Vincent, Matej Kastelic, Zaln Borsos, Brian McWilliams, Victor Ungureanu, Olivier Bachem, Olivier Pietquin, Matthieu Geist, Lonard Hussenot, Neil Zeghidour, Andrea Agostinelli
2024ICMLNash Learning from Human Feedback.Rmi Munos, Michal Valko, Daniele Calandriello, Mohammad Gheshlaghi Azar, Mark Rowland, Daniel Guo, Yunhao Tang, Matthieu Geist, Thomas Mesnard, Cme Fiegel, Andrea Michi, Marco Selvi, Sertan Girgin, Nikola Momchev, Olivier Bachem, Daniel J. Mankowitz, Doina Precup, Bilal Piot
2024IROSDRIFT: Deep Reinforcement Learning for Intelligent Floating Platforms Trajectories.Matteo El Hariry, Antoine Richard, Vivek Muralidharan, Matthieu Geist, Miguel A. Olivares-Mndez
2024UAITowards Minimax Optimality of Model-based Robust Reinforcement Learning.Pierre Clavier, Erwan Le Pennec, Matthieu Geist
2023ACLFactually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback.Paul Roit, Johan Ferret, Lior Shani, Roee Aharoni, Geoffrey Cideron, Robert Dadashi, Matthieu Geist, Sertan Girgin, Lonard Hussenot, Orgad Keller, Nikola Momchev, Sabela Ramos Garea, Piotr Stanczyk, Nino Vieillard, Olivier Bachem, Gal Elidan, Avinatan Hassidim, Olivier Pietquin, Idan Szpektor
2023ICLRExtreme Q-Learning: MaxEnt RL without Entropy.Divyansh Garg, Joey Hejna, Matthieu Geist, Stefano Ermon
2023ICMLA Connection between One-Step RL and Critic Regularization in Reinforcement Learning.Benjamin Eysenbach, Matthieu Geist, Sergey Levine, Ruslan Salakhutdinov
2023ICMLRegularization and Variance-Weighted Regression Achieves Minimax Optimality in Linear MDPs: Theory and Practice.Toshinori Kitamura, Tadashi Kozuno, Yunhao Tang, Nino Vieillard, Michal Valko, Wenhao Yang, Jincheng Mei, Pierre Mnard, Mohammad Gheshlaghi Azar, Rmi Munos, Olivier Pietquin, Matthieu Geist, Csaba Szepesvri, Wataru Kumagai, Yutaka Matsuo
2023ICMLPolicy Mirror Ascent for Efficient and Independent Learning in Mean Field Games.Batuhan Yardim, Semih Cayci, Matthieu Geist, Niao He
2023ICRAPose-graph SLAM Using Multi-order Ultrasonic Echoes and Beamforming for Long-range Inspection Robots.Othmane-Latif Ouabi, Neil Zeghidour, Nico F. Declercq, Matthieu Geist, Cdric Pradalier
2022AAAIGeneralization in Mean Field Games by Learning Master Policies.Sarah Perrin, Mathieu Laurire, Julien Prolat, Romuald lie, Matthieu Geist, Olivier Pietquin
2022AAAIOffline Reinforcement Learning as Anti-exploration.Shideh Rezaeifar, Robert Dadashi, Nino Vieillard, Lonard Hussenot, Olivier Bachem, Olivier Pietquin, Matthieu Geist
2022AISTATSA general class of surrogate functions for stable and efficient reinforcement learning.Sharan Vaswani, Olivier Bachem, Simone Totaro, Robert Mller, Shivam Garg, Matthieu Geist, Marlos C. Machado, Pablo Samuel Castro, Nicolas Le Roux
2022AISTATSImplicitly Regularized RL with Implicit Q-values.Nino Vieillard, Marcin Andrychowicz, Anton Raichuk, Olivier Pietquin, Matthieu Geist
2022ICMLContinuous Control with Action Quantization from Demonstrations.Robert Dadashi, Lonard Hussenot, Damien Vincent, Sertan Girgin, Anton Raichuk, Matthieu Geist, Olivier Pietquin
2022ICMLLarge Batch Experience Replay.Thibault Lahire, Matthieu Geist, Emmanuel Rachelson
2022ICMLScalable Deep Reinforcement Learning Algorithms for Mean Field Games.Mathieu Laurire, Sarah Perrin, Sertan Girgin, Paul Muller, Ayush Jain, Theophile Cabannes, Georgios Piliouras, Julien Prolat, Romuald Elie, Olivier Pietquin, Matthieu Geist
2022ICRACombined Grid and Feature-based Mapping of Metal Structures with Ultrasonic Guided Waves.Othmane-Latif Ouabi, Ayoub Ridani, Pascal Pomarede, Neil Zeghidour, Nico F. Declercq, Matthieu Geist, Cdric Pradalier
2021CoRLLearning Behaviors through Physics-driven Latent Imagination.Antoine Richard, Stphanie Aravecchia, Matthieu Geist, Cdric Pradalier
2021ICLRWhat Matters for On-Policy Deep Actor-Critic Methods? A Large-Scale Study.Marcin Andrychowicz, Anton Raichuk, Piotr Stanczyk, Manu Orsini, Sertan Girgin, Raphal Marinier, Lonard Hussenot, Matthieu Geist, Olivier Pietquin, Marcin Michalski, Sylvain Gelly, Olivier Bachem
2021ICLRPrimal Wasserstein Imitation Learning.Robert Dadashi, Lonard Hussenot, Matthieu Geist, Olivier Pietquin
2021ICLRAdversarially Guided Actor-Critic.Yannis Flet-Berliac, Johan Ferret, Olivier Pietquin, Philippe Preux, Matthieu Geist
2021ICMLOffline Reinforcement Learning with Pseudometric Learning.Robert Dadashi, Shideh Rezaeifar, Nino Vieillard, Lonard Hussenot, Olivier Pietquin, Matthieu Geist
2021ICMLHyperparameter Selection for Imitation Learning.Lonard Hussenot, Marcin Andrychowicz, Damien Vincent, Robert Dadashi, Anton Raichuk, Sabela Ramos, Nikola Momchev, Sertan Girgin, Raphal Marinier, Lukasz Stafiniak, Manu Orsini, Olivier Bachem, Matthieu Geist, Olivier Pietquin
2021IJCAIMean Field Games Flock! The Reinforcement Learning Way.Sarah Perrin, Mathieu Laurire, Julien Prolat, Matthieu Geist, Romuald lie, Olivier Pietquin
2020AAAIOn the Convergence of Model Free Learning in Mean Field Games.Romuald Elie, Julien Prolat, Mathieu Laurire, Matthieu Geist, Olivier Pietquin
2020AAAIDeep Conservative Policy Iteration.Nino Vieillard, Olivier Pietquin, Matthieu Geist
2020ACMLFoolproof Cooperative Learning.Alexis Jacq, Julien Prolat, Matthieu Geist, Olivier Pietquin
2020AISTATSMomentum in Reinforcement Learning.Nino Vieillard, Bruno Scherrer, Olivier Pietquin, Matthieu Geist
2020IJCAISelf-Attentional Credit Assignment for Transfer in Reinforcement Learning.Johan Ferret, Raphal Marinier, Matthieu Geist, Olivier Pietquin
2020ICRAImage-Based Place Recognition on Bucolic Environment Across Seasons From Semantic Edge Description.Assia Benbihi, Stphanie Arravechia, Matthieu Geist, Cdric Pradalier
2019CoDITDeep Reinforcement Learning-based Continuous Control for Multicopter Systems.Anush Manukyan, Miguel A. Olivares-Mndez, Matthieu Geist, Holger Voos
2019ICCVELF: Embedded Localisation of Features in Pre-Trained CNN.Assia Benbihi, Matthieu Geist, Cdric Pradalier
2019ICMLA Theory of Regularized Markov Decision Processes.Matthieu Geist, Bruno Scherrer, Olivier Pietquin
2019ICMLLearning from a Learner.Alexis Jacq, Matthieu Geist, Ana Paiva, Olivier Pietquin
2019ICONIPSemi-supervised Domain Adaptation with Representation Learning for Semantic Segmentation Across Time.Assia Benbihi, Matthieu Geist, Cdric Pradalier
2019ISCCLearning Sensor Placement from Demonstration for UAV networks.Assia Benbihi, Matthieu Geist, Cdric Pradalier
2019UICImage-Based Text Classification using 2D Convolutional Neural Networks.Erinc Merdivan, Anastasios Vafeiadis, Dimitrios Kalatzis, Sten Hanke, Joahannes Kroph, Konstantinos Votis, Dimitrios Giakoumis, Dimitrios Tzovaras, Liming Chen, Raouf Hamzaoui, Matthieu Geist
2018PERCOMA Deep Learning Approach for Privacy Preservation in Assisted Living.Ismini Psychoula, Erinc Merdivan, Deepika Singh, Liming Chen, Feng Chen, Sten Hanke, Johannes Kropf, Andreas Holzinger, Matthieu Geist
2016ICMLSoftened Approximate Policy Iteration for Markov Games.Julien Prolat, Bilal Piot, Matthieu Geist, Bruno Scherrer, Olivier Pietquin
2015ICMLImitation Learning Applied to Embodied Conversational Agents.Bilal Piot, Olivier Pietquin, Matthieu Geist
2015IJCAIInverse Reinforcement Learning in Relational Domains.Thibaut Munzer, Bilal Piot, Matthieu Geist, Olivier Pietquin, Manuel Lopes
2014InterspeechPredicting when to laugh with structured classification.Bilal Piot, Olivier Pietquin, Matthieu Geist
2013ICASSPRandom projections: A remedy for overfitting issues in time series prediction with echo state networks.Lucie Daubigney, Matthieu Geist, Olivier Pietquin
2013InterspeechParticle swarm optimisation of spoken dialogue system strategies.Lucie Daubigney, Matthieu Geist, Olivier Pietquin
2013SIGdialModel-free POMDP optimisation of tutoring systems with echo-state networks.Lucie Daubigney, Matthieu Geist, Olivier Pietquin
2012ICAISCMonte-Carlo Swarm Policy Search.Jrmy Fix, Matthieu Geist
2012ICASSPClustering behaviors of Spoken Dialogue Systems users.Senthilkumar Chandramohan, Matthieu Geist, Fabrice Lefvre, Olivier Pietquin
2012ICASSPOff-policy learning in large-scale POMDP-based dialogue systems.Lucie Daubigney, Matthieu Geist, Olivier Pietquin
2012ICMLA Dantzig Selector Approach to Temporal Difference Learning.Matthieu Geist, Bruno Scherrer, Alessandro Lazaric, Mohammad Ghavamzadeh
2012ICMLApproximate Modified Policy Iteration.Bruno Scherrer, Victor Gabillon, Mohammad Ghavamzadeh, Matthieu Geist
2011FUSIONPerformance evaluation for particle filters.Remi Chou, Yvo Boers, Martin Podt, Matthieu Geist
2011ICMLAA Non-parametric Approach to Approximate Dynamic Programming.Hadrien Glaude, Fadi Akrimi, Matthieu Geist, Olivier Pietquin
2011IJCAISample Efficient On-Line Learning of Optimal Dialogue Policies with Kalman Temporal Differences.Olivier Pietquin, Matthieu Geist, Senthilkumar Chandramohan
2011InterspeechUser Simulation in Dialogue Systems Using Inverse Reinforcement Learning.Senthilkumar Chandramohan, Matthieu Geist, Fabrice Lefvre, Olivier Pietquin
2011InterspeechUncertainty Management for On-Line Optimisation of a POMDP-Based Large-Scale Spoken Dialogue System.Lucie Daubigney, Milica Gasic, Senthilkumar Chandramohan, Matthieu Geist, Olivier Pietquin, Steve J. Young
2010InterspeechOptimizing spoken dialogue management with fitted value iteration.Senthilkumar Chandramohan, Matthieu Geist, Olivier Pietquin
2010MDAIRevisiting Natural Actor-Critics with Value Function Approximation.Matthieu Geist, Olivier Pietquin
2010SIGdialSparse Approximate Dynamic Programming for Dialog Management.Senthilkumar Chandramohan, Matthieu Geist, Olivier Pietquin
2009ESANNKernelizing Vector Quantization Algorithms.Matthieu Geist, Olivier Pietquin, Gabriel Fricout
2009ICONIPTracking in Reinforcement Learning.Matthieu Geist, Olivier Pietquin, Gabriel Fricout