| 2025 | ICML | Rejecting Hallucinated State Targets during Planning. | Harry Zhao, Tristan Sylvain, Romain Laroche, Doina Precup, Yoshua Bengio |
| 2025 | ICML | Learning Fused State Representations for Control from Multi-View Observations. | Zeyu Wang, Yao-Hui Li, Xin Li, Hongyu Zang, Romain Laroche, Riashat Islam |
| 2024 | ICLR | Consciousness-Inspired Spatio-Temporal Abstractions for Better Generalization in Reinforcement Learning. | Harry Zhao, Safa Alver, Harm van Seijen, Romain Laroche, Doina Precup, Yoshua Bengio |
| 2024 | ICML | Think Before You Act: Decision Transformers with Working Memory. | Jikun Kang, Romain Laroche, Xingdi Yuan, Adam Trischler, Xue Liu, Jie Fu |
| 2023 | ICLR | Harnessing Mixed Offline Reinforcement Learning Datasets via Trajectory Weighting. | Zhang-Wei Hong, Pulkit Agrawal, Remi Tachet des Combes, Romain Laroche |
| 2023 | ICLR | Behavior Prior Representation learning for Offline Reinforcement Learning. | Hongyu Zang, Xin Li, Jie Yu, Chen Liu, Riashat Islam, Remi Tachet des Combes, Romain Laroche |
| 2023 | ICML | On the Occupancy Measure of Non-Markovian Policies in Continuous MDPs. | Romain Laroche, Remi Tachet des Combes |
| 2023 | ICML | On the Convergence of SARSA with Linear Function Approximation. | Shangtong Zhang, Remi Tachet des Combes, Romain Laroche |
| 2022 | AISTATS | Beyond the Policy Gradient Theorem for Efficient Policy Updates in Actor-Critic Algorithms. | Romain Laroche, Remi Tachet des Combes |
| 2021 | CoNLL | The Emergence of the Shape Bias Results from Communicative Efficiency. | Eva Portelance, Michael C. Frank, Dan Jurafsky, Alessandro Sordoni, Romain Laroche |
| 2020 | IJCAI | Reinforcement Learning Framework for Deep Brain Stimulation Study. | Dmitrii Krylov, Remi Tachet des Combes, Romain Laroche, Michael Rosenblum, Dmitry V. Dylov |
| 2019 | ICML | Decentralized Exploration in Multi-Armed Bandits. | Raphal Fraud, Rda Alami, Romain Laroche |
| 2019 | ICML | Safe Policy Improvement with Baseline Bootstrapping. | Romain Laroche, Paul Trichelair, Remi Tachet des Combes |
| 2018 | AAAI | On Value Function Representation of Long Horizon Problems. | Lucas Lehnert, Romain Laroche, Harm van Seijen |
| 2018 | ICLR | Reinforcement Learning Algorithm Selection. | Romain Laroche, Raphal Fraud |
| 2018 | ICLR | In reinforcement learning, all objective functions are not equal. | Romain Laroche, Harm van Seijen |
| 2017 | AAAI | Transfer Reinforcement Learning with Shared Dynamics. | Romain Laroche, Merwan Barlier |
| 2016 | IJCAI | Reinforcement Learning for Turn-Taking Management in Incremental Spoken Dialogue Systems. | Hatim Khouzaimi, Romain Laroche, Fabrice Lefvre |
| 2016 | Interspeech | A Stochastic Model for Computer-Aided Human-Human Dialogue. | Merwan Barlier, Romain Laroche, Olivier Pietquin |
| 2015 | EMNLP | Turn-taking phenomena in incremental dialogue systems. | Hatim Khouzaimi, Romain Laroche, Fabrice Lefvre |
| 2015 | HCI | Dialogue Efficiency Evaluation of Turn-Taking Phenomena in a Multi-layer Incremental Simulated Environment. | Hatim Khouzaimi, Romain Laroche, Fabrice Lefvre |
| 2015 | ICIN | Content finder AssistanT. | Romain Laroche |
| 2015 | SIGdial | Human-Machine Dialogue as a Stochastic Game. | Merwan Barlier, Julien Prolat, Romain Laroche, Olivier Pietquin |
| 2015 | SIGdial | Optimising Turn-Taking Strategies With Reinforcement Learning. | Hatim Khouzaimi, Romain Laroche, Fabrice Lefvre |
| 2014 | ICASSP | Ordinal regression for interaction quality prediction. | Layla El Asri, Hatim Khouzaimi, Romain Laroche, Olivier Pietquin |
| 2014 | ICONIP | Contextual Bandit for Active Learning: Active Thompson Sampling. | Djallel Bouneffouf, Romain Laroche, Tanguy Urvoy, Raphal Fraud, Robin Allesiardo |
| 2014 | LREC | NASTIA: Negotiating Appointment Setting Interface. | Layla El Asri, Rmi Lemonnier, Romain Laroche, Olivier Pietquin, Hatim Khouzaimi |
| 2014 | LREC | DINASTI: Dialogues with a Negotiating Appointment Setting Interface. | Layla El Asri, Romain Laroche, Olivier Pietquin |
| 2014 | SIGdial | An easy method to make dialogue systems incremental. | Hatim Khouzaimi, Romain Laroche, Fabrice Lefvre |
| 2013 | SIGdial | Will my Spoken Dialogue System be a Slow Learner ? | Layla El Asri, Romain Laroche |
| 2010 | Interspeech | Enhanced monitoring tools and online dialogue optimisation merged into a new spoken dialogue system design experience. | Romain Laroche, Philippe Bretier, Ghislain Putois |
| 2010 | Interspeech | Optimising a handcrafted dialogue system design. | Romain Laroche, Ghislain Putois, Philippe Bretier |
| 2010 | SIGdial | Enhanced Monitoring Tools and Online Dialogue Optimisation Merged into a New Spoken Dialogue System Design Experience. | Ghislain Putois, Romain Laroche, Philippe Bretier |
| 2009 | Interspeech | Hybridisation of expertise and reinforcement learning in dialogue systems. | Romain Laroche, Ghislain Putois, Philippe Bretier, Bernadette Bouchon-Meunier |