| 2025 | ICML | MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces. | Loris Gaven, Thomas Carta, Clment Romac, Cdric Colas, Sylvain Lamprier, Olivier Sigaud, Pierre-Yves Oudeyer |
| 2025 | IROS | RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms. | Zakariae El Asri, Ibrahim Laiche, Clment Rambour, Olivier Sigaud, Nicolas Thome |
| 2025 | NAACL | Reinforcement Learning for Aligning Large Language Models Agents with Interactive Environments: Quantifying and Mitigating Prompt Overfitting. | Mohamed Salim Aissi, Clment Romac, Thomas Carta, Sylvain Lamprier, Pierre-Yves Oudeyer, Olivier Sigaud, Laure Soulier, Nicolas Thome |
| 2024 | ICML | Bridging Environments and Language with Rendering Functions and Vision-Language Models. | Tho Cachet, Christopher R. Dance, Olivier Sigaud |
| 2023 | ICML | Grounding Large Language Models in Interactive Environments with Online Reinforcement Learning. | Thomas Carta, Clment Romac, Thomas Wolf, Sylvain Lamprier, Olivier Sigaud, Pierre-Yves Oudeyer |
| 2023 | ICML | Stein Variational Goal Generation for adaptive Exploration in Multi-Goal Reinforcement Learning. | Nicolas Castanet, Olivier Sigaud, Sylvain Lamprier |
| 2022 | GECCO | Diversity policy gradient for sample efficient quality-diversity optimization. | Thomas Pierrot, Valentin Mac, Flix Chalumeau, Arthur Flajolet, Geoffrey Cideron, Karim Beguir, Antoine Cully, Olivier Sigaud, Nicolas Perrin-Gilbert |
| 2022 | ICIP | Neural Architecture Search for Fracture Classification. | Alos Pourchot, Kvin Bailly, Alexis Ducarouge, Olivier Sigaud |
| 2022 | IROS | Divide & Conquer Imitation Learning. | Alexandre Chenu, Nicolas Perrin-Gilbert, Olivier Sigaud |
| 2021 | ICANN | Selection-Expansion: A Unifying Framework for Motion-Planning and Diversity Search Algorithms. | Alexandre Chenu, Nicolas Perrin, Stphane Doncieux, Olivier Sigaud |
| 2021 | ICANN | First-Order and Second-Order Variants of the Gradient Descent in a Unified Framework. | Thomas Pierrot, Nicolas Perrin-Gilbert, Olivier Sigaud |
| 2021 | ICLR | Grounding Language to Autonomously-Acquired Skills via Goal Generation. | Ahmed Akakzia, Cdric Colas, Pierre-Yves Oudeyer, Mohamed Chetouani, Olivier Sigaud |
| 2020 | ICANN | PBCS: Efficient Exploration and Exploitation Using a Synergy Between Reinforcement Learning and Motion Planning. | Guillaume Matheron, Nicolas Perrin, Olivier Sigaud |
| 2020 | ICANN | Understanding Failures of Deterministic Actor-Critic with Continuous Action Spaces and Sparse Rewards. | Guillaume Matheron, Nicolas Perrin, Olivier Sigaud |
| 2020 | RO-MAN | TIRL: Enriching Actor-Critic RL with non-expert human teachers and a Trust Model. | Flix Rutard, Olivier Sigaud, Mohamed Chetouani |
| 2019 | ICLR | A Hitchhiker's Guide to Statistical Comparisons of Reinforcement Learning Algorithms. | Cdric Colas, Olivier Sigaud, Pierre-Yves Oudeyer |
| 2019 | ICLR | CEM-RL: Combining evolutionary and gradient-based methods for policy search. | Alos Pourchot, Olivier Sigaud |
| 2019 | ICML | CURIOUS: Intrinsically Motivated Modular Multi-Goal Reinforcement Learning. | Cdric Colas, Pierre-Yves Oudeyer, Olivier Sigaud, Pierre Fournier, Mohamed Chetouani |
| 2018 | ICLR | Unsupervised Learning of Goal Spaces for Intrinsically Motivated Goal Exploration. | Alexandre Pr, Sbastien Forestier, Olivier Sigaud, Pierre-Yves Oudeyer |
| 2018 | ICML | GEP-PG: Decoupling Exploration and Exploitation in Deep Reinforcement Learning Algorithms. | Cdric Colas, Olivier Sigaud, Pierre-Yves Oudeyer |
| 2017 | IJCAI | Tensor Based Knowledge Transfer Across Skill Categories for Robot Control. | Chenyang Zhao, Timothy M. Hospedales, Freek Stulp, Olivier Sigaud |
| 2016 | RO-MAN | Training a robot with evaluative feedback and unlabeled guidance signals. | Anis Najar, Olivier Sigaud, Mohamed Chetouani |
| 2015 | GECCO | Socially Guided XCS: Using Teaching Signals to Boost Learning. | Anis Najar, Olivier Sigaud, Mohamed Chetouani |
| 2015 | IROS | Variance modulated task prioritization in Whole-Body Control. | Ryan Lober, Vincent Padois, Olivier Sigaud |
| 2013 | ICML | Gated Autoencoders with Tied Input Weights. | Alain Droniou, Olivier Sigaud |
| 2012 | ICML | Path Integral Policy Improvement with Covariance Matrix Adaptation. | Freek Stulp, Olivier Sigaud |
| 2012 | IROS | Autonomous online learning of velocity kinematics on the iCub: A comparative study. | Alain Droniou, Serena Ivaldi, Vincent Padois, Olivier Sigaud |
| 2011 | GECCO | XCSF with local deletion: preventing detrimental forgetting. | Martin V. Butz, Olivier Sigaud |
| 2011 | GECCO | Learning cost-efficient control policies with XCSF: generalization capabilities and further improvement. | Didier Marin, Jrmie Decock, Lionel Rigoux, Olivier Sigaud |
| 2010 | GECCO | A comparative study: function approximation with LWPR and XCSF. | Patrick O. Stalph, Jrmie Rubinsztajn, Olivier Sigaud, Martin V. Butz |
| 2009 | ICRA | Transfer of knowledge for a climbing Virtual Human: A reinforcement learning approach. | Benoit Libeau, Alain Micaelli, Olivier Sigaud |
| 2009 | IROS | Control of redundant robots using learned models: An operational space control approach. | Camille Salan, Vincent Padois, Olivier Sigaud |
| 2006 | ICML | Learning the structure of Factored Markov Decision Processes in reinforcement learning problems. | Thomas Degris, Olivier Sigaud, Pierre-Henri Wuillemin |
| 2006 | UAI | Chi-square Tests Driven Method for Learning the Structure of Factored MDPs. | Thomas Degris, Olivier Sigaud, Pierre-Henri Wuillemin |
| 2005 | GECCO | ATNoSFERES revisited. | Samuel Landau, Olivier Sigaud, Marc Schoenauer |
| 2004 | GECCO | Improving MACS Thanks to a Comparison with 2TBNs. | Olivier Sigaud, Thierry Gourdin, Pierre-Henri Wuillemin |
| 2003 | GECCO | Designing Efficient Exploration with MACS: Modules and Function Approximation. | Pierre Grard, Olivier Sigaud |
| 2002 | GECCO | A Comparison Between ATNoSFERES And XCSM. | Samuel Landau, Sbastien Picault, Olivier Sigaud, Pierre Grard |