Pierre-Luc Bacon
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
24
Venues
5
Active years
2015–2025
Best venue rank
A*
Where they publish
Papers
24 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICLR | MaestroMotif: Skill Design from Artificial Intelligence Feedback. | Martin Klissarov, Mikael Henaff, Roberta Raileanu, Shagun Sodhani, Pascal Vincent, Amy Zhang, Pierre-Luc Bacon, Doina Precup, Marlos C. Machado, Pierluca D'Oro |
| 2025 | ICML | Scaling Trends in Language Model Robustness. | Nikolaus H. R. Howe, Ian R. McKenzie, Oskar John Hollinsworth, Michal Zajac, Tom Tseng, Aaron David Tucker, Pierre-Luc Bacon, Adam Gleave |
| 2025 | ICML | Network Sparsity Unlocks the Scaling Potential of Deep Reinforcement Learning. | Guozheng Ma, Lu Li, Zilin Wang, Li Shen, Pierre-Luc Bacon, Dacheng Tao |
| 2024 | AISTATS | Maximum entropy GFlowNets with soft Q-learning. | Sobhan Mohammadpour, Emmanuel Bengio, Emma Frejinger, Pierre-Luc Bacon |
| 2024 | ICLR | Course Correcting Koopman Representations. | Mahan Fathi, Clement Gehring, Jonathan Pilault, David Kanaa, Pierre-Luc Bacon, Ross Goroshin |
| 2024 | ICLR | Motif: Intrinsic Motivation from Artificial Intelligence Feedback. | Martin Klissarov, Pierluca D'Oro, Shagun Sodhani, Roberta Raileanu, Pierre-Luc Bacon, Pascal Vincent, Amy Zhang, Mikael Henaff |
| 2024 | ICLR | Decoupling regularization from the action space. | Sobhan Mohammadpour, Emma Frejinger, Pierre-Luc Bacon |
| 2024 | ICLR | Bridging State and History Representations: Understanding Self-Predictive RL. | Tianwei Ni, Benjamin Eysenbach, Erfan Seyedsalehi, Michel Ma, Clement Gehring, Aditya Mahajan, Pierre-Luc Bacon |
| 2024 | ICML | Do Transformer World Models Give Better Policy Gradients? | Michel Ma, Tianwei Ni, Clement Gehring, Pierluca D'Oro, Pierre-Luc Bacon |
| 2023 | ICLR | Sample-Efficient Reinforcement Learning by Breaking the Replay Ratio Barrier. | Pierluca D'Oro, Max Schwarzer, Evgenii Nikishin, Pierre-Luc Bacon, Marc G. Bellemare, Aaron C. Courville |
| 2022 | AAAI | Control-Oriented Model-Based Reinforcement Learning with Implicit Differentiation. | Evgenii Nikishin, Romina Abachi, Rishabh Agarwal, Pierre-Luc Bacon |
| 2022 | ICLR | Continuous-Time Meta-Learning with Forward Mode Differentiation. | Tristan Deleu, David Kanaa, Leo Feng, Giancarlo Kerg, Yoshua Bengio, Guillaume Lajoie, Pierre-Luc Bacon |
| 2022 | ICML | The Primacy Bias in Deep Reinforcement Learning. | Evgenii Nikishin, Max Schwarzer, Pierluca D'Oro, Pierre-Luc Bacon, Aaron C. Courville |
| 2022 | ICML | Direct Behavior Specification via Constrained Reinforcement Learning. | Julien Roy, Roger Girgis, Joshua Romoff, Pierre-Luc Bacon, Christopher J. Pal |
| 2020 | AAAI | Options of Interest: Temporal Abstraction with Interest Functions. | Khimya Khetarpal, Martin Klissarov, Maxime Chevalier-Boisvert, Pierre-Luc Bacon, Doina Precup |
| 2020 | ICML | Understanding the Curse of Horizon in Off-Policy Evaluation via Conditional Importance Sampling. | Yao Liu, Pierre-Luc Bacon, Emma Brunskill |
| 2018 | AAAI | OptionGAN: Learning Joint Reward-Policy Options Using Generative Adversarial Inverse Reinforcement Learning. | Peter Henderson, Wei-Di Chang, Pierre-Luc Bacon, David Meger, Joelle Pineau, Doina Precup |
| 2018 | AAAI | When Waiting Is Not an Option: Learning Options With a Deliberation Cost. | Jean Harb, Pierre-Luc Bacon, Martin Klissarov, Doina Precup |
| 2018 | AAAI | Learning With Options That Terminate Off-Policy. | Anna Harutyunyan, Peter Vrancx, Pierre-Luc Bacon, Doina Precup, Ann Now |
| 2018 | AAAI | Learning Robust Options. | Daniel J. Mankowitz, Timothy A. Mann, Pierre-Luc Bacon, Doina Precup, Shie Mannor |
| 2018 | ICML | Convergent TREE BACKUP and RETRACE with Function Approximation. | Ahmed Touati, Pierre-Luc Bacon, Doina Precup, Pascal Vincent |
| 2017 | AAAI | The Option-Critic Architecture. | Pierre-Luc Bacon, Jean Harb, Doina Precup |
| 2015 | ICML | Analyzing Open Data from the City of Montreal. | Joelle Pineau, Pierre-Luc Bacon |
| 2015 | UAI | Learning and Planning with Timing Information in Markov Decision Processes. | Pierre-Luc Bacon, Borja Balle, Doina Precup |