Bernardo vila Pires
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
11
Venues
4
Active years
2012–2025
Best venue rank
A*
Where they publish
Papers
11 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | AISTATS | A Unifying Framework for Action-Conditional Self-Predictive Reinforcement Learning. | Khimya Khetarpal, Zhaohan Daniel Guo, Bernardo vila Pires, Yunhao Tang, Clare Lyle, Mark Rowland, Nicolas Heess, Diana L. Borsa, Arthur Guez, Will Dabney |
| 2024 | ICML | Human Alignment of Large Language Models through Online Preference Optimisation. | Daniele Calandriello, Zhaohan Daniel Guo, Rmi Munos, Mark Rowland, Yunhao Tang, Bernardo vila Pires, Pierre Harvey Richemond, Charline Le Lan, Michal Valko, Tianqi Liu, Rishabh Joshi, Zeyu Zheng, Bilal Piot |
| 2024 | ICML | Generalized Preference Optimization: A Unified Approach to Offline Alignment. | Yunhao Tang, Zhaohan Daniel Guo, Zeyu Zheng, Daniele Calandriello, Rmi Munos, Mark Rowland, Pierre Harvey Richemond, Michal Valko, Bernardo vila Pires, Bilal Piot |
| 2023 | ICML | Understanding Plasticity in Neural Networks. | Clare Lyle, Zeyu Zheng, Evgenii Nikishin, Bernardo vila Pires, Razvan Pascanu, Will Dabney |
| 2023 | ICML | Understanding Self-Predictive Learning for Reinforcement Learning. | Yunhao Tang, Zhaohan Daniel Guo, Pierre Harvey Richemond, Bernardo vila Pires, Yash Chandak, Rmi Munos, Mark Rowland, Mohammad Gheshlaghi Azar, Charline Le Lan, Clare Lyle, Andrs Gyrgy, Shantanu Thakoor, Will Dabney, Bilal Piot, Daniele Calandriello, Michal Valko |
| 2023 | ICML | DoMo-AC: Doubly Multi-step Off-policy Actor-Critic Algorithm. | Yunhao Tang, Tadashi Kozuno, Mark Rowland, Anna Harutyunyan, Rmi Munos, Bernardo vila Pires, Michal Valko |
| 2020 | ICML | Bootstrap Latent-Predictive Representations for Multitask Reinforcement Learning. | Zhaohan Daniel Guo, Bernardo vila Pires, Bilal Piot, Jean-Bastien Grill, Florent Altch, Rmi Munos, Mohammad Gheshlaghi Azar |
| 2016 | COLT | Policy Error Bounds for Model-Based Reinforcement Learning with Factored Linear Models. | Bernardo vila Pires |
| 2015 | AAAI | Pathological Effects of Variance on Classification-Based Policy Iteration. | Bernardo vila Pires, Csaba Szepesvri |
| 2013 | ICML | Cost-sensitive Multiclass Classification Risk Bounds. | Bernardo vila Pires, Csaba Szepesvri, Mohammad Ghavamzadeh |
| 2012 | ICML | Statistical linear estimation with penalized estimators: an application to reinforcement learning. | Bernardo vila Pires, Csaba Szepesvri |