Zhaohan Daniel Guo
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
10
Venues
3
Active years
2016–2025
Best venue rank
A*
Where they publish
Papers
10 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | AISTATS | A Unifying Framework for Action-Conditional Self-Predictive Reinforcement Learning. | Khimya Khetarpal, Zhaohan Daniel Guo, Bernardo vila Pires, Yunhao Tang, Clare Lyle, Mark Rowland, Nicolas Heess, Diana L. Borsa, Arthur Guez, Will Dabney |
| 2024 | AISTATS | A General Theoretical Paradigm to Understand Learning from Human Preferences. | Mohammad Gheshlaghi Azar, Zhaohan Daniel Guo, Bilal Piot, Rmi Munos, Mark Rowland, Michal Valko, Daniele Calandriello |
| 2024 | ICML | Human Alignment of Large Language Models through Online Preference Optimisation. | Daniele Calandriello, Zhaohan Daniel Guo, Rmi Munos, Mark Rowland, Yunhao Tang, Bernardo vila Pires, Pierre Harvey Richemond, Charline Le Lan, Michal Valko, Tianqi Liu, Rishabh Joshi, Zeyu Zheng, Bilal Piot |
| 2024 | ICML | Generalized Preference Optimization: A Unified Approach to Offline Alignment. | Yunhao Tang, Zhaohan Daniel Guo, Zeyu Zheng, Daniele Calandriello, Rmi Munos, Mark Rowland, Pierre Harvey Richemond, Michal Valko, Bernardo vila Pires, Bilal Piot |
| 2023 | ICML | Representations and Exploration for Deep Reinforcement Learning using Singular Value Decomposition. | Yash Chandak, Shantanu Thakoor, Zhaohan Daniel Guo, Yunhao Tang, Rmi Munos, Will Dabney, Diana L. Borsa |
| 2023 | ICML | Understanding Self-Predictive Learning for Reinforcement Learning. | Yunhao Tang, Zhaohan Daniel Guo, Pierre Harvey Richemond, Bernardo vila Pires, Yash Chandak, Rmi Munos, Mark Rowland, Mohammad Gheshlaghi Azar, Charline Le Lan, Clare Lyle, Andrs Gyrgy, Shantanu Thakoor, Will Dabney, Bilal Piot, Daniele Calandriello, Michal Valko |
| 2020 | ICLR | Never Give Up: Learning Directed Exploration Strategies. | Adri Puigdomnech Badia, Pablo Sprechmann, Alex Vitvitskyi, Zhaohan Daniel Guo, Bilal Piot, Steven Kapturowski, Olivier Tieleman, Martn Arjovsky, Alexander Pritzel, Andrew Bolt, Charles Blundell |
| 2020 | ICML | Agent57: Outperforming the Atari Human Benchmark. | Adri Puigdomnech Badia, Bilal Piot, Steven Kapturowski, Pablo Sprechmann, Alex Vitvitskyi, Zhaohan Daniel Guo, Charles Blundell |
| 2020 | ICML | Bootstrap Latent-Predictive Representations for Multitask Reinforcement Learning. | Zhaohan Daniel Guo, Bernardo vila Pires, Bilal Piot, Jean-Bastien Grill, Florent Altch, Rmi Munos, Mohammad Gheshlaghi Azar |
| 2016 | AISTATS | A PAC RL Algorithm for Episodic POMDPs. | Zhaohan Daniel Guo, Shayan Doroudi, Emma Brunskill |