Skip to content

Zhaohan Daniel Guo

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

10

Venues

3

Active years

2016–2025

Best venue rank

A*

Where they publish

Papers

10 indexed papers, newest first.

YearVenueTitleAuthors
2025AISTATSA Unifying Framework for Action-Conditional Self-Predictive Reinforcement Learning.Khimya Khetarpal, Zhaohan Daniel Guo, Bernardo vila Pires, Yunhao Tang, Clare Lyle, Mark Rowland, Nicolas Heess, Diana L. Borsa, Arthur Guez, Will Dabney
2024AISTATSA General Theoretical Paradigm to Understand Learning from Human Preferences.Mohammad Gheshlaghi Azar, Zhaohan Daniel Guo, Bilal Piot, Rmi Munos, Mark Rowland, Michal Valko, Daniele Calandriello
2024ICMLHuman Alignment of Large Language Models through Online Preference Optimisation.Daniele Calandriello, Zhaohan Daniel Guo, Rmi Munos, Mark Rowland, Yunhao Tang, Bernardo vila Pires, Pierre Harvey Richemond, Charline Le Lan, Michal Valko, Tianqi Liu, Rishabh Joshi, Zeyu Zheng, Bilal Piot
2024ICMLGeneralized Preference Optimization: A Unified Approach to Offline Alignment.Yunhao Tang, Zhaohan Daniel Guo, Zeyu Zheng, Daniele Calandriello, Rmi Munos, Mark Rowland, Pierre Harvey Richemond, Michal Valko, Bernardo vila Pires, Bilal Piot
2023ICMLRepresentations and Exploration for Deep Reinforcement Learning using Singular Value Decomposition.Yash Chandak, Shantanu Thakoor, Zhaohan Daniel Guo, Yunhao Tang, Rmi Munos, Will Dabney, Diana L. Borsa
2023ICMLUnderstanding Self-Predictive Learning for Reinforcement Learning.Yunhao Tang, Zhaohan Daniel Guo, Pierre Harvey Richemond, Bernardo vila Pires, Yash Chandak, Rmi Munos, Mark Rowland, Mohammad Gheshlaghi Azar, Charline Le Lan, Clare Lyle, Andrs Gyrgy, Shantanu Thakoor, Will Dabney, Bilal Piot, Daniele Calandriello, Michal Valko
2020ICLRNever Give Up: Learning Directed Exploration Strategies.Adri Puigdomnech Badia, Pablo Sprechmann, Alex Vitvitskyi, Zhaohan Daniel Guo, Bilal Piot, Steven Kapturowski, Olivier Tieleman, Martn Arjovsky, Alexander Pritzel, Andrew Bolt, Charles Blundell
2020ICMLAgent57: Outperforming the Atari Human Benchmark.Adri Puigdomnech Badia, Bilal Piot, Steven Kapturowski, Pablo Sprechmann, Alex Vitvitskyi, Zhaohan Daniel Guo, Charles Blundell
2020ICMLBootstrap Latent-Predictive Representations for Multitask Reinforcement Learning.Zhaohan Daniel Guo, Bernardo vila Pires, Bilal Piot, Jean-Bastien Grill, Florent Altch, Rmi Munos, Mohammad Gheshlaghi Azar
2016AISTATSA PAC RL Algorithm for Episodic POMDPs.Zhaohan Daniel Guo, Shayan Doroudi, Emma Brunskill