Skip to content

Canzhe Zhao

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

11

Venues

10

Active years

2021–2025

Best venue rank

A*

Where they publish

Papers

11 indexed papers, newest first.

YearVenueTitleAuthors
2025AAAILogarithmic Regret for Linear Markov Decision Processes with Adversarial Corruptions.Canzhe Zhao, Xiangcheng Zhang, Baoxiang Wang, Shuai Li
2025ICMLLearning Imperfect Information Extensive-form Games with Last-iterate Convergence under Bandit Feedback.Canzhe Zhao, Yutian Cheng, Jing Dong, Baoxiang Wang, Shuai Li
2025UAITowards Provably Efficient Learning of Imperfect Information Extensive-Form Games with Linear Function Approximation.Canzhe Zhao, Shuze Chen, Weiming Liu, Haobo Fu, Qiang Fu, Shuai Li
2023COLTBest-of-three-worlds Analysis for Linear Bandits with Follow-the-regularized-leader Algorithm.Fang Kong, Canzhe Zhao, Shuai Li
2023ICLRLearning Adversarial Linear Mixture Markov Decision Processes with Bandit Feedback and Unknown Transition.Canzhe Zhao, Ruofeng Yang, Baoxiang Wang, Shuai Li
2023IJCAIDPMAC: Differentially Private Communication for Cooperative Multi-Agent Reinforcement Learning.Canzhe Zhao, Yanjie Ze, Jing Dong, Baoxiang Wang, Shuai Li
2023WSDMDifferentially Private Temporal Difference Learning with Stochastic Nonconvex-Strongly-Concave Optimization.Canzhe Zhao, Yanjie Ze, Jing Dong, Baoxiang Wang, Shuai Li
2022AAAISimultaneously Learning Stochastic and Adversarial Bandits under the Position-Based Model.Cheng Chen, Canzhe Zhao, Shuai Li
2022WWWKnowledge-aware Conversational Preference Elicitation with Bandit Feedback.Canzhe Zhao, Tong Yu, Zhihui Xie, Shuai Li
2021CIKMClustering of Conversational Bandits for User Preference Learning and Elicitation.Junda Wu, Canzhe Zhao, Tong Yu, Jingyang Li, Shuai Li
2021SIGIRComparison-based Conversational Recommender System with Relative Bandit Feedback.Zhihui Xie, Tong Yu, Canzhe Zhao, Shuai Li