Skip to content

Kaiqing Zhang

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

28

Venues

10

Active years

2015–2026

Best venue rank

A*

Where they publish

Papers

28 indexed papers, newest first.

YearVenueTitleAuthors
2026COLTRegret Minimization with Adaptive Opponents in Repeated Games.Mingyang Liu, Asuman Ozdaglar, Tiancheng Yu, Kaiqing Zhang
2025AAAIFoundations of Multi-Agent Learning in Dynamic Environments: Where Reinforcement Learning Meets Strategic Decision-Making.Kaiqing Zhang
2025ACLMAPoRL: Multi-Agent Post-Co-Training for Collaborative Large Language Models with Reinforcement Learning.Chanwoo Park, Seungju Han, Xingzhi Guo, Asuman E. Ozdaglar, Kaiqing Zhang, Joo-Kyung Kim
2025ICLRDo LLM Agents Have Regret? A Case Study in Online Learning and Games.Chanwoo Park, Xiangyu Liu, Asuman E. Ozdaglar, Kaiqing Zhang
2025IPCCCUnsupervised Indoor Relative Positioning and Group Proximity Prediction Based on CIR Graph Convolutional Network and MDS Embedding.Kaiqing Zhang, Xuemin Hong, Ao Peng
2024ICLRRobot Fleet Learning via Policy Merging.Lirui Wang, Kaiqing Zhang, Allan Zhou, Max Simchowitz, Russ Tedrake
2023AISTATSByzantine-Robust Online and Offline Distributed Reinforcement Learning.Yiding Chen, Xuezhou Zhang, Kaiqing Zhang, Mengdi Wang, Xiaojin Zhu
2023AISTATSSymmetric (Optimistic) Natural Policy Gradient for Multi-Agent Learning with Parameter Convergence.Sarath Pattathil, Kaiqing Zhang, Asuman E. Ozdaglar
2023COLTBreaking the Curse of Multiagents in a Large State Space: RL in Markov Games with Independent Linear Function Approximation.Qiwen Cui, Kaiqing Zhang, Simon S. Du
2023COLTThe Complexity of Markov Equilibrium in Stochastic Games.Constantinos Daskalakis, Noah Golowich, Kaiqing Zhang
2023COLTTackling Combinatorial Distribution Shift: A Matrix Completion Perspective.Max Simchowitz, Abhishek Gupta, Kaiqing Zhang
2023ICLRThe Power of Regularization in Solving Extensive-Form Games.Mingyang Liu, Asuman E. Ozdaglar, Tiancheng Yu, Kaiqing Zhang
2023ICLRLearning to Extrapolate: A Transductive Approach.Aviv Netanyahu, Abhishek Gupta, Max Simchowitz, Kaiqing Zhang, Pulkit Agrawal
2023ICLRDoes Learning from Decentralized Non-IID Unlabeled Data Benefit from Self Supervision?Lirui Wang, Kaiqing Zhang, Yunzhu Li, Yonglong Tian, Russ Tedrake
2023ICMLPartially Observable Multi-agent RL with (Quasi-)Efficiency: The Blessing of Information Sharing.Xiangyu Liu, Kaiqing Zhang
2023ICMLRevisiting the Linear-Programming Framework for Offline RL with General Function Approximation.Asuman E. Ozdaglar, Sarath Pattathil, Jiawei Zhang, Kaiqing Zhang
2022ICMLIndependent Policy Gradient for Large-Scale Markov Potential Games: Sharper Rates, Function Approximation, and Game-Agnostic Convergence.Dongsheng Ding, Chen-Yu Wei, Kaiqing Zhang, Mihailo R. Jovanovic
2022ICMLOn Improving Model-Free Algorithms for Decentralized Multi-Agent Reinforcement Learning.Weichao Mao, Lin Yang, Kaiqing Zhang, Tamer Basar
2022ICMLDo Differentiable Simulators Give Better Policy Gradients?Hyung Ju Terry Suh, Max Simchowitz, Kaiqing Zhang, Russ Tedrake
2021AAAIDecentralized Policy Gradient Descent Ascent for Safe Multi-Agent Reinforcement Learning.Songtao Lu, Kaiqing Zhang, Tianyi Chen, Tamer Basar, Lior Horesh
2021ICLRLearning Safe Multi-agent Control with Decentralized Neural Barrier Certificates.Zengyi Qin, Kaiqing Zhang, Yuxiao Chen, Jingkai Chen, Chuchu Fan
2021ICMLNear-Optimal Model-Free Reinforcement Learning in Non-Stationary Episodic MDPs.Weichao Mao, Kaiqing Zhang, Ruihao Zhu, David Simchi-Levi, Tamer Basar
2021ICMLReinforcement Learning for Cost-Aware Markov Decision Processes.Wesley Suttle, Kaiqing Zhang, Zhuoran Yang, Ji Liu, David N. Kraemer
2019CISSPolicy Search in Infinite-Horizon Discounted Reinforcement Learning: Advances through Connections to Non-Convex Optimization : Invited Presentation.Kaiqing Zhang, Alec Koppel, Hao Zhu, Tamer Basar
2018AISTATSNonlinear Structured Signal Estimation in High Dimensions via Iterative Hard Thresholding.Kaiqing Zhang, Zhuoran Yang, Zhaoran Wang
2018ICMLFully Decentralized Multi-Agent Reinforcement Learning with Networked Agents.Kaiqing Zhang, Zhuoran Yang, Han Liu, Tong Zhang, Tamer Basar
2015GLOBECOMSequential Detection Aided Modulation Classification in Cognitive Radio Networks.Lubing Han, Feifei Gao, Kaiqing Zhang, Shun Zhang
2015PIMRCSpectrum prediction and channel selection for sensing-based spectrum sharing scheme using online learning techniques.Zhao Zhang, Kaiqing Zhang, Feifei Gao, Shun Zhang