| 2026 | AAAI | An MRP Formulation for Supervised Learning: Generalized Temporal Difference Learning Models (Abstract Reprint). | Yangchen Pan, Junfeng Wen, Chenjun Xiao, Philip Torr |
| 2025 | ICML | PANDAS: Improving Many-shot Jailbreaking via Positive Affirmation, Negative Demonstration, and Adaptive Sampling. | Avery Ma, Yangchen Pan, Amir-massoud Farahmand |
| 2024 | ECCV | Improving Adversarial Transferability via Model Alignment. | Avery Ma, Amir-massoud Farahmand, Yangchen Pan, Philip Torr, Jindong Gu |
| 2024 | ICML | Position: Reinforcement Learning in Dynamic Treatment Regimes Needs Critical Reexamination. | Zhiyao Luo, Yangchen Pan, Peter J. Watkinson, Tingting Zhu |
| 2023 | ICLR | Greedy Actor-Critic: A New Conditional Cross-Entropy Method for Policy Improvement. | Samuel Neumann, Sungsu Lim, Ajin George Joseph, Yangchen Pan, Adam White, Martha White |
| 2023 | ICLR | The In-Sample Softmax for Offline Reinforcement Learning. | Chenjun Xiao, Han Wang, Yangchen Pan, Adam White, Martha White |
| 2023 | UAI | Conditionally optimistic exploration for cooperative deep multi-agent reinforcement learning. | Xutong Zhao, Yangchen Pan, Chenjun Xiao, Sarath Chandar, Janarthanan Rajendran |
| 2022 | AISTATS | An Alternate Policy Gradient Estimator for Softmax Policies. | Shivam Garg, Samuele Tosatto, Yangchen Pan, Martha White, Rupam Mahmood |
| 2022 | UAI | Understanding and mitigating the limitations of prioritized experience replay. | Yangchen Pan, Jincheng Mei, Amir-massoud Farahmand, Martha White, Hengshuai Yao, Mohsen Rohani, Jun Luo |
| 2021 | ICLR | Fuzzy Tiling Activations: A Simple Approach to Learning Sparse Representations Online. | Yangchen Pan, Kirby Banman, Martha White |
| 2020 | ICLR | Maxmin Q-learning: Controlling the Estimation Bias of Q-learning. | Qingfeng Lan, Yangchen Pan, Alona Fyshe, Martha White |
| 2020 | ICLR | Frequency-based Search-control in Dyna. | Yangchen Pan, Jincheng Mei, Amir-massoud Farahmand |
| 2019 | IJCAI | Hill Climbing on Value Estimates for Search-control in Dyna. | Yangchen Pan, Hengshuai Yao, Amir-massoud Farahmand, Martha White |
| 2018 | ICML | Reinforcement Learning with Function-Valued Action Spaces for Partial Differential Equation Control. | Yangchen Pan, Amir-massoud Farahmand, Martha White, Saleh Nabi, Piyush Grover, Daniel Nikovski |
| 2018 | IJCAI | Organizing Experience: a Deeper Look at Replay Mechanisms for Sample-Based Planning in Continuous State Domains. | Yangchen Pan, Muhammad Zaheer, Adam White, Andrew Patterson, Martha White |
| 2017 | AAAI | Accelerated Gradient Temporal Difference Learning. | Yangchen Pan, Adam White, Martha White |
| 2017 | ICML | Adapting Kernel Representations Online Using Submodular Maximization. | Matthew Schlegel, Yangchen Pan, Jiecao Chen, Martha White |
| 2017 | UAI | Effective sketching methods for value function approximation. | Yangchen Pan, Erfan Sadeqi Azer, Martha White |
| 2016 | IJCAI | Incremental Truncated LSTD. | Clement Gehring, Yangchen Pan, Martha White |