| 2025 | ICLR | Correlated Proxies: A New Definition and Improved Mitigation for Reward Hacking. | Cassidy Laidlaw, Shivam Singhal, Anca D. Dragan |
| 2025 | ICLR | Iterative Label Refinement Matters More than Preference Optimization under Weak Supervision. | Yaowen Ye, Cassidy Laidlaw, Jacob Steinhardt |
| 2025 | ICML | AssistanceZero: Scalably Solving Assistance Games. | Cassidy Laidlaw, Eli Bronstein, Timothy Guo, Dylan Feng, Lukas Berglund, Justin Svegliato, Stuart Russell, Anca D. Dragan |
| 2024 | ICLR | Toward Computationally Efficient Inverse Reinforcement Learning via Reward Shaping. | Lauren H. Cooke, Harvey Klyne, Edwin Zhang, Cassidy Laidlaw, Milind Tambe, Finale Doshi-Velez |
| 2024 | ICLR | The Effective Horizon Explains Deep RL Performance in Stochastic Environments. | Cassidy Laidlaw, Banghua Zhu, Stuart Russell, Anca D. Dragan |
| 2024 | ICLR | Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF. | Anand Siththaranjan, Cassidy Laidlaw, Dylan Hadfield-Menell |
| 2022 | ICLR | The Boltzmann Policy Distribution: Accounting for Systematic Suboptimality in Human Models. | Cassidy Laidlaw, Anca D. Dragan |
| 2021 | ICLR | Perceptual Adversarial Robustness: Defense Against Unseen Threat Models. | Cassidy Laidlaw, Sahil Singla, Soheil Feizi |
| 2019 | CVPR | Capture, Learning, and Synthesis of 3D Speaking Styles. | Daniel Cudeiro, Timo Bolkart, Cassidy Laidlaw, Anurag Ranjan, Michael J. Black |