| 2025 | ICLR | q-exponential family for policy optimization. | Lingwei Zhu, Haseeb Shah, Han Wang, Yukie Nagai, Martha White |
| 2025 | ICML | Position: Lifetime tuning is incompatible with continual reinforcement learning. | Golnaz Mesbahi, Parham Mohammad Panahi, Olya Mastikhina, Steven Tang, Martha White, Adam White |
| 2024 | AAAI | Exploiting Action Impact Regularity and Exogenous State Variables for Offline Reinforcement Learning (Abstract Reprint). | Vincent Liu, James R. Wright, Martha White |
| 2024 | ICML | Averaging n-step Returns Reduces Variance in Reinforcement Learning. | Brett Daley, Martha White, Marlos C. Machado |
| 2024 | ICML | Position: Benchmarking is Limited in Reinforcement Learning Research. | Scott M. Jordan, Adam White, Bruno Castro da Silva, Martha White, Philip S. Thomas |
| 2023 | AISTATS | Asymptotically Unbiased Off-Policy Policy Evaluation when Reusing Old Data in Nonstationary Environments. | Vincent Liu, Yash Chandak, Philip S. Thomas, Martha White |
| 2023 | ICLR | Greedy Actor-Critic: A New Conditional Cross-Entropy Method for Policy Improvement. | Samuel Neumann, Sungsu Lim, Ajin George Joseph, Yangchen Pan, Adam White, Martha White |
| 2023 | ICLR | The In-Sample Softmax for Offline Reinforcement Learning. | Chenjun Xiao, Han Wang, Yangchen Pan, Adam White, Martha White |
| 2023 | ICML | Trajectory-Aware Eligibility Traces for Off-Policy Reinforcement Learning. | Brett Daley, Martha White, Christopher Amato, Marlos C. Machado |
| 2022 | AISTATS | An Alternate Policy Gradient Estimator for Softmax Policies. | Shivam Garg, Samuele Tosatto, Yangchen Pan, Martha White, Rupam Mahmood |
| 2022 | ICLR | Resonance in Weight Space: Covariate Shift Can Drive Divergence of SGD with Momentum. | Kirby Banman, Liam Peet-Pare, Nidhi Hegde, Alona Fyshe, Martha White |
| 2022 | ICML | A Temporal-Difference Approach to Policy Gradient Estimation. | Samuele Tosatto, Andrew Patterson, Martha White, Rupam Mahmood |
| 2022 | UAI | Understanding and mitigating the limitations of prioritized experience replay. | Yangchen Pan, Jincheng Mei, Amir-massoud Farahmand, Martha White, Hengshuai Yao, Mohsen Rohani, Jun Luo |
| 2021 | ICLR | Fuzzy Tiling Activations: A Simple Approach to Learning Sparse Representations Online. | Yangchen Pan, Kirby Banman, Martha White |
| 2020 | EMNLP | From Language to Language-ish: How Brain-Like is an LSTM's Representation of Atypical Language Stimuli? | Maryam Hashemzadeh, Greta Kaufeld, Martha White, Andrea E. Martin, Alona Fyshe |
| 2020 | ICLR | Maxmin Q-learning: Controlling the Estimation Bias of Q-learning. | Qingfeng Lan, Yangchen Pan, Alona Fyshe, Martha White |
| 2020 | ICLR | Training Recurrent Neural Networks Online by Learning Explicit State Variables. | Somjit Nath, Vincent Liu, Alan Chan, Xin Li, Adam White, Martha White |
| 2020 | ICML | Selective Dyna-Style Planning Under Limited Model Capacity. | Zaheer Abbas, Samuel Sokota, Erin Talvitie, Martha White |
| 2020 | ICML | Optimizing for the Future in Non-Stationary MDPs. | Yash Chandak, Georgios Theocharous, Shiv Shankar, Martha White, Sridhar Mahadevan, Philip S. Thomas |
| 2020 | ICML | Gradient Temporal-Difference Learning with Regularized Corrections. | Sina Ghiassian, Andrew Patterson, Shivam Garg, Dhawal Gupta, Adam White, Martha White |
| 2019 | AAAI | Meta-Descent for Online, Continual Prediction. | Andrew Jacobsen, Matthew Schlegel, Cameron Linke, Thomas Degris, Adam White, Martha White |
| 2019 | AAAI | The Utility of Sparse Representations for Control in Reinforcement Learning. | Vincent Liu, Raksha Kumaraswamy, Lei Le, Martha White |
| 2019 | ICLR | Two-Timescale Networks for Nonlinear Value Function Approximation. | Wesley Chung, Somjit Nath, Ajin Joseph, Martha White |
| 2019 | IJCAI | Hill Climbing on Value Estimates for Search-control in Dyna. | Yangchen Pan, Hengshuai Yao, Amir-massoud Farahmand, Martha White |
| 2019 | IJCAI | Planning with Expectation Models. | Yi Wan, Muhammad Zaheer, Adam White, Martha White, Richard S. Sutton |
| 2018 | ICML | Improving Regression Performance with Distributional Losses. | Ehsan Imani, Martha White |
| 2018 | ICML | Reinforcement Learning with Function-Valued Action Spaces for Partial Differential Equation Control. | Yangchen Pan, Amir-massoud Farahmand, Martha White, Saleh Nabi, Piyush Grover, Daniel Nikovski |
| 2018 | IJCAI | Organizing Experience: a Deeper Look at Replay Mechanisms for Sample-Based Planning in Continuous State Domains. | Yangchen Pan, Muhammad Zaheer, Adam White, Andrew Patterson, Martha White |
| 2018 | UAI | High-confidence error estimates for learned value functions. | Touqir Sajed, Wesley Chung, Martha White |
| 2018 | UAI | Comparing Direct and Indirect Temporal-Difference Methods for Estimating the Variance of the Return. | Craig Sherstan, Dylan R. Ashley, Brendan Bennett, Kenny Young, Adam White, Martha White, Richard S. Sutton |
| 2017 | AAAI | Recovering True Classifier Performance in Positive-Unlabeled Learning. | Shantanu Jain, Martha White, Predrag Radivojac |
| 2017 | AAAI | Accelerated Gradient Temporal Difference Learning. | Yangchen Pan, Adam White, Martha White |
| 2017 | ICML | Adapting Kernel Representations Online Using Submodular Maximization. | Matthew Schlegel, Yangchen Pan, Jiecao Chen, Martha White |
| 2017 | ICML | Unifying Task Specification in Reinforcement Learning. | Martha White |
| 2017 | IJCAI | Learning Sparse Representations in Reinforcement Learning with Sparse Coding. | Lei Le, Raksha Kumaraswamy, Martha White |
| 2017 | UAI | Effective sketching methods for value function approximation. | Yangchen Pan, Erfan Sadeqi Azer, Martha White |
| 2016 | IJCAI | Incremental Truncated LSTD. | Clement Gehring, Yangchen Pan, Martha White |
| 2015 | AAAI | Optimal Estimation of Multivariate ARMA Models. | Martha White, Junfeng Wen, Michael Bowling, Dale Schuurmans |
| 2013 | DCC | Partition Tree Weighting. | Joel Veness, Martha White, Michael Bowling, Andrs Gyrgy |
| 2012 | ICML | Linear Off-Policy Actor-Critic. | Thomas Degris, Martha White, Richard S. Sutton |
| 2011 | AAAI | Convex Sparse Coding, Subspace Learning, and Semi-Supervised Extensions. | Xinhua Zhang, Yaoliang Yu, Martha White, Ruitong Huang, Dale Schuurmans |
| 2009 | ICML | Optimal reverse prediction: a unified perspective on supervised, unsupervised and semi-supervised learning. | Linli Xu, Martha White, Dale Schuurmans |
| 2009 | IJCAI | Learning a Value Analysis Tool for Agent Evaluation. | Martha White, Michael H. Bowling |