Skip to content

Martha White

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

43

Venues

8

Active years

2009–2025

Best venue rank

A*

Where they publish

Papers

43 indexed papers, newest first.

YearVenueTitleAuthors
2025ICLRq-exponential family for policy optimization.Lingwei Zhu, Haseeb Shah, Han Wang, Yukie Nagai, Martha White
2025ICMLPosition: Lifetime tuning is incompatible with continual reinforcement learning.Golnaz Mesbahi, Parham Mohammad Panahi, Olya Mastikhina, Steven Tang, Martha White, Adam White
2024AAAIExploiting Action Impact Regularity and Exogenous State Variables for Offline Reinforcement Learning (Abstract Reprint).Vincent Liu, James R. Wright, Martha White
2024ICMLAveraging n-step Returns Reduces Variance in Reinforcement Learning.Brett Daley, Martha White, Marlos C. Machado
2024ICMLPosition: Benchmarking is Limited in Reinforcement Learning Research.Scott M. Jordan, Adam White, Bruno Castro da Silva, Martha White, Philip S. Thomas
2023AISTATSAsymptotically Unbiased Off-Policy Policy Evaluation when Reusing Old Data in Nonstationary Environments.Vincent Liu, Yash Chandak, Philip S. Thomas, Martha White
2023ICLRGreedy Actor-Critic: A New Conditional Cross-Entropy Method for Policy Improvement.Samuel Neumann, Sungsu Lim, Ajin George Joseph, Yangchen Pan, Adam White, Martha White
2023ICLRThe In-Sample Softmax for Offline Reinforcement Learning.Chenjun Xiao, Han Wang, Yangchen Pan, Adam White, Martha White
2023ICMLTrajectory-Aware Eligibility Traces for Off-Policy Reinforcement Learning.Brett Daley, Martha White, Christopher Amato, Marlos C. Machado
2022AISTATSAn Alternate Policy Gradient Estimator for Softmax Policies.Shivam Garg, Samuele Tosatto, Yangchen Pan, Martha White, Rupam Mahmood
2022ICLRResonance in Weight Space: Covariate Shift Can Drive Divergence of SGD with Momentum.Kirby Banman, Liam Peet-Pare, Nidhi Hegde, Alona Fyshe, Martha White
2022ICMLA Temporal-Difference Approach to Policy Gradient Estimation.Samuele Tosatto, Andrew Patterson, Martha White, Rupam Mahmood
2022UAIUnderstanding and mitigating the limitations of prioritized experience replay.Yangchen Pan, Jincheng Mei, Amir-massoud Farahmand, Martha White, Hengshuai Yao, Mohsen Rohani, Jun Luo
2021ICLRFuzzy Tiling Activations: A Simple Approach to Learning Sparse Representations Online.Yangchen Pan, Kirby Banman, Martha White
2020EMNLPFrom Language to Language-ish: How Brain-Like is an LSTM's Representation of Atypical Language Stimuli?Maryam Hashemzadeh, Greta Kaufeld, Martha White, Andrea E. Martin, Alona Fyshe
2020ICLRMaxmin Q-learning: Controlling the Estimation Bias of Q-learning.Qingfeng Lan, Yangchen Pan, Alona Fyshe, Martha White
2020ICLRTraining Recurrent Neural Networks Online by Learning Explicit State Variables.Somjit Nath, Vincent Liu, Alan Chan, Xin Li, Adam White, Martha White
2020ICMLSelective Dyna-Style Planning Under Limited Model Capacity.Zaheer Abbas, Samuel Sokota, Erin Talvitie, Martha White
2020ICMLOptimizing for the Future in Non-Stationary MDPs.Yash Chandak, Georgios Theocharous, Shiv Shankar, Martha White, Sridhar Mahadevan, Philip S. Thomas
2020ICMLGradient Temporal-Difference Learning with Regularized Corrections.Sina Ghiassian, Andrew Patterson, Shivam Garg, Dhawal Gupta, Adam White, Martha White
2019AAAIMeta-Descent for Online, Continual Prediction.Andrew Jacobsen, Matthew Schlegel, Cameron Linke, Thomas Degris, Adam White, Martha White
2019AAAIThe Utility of Sparse Representations for Control in Reinforcement Learning.Vincent Liu, Raksha Kumaraswamy, Lei Le, Martha White
2019ICLRTwo-Timescale Networks for Nonlinear Value Function Approximation.Wesley Chung, Somjit Nath, Ajin Joseph, Martha White
2019IJCAIHill Climbing on Value Estimates for Search-control in Dyna.Yangchen Pan, Hengshuai Yao, Amir-massoud Farahmand, Martha White
2019IJCAIPlanning with Expectation Models.Yi Wan, Muhammad Zaheer, Adam White, Martha White, Richard S. Sutton
2018ICMLImproving Regression Performance with Distributional Losses.Ehsan Imani, Martha White
2018ICMLReinforcement Learning with Function-Valued Action Spaces for Partial Differential Equation Control.Yangchen Pan, Amir-massoud Farahmand, Martha White, Saleh Nabi, Piyush Grover, Daniel Nikovski
2018IJCAIOrganizing Experience: a Deeper Look at Replay Mechanisms for Sample-Based Planning in Continuous State Domains.Yangchen Pan, Muhammad Zaheer, Adam White, Andrew Patterson, Martha White
2018UAIHigh-confidence error estimates for learned value functions.Touqir Sajed, Wesley Chung, Martha White
2018UAIComparing Direct and Indirect Temporal-Difference Methods for Estimating the Variance of the Return.Craig Sherstan, Dylan R. Ashley, Brendan Bennett, Kenny Young, Adam White, Martha White, Richard S. Sutton
2017AAAIRecovering True Classifier Performance in Positive-Unlabeled Learning.Shantanu Jain, Martha White, Predrag Radivojac
2017AAAIAccelerated Gradient Temporal Difference Learning.Yangchen Pan, Adam White, Martha White
2017ICMLAdapting Kernel Representations Online Using Submodular Maximization.Matthew Schlegel, Yangchen Pan, Jiecao Chen, Martha White
2017ICMLUnifying Task Specification in Reinforcement Learning.Martha White
2017IJCAILearning Sparse Representations in Reinforcement Learning with Sparse Coding.Lei Le, Raksha Kumaraswamy, Martha White
2017UAIEffective sketching methods for value function approximation.Yangchen Pan, Erfan Sadeqi Azer, Martha White
2016IJCAIIncremental Truncated LSTD.Clement Gehring, Yangchen Pan, Martha White
2015AAAIOptimal Estimation of Multivariate ARMA Models.Martha White, Junfeng Wen, Michael Bowling, Dale Schuurmans
2013DCCPartition Tree Weighting.Joel Veness, Martha White, Michael Bowling, Andrs Gyrgy
2012ICMLLinear Off-Policy Actor-Critic.Thomas Degris, Martha White, Richard S. Sutton
2011AAAIConvex Sparse Coding, Subspace Learning, and Semi-Supervised Extensions.Xinhua Zhang, Yaoliang Yu, Martha White, Ruitong Huang, Dale Schuurmans
2009ICMLOptimal reverse prediction: a unified perspective on supervised, unsupervised and semi-supervised learning.Linli Xu, Martha White, Dale Schuurmans
2009IJCAILearning a Value Analysis Tool for Agent Evaluation.Martha White, Michael H. Bowling