Skip to content

Ronald Parr

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

42

Venues

9

Active years

1993–2024

Best venue rank

A*

Where they publish

Papers

42 indexed papers, newest first.

YearVenueTitleAuthors
2024ICMLPosition: Amazing Things Come From Having Many Good Models.Cynthia Rudin, Chudi Zhong, Lesia Semenova, Margo I. Seltzer, Ronald Parr, Jiachang Liu, Srikar Katta, Jon Donnelly, Harry Chen, Zachery Boner
2019ICMLRevisiting the Softmax Bellman Operator: New Benefits and New Perspective.Zhao Song, Ronald Parr, Lawrence Carin
2016AAAIDistance Minimization for Reward Learning from Scored Trajectories.Benjamin Burchfiel, Carlo Tomasi, Ronald Parr
2016AAAIEfficient PAC-Optimal Exploration in Concurrent, Continuous State MDPs with Delayed Updates.Jason Pazis, Ronald Parr
2014ICRAUnsupervised discovery of object classes with a mobile robot.Julian Mason, Bhaskara Marthi, Ronald Parr
2013AAAIPAC Optimal Exploration in Continuous Space Markov Decision Processes.Jason Pazis, Ronald Parr
2013AAAISample Complexity and Performance Bounds for Non-Parametric Approximate Linear Programming.Jason Pazis, Ronald Parr
2012AAAIComputing Optimal Strategies to Commit to in Stochastic Games.Joshua Letchford, Liam MacDermed, Vincent Conitzer, Ronald Parr, Charles L. Isbell Jr.
2012ICMLGreedy Algorithms for Sparse Reinforcement Learning.Christopher Painter-Wakefield, Ronald Parr
2012IROSObject disappearance for object discovery.Julian Mason, Bhaskara Marthi, Ronald Parr
2012UAIValue Function Approximation in Noisy Environments Using Locally Smoothed Regularized Approximate Linear Programs.Gavin Taylor, Ronald Parr
2011AAAINon-Parametric Approximate Linear Programming for MDPs.Jason Pazis, Ronald Parr
2011ICMLGeneralized Value Functions for Large Action Sets.Jason Pazis, Ronald Parr
2011IJCAISecurity Games with Multiple Attacker Resources.Dmytro Korzhyk, Vincent Conitzer, Ronald Parr
2011ICRATextured occupancy grids for monocular localization without features.Julian Mason, Susanna Ricco, Ronald Parr
2010AAAIComplexity of Computing Optimal Stackelberg Strategies in Security Resource Allocation Games.Dmytro Korzhyk, Vincent Conitzer, Ronald Parr
2010ICMLFeature Selection Using Regularization in Approximate Linear Programs for Markov Decision Processes.Marek Petrik, Gavin Taylor, Ronald Parr, Shlomo Zilberstein
2009ICMLKernelized value function approximation for reinforcement learning.Gavin Taylor, Ronald Parr
2009IJCAIMulti-Step Multi-Sensor Hider-Seeker Games.Erik Halvorson, Vincent Conitzer, Ronald Parr
2008ICMLAn analysis of linear models, linear value-function approximation, and feature selection for reinforcement learning.Ronald Parr, Lihong Li, Gavin Taylor, Christopher Painter-Wakefield, Michael L. Littman
2008ISAIMPlanning Aims for a Network of Horizontal and Overhead Sensors.Erik Halvorson, Ronald Parr
2008WAFRPlanning Aims for a Network of Horizontal and Overhead Sensors.Erik Halvorson, Ronald Parr
2007AAAIPoint-Based Policy Iteration.Shihao Ji, Ronald Parr, Hui Li, Xuejun Liao, Lawrence Carin
2007ICMLAnalyzing feature generation for value-function approximation.Ronald Parr, Christopher Painter-Wakefield, Lihong Li, Michael L. Littman
2006UAIEfficient Selection of Disambiguating Actions for Stereo Vision.Monika Schaeffer, Ronald Parr
2004ICMLLearning probabilistic motion models for mobile robots.Austin I. Eliazar, Ronald Parr
2004ICRADP-SLAM 2.0.Austin I. Eliazar, Ronald Parr
2003ICMLReinforcement Learning as Classification: Leveraging Modern Classifiers.Michail G. Lagoudakis, Ronald Parr
2003IJCAIDP-SLAM: Fast, Robust Simultaneous Localization and Mapping Without Predetermined Landmarks.Austin I. Eliazar, Ronald Parr
2003IJCAIApproximate Policy Iteration using Large-Margin Classifiers.Michail G. Lagoudakis, Ronald Parr
2002ICMLCoordinated Reinforcement Learning.Carlos Guestrin, Michail G. Lagoudakis, Ronald Parr
2002UAIValue Function Approximation in Zero-Sum Markov Games.Michail G. Lagoudakis, Ronald Parr
2002VLDBXPathLearner: An On-line Self-Tuning Markov Histogram for XML Path Selectivity Estimation.Lipyeow Lim, Min Wang, Sriram Padmanabhan, Jeffrey Scott Vitter, Ronald Parr
2001IJCAIMax-norm Projections for Factored MDPs.Carlos Guestrin, Daphne Koller, Ronald Parr
2001UAIInference in Hybrid Networks: Theoretical Limits and Practical Algorithms.Uri Lerner, Ronald Parr
2000AAAIMaking Rational Decisions Using Adaptive Utility Elicitation.Urszula Chajewska, Daphne Koller, Ronald Parr
2000AAAIBayesian Fault Detection and Diagnosis in Dynamic Systems.Uri Lerner, Ronald Parr, Daphne Koller, Gautam Biswas
2000UAIPolicy Iteration for Factored MDPs.Daphne Koller, Ronald Parr
1999IJCAIComputing Factored Value Functions for Policies in Structured MDPs.Daphne Koller, Ronald Parr
1998UAIFlexible Decomposition Algorithms for Weakly Coupled Markov Decision Problems.Ronald Parr
1995IJCAIApproximating Optimal Policies for Partially Observable Stochastic Domains.Ronald Parr, Stuart Russell
1993IJCAIProvably Bounded Optimal Agents.Stuart J. Russell, Devika Subramanian, Ronald Parr