Skip to content

Stuart Russell

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

118

Venues

25

Active years

1986–2025

Best venue rank

A*

Where they publish

Papers

118 indexed papers, newest first.

YearVenueTitleAuthors
2025AAAIThe Partially Observable Off-Switch Game.Andrew Garber, Rohan Subramani, Linus Luu, Mark Bedaywi, Stuart Russell, Scott Emmons
2025ICLRMonitoring Latent World States in Language Models with Propositional Probes.Jiahai Feng, Stuart Russell, Jacob Steinhardt
2025ICLRDiffusion On Syntax Trees For Program Synthesis.Shreyas Kapur, Erik Jenner, Stuart Russell
2025ICLRBAMDP Shaping: a Unified Framework for Intrinsic Motivation and Reward Shaping.Aly Lidayan, Michael D. Dennis, Stuart Russell
2025ICMLObservation Interference in Partially Observable Assistance Games.Scott Emmons, Caspar Oesterheld, Vincent Conitzer, Stuart Russell
2025ICMLExtractive Structures Learned in Pretraining Enable Generalization on Finetuned Facts.Jiahai Feng, Stuart Russell, Jacob Steinhardt
2025ICMLAssistanceZero: Scalably Solving Assistance Games.Cassidy Laidlaw, Eli Bronstein, Timothy Guo, Dylan Feng, Lukas Berglund, Justin Svegliato, Stuart Russell, Anca D. Dragan
2025ICMLAvoiding Catastrophe in Online Learning by Asking for Help.Benjamin Plaut, Hanlin Zhu, Stuart Russell
2025UAIRL, but don't do anything I wouldn't do.Michael K. Cohen, Marcus Hutter, Yoshua Bengio, Stuart Russell
2024CoRLTrajectory Improvement and Reward Learning from Comparative Language Feedback.Zhaojing Yang, Miru Jun, Jeremy Tien, Stuart Russell, Anca D. Dragan, Erdem Biyik
2024ICLRThe Effective Horizon Explains Deep RL Performance in Stochastic Environments.Cassidy Laidlaw, Banghua Zhu, Stuart Russell, Anca D. Dragan
2024ICLRTensor Trust: Interpretable Prompt Injection Attacks from an Online Game.Sam Toyer, Olivia Watkins, Ethan Adrian Mendes, Justin Svegliato, Luke Bailey, Tiffany Wang, Isaac Ong, Karim Elmaaroufi, Pieter Abbeel, Trevor Darrell, Alan Ritter, Stuart Russell
2024ICLROn Representation Complexity of Model-based and Model-free Reinforcement Learning.Hanlin Zhu, Baihe Huang, Stuart Russell
2024ICMLImage Hijacks: Adversarial Images can Control Generative Models at Runtime.Luke Bailey, Euan Ong, Stuart Russell, Scott Emmons
2024ICMLAI Alignment with Changing and Influenceable Reward Functions.Micah Carroll, Davis Foote, Anand Siththaranjan, Stuart Russell, Anca D. Dragan
2024ICMLPosition: Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback.Vincent Conitzer, Rachel Freedman, Jobst Heitzig, Wesley H. Holliday, Bob M. Jacobs, Nathan Lambert, Milan Moss, Eric Pacuit, Stuart Russell, Hailey Schoelkopf, Emanuel Tewolde, William S. Zwicker
2024ICRAEthically Compliant Autonomous Systems under Partial Observability.Qingyuan Lu, Justin Svegliato, Samer B. Nashed, Shlomo Zilberstein, Stuart Russell
2023AAAIActive Reward Learning from Multiple Teachers.Peter Barnett, Rachel Freedman, Justin Svegliato, Stuart Russell
2023AISTATSSMCP3: Sequential Monte Carlo with Probabilistic Program Proposals.Alexander K. Lew, George Matheos, Tan Zhi-Xuan, Matin Ghavamizadeh, Nishad Gothoskar, Stuart Russell, Vikash K. Mansinghka
2023ICLROptimal Conservative Offline RL with General Function Approximation via Augmented Lagrangian.Paria Rashidinejad, Hanlin Zhu, Kunhe Yang, Stuart Russell, Jiantao Jiao
2023ICMLWho Needs to Know? Minimal Knowledge for Optimal Coordination.Niklas Lauffer, Ameesh Shah, Micah Carroll, Michael D. Dennis, Stuart Russell
2023ICMLInvariance in Policy Optimisation and Partial Identifiability in Reward Learning.Joar Max Viktor Skalse, Matthew Farrugia-Roberts, Stuart Russell, Alessandro Abate, Adam Gleave
2023ICMLAdversarial Policies Beat Superhuman Go AIs.Tony Tong Wang, Adam Gleave, Tom Tseng, Kellin Pelrine, Nora Belrose, Joseph Miller, Michael D. Dennis, Yawen Duan, Viktor Pogrebniak, Sergey Levine, Stuart Russell
2023IROSFormal Composition of Robotic Systems as Contract Programs.Mason Nakamura, Justin Svegliato, Samer B. Nashed, Shlomo Zilberstein, Stuart Russell
2022ICLRCross-Domain Imitation Learning via Optimal Transport.Arnaud Fickinger, Samuel Cohen, Stuart Russell, Brandon Amos
2022ICMLEstimating and Penalizing Induced Preference Shifts in Recommender Systems.Micah D. Carroll, Anca D. Dragan, Stuart Russell, Dylan Hadfield-Menell
2022ICMLFor Learning in Symmetric Teams, Local Optima are Global Nash Equilibria.Scott Emmons, Caspar Oesterheld, Andrew Critch, Vincent Conitzer, Stuart Russell
2022IROSSelecting the Partial State Abstractions of MDPs: A Metareasoning Approach with Deep Reinforcement Learning.Samer B. Nashed, Justin Svegliato, Abhinav Bhatia, Stuart Russell, Shlomo Zilberstein
2022IUIProvably Beneficial Artificial Intelligence.Stuart Russell
2021ICLRQuantifying Differences in Reward Functions.Adam Gleave, Michael Dennis, Shane Legg, Stuart Russell, Jan Leike
2021RecSysEstimating and Penalizing Preference Shift in Recommender Systems.Micah Carroll, Dylan Hadfield-Menell, Stuart Russell, Anca D. Dragan
2020ICLRAdversarial Policies: Attacking Deep Reinforcement Learning.Adam Gleave, Michael Dennis, Cody Wild, Neel Kant, Sergey Levine, Stuart Russell
2019AAAIRobust Multi-Agent Reinforcement Learning via Minimax Deep Deterministic Policy Gradient.Shihui Li, Yi Wu, Xinyue Cui, Honghua Dong, Fei Fang, Stuart Russell
2019ICCVBayesian Relational Memory for Semantic Visual Navigation.Yi Wu, Yuxin Wu, Aviv Tamar, Stuart Russell, Georgia Gkioxari, Yuandong Tian
2018ICMLAn Efficient, Generalized Bellman Update For Cooperative Inverse Reinforcement Learning.Dhruv Malik, Malayandi Palaniappan, Jaime F. Fisac, Dylan Hadfield-Menell, Stuart Russell, Anca D. Dragan
2018ICMLDiscrete-Continuous Mixtures in Probabilistic Programming: Generalized Semantics and Inference Algorithms.Yi Wu, Siddharth Srivastava, Nicholas Hay, Simon S. Du, Stuart Russell
2017AAAIA Nearly-Black-Box Online Algorithm for Joint Parameter and State Estimation in Temporal Models.Yusuf Bugra Erol, Yi Wu, Lei Li, Stuart Russell
2017AAAIThe Off-Switch Game.Dylan Hadfield-Menell, Anca D. Dragan, Pieter Abbeel, Stuart Russell
2017AISTATSSignal-based Bayesian Seismic Monitoring.David A. Moore, Stuart Russell
2017EMNLPAdversarial Training for Relation Extraction.Yi Wu, David Bamman, Stuart Russell
2017IJCAIEfficient Reinforcement Learning with Hierarchies of Machines by Leveraging Internal Transitions.Aijun Bai, Stuart Russell
2017IJCAIThe Off-Switch Game.Dylan Hadfield-Menell, Anca D. Dragan, Pieter Abbeel, Stuart Russell
2017IJCAIShould Robots be Obedient?Smitha Milli, Dylan Hadfield-Menell, Anca D. Dragan, Stuart Russell
2017RoboCupConcurrent Hierarchical Reinforcement Learning for RoboCup Keepaway.Aijun Bai, Stuart Russell, Xiaoping Chen
2016AAAIMetaphysics of Planning Domain Descriptions.Siddharth Srivastava, Stuart Russell, Alessandro Pinto
2016IJCAIMarkovian State and Action Abstractions for MDPs via Hierarchical MCTS.Aijun Bai, Siddharth Srivastava, Stuart Russell
2016IJCAISwift: Compiled Inference for Probabilistic Programming Languages.Yi Wu, Lei Li, Stuart Russell, Rastislav Bodk
2016IROSSequential quadratic programming for task plan optimization.Dylan Hadfield-Menell, Christopher Lin, Rohan Chitnis, Stuart Russell, Pieter Abbeel
2015AAAITractability of Planning with Loops.Siddharth Srivastava, Shlomo Zilberstein, Abhishek Gupta, Pieter Abbeel, Stuart Russell
2015UAIMultitasking: Optimal Planning for Bandit Superprocesses.Dylan Hadfield-Menell, Stuart Russell
2015UAIA Smart-Dumb/Dumb-Smart Algorithm for Efficient Split-Merge MCMC.Wei Wang, Stuart Russell
2014IPMUUnifying Logic and Probability: A New Dawn for AI?Stuart Russell
2014ICRACombined task and motion planning through an extensible planner-independent interface layer.Siddharth Srivastava, Eugene Fang, Lorenzo Riano, Rohan Chitnis, Stuart Russell, Pieter Abbeel
2014UAIFast Gaussian Process Posteriors with Product Trees.David A. Moore, Stuart Russell
2014UAIFirst-Order Open-Universe POMDPs.Siddharth Srivastava, Stuart Russell, Paul Ruan, Xiang Cheng
2013AISTATSDynamic Scaled Sampling for Deterministic Constraints.Lei Li, Bharath Ramsundar, Stuart Russell
2013CHIWriting and sketching in the air, recognizing and controlling on the fly.Sharad Vikram, Lei Li, Stuart Russell
2013ICMLThe Extended Parameter Filter.Yusuf Erol, Lei Li, Bharath Ramsundar, Stuart Russell
2013UAIProduct Trees for Gaussian Process Covariance in Sublinear Time.David A. Moore, Stuart Russell
2012UAISelecting Computations: Theory and Applications.Nicholas Hay, Stuart Russell, David Tolpin, Solomon Eyal Shimony
2011AAAIGlobal Seismic Monitoring: A Bayesian Approach.Nimar S. Arora, Stuart Russell, Paul Kidwell, Erik B. Sudderth
2011EDMPartially Observable Sequential Decision Making for Problem Selection in an Intelligent Tutoring System.Emma Brunskill, Stuart Russell
2011IJCAIBounded Intention Planning.Jason Andrew Wolfe, Stuart Russell
2011UAIA temporally abstracted Viterbi algorithm.Shaunak Chatterjee, Stuart Russell
2010AAAIAutomatic Inference in BLOG.Nimar S. Arora, Stuart Russell, Erik B. Sudderth
2010AAAIHierarchical Planning for Mobile Manipulation.Jason Andrew Wolfe, Bhaskara Marthi, Stuart Russell
2010UAIGibbs Sampling in Open-Universe Stochastic Languages.Nimar S. Arora, Rodrigo de Salvo Braz, Erik B. Sudderth, Stuart Russell
2010UAIRAPID: A Reachable Anytime Planner for Imprecisely-sensed Domains.Emma Brunskill, Stuart Russell
2008UAIImproving Gradient Estimation by Incorporating Sensor Data.Gregory Lawrence, Stuart Russell
2006ILPFirst-Order Probabilistic Languages: Into the Unknown.Brian Milch, Stuart Russell
2006UAIA Compact, Hierarchical Q-function Decomposition.Bhaskara Marthi, Stuart Russell, David Andre
2006UAIGeneral-Purpose MCMC Inference over Relational Structures.Brian Milch, Stuart Russell
2005AISTATSApproximate Inference for Infinite Contingent Bayesian Networks.Brian Milch, Bhaskara Marthi, David A. Sontag, Stuart Russell, Daniel L. Ong, Andrey Kolobov
2005IJCAIConcurrent Hierarchical Reinforcement Learning.Bhaskara Marthi, Stuart Russell, David Latham, Carlos Guestrin
2005IJCAIBLOG: Probabilistic Models with Unknown Objects.Brian Milch, Bhaskara Marthi, Stuart Russell, David A. Sontag, Daniel L. Ong, Andrey Kolobov
2005IJCAIEfficient belief-state AND-OR search, with application to Kriegspiel.Stuart Russell, Jason Andrew Wolfe
2003ICMLQ-Decomposition for Reinforcement Learning Agents.Stuart Russell, Andrew Zimdars
2003IJCAILogical Filtering.Eyal Amir, Stuart Russell
2003UAIEfficient Gradient Estimation for Motor Control Learning.Gregory Lawrence, Noah J. Cowan, Stuart Russell
2003UAIA generalized mean field algorithm for variational inference in exponential families.Eric P. Xing, Michael I. Jordan, Stuart Russell
2002UAIDecayed MCMC Filtering.Bhaskara Marthi, Hanna Pasula, Stuart Russell, Yuval Peres
2001AISTATSOnline Bagging and Boosting.Nikunj C. Oza, Stuart Russell
2001IJCAIApproximate inference for first-order probabilistic languages.Hanna Pasula, Stuart Russell
2001KDDExperimental comparisons of online and batch versions of bagging and boosting.Nikunj C. Oza, Stuart Russell
2001UAIVariational MCMC.Nando de Freitas, Pedro A. d. F. R. Hjen-Srensen, Stuart Russell
2000ICMLAlgorithms for Inverse Reinforcement Learning.Andrew Y. Ng, Stuart Russell
2000UAIRao-Blackwellised Particle Filtering for Dynamic Bayesian Networks.Arnaud Doucet, Nando de Freitas, Kevin P. Murphy, Stuart Russell
1999DISExpressive Probability Models in Science.Stuart Russell
1999ICMLPolicy Invariance Under Reward Transformations: Theory and Application to Reward Shaping.Andrew Y. Ng, Daishi Harada, Stuart Russell
1999IJCAIConvergence of Reinforcement Learning with General Function Approximators.Vassilis A. Papavassiliou, Stuart Russell
1999IJCAITracking Many Objects with Many Sensors.Hanna Pasula, Stuart Russell, Michael Ostland, Yaacov Ritov
1998AAAIBayesian Q-Learning.Richard Dearden, Nir Friedman, Stuart Russell
1998AAAISpeech Recognition with Dynamic Bayesian Networks.Geoffrey Zweig, Stuart Russell
1998COLTLearning Agents for Uncertain Environments (Extended Abstract).Stuart Russell
1998InterspeechProbabilistic modeling with Bayesian networks for automatic speech recognition.Geoffrey Zweig, Stuart Russell
1998UAILearning the Structure of Dynamic Probabilistic Networks.Nir Friedman, Kevin P. Murphy, Stuart Russell
1997IJCAISpace-Efficient Inference in Dynamic Probabilistic Networks.John Binder, Kevin P. Murphy, Stuart Russell
1997IJCAIChallenge: What is the Impact of Bayesian Networks on Learning?Nir Friedman, Moiss Goldszmidt, David Heckerman, Stuart Russell
1997IJCAIObject Identification in a Bayesian Context.Timothy Huang, Stuart Russell
1997UAIImage Segmentation in Video Sequences: A Probabilistic Approach.Nir Friedman, Stuart Russell
1996KITools for Autonomous Agents (Abstract).Stuart Russell
1995IJCAIThe BATmobile: Towards a Bayesian Automated Taxi.Jeff Forbes, Timothy Huang, Keiji Kanazawa, Stuart Russell
1995IJCAIApproximating Optimal Policies for Partially Observable Stochastic Domains.Ronald Parr, Stuart Russell
1995IJCAIRationality and Intelligence.Stuart Russell
1995IJCAILocal Learning in Probabilistic Networks with Hidden Variables.Stuart Russell, John Binder, Daphne Koller, Keiji Kanazawa
1995UAIStochastic simulation algorithms for dynamic probabilistic networks.Keiji Kanazawa, Daphne Koller, Stuart Russell
1994AAAIAutomatic Symbolic Traffic Scene Analysis Using Belief Networks.Timothy Huang, Daphne Koller, Jitendra Malik, Gary H. Ogasawara, Bobby S. Rao, Stuart Russell, Joseph Weber
1994AAAIControl Strategies for a Stochastic Planner.Jonathan Tash, Stuart Russell
1994ICPRTowards robust automatic traffic scene analysis in real-time.Daphne Koller, Joseph Weber, Timothy Huang, Jitendra Malik, Gary H. Ogasawara, Stuart Russell, Bobby S. Rao
1993ICMLDecision Theoretic Subsampling for Induction on Large Databases.Ron Musick, Jason Catlett, Stuart Russell
1993IJCAIPlanning Using Multiple Execution Architectures.Gary H. Ogasawara, Stuart Russell
1993IJCAIAnytime Sensing Planning and Action: A Practical Model for Robot Control.Shlomo Zilberstein, Stuart Russell
1992AAAIHow Long Will It Take?Ron Musick, Stuart Russell
1992COLTPAC-Learnability of Determinate Logic Programs.Saso Dzeroski, Stephen H. Muggleton, Stuart Russell
1992ECAIEfficient Memory-Bounded Search Methods.Stuart Russell
1989IJCAIOn Optimal Game-Tree Search using Rational Meta-Reasoning.Stuart Russell, Eric Wefald
1989UAIAutomated Construction of Sparse Bayesian Networks from Unstructured Probabilistic Models and Domain Information.Sampath Srinivas, Stuart Russell, Alice M. Agogino
1986AAAIPreliminary Steps Toward the Automation of Induction.Stuart Russell