Skip to content

Shimon Whiteson

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

96

Venues

21

Active years

2003–2025

Best venue rank

A*

Where they publish

Papers

96 indexed papers, newest first.

YearVenueTitleAuthors
2025IROSIGDrivSim: A Benchmark for the Imitation Gap in Autonomous Driving.Clmence Grislain, Risto Vuorio, Cong Lu, Shimon Whiteson
2024CoRLRate-Informed Discovery via Bayesian Adaptive Multifidelity Sampling.Aman Sinha, Payam Nikdel, Supratik Paul, Shimon Whiteson
2024ICLRDiscovering Temporally-Aware Reinforcement Learning Algorithms.Matthew Thomas Jackson, Chris Lu, Louis Kirsch, Robert Tjarko Lange, Shimon Whiteson, Jakob Nicolaus Foerster
2024ICMLBayesian Exploration Networks.Mattie Fellows, Brandon Kaplowitz, Christian Schrder de Witt, Shimon Whiteson
2024ICMLDistilling Morphology-Conditioned Hypernetworks for Efficient Universal Morphology Control.Zheng Xiong, Risto Vuorio, Jacob Beck, Matthieu Zimmer, Kun Shao, Shimon Whiteson
2024ICRAUniGen: Unified Modeling of Initial Agent States and Trajectories for Generating Autonomous Driving Scenarios.Reza Mahjourian, Rongbing Mu, Valerii Likhosherstov, Paul Mougin, Xiukun Huang, Joo V. Messias, Shimon Whiteson
2023ICLRCheap Talk Discovery and Utilization in Multi-Agent Reinforcement Learning.Yat Long Lo, Christian Schrder de Witt, Samuel Sokota, Jakob Nicolaus Foerster, Shimon Whiteson
2023ICMLWhy Target Networks Stabilise Temporal Difference Methods.Mattie Fellows, Matthew J. A. Smith, Shimon Whiteson
2023ICMLUniversal Morphology Control via Contextual Modulation.Zheng Xiong, Jacob Beck, Shimon Whiteson
2023IROSHierarchical Imitation Learning for Stochastic Environments.Maximilian Igl, Punit Shah, Paul Mougin, Sirish Srinivasan, Tarun Gupta, Brandyn White, Kyriacos Shiarlis, Shimon Whiteson
2023IROSImitation Is Not Enough: Robustifying Imitation with Reinforcement Learning for Challenging Driving Scenarios.Yiren Lu, Justin Fu, George Tucker, Xinlei Pan, Eli Bronstein, Rebecca Roelofs, Benjamin Sapp, Brandyn White, Aleksandra Faust, Shimon Whiteson, Dragomir Anguelov, Sergey Levine
2022AAAIDeterministic and Discriminative Imitation (D2-Imitation): Revisiting Adversarial Imitation for Sample Efficiency.Mingfei Sun, Sam Devlin, Katja Hofmann, Shimon Whiteson
2022CoRLHypernetworks in Meta-Reinforcement Learning.Jacob Beck, Matthew Thomas Jackson, Risto Vuorio, Shimon Whiteson
2022CoRLEmbedding Synthetic Off-Policy Experience for Autonomous Driving via Zero-Shot Curricula.Eli Bronstein, Sirish Srinivasan, Supratik Paul, Aman Sinha, Matthew O'Kelly, Payam Nikdel, Shimon Whiteson
2022CoRLParticle-Based Score Estimation for State Space Model Learning in Autonomous Driving.Angad Singh, Omar Makhlouf, Maximilian Igl, Joo V. Messias, Arnaud Doucet, Shimon Whiteson
2022ICMLGeneralized Beliefs for Cooperative AI.Darius Muglich, Luisa M. Zintgraf, Christian A. Schrder de Witt, Shimon Whiteson, Jakob N. Foerster
2022ICMLCommunicating via Markov Decision Processes.Samuel Sokota, Christian A. Schrder de Witt, Maximilian Igl, Luisa M. Zintgraf, Philip H. S. Torr, Martin Strohmeier, J. Zico Kolter, Shimon Whiteson, Jakob N. Foerster
2022ICRASymphony: Learning Realistic and Diverse Agents for Autonomous Driving Simulation.Maximilian Igl, Daewoo Kim, Alex Kuefler, Paul Mougin, Punit Shah, Kyriacos Shiarlis, Dragomir Anguelov, Mark Palatucci, Brandyn White, Shimon Whiteson
2022IROSHierarchical Model-Based Imitation Learning for Planning in Autonomous Driving.Eli Bronstein, Mark Palatucci, Dominik Notz, Brandyn White, Alex Kuefler, Yiren Lu, Supratik Paul, Payam Nikdel, Paul Mougin, Hongge Chen, Justin Fu, Austin Abrams, Punit Shah, Evan Racah, Benjamin Frenkel, Shimon Whiteson, Dragomir Anguelov
2021AAAIMean-Variance Policy Iteration for Risk-Averse Reinforcement Learning.Shangtong Zhang, Bo Liu, Shimon Whiteson
2021ICLRRODE: Learning Roles to Decompose Multi-Agent Tasks.Tonghan Wang, Tarun Gupta, Anuj Mahajan, Bei Peng, Shimon Whiteson, Chongjie Zhang
2021ICLRTransient Non-stationarity and Generalisation in Deep Reinforcement Learning.Maximilian Igl, Gregory Farquhar, Jelena Luketina, Wendelin Boehmer, Shimon Whiteson
2021ICLRMy Body is a Cage: the Role of Morphology in Graph-Based Incompatible Control.Vitaly Kurin, Maximilian Igl, Tim Rocktschel, Wendelin Boehmer, Shimon Whiteson
2021ICMLUneVEn: Universal Value Exploration for Multi-Agent Reinforcement Learning.Tarun Gupta, Anuj Mahajan, Bei Peng, Wendelin Boehmer, Shimon Whiteson
2021ICMLRandomized Entity-wise Factorization for Multi-Agent Reinforcement Learning.Shariq Iqbal, Christian A. Schrder de Witt, Bei Peng, Wendelin Boehmer, Shimon Whiteson, Fei Sha
2021ICMLTesseract: Tensorised Actors for Multi-Agent Reinforcement Learning.Anuj Mahajan, Mikayel Samvelyan, Lei Mao, Viktor Makoviychuk, Animesh Garg, Jean Kossaifi, Shimon Whiteson, Yuke Zhu, Animashree Anandkumar
2021ICMLAverage-Reward Off-Policy Policy Evaluation with Function Approximation.Shangtong Zhang, Yi Wan, Richard S. Sutton, Shimon Whiteson
2021ICMLBreaking the Deadly Triad with a Target Network.Shangtong Zhang, Hengshuai Yao, Shimon Whiteson
2021ICMLExploration in Approximate Hyper-State Space for Meta Reinforcement Learning.Luisa M. Zintgraf, Leo Feng, Cong Lu, Maximilian Igl, Kristian Hartikainen, Katja Hofmann, Shimon Whiteson
2021IJCAIDeep Residual Reinforcement Learning (Extended Abstract).Shangtong Zhang, Wendelin Boehmer, Shimon Whiteson
2020ICLROptimistic Exploration even with a Pessimistic Initialisation.Tabish Rashid, Bei Peng, Wendelin Boehmer, Shimon Whiteson
2020ICLRVariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning.Luisa M. Zintgraf, Kyriacos Shiarlis, Maximilian Igl, Sebastian Schulze, Yarin Gal, Katja Hofmann, Shimon Whiteson
2020ICMLDeep Coordination Graphs.Wendelin Boehmer, Vitaly Kurin, Shimon Whiteson
2020ICMLGrowing Action Spaces.Gregory Farquhar, Laura Gustafson, Zeming Lin, Shimon Whiteson, Nicolas Usunier, Gabriel Synnaeve
2020ICMLGradientDICE: Rethinking Generalized Offline Estimation of Stationary Values.Shangtong Zhang, Bo Liu, Shimon Whiteson
2020ICMLProvably Convergent Two-Timescale Off-Policy Actor-Critic with Function Approximation.Shangtong Zhang, Bo Liu, Hengshuai Yao, Shimon Whiteson
2020UAIMultitask Soft Option Learning.Maximilian Igl, Andrew Gambardella, Jinke He, Nantas Nardelli, N. Siddharth, Wendelin Boehmer, Shimon Whiteson
2019ICLRStable Opponent Shaping in Differentiable Games.Alistair Letcher, Jakob N. Foerster, David Balduzzi, Tim Rocktschel, Shimon Whiteson
2019ICMLBayesian Action Decoder for Deep Multi-Agent Reinforcement Learning.Jakob N. Foerster, H. Francis Song, Edward Hughes, Neil Burch, Iain Dunning, Shimon Whiteson, Matthew M. Botvinick, Michael Bowling
2019ICMLA Baseline for Any Order Gradient Estimation in Stochastic Computation Graphs.Jingkai Mao, Jakob N. Foerster, Tim Rocktschel, Maruan Al-Shedivat, Gregory Farquhar, Shimon Whiteson
2019ICMLFingerprint Policy Optimisation for Robust Reinforcement Learning.Supratik Paul, Michael A. Osborne, Shimon Whiteson
2019ICMLFast Context Adaptation via Meta-Learning.Luisa M. Zintgraf, Kyriacos Shiarlis, Vitaly Kurin, Katja Hofmann, Shimon Whiteson
2019IJCAIA Survey of Reinforcement Learning Informed by Natural Language.Jelena Luketina, Nantas Nardelli, Gregory Farquhar, Jakob N. Foerster, Jacob Andreas, Edward Grefenstette, Shimon Whiteson, Tim Rocktschel
2019ICRALearning From Demonstration in the Wild.Feryal M. P. Behbahani, Kyriacos Shiarlis, Xi Chen, Vitaly Kurin, Sudhanshu Kasewa, Ciprian Stirbu, Joo Gomes, Supratik Paul, Frans A. Oliehoek, Joo V. Messias, Shimon Whiteson
2018AAAIExpected Policy Gradients.Kamil Ciosek, Shimon Whiteson
2018AAAICounterfactual Multi-Agent Policy Gradients.Jakob N. Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, Shimon Whiteson
2018AAAIAlternating Optimisation and Quadrature for Robust Control.Supratik Paul, Konstantinos I. Chatzilygeroudis, Kamil Ciosek, Jean-Baptiste Mouret, Michael A. Osborne, Shimon Whiteson
2018ICLRTreeQN and ATreeC: Differentiable Tree-Structured Models for Deep Reinforcement Learning.Gregory Farquhar, Tim Rocktschel, Maximilian Igl, Shimon Whiteson
2018ICLRDiCE: The Infinitely Differentiable Monte-Carlo Estimator.Jakob N. Foerster, Gregory Farquhar, Maruan Al-Shedivat, Tim Rocktschel, Eric P. Xing, Shimon Whiteson
2018ICMLFourier Policy Gradients.Matthew Fellows, Kamil Ciosek, Shimon Whiteson
2018ICMLDiCE: The Infinitely Differentiable Monte Carlo Estimator.Jakob N. Foerster, Gregory Farquhar, Maruan Al-Shedivat, Tim Rocktschel, Eric P. Xing, Shimon Whiteson
2018ICMLDeep Variational Reinforcement Learning for POMDPs.Maximilian Igl, Luisa M. Zintgraf, Tuan Anh Le, Frank Wood, Shimon Whiteson
2018ICMLQMIX: Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning.Tabish Rashid, Mikayel Samvelyan, Christian Schrder de Witt, Gregory Farquhar, Jakob N. Foerster, Shimon Whiteson
2018ICMLTACO: Learning Task Decomposition via Temporal Alignment for Control.Kyriacos Shiarlis, Markus Wulfmeier, Sasha Salter, Shimon Whiteson, Ingmar Posner
2017AAAIOFFER: Off-Environment Reinforcement Learning.Kamil Andrzej Ciosek, Shimon Whiteson
2017BMVCIntro to Reinforcement Learning.Shimon Whiteson
2017ICMLStabilising Experience Replay for Deep Multi-Agent Reinforcement Learning.Jakob N. Foerster, Nantas Nardelli, Gregory Farquhar, Triantafyllos Afouras, Philip H. S. Torr, Pushmeet Kohli, Shimon Whiteson
2017IROSAcquiring social interaction behaviours for telepresence robots via deep learning from demonstration.Kyriacos Shiarlis, Joo V. Messias, Shimon Whiteson
2017ICRARapidly exploring learning trees.Kyriacos Shiarlis, Joo V. Messias, Shimon Whiteson
2017UAIReal-Time Resource Allocation for Tracking Systems.Yash Satsangi, Shimon Whiteson, Frans A. Oliehoek, Henri Bouma
2016IJCAIPAC Greedy Maximization with Efficient Bounds on Information Gain for Sensor Selection.Yash Satsangi, Shimon Whiteson, Frans A. Oliehoek
2016WSDMMultileave Gradient Descent for Fast Online Learning to Rank.Anne Schuth, Harrie Oosterhuis, Shimon Whiteson, Maarten de Rijke
2015AAAIExploiting Submodular Value Functions for Faster Dynamic Sensor Selection.Yash Satsangi, Shimon Whiteson, Frans A. Oliehoek
2015ESANNPareto Local Search for MOMDP Planning.Chiel Kooijman, Maarten de Waard, Maarten Inja, Diederik M. Roijers, Shimon Whiteson
2015IJCAIPoint-Based Planning for Multi-Objective POMDPs.Diederik Marijn Roijers, Shimon Whiteson, Frans A. Oliehoek
2015SIGIRBayesian Ranker Comparison Based on Historical User Interactions.Artem Grotov, Shimon Whiteson, Maarten de Rijke
2015WSDMMergeRUCB: A Method for Large-Scale Online Ranker Evaluation.Masrour Zoghi, Shimon Whiteson, Maarten de Rijke
2014CIKMMultileaved Comparisons for Fast Online Evaluation.Anne Schuth, Floor Sietsma, Shimon Whiteson, Damien Lefortier, Maarten de Rijke
2014ECIROptimizing Base Rankers Using Clicks - A Case Study Using BM25.Anne Schuth, Floor Sietsma, Shimon Whiteson, Maarten de Rijke
2014FDGDesign criteria for challenge balancing of personalised game spaces.Sander Bakkes, Shimon Whiteson
2014ICMLRelative Upper Confidence Bound for the K-Armed Dueling Bandit Problem.Masrour Zoghi, Shimon Whiteson, Rmi Munos, Maarten de Rijke
2014PPSNQueued Pareto Local Search for Multi-Objective Optimization.Maarten Inja, Chiel Kooijman, Maarten de Waard, Diederik M. Roijers, Shimon Whiteson
2014WSDMRelative confidence sampling for efficient on-line ranker evaluation.Masrour Zoghi, Shimon Whiteson, Maarten de Rijke, Rmi Munos
2013CIKMLerot: an online learning to rank framework.Anne Schuth, Katja Hofmann, Shimon Whiteson, Maarten de Rijke
2013GECCOCritical factors in the performance of hyperNEAT.Thomas G. van den Berg, Shimon Whiteson
2013WSDMReusing historical interaction data for faster online learning to rank for IR.Katja Hofmann, Anne Schuth, Shimon Whiteson, Maarten de Rijke
2012AAMASV-MAX: tempered optimism for better PAC reinforcement learning.Karun Rao, Shimon Whiteson
2012CIKMEstimating interleaved comparison outcomes from historical click data.Katja Hofmann, Shimon Whiteson, Maarten de Rijke
2012UAIExploiting Structure in Cooperative Bayesian Games.Frans A. Oliehoek, Shimon Whiteson, Matthijs T. J. Spaan
2011CIKMA probabilistic method for inferring preferences from clicks.Katja Hofmann, Shimon Whiteson, Maarten de Rijke
2011ECIRBalancing Exploration and Exploitation in Learning to Rank Online.Katja Hofmann, Shimon Whiteson, Maarten de Rijke
2011GECCOCritical factors in the performance of novelty search.Steijn Kistemaker, Shimon Whiteson
2010GECCOMulti-task evolutionary shaping without pre-specified representations.Matthijs Snel, Shimon Whiteson
2009FUSIONIntegrating distributed Bayesian inference and reinforcement learning for sensor management.Corrado Grappiolo, Shimon Whiteson, Gregor Pavlin, Bram Bakker
2009GECCONeuroevolutionary reinforcement learning for generalized helicopter control.Rogier Koppejan, Shimon Whiteson
2009ICMLAAutomatic Feature Selection for Model-Based Reinforcement Learning in Factored MDPs.Mark Kroon, Shimon Whiteson
2009ISDAPostponed Updates for Temporal-Difference Reinforcement Learning.Harm van Seijen, Shimon Whiteson
2007AAAITemporal Difference and Policy Search Methods for Reinforcement Learning: An Empirical Comparison.Matthew E. Taylor, Shimon Whiteson, Peter Stone
2007AAAIStochastic Optimization for Collision Selection in High Energy Physics.Shimon Whiteson, Daniel Whiteson
2006AAAISample-Efficient Evolutionary Function Approximation for Reinforcement Learning.Shimon Whiteson, Peter Stone
2006GECCOComparing evolutionary and temporal difference methods in a reinforcement learning domain.Matthew E. Taylor, Shimon Whiteson, Peter Stone
2006GECCOOn-line evolutionary computation for reinforcement learning in stochastic domains.Shimon Whiteson, Peter Stone
2005AAAIImproving Reinforcement Learning Function Approximators via Neuroevolution.Shimon Whiteson
2005GECCOAutomatic feature selection in neuroevolution.Shimon Whiteson, Peter Stone, Kenneth O. Stanley, Risto Miikkulainen, Nate Kohl
2004AAAITowards Autonomic Computing: Adaptive Job Routing and Scheduling.Shimon Whiteson, Peter Stone
2003GECCOEvolving Keepaway Soccer Players through Task Decomposition.Shimon Whiteson, Nate Kohl, Risto Miikkulainen, Peter Stone