Skip to content

Herke van Hoof

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

36

Venues

15

Active years

2012–2025

Best venue rank

A*

Where they publish

Papers

36 indexed papers, newest first.

YearVenueTitleAuthors
2025AAAIData Augmentation for Instruction Following Policies via Trajectory Segmentation.Niklas Hpner, Ilaria Tiddi, Herke van Hoof
2024CIKMMitigating Exposure Bias in Online Learning to Rank Recommendation: A Novel Reward Model for Cascading Bandits.Masoud Mansoury, Bamshad Mobasher, Herke van Hoof
2024ICAPSPlanning with a Learned Policy Basis to Optimally Solve Complex Tasks.David Kuric, Guillermo Infante, Vicen Gmez, Anders Jonsson, Herke van Hoof
2024SIGIRGoing Beyond Popularity and Positivity Bias: Correcting for Multifactorial Bias in Recommender Systems.Jin Huang, Harrie Oosterhuis, Masoud Mansoury, Herke van Hoof, Maarten de Rijke
2023ICLRBridge the Inference Gaps of Neural Processes via Expectation Maximization.Qi Wang, Marco Federici, Herke van Hoof
2022AAAIFast and Data Efficient Reinforcement Learning from Pixels via Non-parametric Value Approximation.Alexander Long, Alan Blair, Herke van Hoof
2022CPAIORDeep Policy Dynamic Programming for Vehicle Routing Problems.Wouter Kool, Herke van Hoof, Joaquim A. S. Gromicho, Max Welling
2022ICLRMulti-Agent MDP Homomorphic Networks.Elise van der Pol, Herke van Hoof, Frans A. Oliehoek, Max Welling
2022ICMLModel-based Meta Reinforcement Learning using Graph Structured Surrogate Models and Amortized Policy Search.Qi Wang, Herke van Hoof
2022IJCAILeveraging Class Abstraction for Commonsense Reinforcement Learning via Residual Policy Gradient Methods.Niklas Hpner, Ilaria Tiddi, Herke van Hoof
2022IJCAIValue Refinement Network (VRN).Jan Whlke, Felix Schmitt, Herke van Hoof
2022IJCNNLogic-based AI for Interpretable Board Game Winner Prediction with Tsetlin Machine.Charul Giri, Ole-Christoffer Granmo, Herke van Hoof, Christian D. Blakely
2021ICMLDeep Coherent Exploration for Continuous Control.Yijie Zhang, Herke van Hoof
2021ICRAHierarchies of Planning and Reinforcement Learning for Robot Navigation.Jan Whlke, Felix Schmitt, Herke van Hoof
2020ICANNSocial Navigation with Human Empowerment Driven Deep Reinforcement Learning.Tessa van der Heiden, Florian Mirus, Herke van Hoof
2020ICLREstimating Gradients for Discrete Random Variables by Sampling without Replacement.Wouter Kool, Herke van Hoof, Max Welling
2020ICMLDoubly Stochastic Variational Inference for Neural Processes with Hierarchical Latent Variables.Qi Wang, Herke van Hoof
2020RecSysKeeping Dataset Biases out of the Simulation: A Debiased Simulator for Reinforcement Learning based Recommender Systems.Jin Huang, Harrie Oosterhuis, Maarten de Rijke, Herke van Hoof
2019ICLRAttention, Learn to Solve Routing Problems!Wouter Kool, Herke van Hoof, Max Welling
2019ICLRBuy 4 REINFORCE Samples, Get a Baseline for Free!Wouter Kool, Herke van Hoof, Max Welling
2019ICMLStochastic Beams and Where To Find Them: The Gumbel-Top-k Trick for Sampling Sequences Without Replacement.Wouter Kool, Herke van Hoof, Max Welling
2019IROSDeep Generative Modeling of LiDAR Data.Lucas Caccia, Herke van Hoof, Aaron C. Courville, Joelle Pineau
2019ICRAUncertainty Aware Learning from Demonstrations in Multiple Contexts using Bayesian Neural Networks.Sanjay Thakur, Herke van Hoof, Juan Camilo Gamboa Higuera, Doina Precup, David Meger
2018EMNLPBanditSum: Extractive Summarization as a Contextual Bandit.Yue Dong, Yikang Shen, Eric Crawford, Herke van Hoof, Jackie Chi Kit Cheung
2018ICMLAddressing Function Approximation Error in Actor-Critic Methods.Scott Fujimoto, Herke van Hoof, David Meger
2018ICMLAn Inference-Based Policy Gradient Method for Learning Options.Matthew J. A. Smith, Herke van Hoof, Joelle Pineau
2018ICRAEager and Memory-Based Non-Parametric Stochastic Search Methods for Learning Control.Victor Barbaros, Herke van Hoof, Abbas Abdolmaleki, David Meger
2017AAAIPolicy Search with High-Dimensional Context Variables.Voot Tangkaratt, Herke van Hoof, Simone Parisi, Gerhard Neumann, Jan Peters, Masashi Sugiyama
2016IROSStable reinforcement learning with autoencoders for tactile and visual data.Herke van Hoof, Nutan Chen, Maximilian Karl, Patrick van der Smagt, Jan Peters
2016IROSActive tactile object exploration with Gaussian processes.Zhengkun Yi, Roberto Calandra, Filipe Veiga, Herke van Hoof, Tucker Hermans, Yilei Zhang, Jan Peters
2015AISTATSLearning of Non-Parametric Control Policies with High-Dimensional State Features.Herke van Hoof, Jan Peters, Gerhard Neumann
2015ICRATowards learning hierarchical skills for multi-phase manipulation tasks.Oliver Kroemer, Christian Daniel, Gerhard Neumann, Herke van Hoof, Jan Peters
2015IROSStabilizing novel objects by learning to predict tactile slip.Filipe Veiga, Herke van Hoof, Jan Peters, Tucker Hermans
2014ICRAPolicy search for learning robot control using sparse data.Bastian Bischoff, Duy Nguyen-Tuong, Herke van Hoof, Andrew McHutchon, Carl E. Rasmussen, Alois C. Knoll, Jan Peters, Marc Peter Deisenroth
2014ICRALearning to predict phases of manipulation tasks as hidden states.Oliver Kroemer, Herke van Hoof, Gerhard Neumann, Jan Peters
2012IROSMaximally informative interaction learning for scene exploration.Herke van Hoof, Oliver Kroemer, Heni Ben Amor, Jan Peters