Skip to content

Bruno Scherrer

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

21

Venues

9

Active years

2002–2020

Best venue rank

A*

Where they publish

Papers

21 indexed papers, newest first.

YearVenueTitleAuthors
2020AISTATSMomentum in Reinforcement Learning.Nino Vieillard, Bruno Scherrer, Olivier Pietquin, Matthieu Geist
2019AAAIHow to Combine Tree-Search Methods in Reinforcement Learning.Yonathan Efroni, Gal Dalal, Bruno Scherrer, Shie Mannor
2019ICMLA Theory of Regularized Markov Decision Processes.Matthieu Geist, Bruno Scherrer, Olivier Pietquin
2018ICMLBeyond the One-Step Greedy Approach in Reinforcement Learning.Yonathan Efroni, Gal Dalal, Bruno Scherrer, Shie Mannor
2016AISTATSOn the Use of Non-Stationary Strategies for Solving Two-Player Zero-Sum Markov Games.Julien Prolat, Bilal Piot, Bruno Scherrer, Olivier Pietquin
2016ICMLSoftened Approximate Policy Iteration for Markov Games.Julien Prolat, Bilal Piot, Matthieu Geist, Bruno Scherrer, Olivier Pietquin
2015ICMLNon-Stationary Approximate Modified Policy Iteration.Boris Lesner, Bruno Scherrer
2015ICMLApproximate Dynamic Programming for Two-Player Zero-Sum Markov Games.Julien Prolat, Bruno Scherrer, Bilal Piot, Olivier Pietquin
2015ICMLOn the Rate of Convergence and Error Bounds for LSTD(\(\lambda\)).Manel Tagorti, Bruno Scherrer
2014ICMLApproximate Policy Iteration Schemes: A Comparison.Bruno Scherrer
2012ICMLA Dantzig Selector Approach to Temporal Difference Learning.Matthieu Geist, Bruno Scherrer, Alessandro Lazaric, Mohammad Ghavamzadeh
2012ICMLApproximate Modified Policy Iteration.Bruno Scherrer, Victor Gabillon, Mohammad Ghavamzadeh, Matthieu Geist
2011ICMLClassification-based Policy Iteration with a Critic.Victor Gabillon, Alessandro Lazaric, Mohammad Ghavamzadeh, Bruno Scherrer
2010ICMLShould one compute the Temporal Difference fix point or minimize the Bellman Residual? The unified oblique projection view.Bruno Scherrer
2010ICMLLeast-Squares Policy Iteration: Bias-Variance Trade-off in Control Problems.Christophe Thiery, Bruno Scherrer
2007CECConvergence and rate of convergence of a foraging ant model.Amine M. Boumaza, Bruno Scherrer
2007ICRAOptimal control subsumes harmonic control.Amine M. Boumaza, Bruno Scherrer
2003ESANNParallel asynchronous distributed computations of optimal control in large state space Markov Decision processes.Bruno Scherrer
2003IJCAIModular self-organization for a long-living autonomous agent.Bruno Scherrer
2002ICTAICooperative Co-Learning: A Model-Based Approach for Solving Multi Agent Reinforcement Problems.Bruno Scherrer, Franois Charpillet
2002SACA heuristic approach for solving decentralized-POMDP: assessment on the pursuit problem.Iadine Chades, Bruno Scherrer, Franois Charpillet