Skip to content

Michal Valko

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

83

Venues

17

Active years

2007–2025

Best venue rank

A*

Where they publish

Papers

83 indexed papers, newest first.

YearVenueTitleAuthors
2025ICMLThe Harder Path: Last Iterate Convergence for Uncoupled Learning in Zero-Sum Games with Bandit Feedback.Cme Fiegel, Pierre Mnard, Tadashi Kozuno, Michal Valko, Vianney Perchet
2024AISTATSA General Theoretical Paradigm to Understand Learning from Human Preferences.Mohammad Gheshlaghi Azar, Zhaohan Daniel Guo, Bilal Piot, Rmi Munos, Mark Rowland, Michal Valko, Daniele Calandriello
2024BMVCPatchhealer: Counterfactual Image Segment Transplants with chest X-ray domain check.Hakan Lane, Michal Valko, Veda Sahaja Bandi, Nandini Lokesh Reddy, Stefan Kramer
2024BMVCCardioCNN: Gamification of counterfactual image transplants.Hakan Lane, Michal Valko, Stefan Kramer
2024ICLRUnlocking the Power of Representations in Long-term Novelty-based Exploration.Alaa Saade, Steven Kapturowski, Daniele Calandriello, Charles Blundell, Pablo Sprechmann, Leopoldo Sarra, Oliver Groth, Michal Valko, Bilal Piot
2024ICLRDemonstration-Regularized RL.Daniil Tiapkin, Denis Belomestny, Daniele Calandriello, Eric Moulines, Alexey Naumov, Pierre Perrault, Michal Valko, Pierre Mnard
2024ICMLHuman Alignment of Large Language Models through Online Preference Optimisation.Daniele Calandriello, Zhaohan Daniel Guo, Rmi Munos, Mark Rowland, Yunhao Tang, Bernardo vila Pires, Pierre Harvey Richemond, Charline Le Lan, Michal Valko, Tianqi Liu, Rishabh Joshi, Zeyu Zheng, Bilal Piot
2024ICMLDecoding-time Realignment of Language Models.Tianlin Liu, Shangmin Guo, Leonardo Bianco, Daniele Calandriello, Quentin Berthet, Felipe Llinares-Lpez, Jessica Hoffmann, Lucas Dixon, Michal Valko, Mathieu Blondel
2024ICMLNash Learning from Human Feedback.Rmi Munos, Michal Valko, Daniele Calandriello, Mohammad Gheshlaghi Azar, Mark Rowland, Daniel Guo, Yunhao Tang, Matthieu Geist, Thomas Mesnard, Cme Fiegel, Andrea Michi, Marco Selvi, Sertan Girgin, Nikola Momchev, Olivier Bachem, Daniel J. Mankowitz, Doina Precup, Bilal Piot
2024ICMLGeneralized Preference Optimization: A Unified Approach to Offline Alignment.Yunhao Tang, Zhaohan Daniel Guo, Zeyu Zheng, Daniele Calandriello, Rmi Munos, Mark Rowland, Pierre Harvey Richemond, Michal Valko, Bernardo vila Pires, Bilal Piot
2023ICMLHalf-Hop: A graph upsampling approach for slowing down message passing.Mehdi Azabou, Venkataramana Ganesh, Shantanu Thakoor, Chi-Heng Lin, Lakshmi Sathidevi, Ran Liu, Michal Valko, Petar Velickovic, Eva L. Dyer
2023ICMLAdapting to game trees in zero-sum imperfect information games.Cme Fiegel, Pierre Mnard, Tadashi Kozuno, Rmi Munos, Vianney Perchet, Michal Valko
2023ICMLCuriosity in Hindsight: Intrinsic Exploration in Stochastic Environments.Daniel Jarrett, Corentin Tallec, Florent Altch, Thomas Mesnard, Rmi Munos, Michal Valko
2023ICMLRegularization and Variance-Weighted Regression Achieves Minimax Optimality in Linear MDPs: Theory and Practice.Toshinori Kitamura, Tadashi Kozuno, Yunhao Tang, Nino Vieillard, Michal Valko, Wenhao Yang, Jincheng Mei, Pierre Mnard, Mohammad Gheshlaghi Azar, Rmi Munos, Olivier Pietquin, Matthieu Geist, Csaba Szepesvri, Wataru Kumagai, Yutaka Matsuo
2023ICMLQuantile Credit Assignment.Thomas Mesnard, Wenqi Chen, Alaa Saade, Yunhao Tang, Mark Rowland, Theophane Weber, Clare Lyle, Audrunas Gruslys, Michal Valko, Will Dabney, Georg Ostrovski, Eric Moulines, Rmi Munos
2023ICMLUnderstanding Self-Predictive Learning for Reinforcement Learning.Yunhao Tang, Zhaohan Daniel Guo, Pierre Harvey Richemond, Bernardo vila Pires, Yash Chandak, Rmi Munos, Mark Rowland, Mohammad Gheshlaghi Azar, Charline Le Lan, Clare Lyle, Andrs Gyrgy, Shantanu Thakoor, Will Dabney, Bilal Piot, Daniele Calandriello, Michal Valko
2023ICMLDoMo-AC: Doubly Multi-step Off-policy Actor-Critic Algorithm.Yunhao Tang, Tadashi Kozuno, Mark Rowland, Anna Harutyunyan, Rmi Munos, Bernardo vila Pires, Michal Valko
2023ICMLVA-learning as a more efficient alternative to Q-learning.Yunhao Tang, Rmi Munos, Mark Rowland, Michal Valko
2023ICMLFast Rates for Maximum Entropy Exploration.Daniil Tiapkin, Denis Belomestny, Daniele Calandriello, Eric Moulines, Rmi Munos, Alexey Naumov, Pierre Perrault, Yunhao Tang, Michal Valko, Pierre Mnard
2022AISTATSMarginalized Operators for Off-policy Reinforcement Learning.Yunhao Tang, Mark Rowland, Rmi Munos, Michal Valko
2022AISTATSAdaptive Multi-Goal Exploration.Jean Tarbouriech, Omar Darwiche Domingues, Pierre Mnard, Matteo Pirotta, Michal Valko, Alessandro Lazaric
2022ICLRLarge-Scale Representation Learning on Graphs via Bootstrapping.Shantanu Thakoor, Corentin Tallec, Mohammad Gheshlaghi Azar, Mehdi Azabou, Eva L. Dyer, Rmi Munos, Petar Velickovic, Michal Valko
2022ICMLScaling Gaussian Process Optimization by Evaluating a Few Unique Candidates Multiple Times.Daniele Calandriello, Luigi Carratino, Alessandro Lazaric, Michal Valko, Lorenzo Rosasco
2022ICMLRetrieval-Augmented Reinforcement Learning.Anirudh Goyal, Abram L. Friesen, Andrea Banino, Theophane Weber, Nan Rosemary Ke, Adri Puigdomnech Badia, Arthur Guez, Mehdi Mirza, Peter Conway Humphreys, Ksenia Konyushkova, Michal Valko, Simon Osindero, Timothy P. Lillicrap, Nicolas Heess, Charles Blundell
2022ICMLFrom Dirichlet to Rubin: Optimistic Exploration in RL without Bonuses.Daniil Tiapkin, Denis Belomestny, Eric Moulines, Alexey Naumov, Sergey Samsonov, Yunhao Tang, Michal Valko, Pierre Mnard
2021AISTATSA Kernel-Based Approach to Non-Stationary Reinforcement Learning in Metric Spaces.Omar Darwiche Domingues, Pierre Mnard, Matteo Pirotta, Emilie Kaufmann, Michal Valko
2021ALTEpisodic Reinforcement Learning in Finite MDPs: Minimax Lower Bounds Revisited.Omar Darwiche Domingues, Pierre Mnard, Emilie Kaufmann, Michal Valko
2021ALTAdaptive Reward-Free Exploration.Emilie Kaufmann, Pierre Mnard, Omar Darwiche Domingues, Anders Jonsson, Edouard Leurent, Michal Valko
2021ALTSample Complexity Bounds for Stochastic Shortest Path with a Generative Model.Jean Tarbouriech, Matteo Pirotta, Michal Valko, Alessandro Lazaric
2021ICCVBroaden Your Views for Self-Supervised Video Learning.Adri Recasens, Pauline Luc, Jean-Baptiste Alayrac, Luyu Wang, Florian Strub, Corentin Tallec, Mateusz Malinowski, Viorica Patraucean, Florent Altch, Michal Valko, Jean-Bastien Grill, Aron van den Oord, Andrew Zisserman
2021ICMLKernel-Based Reinforcement Learning: A Finite-Time Analysis.Omar Darwiche Domingues, Pierre Mnard, Matteo Pirotta, Emilie Kaufmann, Michal Valko
2021ICMLOnline A-Optimal Design and Active Linear Regression.Xavier Fontaine, Pierre Perrault, Michal Valko, Vianney Perchet
2021ICMLRevisiting Peng's Q(λ) for Modern Reinforcement Learning.Tadashi Kozuno, Yunhao Tang, Mark Rowland, Rmi Munos, Steven Kapturowski, Will Dabney, Michal Valko, David Abel
2021ICMLFast active learning for pure exploration in reinforcement learning.Pierre Mnard, Omar Darwiche Domingues, Anders Jonsson, Emilie Kaufmann, Edouard Leurent, Michal Valko
2021ICMLUCB Momentum Q-learning: Correcting the bias without forgetting.Pierre Mnard, Omar Darwiche Domingues, Xuedong Shang, Michal Valko
2021ICMLTaylor Expansion of Discount Factors.Yunhao Tang, Mark Rowland, Rmi Munos, Michal Valko
2020AISTATSDerivative-Free & Order-Robust Optimisation.Haitham Ammar, Victor Gabillon, Rasul Tutunov, Michal Valko
2020AISTATSAdaptive multi-fidelity optimization with fast learning rates.Cme Fiegel, Victor Gabillon, Michal Valko
2020AISTATSA single algorithm for both restless and rested rotting bandits.Julien Seznec, Pierre Mnard, Alessandro Lazaric, Michal Valko
2020AISTATSFixed-confidence guarantees for Bayesian best-arm identification.Xuedong Shang, Rianne de Heide, Pierre Mnard, Emilie Kaufmann, Michal Valko
2020COLTCovariance-adapting algorithm for semi-bandits with application to sparse outcomes.Pierre Perrault, Michal Valko, Vianney Perchet
2020ICMLNear-linear time Gaussian process optimization with adaptive batching and resparsification.Daniele Calandriello, Luigi Carratino, Alessandro Lazaric, Michal Valko, Lorenzo Rosasco
2020ICMLGamification of Pure Exploration for Linear Bandits.Rmy Degenne, Pierre Mnard, Xuedong Shang, Michal Valko
2020ICMLMonte-Carlo Tree Search as Regularized Policy Optimization.Jean-Bastien Grill, Florent Altch, Yunhao Tang, Thomas Hubert, Michal Valko, Ioannis Antonoglou, Rmi Munos
2020ICMLStochastic bandits with arm-dependent delays.Anne Gael Manegueu, Claire Vernade, Alexandra Carpentier, Michal Valko
2020ICMLBudgeted Online Influence Maximization.Pierre Perrault, Jennifer Healey, Zheng Wen, Michal Valko
2020ICMLImproved Sleeping Bandits with Stochastic Action Sets and Adversarial Rewards.Aadirupa Saha, Pierre Gaillard, Michal Valko
2020ICMLTaylor Expansion Policy Optimization.Yunhao Tang, Michal Valko, Rmi Munos
2020ICMLNo-Regret Exploration in Goal-Oriented Reinforcement Learning.Jean Tarbouriech, Evrard Garcelon, Michal Valko, Matteo Pirotta, Alessandro Lazaric
2019AISTATSActive multiple matrix completion with adaptive confidence sets.Andrea Locatelli, Alexandra Carpentier, Michal Valko
2019AISTATSFinding the bandit in a graph: Sequential search-and-stop.Pierre Perrault, Vianney Perchet, Michal Valko
2019AISTATSRotting bandits are no harder than stochastic ones.Julien Seznec, Andrea Locatelli, Alexandra Carpentier, Alessandro Lazaric, Michal Valko
2019ALTA simple parameter-free and adaptive approach to optimization under a minimal local smoothness assumption.Peter L. Bartlett, Victor Gabillon, Michal Valko
2019ALTGeneral parallel optimization a without metric.Xuedong Shang, Emilie Kaufmann, Michal Valko
2019COLTGaussian Process Optimization with Adaptive Sketching: Scalable and No Regret.Daniele Calandriello, Luigi Carratino, Alessandro Lazaric, Michal Valko, Lorenzo Rosasco
2019ICMLScale-free adaptive planning for deterministic dynamics & discounted rewards.Peter L. Bartlett, Victor Gabillon, Jennifer Healey, Michal Valko
2019ICMLExploiting structure of uncertainty for efficient matroid semi-bandits.Pierre Perrault, Vianney Perchet, Michal Valko
2018COLTBest of both worlds: Stochastic & adversarial best-arm identification.Yasin Abbasi-Yadkori, Peter L. Bartlett, Victor Gabillon, Alan Malek, Michal Valko
2018ECCVCompressing the Input for CNNs with the First-Order Scattering Transform.Edouard Oyallon, Eugene Belilovsky, Sergey Zagoruyko, Michal Valko
2018ICMLImproved Large-Scale Graph Learning through Ridge Spectral Sparsification.Daniele Calandriello, Ioannis Koutis, Alessandro Lazaric, Michal Valko
2018ITSPreface.Fabrice Popineau, Michal Valko, Jill-Jnn Vie
2017AISTATSDistributed Adaptive Sampling for Kernel Matrix Approximation.Daniele Calandriello, Alessandro Lazaric, Michal Valko
2017AISTATSTrading off Rewards and Errors in Multi-Armed Bandits.Akram Erraqabi, Alessandro Lazaric, Michal Valko, Emma Brunskill, Yun-En Liu
2017ICMLSecond-Order Kernel Online Convex Optimization with Adaptive Sketching.Daniele Calandriello, Alessandro Lazaric, Michal Valko
2017ICMLZonotope Hit-and-run for Efficient Sampling from Projection DPPs.Guillaume Gautier, Rmi Bardenet, Michal Valko
2016AISTATSRevealing Graph Bandits for Maximizing Local Influence.Alexandra Carpentier, Michal Valko
2016AISTATSOnline Learning with Noisy Side Observations.Toms Kock, Gergely Neu, Michal Valko
2016ICMLPliable Rejection Sampling.Akram Erraqabi, Michal Valko, Alexandra Carpentier, Odalric-Ambrym Maillard
2016UAIAnalysis of Nystrm method with sequential ridge leverage scores.Daniele Calandriello, Alessandro Lazaric, Michal Valko
2016UAIOnline learning with Erdos-Renyi side-observation graphs.Toms Kock, Gergely Neu, Michal Valko
2015ICMLSimple regret for infinitely many armed bandits.Alexandra Carpentier, Michal Valko
2015ICMLCheap Bandits.Manjesh Kumar Hanawal, Venkatesh Saligrama, Michal Valko, Rmi Munos
2015IJCAIMaximum Entropy Semi-Supervised Inverse Reinforcement Learning.Julien Audiffren, Michal Valko, Alessandro Lazaric, Mohammad Ghavamzadeh
2014AAAISpectral Thompson Sampling.Toms Kock, Michal Valko, Rmi Munos, Shipra Agrawal
2014CECBandits attack function optimization.Philippe Preux, Rmi Munos, Michal Valko
2014ICMLSpectral Bandits for Smooth Graph Functions.Michal Valko, Rmi Munos, Branislav Kveton, Toms Kock
2013ICMLStochastic Simultaneous Optimistic Optimization.Michal Valko, Alexandra Carpentier, Rmi Munos
2013UAIFinite-Time Analysis of Kernelised Contextual Bandits.Michal Valko, Nathaniel Korda, Rmi Munos, Ilias N. Flaounas, Nello Cristianini
2011ICDMConditional Anomaly Detection with Soft Harmonic Functions.Michal Valko, Branislav Kveton, Hamed Valizadegan, Gregory F. Cooper, Milos Hauskrecht
2010CVPROnline semi-supervised perception: Real-time learning without explicit feedback.Branislav Kveton, Matthai Philipose, Michal Valko, Ling Huang
2010UAIOnline Semi-Supervised Learning on Quantized Graphs.Michal Valko, Branislav Kveton, Ling Huang, Daniel Ting
2008FlAIRSDistance Metric Learning for Conditional Anomaly Detection.Michal Valko, Milos Hauskrecht
2007AMIAEvidence-based Anomaly Detection in Clinical Domains.Milos Hauskrecht, Michal Valko, Branislav Kveton, Shyam Visweswaran, Gregory F. Cooper