Ian Osband
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
12
Venues
4
Active years
2016–2023
Best venue rank
A*
Where they publish
Papers
12 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2023 | UAI | Approximate Thompson Sampling via Epistemic Neural Networks. | Ian Osband, Zheng Wen, Seyed Mohammad Asghari, Vikranth Dwaracherla, Morteza Ibrahimi, Xiuyuan Lu, Benjamin Van Roy |
| 2022 | UAI | Evaluating high-order predictive distributions in deep learning. | Ian Osband, Zheng Wen, Seyed Mohammad Asghari, Vikranth Dwaracherla, Xiuyuan Lu, Benjamin Van Roy |
| 2021 | UAI | Matrix games with bandit feedback. | Brendan O'Donoghue, Tor Lattimore, Ian Osband |
| 2020 | ICLR | Hypermodels for Exploration. | Vikranth Dwaracherla, Xiuyuan Lu, Morteza Ibrahimi, Ian Osband, Zheng Wen, Benjamin Van Roy |
| 2020 | ICLR | Making Sense of Reinforcement Learning and Probabilistic Inference. | Brendan O'Donoghue, Ian Osband, Catalin Ionescu |
| 2020 | ICLR | Behaviour Suite for Reinforcement Learning. | Ian Osband, Yotam Doron, Matteo Hessel, John Aslanides, Eren Sezener, Andre Saraiva, Katrina McKinney, Tor Lattimore, Csaba Szepesvri, Satinder Singh, Benjamin Van Roy, Richard S. Sutton, David Silver, Hado van Hasselt |
| 2018 | AAAI | Deep Q-learning From Demonstrations. | Todd Hester, Matej Vecerk, Olivier Pietquin, Marc Lanctot, Tom Schaul, Bilal Piot, Dan Horgan, John Quan, Andrew Sendonaris, Ian Osband, Gabriel Dulac-Arnold, John P. Agapiou, Joel Z. Leibo, Audrunas Gruslys |
| 2018 | ICLR | Noisy Networks For Exploration. | Meire Fortunato, Mohammad Gheshlaghi Azar, Bilal Piot, Jacob Menick, Matteo Hessel, Ian Osband, Alex Graves, Volodymyr Mnih, Rmi Munos, Demis Hassabis, Olivier Pietquin, Charles Blundell, Shane Legg |
| 2018 | ICML | The Uncertainty Bellman Equation and Exploration. | Brendan O'Donoghue, Ian Osband, Rmi Munos, Volodymyr Mnih |
| 2017 | ICML | Minimax Regret Bounds for Reinforcement Learning. | Mohammad Gheshlaghi Azar, Ian Osband, Rmi Munos |
| 2017 | ICML | Why is Posterior Sampling Better than Optimism for Reinforcement Learning? | Ian Osband, Benjamin Van Roy |
| 2016 | ICML | Generalization and Exploration via Randomized Value Functions. | Ian Osband, Benjamin Van Roy, Zheng Wen |