Benjamin Van Roy
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
24
Venues
11
Active years
2000–2026
Best venue rank
A*
Where they publish
Papers
24 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | AAAI | Misalignment from Treating Means as Ends. | Henrik Marklund, Alex Infanger, Benjamin Van Roy |
| 2024 | ICML | Efficient Exploration for LLMs. | Vikranth Dwaracherla, Seyed Mohammad Asghari, Botao Hao, Benjamin Van Roy |
| 2024 | ICML | An Information-Theoretic Analysis of In-Context Learning. | Hong Jun Jeon, Jason D. Lee, Qi Lei, Benjamin Van Roy |
| 2023 | AISTATS | Nonstationary Bandit Learning via Predictive Sampling. | Yueyang Liu, Benjamin Van Roy, Kuang Xu |
| 2023 | CIKM | Scalable Neural Contextual Bandit for Recommender Systems. | Zheqing Zhu, Benjamin Van Roy |
| 2023 | ICML | Leveraging Demonstrations to Improve Online Learning: Quality Matters. | Botao Hao, Rahul Jain, Tor Lattimore, Benjamin Van Roy, Zheng Wen |
| 2023 | RecSys | Deep Exploration for Recommendation Systems. | Zheqing Zhu, Benjamin Van Roy |
| 2023 | UAI | Approximate Thompson Sampling via Epistemic Neural Networks. | Ian Osband, Zheng Wen, Seyed Mohammad Asghari, Vikranth Dwaracherla, Morteza Ibrahimi, Xiuyuan Lu, Benjamin Van Roy |
| 2022 | UAI | Evaluating high-order predictive distributions in deep learning. | Ian Osband, Zheng Wen, Seyed Mohammad Asghari, Vikranth Dwaracherla, Xiuyuan Lu, Benjamin Van Roy |
| 2021 | ICML | Deciding What to Learn: A Rate-Distortion Approach. | Dilip Arumugam, Benjamin Van Roy |
| 2020 | ICLR | Hypermodels for Exploration. | Vikranth Dwaracherla, Xiuyuan Lu, Morteza Ibrahimi, Ian Osband, Zheng Wen, Benjamin Van Roy |
| 2020 | ICLR | Behaviour Suite for Reinforcement Learning. | Ian Osband, Yotam Doron, Matteo Hessel, John Aslanides, Eren Sezener, Andre Saraiva, Katrina McKinney, Tor Lattimore, Csaba Szepesvri, Satinder Singh, Benjamin Van Roy, Richard S. Sutton, David Silver, Hado van Hasselt |
| 2019 | COLT | On the Performance of Thompson Sampling on Logistic Bandits. | Shi Dong, Tengyu Ma, Benjamin Van Roy |
| 2018 | ICML | Coordinated Exploration in Concurrent Reinforcement Learning. | Maria Dimakopoulou, Benjamin Van Roy |
| 2017 | ICML | Why is Posterior Sampling Better than Optimism for Reinforcement Learning? | Ian Osband, Benjamin Van Roy |
| 2016 | ICML | Generalization and Exploration via Randomized Value Functions. | Ian Osband, Benjamin Van Roy, Zheng Wen |
| 2009 | RecSys | Manipulation-resistant collaborative filtering systems. | Benjamin Van Roy, Xiang Yan |
| 2008 | SIGCOMM | Reputation markets. | Xiang Yan, Benjamin Van Roy |
| 2007 | ISIT | Capacity and Zero-Error Capacity of the Chemical Channel with Feedback. | Haim H. Permuter, Paul Cuff, Benjamin Van Roy, Tsachy Weissman |
| 2005 | ISIT | A universal scheme for learning. | Vivek F. Farias, Ciamac C. Moallemi, Benjamin Van Roy, Tsachy Weissman |
| 2004 | WAW | Making Eigenvector-Based Reputation Systems Robust to Collusion. | Hui Zhang, Ashish Goel, Ramesh Govindan, Kahn Mason, Benjamin Van Roy |
| 2001 | ICML | A Generalized Kalman Filter for Fixed Point Approximation and Efficient Temporal Difference Learning. | David Choi, Benjamin Van Roy |
| 2001 | UAI | A Tractable POMDP for Dynamic Sequencing with Applications to Personalized Internet Content Provision. | Paat Rusmevichientong, Benjamin Van Roy |
| 2000 | ICML | Fixed Points of Approximate Value Iteration and Temporal-Difference Learning. | Daniela Pucci de Farias, Benjamin Van Roy |