| 2025 | ICML | Stable Offline Value Function Learning with Bisimulation-based Representations. | Brahma S. Pavse, Yudong Chen, Qiaomin Xie, Josiah P. Hanna |
| 2025 | ICML | Demystifying the Paradox of Importance Sampling with an Estimated History-Dependent Behavior Policy in Off-Policy Evaluation. | Hongyi Zhou, Josiah P. Hanna, Jin Zhu, Ying Yang, Chengchun Shi |
| 2025 | ICRA | Multi-Robot Collaboration Through Reinforcement Learning and Abstract Simulation. | Adam Labiosa, Josiah P. Hanna |
| 2025 | ICRA | Reinforcement Learning Within the Classical Robotics Stack: A Case Study in Robot Soccer. | Adam Labiosa, Zhihan Wang, Siddhant Agarwal, William Cong, Geethika Hemkumar, Abhinav Narayan Harish, Benjamin Hong, Josh Kelle, Chen Li, Yuhao Li, Zisen Shao, Peter Stone, Josiah P. Hanna |
| 2024 | AAAI | Scaling Offline Evaluation of Reinforcement Learning Agents through Abstraction. | Josiah P. Hanna |
| 2024 | AISTATS | SPEED: Experimental Design for Policy Evaluation in Linear Heteroscedastic Bandits. | Subhojyoti Mukherjee, Qiaomin Xie, Josiah P. Hanna, Robert D. Nowak |
| 2024 | ECCV | Reinforcement Learning via Auxiliary Task Distillation. | Abhinav Narayan Harish, Larry Heck, Josiah P. Hanna, Zsolt Kira, Andrew Szot |
| 2024 | ICLR | Understanding when Dynamics-Invariant Data Augmentations Benefit Model-free Reinforcement Learning Updates. | Nicholas Corrado, Josiah P. Hanna |
| 2024 | ICML | SaVeR: Optimal Data Collection Strategy for Safe Policy Evaluation in Tabular MDP. | Subhojyoti Mukherjee, Josiah P. Hanna, Robert D. Nowak |
| 2024 | ICML | Learning to Stabilize Online Reinforcement Learning in Unbounded State Spaces. | Brahma S. Pavse, Matthew Zurek, Yudong Chen, Qiaomin Xie, Josiah P. Hanna |
| 2023 | AAAI | Scaling Marginalized Importance Sampling to High-Dimensional State-Spaces via State Abstraction. | Brahma S. Pavse, Josiah P. Hanna |
| 2023 | ICLR | Temporal Disentanglement of Representations for Improved Generalisation in Reinforcement Learning. | Mhairi Dunion, Trevor McInroe, Kevin Sebastian Luck, Josiah P. Hanna, Stefano V. Albrecht |
| 2022 | UAI | ReVar: Strengthening policy evaluation via reduced variance sampling. | Subhojyoti Mukherjee, Josiah P. Hanna, Robert D. Nowak |
| 2021 | IROS | A Joint Imitation-Reinforcement Learning Framework for Reduced Baseline Regret. | Sheelabhadra Dey, Sumedh Pendurkar, Guni Sharon, Josiah P. Hanna |
| 2021 | IROS | Interpretable Goal Recognition in the Presence of Occluded Factors for Autonomous Vehicles. | Josiah P. Hanna, Arrasy Rahman, Elliot Fosong, Francisco Eiras, Mihai Dobre, John Redford, Subramanian Ramamoorthy, Stefano V. Albrecht |
| 2021 | PAAMS | Towards Quantum-Secure Authentication and Key Agreement via Abstract Multi-Agent Interaction. | Ibrahim Ahmed, Josiah P. Hanna, Elliot Fosong, Stefano V. Albrecht |
| 2020 | IROS | Stochastic Grounded Action Transformation for Robot Learning in Simulation. | Siddharth Desai, Haresh Karnan, Josiah P. Hanna, Garrett Warnell, Peter Stone |
| 2020 | IROS | Reinforced Grounded Action Transformation for Sim-to-Real Transfer. | Haresh Karnan, Siddharth Desai, Josiah P. Hanna, Garrett Warnell, Peter Stone |
| 2019 | AAAI | Selecting Compliant Agents for Opt-in Micro-Tolling. | Josiah P. Hanna, Guni Sharon, Stephen D. Boyles, Peter Stone |
| 2018 | AAAI | DyETC: Dynamic Electronic Toll Collection for Traffic Congestion Alleviation. | Haipeng Chen, Bo An, Guni Sharon, Josiah P. Hanna, Peter Stone, Chunyan Miao, Yeng Chai Soh |
| 2017 | AAAI | Grounded Action Transformation for Robot Learning in Simulation. | Josiah P. Hanna, Peter Stone |
| 2017 | AAAI | Grounded Action Transformation for Robot Learning in Simulation. | Josiah P. Hanna, Peter Stone |
| 2017 | AAAI | Bootstrapping with Models: Confidence Intervals for Off-Policy Evaluation. | Josiah P. Hanna, Peter Stone, Scott Niekum |
| 2017 | ICML | Data-Efficient Policy Evaluation Through Behavior Policy Search. | Josiah P. Hanna, Philip S. Thomas, Peter Stone, Scott Niekum |