Skip to content

Joey Hong

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

16

Venues

5

Active years

2017–2025

Best venue rank

A*

Where they publish

Papers

16 indexed papers, newest first.

YearVenueTitleAuthors
2025ICLRQ-SFT: Q-Learning for Language Models via Supervised Fine-Tuning.Joey Hong, Anca D. Dragan, Sergey Levine
2025ICMLLMRL Gym: Benchmarks for Multi-Turn Reinforcement Learning with Language Models.Marwa Abdulhai, Isadora White, Charlie Victor Snell, Charles Sun, Joey Hong, Yuexiang Zhai, Kelvin Xu, Sergey Levine
2024ICLROffline RL with Observation Histories: Analyzing and Improving Sample Complexity.Joey Hong, Anca D. Dragan, Sergey Levine
2024ICLRExeDec: Execution Decomposition for Compositional Generalization in Neural Program Synthesis.Kensen Shi, Joey Hong, Yinlin Deng, Pengcheng Yin, Manzil Zaheer, Charles Sutton
2024ICMLLearning to Explore in POMDPs with Informational Rewards.Annie Xie, Logan M. Bhamidipaty, Evan Zheran Liu, Joey Hong, Sergey Levine, Chelsea Finn
2023ICLROn the Sensitivity of Reward Inference to Misspecified Human Models.Joey Hong, Kush Bhatia, Anca D. Dragan
2023ICLRConfidence-Conditioned Value Functions for Offline Reinforcement Learning.Joey Hong, Aviral Kumar, Sergey Levine
2023ICMLMulti-Task Off-Policy Learning from Bandit Feedback.Joey Hong, Branislav Kveton, Manzil Zaheer, Sumeet Katariya, Mohammad Ghavamzadeh
2022AISTATSHierarchical Bayesian Bandits.Joey Hong, Branislav Kveton, Manzil Zaheer, Mohammad Ghavamzadeh
2022AISTATSThompson Sampling with a Mixture Prior.Joey Hong, Branislav Kveton, Manzil Zaheer, Mohammad Ghavamzadeh, Craig Boutilier
2022ICLRShould I Run Offline Reinforcement Learning or Behavioral Cloning?Aviral Kumar, Joey Hong, Anikait Singh, Sergey Levine
2022ICMLDeep Hierarchy in Bandits.Joey Hong, Branislav Kveton, Sumeet Katariya, Manzil Zaheer, Mohammad Ghavamzadeh
2021AISTATSNon-Stationary Off-Policy Optimization.Joey Hong, Branislav Kveton, Manzil Zaheer, Yinlam Chow, Amr Ahmed
2021ICMLLatent Programmer: Discrete Latent Codes for Program Synthesis.Joey Hong, David Dohan, Rishabh Singh, Charles Sutton, Manzil Zaheer
2019CVPRRules of the Road: Predicting Driving Behavior With a Convolutional Model of Semantic Interactions.Joey Hong, Benjamin Sapp, James Philbin
2017IRIEnsemble Maximum Entropy Classification and Linear Regression for Author Age Prediction.Joey Hong, Chris A. Mattmann, Paul M. Ramirez