Skip to content

Kelvin Xu

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

13

Venues

4

Active years

2015–2025

Best venue rank

A*

Where they publish

Papers

13 indexed papers, newest first.

YearVenueTitleAuthors
2025ICLRScaling LLM Test-Time Compute Optimally Can be More Effective than Scaling Parameters for Reasoning.Charlie Victor Snell, Jaehoon Lee, Kelvin Xu, Aviral Kumar
2025ICMLLMRL Gym: Benchmarks for Multi-Turn Reinforcement Learning with Language Models.Marwa Abdulhai, Isadora White, Charlie Victor Snell, Charles Sun, Joey Hong, Yuexiang Zhai, Kelvin Xu, Sergey Levine
2024ICLRSmall-scale proxies for large-scale Transformer training instabilities.Mitchell Wortsman, Peter J. Liu, Lechao Xiao, Katie E. Everett, Alexander A. Alemi, Ben Adlam, John D. Co-Reyes, Izzeddin Gur, Abhishek Kumar, Roman Novak, Jeffrey Pennington, Jascha Sohl-Dickstein, Kelvin Xu, Jaehoon Lee, Justin Gilmer, Simon Kornblith
2023ICRADexterous Manipulation from Images: Autonomous Real-World RL via Substep Guidance.Kelvin Xu, Zheyuan Hu, Ria Doshi, Aaron Rovinsky, Vikash Kumar, Abhishek Gupta, Sergey Levine
2022ICLRAutonomous Reinforcement Learning: Formalism and Benchmarking.Archit Sharma, Kelvin Xu, Nikhil Sardana, Abhishek Gupta, Karol Hausman, Sergey Levine, Chelsea Finn
2021ICRAReset-Free Reinforcement Learning via Multi-Task Learning: Learning Dexterous Manipulation Behaviors without Human Intervention.Abhishek Gupta, Justin Yu, Tony Z. Zhao, Vikash Kumar, Aaron Rovinsky, Kelvin Xu, Thomas Devlin, Sergey Levine
2020ICLRMeta-Dataset: A Dataset of Datasets for Learning to Learn from Few Examples.Eleni Triantafillou, Tyler Zhu, Vincent Dumoulin, Pascal Lamblin, Utku Evci, Kelvin Xu, Ross Goroshin, Carles Gelada, Kevin Swersky, Pierre-Antoine Manzagol, Hugo Larochelle
2019ICMLLearning a Prior over Intent via Meta-Inverse Reinforcement Learning.Kelvin Xu, Ellis Ratner, Anca D. Dragan, Sergey Levine, Chelsea Finn
2019VCIPPrivacy-Preserving Fall Detection with Deep Learning on mmWave Radar Signal.Yangfan Sun, Renlong Hang, Zhu Li, Mouqing Jin, Kelvin Xu
2018ICLRTrust-PCL: An Off-Policy Trust Region Method for Continuous Control.Ofir Nachum, Mohammad Norouzi, Kelvin Xu, Dale Schuurmans
2017ICLRAn Actor-Critic Algorithm for Sequence Prediction.Dzmitry Bahdanau, Philemon Brakel, Kelvin Xu, Anirudh Goyal, Ryan Lowe, Joelle Pineau, Aaron C. Courville, Yoshua Bengio
2017ICLRUnsupervised Perceptual Rewards for Imitation Learning.Pierre Sermanet, Kelvin Xu, Sergey Levine
2015ICMLShow, Attend and Tell: Neural Image Caption Generation with Visual Attention.Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron C. Courville, Ruslan Salakhutdinov, Richard S. Zemel, Yoshua Bengio