Skip to content

Vincent Zhuang

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

7

Venues

5

Active years

2017–2025

Best venue rank

A*

Where they publish

Papers

7 indexed papers, newest first.

YearVenueTitleAuthors
2025ICLRInference-Aware Fine-Tuning for Best-of-N Sampling in Large Language Models.Yinlam Chow, Guy Tennenholtz, Izzeddin Gur, Vincent Zhuang, Bo Dai, Aviral Kumar, Rishabh Agarwal, Sridhar Thiagarajan, Craig Boutilier, Aleksandra Faust
2025ICLRTraining Language Models to Self-Correct via Reinforcement Learning.Aviral Kumar, Vincent Zhuang, Rishabh Agarwal, Yi Su, John D. Co-Reyes, Avi Singh, Kate Baumli, Shariq Iqbal, Colton Bishop, Rebecca Roelofs, Lei M. Zhang, Kay McKinney, Disha Shrivastava, Cosmin Paduraru, George Tucker, Doina Precup, Feryal M. P. Behbahani, Aleksandra Faust
2025ICLRMotion Control of High-Dimensional Musculoskeletal Systems with Hierarchical Model-Based Planning.Yunyue Wei, Shanning Zhuang, Vincent Zhuang, Yanan Sui
2024IROSThe Design of the Barkour Benchmark for Robot Agility.Wenhao Yu, Ken Caluwaerts, Atil Iscen, J. Chase Kew, Tingnan Zhang, Daniel Freeman, Lisa Lee, Stefano Saliceti, Vincent Zhuang, Nathan Batchelor, Steven Bohez, Federico Casarini, Jos Enrique Chen, Erwin Coumans, Adil Dostmohamed, Gabriel Dulac-Arnold, Alejandro Escontrela, Erik Frey, Roland Hafner, Deepali Jain, Bauyrjan Jyenis, Yuheng Kuang, Tsang-Wei Edward Lee, Ofir Nachum, Ken Oslund, Francesco Romano, Fereshteh Sadeghi, Baruch Tabanpour, Daniel Zheng, Michael Neunert, Raia Hadsell, Nicolas Heess, Francesco Nori, Jeff Seto, Carolina Parada, Vikas Sindhwani, Vincent Vanhoucke, Jie Tan, Kuang-Huei Lee
2021AISTATSNo-Regret Reinforcement Learning with Heavy-Tailed Rewards.Vincent Zhuang, Yanan Sui
2018ICMLStagewise Safe Bayesian Optimization with Gaussian Processes.Yanan Sui, Vincent Zhuang, Joel W. Burdick, Yisong Yue
2017UAIMulti-dueling Bandits with Dependent Arms.Yanan Sui, Vincent Zhuang, Joel W. Burdick, Yisong Yue