Skip to content

Shengyi Huang

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

3

Venues

2

Active years

2022–2025

Best venue rank

A*

Where they publish

Papers

3 indexed papers, newest first.

YearVenueTitleAuthors
2025ICLRAsynchronous RLHF: Faster and More Efficient Off-Policy RL for Language Models.Michael Noukhovitch, Shengyi Huang, Sophie Xhonneux, Arian Hosseini, Rishabh Agarwal, Aaron C. Courville
2024ICLRCleanba: A Reproducible and Efficient Distributed Reinforcement Learning Platform.Shengyi Huang, Jiayi Weng, Rujikorn Charakorn, Min Lin, Zhongwen Xu, Santiago Ontan
2022FlAIRSA Closer Look at Invalid Action Masking in Policy Gradient Algorithms.Shengyi Huang, Santiago Ontan