Shengyi Huang
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
3
Venues
2
Active years
2022–2025
Best venue rank
A*
Where they publish
Papers
3 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICLR | Asynchronous RLHF: Faster and More Efficient Off-Policy RL for Language Models. | Michael Noukhovitch, Shengyi Huang, Sophie Xhonneux, Arian Hosseini, Rishabh Agarwal, Aaron C. Courville |
| 2024 | ICLR | Cleanba: A Reproducible and Efficient Distributed Reinforcement Learning Platform. | Shengyi Huang, Jiayi Weng, Rujikorn Charakorn, Min Lin, Zhongwen Xu, Santiago Ontan |
| 2022 | FlAIRS | A Closer Look at Invalid Action Masking in Policy Gradient Algorithms. | Shengyi Huang, Santiago Ontan |