Skip to content

Ziran Yang

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

1

Venues

1

Active years

2025–2025

Best venue rank

A*

Where they publish

Papers

1 indexed papers, newest first.

YearVenueTitleAuthors
2025ACLOffline Reinforcement Learning for LLM Multi-step Reasoning.Huaijie Wang, Shibo Hao, Hanze Dong, Shenao Zhang, Yilin Bao, Ziran Yang, Yi Wu