Skip to content

GAPO: Learning Preferential Prompt through Generative Adversarial Policy Optimization.

Zhouhong Gu, Xingzhou Chen, Xiaoran Shi, Tao Wang, Suhang Zheng, Tianyu Li, Hongwei Feng, Yanghua Xiao

VenueA*ACL
Year2025
ProceedingsACL (1)

Browse the full ACL paper archive.