Skip to content

PIRA: Preference-Oriented Instruction-Tuned Reward Models with Dual Aggregation.

Yongfu Xue

VenueAEACL
Year2026
ProceedingsEACL (Findings)

Browse the full EACL paper archive.