Skip to content

PersoDPO: Scalable Preference Optimization for Instruction-Adherent, Persona-Grounded Dialogue via Multi-LLM Evaluation.

Saleh Afzoon, MohammadHossein Ahmadi, Usman Naseem, Amin Beheshti

VenueBWISE
Year2025
ProceedingsWISE (2)

Browse the full WISE paper archive.