Skip to content

RAIDEN Benchmark: Evaluating Role-playing Conversational Agents with Measurement-Driven Custom Dialogues.

Bowen Wu, Kaili Sun, Ziwei Bai, Ying Li, Baoxun Wang

VenueBCOLING
Year2025
ProceedingsCOLING

Browse the full COLING paper archive.