Skip to content

PPTC-R benchmark: Towards Evaluating the Robustness of Large Language Models for PowerPoint Task Completion.

Zekai Zhang, Yiduo Guo, Yaobo Liang, Dongyan Zhao, Nan Duan

VenueA*EMNLP
Year2024
ProceedingsEMNLP (Findings)

Browse the full EMNLP paper archive.