Skip to content

Examining the robustness of LLM evaluation to the distributional assumptions of benchmarks.

Charlotte Siska, Katerina Marazopoulou, Melissa Ailem, James Bono

VenueA*ACL
Year2024
ProceedingsACL (1)

Browse the full ACL paper archive.