Skip to content

Position: AI Competitions Provide the Gold Standard for Empirical Rigor in GenAI Evaluation.

D. Sculley, William Cukierski, Phil Culliton, Sohier Dane, Maggie Demkin, Ryan Holbrook, Addison Howard, Paul Mooney, Walter Reade, Meg Risdal, Nate Keating

VenueA*ICML
Year2025
ProceedingsICML (Position Papers)

Browse the full ICML paper archive.