Skip to content

Position: Medical Large Language Model Benchmarks Should Prioritize Construct Validity.

Ahmed Alaa, Thomas Hartvigsen, Niloufar Golchini, Shiladitya Dutta, Frances Dean, Inioluwa Deborah Raji, Travis Zack

VenueA*ICML
Year2025
ProceedingsICML (Position Papers)

Browse the full ICML paper archive.