Skip to content

metabench - A Sparse Benchmark of Reasoning and Knowledge in Large Language Models.

Alexander Kipnis, Konstantinos Voudouris, Luca M. Schulze Buschoff, Eric Schulz

VenueA*ICLR
Year2025
ProceedingsICLR

Browse the full ICLR paper archive.