Skip to content

Pre-trained Models Perform the Best When Token Distributions Follow Zipf's Law.

Yanjin He, Qingkai Zeng, Meng Jiang

VenueA*EMNLP
Year2025
ProceedingsEMNLP

Browse the full EMNLP paper archive.