Skip to content

Never Train from Scratch: Fair Comparison of Long-Sequence Models Requires Data-Driven Priors.

Ido Amos, Jonathan Berant, Ankit Gupta

VenueA*ICLR
Year2024
ProceedingsICLR

Browse the full ICLR paper archive.