Skip to content

Mask More and Mask Later: Efficient Pre-training of Masked Language Models by Disentangling the [MASK] Token.

Baohao Liao, David Thulke, Sanjika Hewavitharana, Hermann Ney, Christof Monz

VenueA*EMNLP
Year2022
ProceedingsEMNLP (Findings)

Browse the full EMNLP paper archive.