Skip to content

Attention Weights in Transformer NMT Fail Aligning Words Between Sequences but Largely Explain Model Predictions.

Javier Ferrando, Marta R. Costa-juss

VenueA*EMNLP
Year2021
ProceedingsEMNLP (Findings)

Browse the full EMNLP paper archive.