| 2024 | ACL | Aligning Speech Segments Beyond Pure Semantics. | Kevin Heffernan, Artyom Kozhevnikov, Loc Barrault, Alexandre Mourachko, Holger Schwenk |
| 2023 | ACL | xSIM++: An Improved Proxy to Bitext Mining Performance for Low-Resource Languages. | Mingda Chen, Kevin Heffernan, Onur elebi, Alexandre Mourachko, Holger Schwenk |
| 2023 | EACL | Multilingual Representation Distillation with Contrastive Learning. | Weiting Tan, Kevin Heffernan, Holger Schwenk, Philipp Koehn |
| 2022 | EMNLP | stopes - Modular Machine Translation Pipelines. | Pierre Andrews, Guillaume Wenzek, Kevin Heffernan, Onur elebi, Anna Y. Sun, Ammar Kamran, Yingzhe Guo, Alexandre Mourachko, Holger Schwenk, Angela Fan |
| 2022 | EMNLP | Bitext Mining Using Distilled Sentence Representations for Low-Resource Languages. | Kevin Heffernan, Onur elebi, Holger Schwenk |
| 2022 | LREC | Problem-solving Recognition in Scientific Text. | Kevin Heffernan, Simone Teufel |
| 2020 | COLING | Homonym normalisation by word sense clustering: a case in Japanese. | Yo Sato, Kevin Heffernan |
| 2020 | LREC | Dialect Clustering with Character-Based Metrics: in Search of the Boundary of Language and Dialect. | Yo Sato, Kevin Heffernan |
| 2018 | LREC | Creating dialect sub-corpora by clustering: a case in Japanese for an adaptive method. | Yo Sato, Kevin Heffernan |
| 2017 | SIGIR | Identifying Problems and Solutions in Scientific Text. | Kevin Heffernan, Simone Teufel |