| 2025 | ACL | Training Bilingual LMs with Data Constraints in the Targeted Language. | Skyler Seto, Maartje ter Hoeve, Richard He Bai, Natalie Schluter, David Grangier |
| 2025 | EMNLP | Assessing the Role of Data Quality in Training Bilingual Language Models. | Skyler Seto, Maartje ter Hoeve, Maureen de Seyssel, David Grangier |
| 2025 | ICLR | No Need to Talk: Asynchronous Mixture of Language Models. | Anastasiia Filippova, Angelos Katharopoulos, David Grangier, Ronan Collobert |
| 2025 | ICLR | Task-Adaptive Pretrained Language Models via Clustered-Importance Sampling. | David Grangier, Simin Fan, Skyler Seto, Pierre Ablin |
| 2025 | ICLR | The AdEMAMix Optimizer: Better, Faster, Older. | Matteo Pagliardini, Pierre Ablin, David Grangier |
| 2025 | ICML | Soup-of-Experts: Pretraining Specialist Models via Parameters Averaging. | Pierre Ablin, Angelos Katharopoulos, Skyler Seto, David Grangier |
| 2025 | ICML | Scaling Laws for Forgetting during Finetuning with Pretraining Data Injection. | Louis Bthune, David Grangier, Dan Busbridge, Eleonora Gualdoni, Marco Cuturi, Pierre Ablin |
| 2024 | ACL | Rephrasing the Web: A Recipe for Compute and Data-Efficient Language Modeling. | Pratyush Maini, Skyler Seto, Richard He Bai, David Grangier, Yizhe Zhang, Navdeep Jaitly |
| 2022 | ACL | A Natural Diet: Towards Improving Naturalness of Machine Translation Output. | Markus Freitag, David Vilar, David Grangier, Colin Cherry, George F. Foster |
| 2022 | ACL | The Trade-offs of Domain Adaptation for Neural Language Models. | David Grangier, Dan Iter |
| 2022 | ICLR | Learning Strides in Convolutional Neural Networks. | Rachid Riad, Olivier Teboul, David Grangier, Neil Zeghidour |
| 2022 | NAACL | On Systematic Style Differences between Unsupervised and Supervised MT and an Application for High-Resource Machine Translation. | Kelly Marchisio, Markus Freitag, David Grangier |
| 2021 | ASRU | Dive: End-to-End Speech Diarization Via Iterative Speaker Embedding. | Neil Zeghidour, Olivier Teboul, David Grangier |
| 2021 | ICASSP | Learning From Heterogeneous Eeg Signals with Differentiable Channel Reordering. | Aaqib Saeed, David Grangier, Olivier Pietquin, Neil Zeghidour |
| 2021 | ICASSP | Contrastive Learning of General-Purpose Audio Representations. | Aaqib Saeed, David Grangier, Neil Zeghidour |
| 2021 | ICLR | Auxiliary Task Update Decomposition: the Good, the Bad and the neutral. | Lucio M. Dery, Yann N. Dauphin, David Grangier |
| 2020 | ACL | Toward Better Storylines with Sentence-Level Language Models. | Daphne Ippolito, David Grangier, Douglas Eck, Chris Callison-Burch |
| 2020 | ACL | Translationese as a Language in "Multilingual" NMT. | Parker Riley, Isaac Caswell, Markus Freitag, David Grangier |
| 2020 | EMNLP | BLEU might be Guilty but References are not Innocent. | Markus Freitag, David Grangier, Isaac Caswell |
| 2019 | ACL | ELI5: Long Form Question Answering. | Angela Fan, Yacine Jernite, Ethan Perez, David Grangier, Jason Weston, Michael Auli |
| 2019 | ACL | Unsupervised Paraphrasing without Translation. | Aurko Roy, David Grangier |
| 2019 | CVPR | 3D Human Pose Estimation in Video With Temporal Convolutions and Semi-Supervised Training. | Dario Pavllo, Christoph Feichtenhofer, David Grangier, Michael Auli |
| 2019 | NAACL | fairseq: A Fast, Extensible Toolkit for Sequence Modeling. | Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, Michael Auli |
| 2018 | BMVC | QuaterNet: A Quaternion-based Recurrent Model for Human Motion. | Dario Pavllo, David Grangier, Michael Auli |
| 2018 | EMNLP | Understanding Back-Translation at Scale. | Sergey Edunov, Myle Ott, Michael Auli, David Grangier |
| 2018 | ICML | Analyzing Uncertainty in Neural Machine Translation. | Myle Ott, Michael Auli, David Grangier, Marc'Aurelio Ranzato |
| 2018 | NAACL | Classical Structured Prediction Losses for Sequence to Sequence Learning. | Sergey Edunov, Myle Ott, Michael Auli, David Grangier, Marc'Aurelio Ranzato |
| 2018 | NAACL | QuickEdit: Editing Text & Translations by Crossing Words Out. | David Grangier, Michael Auli |
| 2017 | ACL | A Convolutional Encoder Model for Neural Machine Translation. | Jonas Gehring, Michael Auli, David Grangier, Yann N. Dauphin |
| 2017 | ICML | Language Modeling with Gated Convolutional Networks. | Yann N. Dauphin, Angela Fan, Michael Auli, David Grangier |
| 2017 | ICML | Convolutional Sequence to Sequence Learning. | Jonas Gehring, Michael Auli, David Grangier, Denis Yarats, Yann N. Dauphin |
| 2017 | ICML | Efficient softmax approximation for GPUs. | Edouard Grave, Armand Joulin, Moustapha Ciss, David Grangier, Herv Jgou |
| 2016 | ACL | Strategies for Training Large Vocabulary Neural Language Models. | Wenlin Chen, David Grangier, Michael Auli |
| 2016 | EMNLP | Neural Text Generation from Structured Data with Application to the Biography Domain. | Rmi Lebret, David Grangier, Michael Auli |
| 2012 | SDM | Learning from Heterogeneous Sources via Gradient Boosting Consensus. | Xiaoxiao Shi, Jean-Franois Paiement, David Grangier, Philip S. Yu |
| 2009 | CIKM | Supervised semantic indexing. | Bing Bai, Jason Weston, David Grangier, Ronan Collobert, Kunihiko Sadamasa, Yanjun Qi, Olivier Chapelle, Kilian Q. Weinberger |
| 2009 | ECIR | Supervised Semantic Indexing. | Bing Bai, Jason Weston, Ronan Collobert, David Grangier |
| 2007 | Interspeech | Learning the inter-frame distance for discriminative template-based keyword detection. | David Grangier, Samy Bengio |
| 2006 | ICANN | A Neural Network to Retrieve Images from Text Queries. | David Grangier, Samy Bengio |
| 2005 | CIKM | Inferring document similarity from hyperlinks. | David Grangier, Samy Bengio |