| 2025 | ICLR | No Need to Talk: Asynchronous Mixture of Language Models. | Anastasiia Filippova, Angelos Katharopoulos, David Grangier, Ronan Collobert |
| 2025 | ICML | Soup-of-Experts: Pretraining Specialist Models via Parameters Averaging. | Pierre Ablin, Angelos Katharopoulos, Skyler Seto, David Grangier |
| 2023 | CVPR | Masked Autoencoding Does Not Help Natural Language Supervision at Scale. | Floris Weers, Vaishaal Shankar, Angelos Katharopoulos, Yinfei Yang, Tom Gunter |
| 2021 | CVPR | Neural Parts: Learning Expressive 3D Shape Abstractions With Invertible Neural Networks. | Despoina Paschalidou, Angelos Katharopoulos, Andreas Geiger, Sanja Fidler |
| 2020 | ICML | Transformers are RNNs: Fast Autoregressive Transformers with Linear Attention. | Angelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, Franois Fleuret |
| 2019 | ICML | Processing Megapixel Images with Deep Attention-Sampling Models. | Angelos Katharopoulos, Franois Fleuret |
| 2018 | ICML | Not All Samples Are Created Equal: Deep Learning with Importance Sampling. | Angelos Katharopoulos, Franois Fleuret |