| 2026 | LREC | Phonetic-based Ranking for Improved Pseudo-Labeling in Low-Resource ASR. | Marco Matassoni, Roberto Gretter, Falavigna Daniele, Mohamed Nabih Ali, Alessio Brutti, Matteo Negri, Mauro Cettolo, Marco Gaido, Sara Papi, Luisa Bentivogli |
| 2025 | Interspeech | Granary: Speech Recognition and Translation Dataset in 25 European Languages. | Nithin Rao Koluguri, Monica Sekoyan, George Zelenfroynd, Sasha Meister, Shuoyang Ding, Sofia Kostandian, He Huang, Nikolay Karpov, Jagadeesh Balam, Vitaly Lavrukhin, Yifan Peng, Sara Papi, Marco Gaido, Alessio Brutti, Boris Ginsburg |
| 2025 | Interspeech | How to Connect Speech Foundation Models and Large Language Models? What Matters and What Does Not. | Francesco Verdini, Pierfrancesco Melucci, Stefano Perna, Francesco Cariaggi, Marco Gaido, Sara Papi, Szymon Mazurek, Marek Kasztelnik, Luisa Bentivogli, Sbastien Bratires, Paolo Merialdo, Simone Scardapane |
| 2025 | NAACL | Prepending or Cross-Attention for Speech-to-Text? An Empirical Comparison. | Tsz Kin Lam, Marco Gaido, Sara Papi, Luisa Bentivogli, Barry Haddow |
| 2024 | ACL | Speech Translation with Speech Foundation Models and Large Language Models: What is There and What is Missing? | Marco Gaido, Sara Papi, Matteo Negri, Luisa Bentivogli |
| 2024 | ACL | SBAAM! Eliminating Transcript Dependency in Automatic Subtitling. | Marco Gaido, Sara Papi, Matteo Negri, Mauro Cettolo, Luisa Bentivogli |
| 2024 | ACL | StreamAtt: Direct Streaming Speech-to-Text Translation with Attention-based Audio History Selection. | Sara Papi, Marco Gaido, Matteo Negri, Luisa Bentivogli |
| 2024 | ACL | When Good and Reproducible Results are a Giant with Feet of Clay: The Importance of Software Quality in NLP. | Sara Papi, Marco Gaido, Andrea Pilzer, Matteo Negri |
| 2024 | COLING | How Do Hyenas Deal with Human Speech? Speech Recognition and Translation with ConfHyena. | Marco Gaido, Sara Papi, Matteo Negri, Luisa Bentivogli |
| 2024 | EMNLP | MOSEL: 950, 000 Hours of Speech Data for Open-Source Speech Foundation Model Training on EU Languages. | Marco Gaido, Sara Papi, Luisa Bentivogli, Alessio Brutti, Mauro Cettolo, Roberto Gretter, Marco Matassoni, Mohamed Nabih Ali, Matteo Negri |
| 2024 | EMNLP | What the Harm? Quantifying the Tangible Impact of Gender Bias in Machine Translation with a Human-centered Study. | Beatrice Savoldi, Sara Papi, Matteo Negri, Ana Guerberof Arenas, Luisa Bentivogli |
| 2024 | ICASSP | Leveraging Timestamp Information for Serialized Joint Streaming Recognition and Translation. | Sara Papi, Peidong Wang, Jun-Kun Chen, Jian Xue, Naoyuki Kanda, Jinyu Li, Yashesh Gaur |
| 2023 | ACL | Attention as a Guide for Simultaneous Speech Translation. | Sara Papi, Matteo Negri, Marco Turchi |
| 2023 | ASRU | Token-Level Serialized Output Training for Joint Streaming ASR and ST Leveraging Textual Alignments. | Sara Papi, Peidong Wang, Jun-Kun Chen, Jian Xue, Jinyu Li, Yashesh Gaur |
| 2023 | EMNLP | Integrating Language Models into Direct Speech Translation: An Inference-Time Solution to Control Gender Inflection. | Dennis Fucci, Marco Gaido, Sara Papi, Mauro Cettolo, Matteo Negri, Luisa Bentivogli |
| 2023 | Interspeech | Joint Speech Translation and Named Entity Recognition. | Marco Gaido, Sara Papi, Matteo Negri, Marco Turchi |
| 2023 | Interspeech | AlignAtt: Using Attention-based Audio-Translation Alignments as a Guide for Simultaneous Speech Translation. | Sara Papi, Marco Turchi, Matteo Negri |
| 2022 | EMNLP | Does Simultaneous Speech Translation need Simultaneous Models? | Sara Papi, Marco Gaido, Matteo Negri, Marco Turchi |
| 2022 | IJCNLP | Dodging the Data Bottleneck: Automatic Subtitling with Automatically Segmented ST Corpora. | Sara Papi, Alina Karakanta, Matteo Negri, Marco Turchi |
| 2021 | EMNLP | Speechformer: Reducing Information Loss in Direct Speech Translation. | Sara Papi, Marco Gaido, Matteo Negri, Marco Turchi |
| 2020 | Interspeech | Mixtures of Deep Neural Experts for Automated Speech Scoring. | Sara Papi, Edmondo Trentin, Roberto Gretter, Marco Matassoni, Daniele Falavigna |