Skip to content

Sara Papi

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

21

Venues

9

Active years

2020–2026

Best venue rank

A*

Where they publish

Papers

21 indexed papers, newest first.

YearVenueTitleAuthors
2026LRECPhonetic-based Ranking for Improved Pseudo-Labeling in Low-Resource ASR.Marco Matassoni, Roberto Gretter, Falavigna Daniele, Mohamed Nabih Ali, Alessio Brutti, Matteo Negri, Mauro Cettolo, Marco Gaido, Sara Papi, Luisa Bentivogli
2025InterspeechGranary: Speech Recognition and Translation Dataset in 25 European Languages.Nithin Rao Koluguri, Monica Sekoyan, George Zelenfroynd, Sasha Meister, Shuoyang Ding, Sofia Kostandian, He Huang, Nikolay Karpov, Jagadeesh Balam, Vitaly Lavrukhin, Yifan Peng, Sara Papi, Marco Gaido, Alessio Brutti, Boris Ginsburg
2025InterspeechHow to Connect Speech Foundation Models and Large Language Models? What Matters and What Does Not.Francesco Verdini, Pierfrancesco Melucci, Stefano Perna, Francesco Cariaggi, Marco Gaido, Sara Papi, Szymon Mazurek, Marek Kasztelnik, Luisa Bentivogli, Sbastien Bratires, Paolo Merialdo, Simone Scardapane
2025NAACLPrepending or Cross-Attention for Speech-to-Text? An Empirical Comparison.Tsz Kin Lam, Marco Gaido, Sara Papi, Luisa Bentivogli, Barry Haddow
2024ACLSpeech Translation with Speech Foundation Models and Large Language Models: What is There and What is Missing?Marco Gaido, Sara Papi, Matteo Negri, Luisa Bentivogli
2024ACLSBAAM! Eliminating Transcript Dependency in Automatic Subtitling.Marco Gaido, Sara Papi, Matteo Negri, Mauro Cettolo, Luisa Bentivogli
2024ACLStreamAtt: Direct Streaming Speech-to-Text Translation with Attention-based Audio History Selection.Sara Papi, Marco Gaido, Matteo Negri, Luisa Bentivogli
2024ACLWhen Good and Reproducible Results are a Giant with Feet of Clay: The Importance of Software Quality in NLP.Sara Papi, Marco Gaido, Andrea Pilzer, Matteo Negri
2024COLINGHow Do Hyenas Deal with Human Speech? Speech Recognition and Translation with ConfHyena.Marco Gaido, Sara Papi, Matteo Negri, Luisa Bentivogli
2024EMNLPMOSEL: 950, 000 Hours of Speech Data for Open-Source Speech Foundation Model Training on EU Languages.Marco Gaido, Sara Papi, Luisa Bentivogli, Alessio Brutti, Mauro Cettolo, Roberto Gretter, Marco Matassoni, Mohamed Nabih Ali, Matteo Negri
2024EMNLPWhat the Harm? Quantifying the Tangible Impact of Gender Bias in Machine Translation with a Human-centered Study.Beatrice Savoldi, Sara Papi, Matteo Negri, Ana Guerberof Arenas, Luisa Bentivogli
2024ICASSPLeveraging Timestamp Information for Serialized Joint Streaming Recognition and Translation.Sara Papi, Peidong Wang, Jun-Kun Chen, Jian Xue, Naoyuki Kanda, Jinyu Li, Yashesh Gaur
2023ACLAttention as a Guide for Simultaneous Speech Translation.Sara Papi, Matteo Negri, Marco Turchi
2023ASRUToken-Level Serialized Output Training for Joint Streaming ASR and ST Leveraging Textual Alignments.Sara Papi, Peidong Wang, Jun-Kun Chen, Jian Xue, Jinyu Li, Yashesh Gaur
2023EMNLPIntegrating Language Models into Direct Speech Translation: An Inference-Time Solution to Control Gender Inflection.Dennis Fucci, Marco Gaido, Sara Papi, Mauro Cettolo, Matteo Negri, Luisa Bentivogli
2023InterspeechJoint Speech Translation and Named Entity Recognition.Marco Gaido, Sara Papi, Matteo Negri, Marco Turchi
2023InterspeechAlignAtt: Using Attention-based Audio-Translation Alignments as a Guide for Simultaneous Speech Translation.Sara Papi, Marco Turchi, Matteo Negri
2022EMNLPDoes Simultaneous Speech Translation need Simultaneous Models?Sara Papi, Marco Gaido, Matteo Negri, Marco Turchi
2022IJCNLPDodging the Data Bottleneck: Automatic Subtitling with Automatically Segmented ST Corpora.Sara Papi, Alina Karakanta, Matteo Negri, Marco Turchi
2021EMNLPSpeechformer: Reducing Information Loss in Direct Speech Translation.Sara Papi, Marco Gaido, Matteo Negri, Marco Turchi
2020InterspeechMixtures of Deep Neural Experts for Automated Speech Scoring.Sara Papi, Edmondo Trentin, Roberto Gretter, Marco Matassoni, Daniele Falavigna