| 2025 | Using gender, phonation and age to interpret automatically discovered speech attributes for explainable speaker recognition. | Carole Millot, Clara Ponchard, Cdric Gendrot, Jean-Franois Bonastre, Orane Dufour |
| 2025 | Fairness in Dysarthric Speech Synthesis: Understanding Intrinsic Bias in Dysarthric Speech Cloning using F5-TTS. | Anuprabha M, Krishna Gurugubelli, Anil Kumar Vuppala |
| 2025 | First Steps Towards Voice Anonymization for Code-Switching Speech. | Sarina Meyer, Ekaterina Kolos, Ngoc Thang Vu |
| 2025 | Concurrent Speech and Auditory Tag Clouds for Non-Visual Web Interaction. | Dhia Eddine Merzougui, Nilesh Tete, Fabrice Maurel, Gal Dias, Mohammed Hasanuzzaman, Aurlien Bournonville, Edgar Madelaine, Thomas Berthelin Le Tellier, Franois Ledoyen, Laure Poutrain-Lejeune, Franois Rioult, Jrmie Pantin |
| 2025 | LASPA: Language Agnostic Speaker Disentanglement with Prefix-Tuned Cross-Attention. | Aditya Srinivas Menon, Raj Prakash Gohil, Kumud Tripathi, Pankaj Wasnik |
| 2025 | Effective Context in Neural Speech Models. | Yen Meng, Sharon Goldwater, Hao Tang |
| 2025 | Federated Learning with Feature Space Separation for Speaker Recognition. | Ying Meng, Zhihua Fang, Liang He |
| 2025 | Multimodal Silent Recognition of Phonemes Using Radar and Optopalatographic Silent Speech Interfaces. | Joo Menezes, Aubin Mouras, Arne-Lukas Fietkau, Dani Kazzy, Peter Birkholz |
| 2025 | Optimizing ASR for Catalan-Spanish Code-Switching: A Comparative Analysis of Methodologies. | Carlos Mena, Pol Serra, Jacobo Romero, Abir Messaoudi, Jos Giraldo, Carme Armentano-Oller, Rodolfo Zevallos, Ivn Meza, Javier Hernando |
| 2025 | Leveraging Geographic Metadata for Dialect-Aware Speech Recognition. | Pouya Mehralian, Hugo Van hamme |
| 2025 | Streaming Sortformer: Speaker Cache-Based Online Speaker Diarization with Arrival-Time Ordering. | Ivan Medennikov, Taejin Park, Weiqing Wang, He Huang, Kunal Dhawan, Jinhan Wang, Jagadeesh Balam, Boris Ginsburg |
| 2025 | Supralaryngeal Kinematics of Implosives in Central Vietnamese: An EMA Study. | Paul McGuire, Kye Shibata, Thanh Viet Cao, Feng-fan Hsieh, Yueh-Chin Chang |
| 2025 | Training Articulatory Inversion Models for Interspeaker Consistency. | Charles McGhee, Mark J. F. Gales, Kate M. Knill |
| 2025 | Modeling Vowel System Typology Using Iterated Confusion Minimization. | John McGahay |
| 2025 | Accessible Delivery of Visual-Acoustic Biofeedback for Speech Sound Disorder. | Tara McAllister, Peter Traver, Amanda Eads, William Haack, Helen Carey, Yi Shan, Wendy Liang, Tae Hong Park |
| 2025 | Web-Based Application for Real-Time Biofeedback of Vocal Resonance in Gender-Affirming Voice Training: Design and Usability Evaluation. | Tara McAllister, Collin Eagen, Yi Shan, Peter Traver, Daphna Harel, Tae Hong Park, Vesna D. Novak |
| 2025 | Real-Time Audio-Visual Speech Enhancement Using Pre-trained Visual Representations. | Teng Aleksandra Ma, Sile Yin, Li-Chia Yang, Shuo Zhang |
| 2025 | Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis. | Paul Mayer, Florian Lux, Alejandro Prez Gonzlez de Martos, Angelina Elizarova, Lindsey Vanderlyn, Dirk Vth, Ngoc Thang Vu |
| 2025 | Clustering-based Hard Negative Sampling for Supervised Contrastive Speaker Verification. | Piotr Masztalski, Michal Romaniuk, Jakub Zak, Mateusz Matuszewski, Konrad Kowalczyk |
| 2025 | Identification of Pathological Pronunciation Profiles in ASR Transcription Errors. | Margot Masson, Isabelle Ferran, Julie Mauclair |
| 2025 | Automatic detection of speech sound disorders in German-speaking children: augmenting the data with typically developed speech. | Darline Monika Marx, Marco Matassoni, Alessio Brutti |
| 2025 | Do you read me? - flow of speech effect on speaker recognition systems. | Alicja Martinek, Joanna Gajewska, Ewelina Bartuzi-Trokielewicz |
| 2025 | Network of acoustic characteristics for the automatic detection of suicide risk from speech. Contribution to the 2025 SpeechWellness challenge by the Semawave team. | Vincent P. Martin, Charles Brazier, Maxime Amblard, Michel Musiol, Jean-Luc Rouas |
| 2025 | Building an Accurate Open-Source Hebrew ASR System through Crowdsourcing. | Yanir Marmor, Yair Lifshitz, Yoad Snapir, Kinneret Misgav |
| 2025 | Multi-Modal Multi-Task Affective States Recognition Based on Label Encoder Fusion. | Maxim Markitantov, Elena Ryumina, Heysem Kaya, Alexey Karpov |