| 2025 | Prolongation in Romanian. | Oana Niculescu, Monica Vasileanu |
| 2025 | Can we train ASR systems on Code-switch without real code-switch data? Case study for Singapore's languages. | Tuan Nguyen, Huy Dat Tran |
| 2025 | Cocktail-Party Audio-Visual Speech Recognition. | Thai-Binh Nguyen, Ngoc-Quan Pham, Alexander Waibel |
| 2025 | Streaming Non-Autoregressive Model for Accent Conversion and Pronunciation Improvement. | Tuan-Nam Nguyen, Ngoc-Quan Pham, Seymanur Akti, Alexander Waibel |
| 2025 | ViCocktail: Automated Multi-Modal Data Collection for Vietnamese Audio-Visual Speech Recognition. | Thai-Binh Nguyen, Thi Van Nguyen, Quoc Truong Do, Chi Mai Luong |
| 2025 | Turing's Echo: Investigating Linguistic Sensitivity of Deepfake Voice Detection via Gamification. | Binh Nguyen, Thai Le |
| 2025 | Thinking Fast and Slow: Robust Speech Recognition via Deep Filter-Tuning. | Dianwen Ng, Kun Zhou, Bin Ma, Eng Siong Chng |
| 2025 | Multimodal Speech, Language and Orofacial Analysis for Remote Assessment of Positive, Negative and Cognitive Symptoms in Schizophrenia. | Michael Neumann, Hardik Kothare, Beverly Insel, Anzalee Khan, Danyah Nadim, Jean-Pierre Lindenmayer, Vikram Ramanarayanan |
| 2025 | On the Production and Perception of a Single Speaker's Gender. | Robin Netzorg, Naomi Carvalho, Andrea Guzman, Lydia Wang, Juliana Francis, Klo Vivienne Garoute, Keith Johnson, Gopala Anumanchipalli |
| 2025 | Scalable Offline ASR for Command-Style Dictation in Courtrooms. | Kumarmanas Nethil, Vaibhav Mishra, Kriti Anandan, Kavya Manohar |
| 2025 | Source Verification for Speech Deepfakes. | Viola Negroni, Davide Salvi, Paolo Bestagini, Stefano Tubaro |
| 2025 | FlowTSE: Target Speaker Extraction with Flow Matching. | Aviv Navon, Aviv Shamsian, Yael Segal-Feldman, Neta Glazer, Gil Hetz, Joseph Keshet |
| 2025 | Developing High-Quality TTS for Punjabi and Urdu: Benchmarking against MMS Models. | Fatima Naseem, Maham Sajid, Farah Adeeba, Sahar Rauf, Asad Mustafa, Sarmad Hussain |
| 2025 | Voice Quality Dimensions as Interpretable Primitives for Speaking Style for Atypical Speech and Affect. | Jaya Narain, Vasudha Kowtha, Colin Lea, Lauren Tooley, Dianna Yee, Vikramjit Mitra, Zifang Huang, Miquel Espi Marques, Jon Huang, Carlos Avendao, Shirley Ren |
| 2025 | Parameter-efficient Fine-tuning of Conformer-based Streaming Speech Recognition into Non-streaming Models. | Yunjae Nam, Jeong U. Han, Kiyeon Kim, Jaemin Lim |
| 2025 | SEED: Speaker Embedding Enhancement Diffusion Model. | Kihyun Nam, Jungwoo Heo, Jee-weon Jung, Gangin Park, Chaeyoung Jung, Ha-Jin Yu, Joon Son Chung |
| 2025 | WCTC-Biasing: Retraining-free Contextual Biasing ASR with Wildcard CTC-based Keyword Spotting and Inter-layer Biasing. | Yu Nakagome, Michael Hentschel |
| 2025 | The Interspeech 2025 Challenge on Speech Emotion Recognition in Naturalistic Conditions. | Abinay Reddy Naini, Lucas Goncalves, Ali N. Salman, Pravin Mote, Ismail Rasim Ulgen, Thomas Thebaud, Laureano Moro-Velzquez, Leibny Paola Garca, Najim Dehak, Berrak Sisman, Carlos Busso |
| 2025 | Eigenvoice Synthesis based on Model Editing for Speaker Generation. | Masato Murata, Koichi Miyazaki, Tomoki Koriyama, Tomoki Toda |
| 2025 | Speaker-agnostic Emotion Vector for Cross-speaker Emotion Intensity Control. | Masato Murata, Koichi Miyazaki, Tomoki Koriyama |
| 2025 | A Cascaded Multimodal Framework for Automatic Social Communication Severity Assessment in Children with Autism Spectrum Disorder. | Jihyun Mun, Sunhee Kim, Minhwa Chung |
| 2025 | Boundary-Conscious Pruning: Hard Set-Aware Model Compression for Efficient Speaker Recognition. | Seongkyu Mun, Jubum Han |
| 2025 | Speech-Based Automatic Chronic Kidney Disease Diagnosis via Transformer Fusion of Glottal and Spectrogram Features. | Jihyun Mun, Minhwa Chung, Sunhee Kim |
| 2025 | Replay Attacks Against Audio Deepfake Detection. | Nicolas M. Mller, Piotr Kawa, Wei Herng Choong, Adriana Stan, Aditya Tirumala Bukkapatnam, Karla Pizzi, Alexander Wagner, Philip Sperl |
| 2025 | Fine-Tuning ASR for Stuttered Speech: Personalized vs. Generalized Approaches. | Dena F. Mujtaba, Nihar R. Mahapatra |