| 2025 | Universal Semantic Disentangled Privacy-preserving Speech Representation Learning. | Biel Tura Vecino, Subhadeep Maji, Aravind Varier, Antonio Bonafonte, Ivan Valles, Michael Owen, Constantinos Papayiannis, Leif Rdel, Grant P. Strimel, Oluwaseyi Feyisetan, Roberto Barra-Chicote, Ariya Rastrow, Volker Leutnant, Trevor Wood |
| 2025 | Legally validated evaluation framework for voice anonymization. | Nathalie Vauquier, Brij Mohan Lal Srivastava, Seyed Ahmad Hosseini, Emmanuel Vincent |
| 2025 | The State Of TTS: A Case Study with Human Fooling Rates. | Praveen Srinivasa Varadhan, Sherry Thomas, Sai Teja M. S., Suvrat Bhooshan, Mitesh M. Khapra |
| 2025 | Clinical Annotations for Automatic Stuttering Severity Assessment. | Ana Rita Valente, Rufael Marew, Hawau Olamide Toyin, Hamdan Al-Ali, Anelise Bohnen, Inma Becerra, Elsa Marta Soares, Gonalo Leal, Hanan Aldarmaki |
| 2025 | Self-supervised learning of speech representations with Dutch archival data. | Nik Vaessen, Roeland Ordelman, David A. van Leeuwen |
| 2025 | Thai Speech Spoofing Detection Dataset with Variations in Speaking Styles. | Ticho Urai, Pachara Boonsarngsuk, Ekapol Chuangsuwanich |
| 2025 | From Pretraining to Performance: Benchmarking Self-Supervised Speech Models for Interspeech-25 SER Challenge. | Drishya Uniyal, Vinayak Abrol |
| 2025 | Weight Factorization and Centralization for Continual Learning in Speech Recognition. | Enes Yavuz Ugan, Ngoc-Quan Pham, Alexander Waibel |
| 2025 | Zero-Shot Learning for Acoustic Event Classification Using an Attribute Vector and Conditional GAN. | Kohei Uehara, Ryoichi Takashima, Tetsuya Takiguchi |
| 2025 | Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model. | Lucas H. Ueda, Joo Lima, Leonardo Marques, Paula Dornhofer Paro Costa |
| 2025 | Lessons Learnt: Revisit Key Training Strategies for Effective Speech Emotion Recognition in the Wild. | Jing-Tong Tzeng, Bo-Hao Su, Ya-Tse Wu, Hsing-Hang Chou, Chi-Chun Lee |
| 2025 | Does English fish sound like French fiche? Perceptual similarity judgments versus acoustic similarity. | Rory Turnbull, Elisa Kiefer, Sharon Peperkamp |
| 2025 | Representing Speech Through Autoregressive Prediction of Cochlear Tokens. | Greta Tuckute, Klemen Kotar, Evelina Fedorenko, Daniel Yamins |
| 2025 | Probing the Robustness Properties of Neural Speech Codecs. | Wei-Cheng Tseng, David Harwath |
| 2025 | Attention Is Not Always the Answer: Optimizing Voice Activity Detection with Simple Feature Fusion. | Kumud Tripathi, Chowdam Venkata Kumar, Pankaj Wasnik |
| 2025 | R2S: Real-to-Synthetic Representation Learning for Training Speech Recognition Models on Synthetic Data. | Minh Tran, Debjyoti Paul, Yutong Pang, Laxmi Pandey, Jinxi Guo, Ke Li, Shun Zhang, Xuedong Zhang, Xin Lei |
| 2025 | Leveraging SSL Speech Features and Mamba for Enhanced DeepFake Detection. | Hoan My Tran, Damien Lolive, David Guennec, Aghilas Sini, Arnaud Delhay, Pierre-Franois Marteau |
| 2025 | Amplifying Artifacts with Speech Enhancement in Voice Anti-spoofing. | Thanapat Trachu, Thanathai Lertpetchpun, Ekapol Chuangsuwanich |
| 2025 | ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis. | Hawau Olamide Toyin, Rufael Marew, Humaid Alblooshi, Samar M. Magdy, Hanan Aldarmaki |
| 2025 | Exploiting Context-dependent Duration Features for Voice Anonymization Attack Systems. | Natalia A. Tomashenko, Emmanuel Vincent, Marc Tommasi |
| 2025 | Improving Respiratory Sound Classification with Architecture-Agnostic Knowledge Distillation from Ensembles. | Miika Toikkanen, June-Woo Kim |
| 2025 | A simple method for predicting Clinical Scores in Huntington's Disease by leveraging ASR's uncertainty on spontaneous speech. | Hadrien Titeux, Quang Tuan Rmy Nguyen, Andres Gil-Salcedo, Anne-Catherine Bachoud-Lvi, Emmanuel Dupoux |
| 2025 | Accessible Real-time Eye-gaze Tracking for Neurocognitive Health Assessment: A Multimodal Web-based Approach. | Daniel Tisdale, Jackson Liscombe, David Pautler, Michael Neumann, Vikram Ramanarayanan |
| 2025 | Articulatory clarity and variability before and after surgery for tongue cancer. | Thomas Tienkamp, Fleur van Ast, Roos van der Veen, Teja Rebernik, Raoul Buurke, Nikki Hoekzema, Katharina Polsterer, Hedwig Sekeres, Rob van Son, Martijn Wieling, Max J. H. Witjes, Sebastiaan A. H. J. de Visscher, Defne Abur |
| 2025 | Discrete Audio Representations for Automated Audio Captioning. | Jingguang Tian, Haoqin Sun, Xinhui Hu, Xinkang Xu |