Skip to content

Interspeech (combined EuroSpeech and ICSLP in 2000)

Interspeech

A

CORE rank

CORE rank (raw)

A

Acceptance rate

48.0% (2024)

Fields of research

Artificial Intelligence

Papers indexed

28,540

1987–2025

Papers per year

19871,180 peak2025

Interspeech papers

28,540 records sourced from DBLP. Search titles, filter by year, sort by recency.

YearTitleAuthors
2025Universal Semantic Disentangled Privacy-preserving Speech Representation Learning.Biel Tura Vecino, Subhadeep Maji, Aravind Varier, Antonio Bonafonte, Ivan Valles, Michael Owen, Constantinos Papayiannis, Leif Rdel, Grant P. Strimel, Oluwaseyi Feyisetan, Roberto Barra-Chicote, Ariya Rastrow, Volker Leutnant, Trevor Wood
2025Legally validated evaluation framework for voice anonymization.Nathalie Vauquier, Brij Mohan Lal Srivastava, Seyed Ahmad Hosseini, Emmanuel Vincent
2025The State Of TTS: A Case Study with Human Fooling Rates.Praveen Srinivasa Varadhan, Sherry Thomas, Sai Teja M. S., Suvrat Bhooshan, Mitesh M. Khapra
2025Clinical Annotations for Automatic Stuttering Severity Assessment.Ana Rita Valente, Rufael Marew, Hawau Olamide Toyin, Hamdan Al-Ali, Anelise Bohnen, Inma Becerra, Elsa Marta Soares, Gonalo Leal, Hanan Aldarmaki
2025Self-supervised learning of speech representations with Dutch archival data.Nik Vaessen, Roeland Ordelman, David A. van Leeuwen
2025Thai Speech Spoofing Detection Dataset with Variations in Speaking Styles.Ticho Urai, Pachara Boonsarngsuk, Ekapol Chuangsuwanich
2025From Pretraining to Performance: Benchmarking Self-Supervised Speech Models for Interspeech-25 SER Challenge.Drishya Uniyal, Vinayak Abrol
2025Weight Factorization and Centralization for Continual Learning in Speech Recognition.Enes Yavuz Ugan, Ngoc-Quan Pham, Alexander Waibel
2025Zero-Shot Learning for Acoustic Event Classification Using an Attribute Vector and Conditional GAN.Kohei Uehara, Ryoichi Takashima, Tetsuya Takiguchi
2025Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model.Lucas H. Ueda, Joo Lima, Leonardo Marques, Paula Dornhofer Paro Costa
2025Lessons Learnt: Revisit Key Training Strategies for Effective Speech Emotion Recognition in the Wild.Jing-Tong Tzeng, Bo-Hao Su, Ya-Tse Wu, Hsing-Hang Chou, Chi-Chun Lee
2025Does English fish sound like French fiche? Perceptual similarity judgments versus acoustic similarity.Rory Turnbull, Elisa Kiefer, Sharon Peperkamp
2025Representing Speech Through Autoregressive Prediction of Cochlear Tokens.Greta Tuckute, Klemen Kotar, Evelina Fedorenko, Daniel Yamins
2025Probing the Robustness Properties of Neural Speech Codecs.Wei-Cheng Tseng, David Harwath
2025Attention Is Not Always the Answer: Optimizing Voice Activity Detection with Simple Feature Fusion.Kumud Tripathi, Chowdam Venkata Kumar, Pankaj Wasnik
2025R2S: Real-to-Synthetic Representation Learning for Training Speech Recognition Models on Synthetic Data.Minh Tran, Debjyoti Paul, Yutong Pang, Laxmi Pandey, Jinxi Guo, Ke Li, Shun Zhang, Xuedong Zhang, Xin Lei
2025Leveraging SSL Speech Features and Mamba for Enhanced DeepFake Detection.Hoan My Tran, Damien Lolive, David Guennec, Aghilas Sini, Arnaud Delhay, Pierre-Franois Marteau
2025Amplifying Artifacts with Speech Enhancement in Voice Anti-spoofing.Thanapat Trachu, Thanathai Lertpetchpun, Ekapol Chuangsuwanich
2025ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis.Hawau Olamide Toyin, Rufael Marew, Humaid Alblooshi, Samar M. Magdy, Hanan Aldarmaki
2025Exploiting Context-dependent Duration Features for Voice Anonymization Attack Systems.Natalia A. Tomashenko, Emmanuel Vincent, Marc Tommasi
2025Improving Respiratory Sound Classification with Architecture-Agnostic Knowledge Distillation from Ensembles.Miika Toikkanen, June-Woo Kim
2025A simple method for predicting Clinical Scores in Huntington's Disease by leveraging ASR's uncertainty on spontaneous speech.Hadrien Titeux, Quang Tuan Rmy Nguyen, Andres Gil-Salcedo, Anne-Catherine Bachoud-Lvi, Emmanuel Dupoux
2025Accessible Real-time Eye-gaze Tracking for Neurocognitive Health Assessment: A Multimodal Web-based Approach.Daniel Tisdale, Jackson Liscombe, David Pautler, Michael Neumann, Vikram Ramanarayanan
2025Articulatory clarity and variability before and after surgery for tongue cancer.Thomas Tienkamp, Fleur van Ast, Roos van der Veen, Teja Rebernik, Raoul Buurke, Nikki Hoekzema, Katharina Polsterer, Hedwig Sekeres, Rob van Son, Martijn Wieling, Max J. H. Witjes, Sebastiaan A. H. J. de Visscher, Defne Abur
2025Discrete Audio Representations for Automated Audio Captioning.Jingguang Tian, Haoqin Sun, Xinhui Hu, Xinkang Xu
201225 of 28,540← PreviousNext →

Comparable venues

Other A*/A conferences filed under the same field of research.