Skip to content

Interspeech (combined EuroSpeech and ICSLP in 2000)

Interspeech

A

CORE rank

CORE rank (raw)

A

Acceptance rate

48.0% (2024)

Fields of research

Artificial Intelligence

Papers indexed

28,540

1987–2025

Papers per year

19871,180 peak2025

Interspeech papers

28,540 records sourced from DBLP. Search titles, filter by year, sort by recency.

YearTitleAuthors
2025Analysis of Avian Biphonic Vocalization Using Computational Modelling.Noumida A, Rajeev Rajan
2025On Enhancing the Performance of Children's ASR Task in Limited Data Scenario.Ankita, Shambhavi, Syed Shahnawazuddin
2025Is it all about race?: A Cross-examination of /s/ in a Multilingual (Nigerian) Context.Oluwasegun Amoniyan
2025A Study of Speech Embedding Similarities Between Australian Aboriginal and High-Resource Languages.Eliathamby Ambikairajah, Jingyao Wu, Ting Dang, Vidhyasaharan Sethu
2025TalTech Systems for the Interspeech 2025 ML-SUPERB 2.0 Challenge.Tanel Alume, Artem Fedorchenko
2025On the Language and Gender Biases in PSTN, VoIP and Neural Audio Codecs.Kemal Altwlkany, Amar Kuric, Emanuel Lacic
2025ReSepNet: A Unified-Light Model for Recursive Speech Separation with Unknown Speaker Count.Hadi Alizadeh, Rahil Mahdian Toroghi, Hassan Zareian
2025A Three-Stage Beamforming with Harmonic Guidance for Multi-Channel Speech Enhancement.Nurali Alip, Tianrui Wang, Rui Cao, Meng Ge, Jingru Lin, Longbiao Wang, Jianwu Dang
2025Defending Speech-enabled LLMs Against Adversarial Jailbreak Threats.Antonios Alexos, Raghuveer Peri, Sai Muralidhar Jayanthi, Metehan Cekic, Srikanth Vishnubhotla, Kyu J. Han, Srikanth Ronanki
2025Evaluating ASR Robustness to Spontaneous Speech Errors: A Study of WhisperX Using a Speech Error Database.John Alderete, Macarious Kin Fung Hui, Aanchan Mohan
2025SpokenNativQA: Multilingual Everyday Spoken Queries for LLMs.Firoj Alam, Md. Arid Hasan, Shammur Absar Chowdhury
2025AfriHuBERT: A self-supervised speech representation model for African languages.Jesujoba O. Alabi, Xuechen Liu, Dietrich Klakow, Junichi Yamagishi
2025MiSTR: Multi-Modal iEEG-to-Speech Synthesis with Transformer-Based Prosody Prediction and Neural Phase Reconstruction.Mohammed Salah Al-Radhi, Gza Nmeth, Branislav Gerazov
2025Towards Better Disentanglement in Non-Autoregressive Zero-Shot Expressive Voice Conversion.Seymanur Akti, Tuan-Nam Nguyen, Alexander Waibel
2025WhisperD: Dementia Speech Recognition and Filler Word Detection with Whisper.Emmanuel Akinrintoyo, Nadine Abdelhalim, Nicole Salomons
2025VoxAging: Continuously Tracking Speaker Aging with a Large-Scale Longitudinal Dataset in English and Mandarin.Zhiqi Ai, Meixuan Bao, Zhiyong Chen, Zhi Yang, Xinnuo Li, Shugong Xu
2025HuBERT-VIC: Improving Noise-Robust Automatic Speech Recognition of Speech Foundation Model via Variance-Invariance-Covariance Regularization.Hyebin Ahn, Kangwook Jang, Hoirin Kim
2025Optimizing CLAP Reward with LLM Feedback for Semantically Aligned and Diverse Automated Audio Captioning.Seyun Ahn, Pil Moo Byun, Won-Gook Choi, Joon-Hyuk Chang
2025A Study on Speech Assessment with Visual Cues.Shafique Ahmed, Ryandhimas E. Zezario, Nasir Saleem, Amir Hussain, Hsin-Min Wang, Yu Tsao
2025Continuous Learning for Children's ASR: Overcoming Catastrophic Forgetting with Elastic Weight Consolidation and Synaptic Intelligence.Edem Ahadzi, Vishwanath Pratap Singh, Tomi Kinnunen, Ville Hautamki
2025Spot and Merge: A Hybrid Context Biasing Approach for Rare Word and Out of Vocabulary Recognition.Jatin Agrawal, Bramhendra Koilakuntla, Srikanth Konjeti
2025Spoken Language Understanding on Unseen Tasks With In-Context Learning.Neeraj Agrawal, Sriram Ganapathy
2025Investigating the Reasoning Abilities of Large Language Models for Understanding Spoken Language in Interpersonal Interactions.Pranjal Aggarwal, Ghritachi Mahajani, Pavan Kumar Malasani, Vaibhav Jamadagni, Caroline J. Wendt, Ehsanul Haque Nirjhar, Theodora Chaspari
2025Domain Adaptation Method and Modality Gap Impact in Audio-Text Models for Prototypical Sound Classification.Emiliano Acevedo, Martn Rocamora, Magdalena Fuentes
2025Bridging ASR and LLMs for Dysarthric Speech Recognition: Benchmarking Self-Supervised and Generative Approaches.Ahmed Aboeitta, Ahmed Sharshar, Youssef Nafea, Shady Shehata
1,1261,150 of 28,540← PreviousNext →

Comparable venues

Other A*/A conferences filed under the same field of research.