Skip to content

Interspeech (combined EuroSpeech and ICSLP in 2000)

Interspeech

A

CORE rank

CORE rank (raw)

A

Acceptance rate

48.0% (2024)

Fields of research

Artificial Intelligence

Papers indexed

28,540

1987–2025

Papers per year

19871,180 peak2025

Interspeech papers

28,540 records sourced from DBLP. Search titles, filter by year, sort by recency.

YearTitleAuthors
2025Exploratory Study of Filled Pauses in Ukrainian Language: Phonetic Properties of Filled Pauses.Anna Havras, Carlos Mendes, Helena Moniz, Gueorgui Hristovsky, Joo Miranda
2025Self-Supervised Models of Speech Processing for Haitian Creole.William N. Havard, Renauld Govain, Benjamin Lecouteux, Emmanuel Schang
2025DnR-nonverbal: Cinematic Audio Source Separation DatasetContaining Non-Verbal Sounds.Takuya Hasumi, Yusuke Fujita
2025Reddit FlairShare: A Human-Annotated Dataset of Gender-Progressive Online Discourse.Carlos Hartmann
2025SepVAC: Multitask Learning of Speaker Separation, Speaker Localization, Microphone Array Localization, and Room Acoustic Parameter Estimation in Various Acoustic Conditions.Roland Hartanto, Sakriani Sakti, Koichi Shinoda
2025Variability in performance across four generations of automatic speaker recognition systems.Lauren Harrington, Vincent Hughes, Philip Harrison, Paul Foulkes, Jessica Wormald, Finnian Kelly, David van der Vloed
2025Can ASR generate valid measures of child reading fluency?Wieke Harmsen, Roeland van Hout, Catia Cucchiarini, Helmer Strik
2025PAST: Phonetic-Acoustic Speech Tokenizer.Nadav Har-Tuv, Or Tal, Yossi Adi
2025L3C-DeepMFC: Low-Latency Low-Complexity Deep Marginal Feedback Cancellation with Closed-Loop Fine Tuning for Hearing Aids.Fengyuan Hao, Brian C. J. Moore, Huiyong Zhang, Xiaodong Li, Chengshi Zheng
2025WTFormer: A Wavelet Conformer Network for MIMO Speech Enhancement with Spatial Cues Peservation.Lu Han, Junqi Zhao, Renhua Peng
2025PAEFF: Precise Alignment and Enhanced Gated Feature Fusion for Face-Voice Association.Abdul Hannan, Muhammad Arslan Manzoor, Shah Nawaz, Muhammad Irzam Liaqat, Markus Schedl, Mubashir Noman
2025An Effective Training Framework for Light-Weight Automatic Speech Recognition Models.Abdul Hannan, Alessio Brutti, Shah Nawaz, Mubashir Noman
2025Fine-tune Before Structured Pruning: Towards Compact and Accurate Self-Supervised Models for Speaker Diarization.Jiangyu Han, Federico Landini, Johan Rohdin, Anna Silnova, Mireia Dez, Jan Cernock, Luks Burget
2025Few-step Adversarial Schrdinger Bridge for Generative Speech Enhancement.Seungu Han, Sungho Lee, Juheon Lee, Kyogu Lee
2025CabinSep: IR-Augmented Mask-Based MVDR for Real-Time In-car Speech Separation with Distributed Heterogeneous Arrays.Runduo Han, Yanxin Hu, Yihui Fu, Zihan Zhang, Yukai Jv, Li Chen, Lei Xie
2025Automatic Speech Recognition for Low-Resourced Middle Eastern Languages.Razhan Hameed, Sina Ahmadi, Hanah Hadi, Rico Sennrich
2025Are loan sequences different from foreign sequences? A perception study with Japanese listeners on coronal obstruent - high front vowel sequences.Silke Hamann, Andrea Alicehajic
2025Relationship between objective and subjective perceptual measures of speech in individuals with head and neck cancer.Bence Mark Halpern, Thomas Tienkamp, Teja Rebernik, Rob J. J. H. van Son, Martijn Wieling, Defne Abur, Tomoki Toda
2025Token-Level Logits Matter: A Closer Look at Speech Foundation Models for Ambiguous Emotion Recognition.Jule Valendo Halim, Siyi Wang, Hong Jia, Ting Dang
2025Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech.Karl El Hajal, Enno Hermann, Sevada Hovsepyan, Mathew Magimai-Doss
2025EzAudio: Enhancing Text-to-Audio Generation with Efficient Diffusion Transformer.Jiarui Hai, Yong Xu, Hao Zhang, Chenxing Li, Helin Wang, Mounya Elhilali, Dong Yu
2025Phonetically-Augmented Discriminative Rescoring for Voice Search Error Correction.Christophe Van Gysel, Maggie Wu, Lyan Verwimp, Caglar Tirkaz, Marco Bertola, Zhihong Lei, Youssef Oualil
2025Deep learning based spatial aliasing reduction in beamforming for audio capture.Mateusz Guzik, Giulio Cengarle, Daniel Arteaga
2025Neurodyne: Neural Pitch Manipulation with Representation Learning and Cycle-Consistency GAN.Yicheng Gu, Chaoren Wang, Zhizheng Wu, Lauri Juvela
2025Audio-Based Classification and Geographic Regression of Austrian Dialects.Lorenz Gutscher, Michael Pucher
851875 of 28,540← PreviousNext →

Comparable venues

Other A*/A conferences filed under the same field of research.