Skip to content

Interspeech (combined EuroSpeech and ICSLP in 2000)

Interspeech

A

CORE rank

CORE rank (raw)

A

Acceptance rate

48.0% (2024)

Fields of research

Artificial Intelligence

Papers indexed

28,540

1987–2025

Papers per year

19871,180 peak2025

Interspeech papers

28,540 records sourced from DBLP. Search titles, filter by year, sort by recency.

YearTitleAuthors
2025GLCLAP: A Novel Contrastive Learning Pre-trained Model for Contextual Biasing in ASR.Yuxiang Kong, Fan Cui, Liyong Guo, Heinrich Dinkel, Lichun Fan, Junbo Zhang, Jian Luan
2025JIS: A Speech Corpus of Japanese Idol Speakers with Various Speaking Styles.Yuto Kondo, Hirokazu Kameoka, Kou Tanaka, Takuhiro Kaneko
2025Can Multimodal Foundation Models Help Analyze Child-Inclusive Autism Diagnostic Videos?Aditya Kommineni, Digbalay Bose, Tiantian Feng, So Hyun Kim, Helen Tager-Flusberg, Somer Bishop, Catherine Lord, Sudarsana Kadiri, Shrikanth Narayanan
2025Towards Classification of Typical and Atypical Disfluencies: A Self Supervised Representation Approach.Priyanka Kommagouni, Pragya Khanna, Vamshiraghusimha Narasinga, Anirudh Bocha, Anil Kumar Vuppala
2025Leveraging Unlabeled Audio for Audio-Text Contrastive Learning via Audio-Composed Text Features.Tatsuya Komatsu, Hokuto Munakata, Yuchi Ishikawa
2025Granary: Speech Recognition and Translation Dataset in 25 European Languages.Nithin Rao Koluguri, Monica Sekoyan, George Zelenfroynd, Sasha Meister, Shuoyang Ding, Sofia Kostandian, He Huang, Nikolay Karpov, Jagadeesh Balam, Vitaly Lavrukhin, Yifan Peng, Sara Papi, Marco Gaido, Alessio Brutti, Boris Ginsburg
2025Improving Generalization of End-to-End ASR through Diversity and Independence Regularization.Ye-Eun Ko, Mun-Hak Lee, Dong-Hyun Kim, Joon-Hyuk Chang
2025What do self-supervised speech models know about Dutch? Analyzing advantages of language-specific pre-training.Marianne de Heer Kloots, Hosein Mohebbi, Charlotte Pouw, Gaofei Shen, Willem H. Zuidema, Martijn Bentum
2025A Practitioner's Guide to Building ASR Models for Low-Resource Languages: A Case Study on Scottish Gaelic.Ondrej Klejch, William Lamb, Peter Bell
2025Open-Set Source Tracing of Audio Deepfake Systems.Nicholas Klein, Hemlata Tak, Elie Khoury
2025Who knows best? Effects of speech disfluencies on incentivized decision-making.Ambika Kirkland, Jens Edlund
2025FairASR: Fair Audio Contrastive Learning for Automatic Speech Recognition.Jongsuk Kim, Jaemyung Yu, Minchan Kwon, Junmo Kim
2025Spatially Weighted Contrastive Learning for Robust Sound Source Localization.Hyun-Soo Kim, Da-Hee Yang, Joon-Hyuk Chang
2025SpeechMLC: Speech Multi-label Classification.Miseul Kim, Seyun Um, Hyeonjin Cha, Hong-Goo Kang
2025Stack Less, Repeat More: A Block Reusing Approach for Progressive Speech Enhancement.Jangyeon Kim, Ui-Hyeop Shin, Jaehyun Ko, Hyung-Min Park
2025Enhancing Audio Deepfake Detection by Improving Representation Similarity of Bonafide Speech.Seung-bin Kim, Hyun-seo Shin, Jungwoo Heo, Chan-yeong Lim, Kyo-Won Koo, Jisoo Son, Sanghyun Hong, Souhwan Jung, Ha-Jin Yu
2025Rethinking Leveraging Pre-Trained Multi-Layer Representations for Speaker Verification.Jin Sob Kim, Hyun Joon Park, Wooseok Shin, Sung Won Han
2025Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries.Minyoung Kim, Sehwan Park, Sungmin Cha, Paul Hongsuck Seo
2025Language-Agnostic Suicidal Risk Detection Using Large Language Models.June-Woo Kim, Wonkyo Oh, Haram Yoon, Sung-Hoon Yoon, Dae-Jin Kim, Dong-Ho Lee, Sang-Yeol Lee, Chan-Mo Yang
2025Towards an Ultra-Low-Delay Neural Audio Coding with Computational Efficiency.Byeong Hyeon Kim, Hyungseob Lim, Inseon Jang, Hong-Goo Kang
2025Quadruple Path Modeling with Latent Feature Transfer for Permutation-free Continuous Speech Separation.Jihyun Kim, Doyeon Kim, Hyewon Han, Jinyoung Lee, Jonguk Yoo, Chang Woo Han, Jeongook Song, Hoon-Young Cho, Hong-Goo Kang
2025Naturalness-Aware Curriculum Learning with Dynamic Temperature for Speech Deepfake Detection.Taewoo Kim, Guisik Kim, Choongsang Cho, Young Han Lee
2025Mamba-based Hybrid Model for Speech Enhancement.Se-Ha Kim, Tae-Gyeong Kim, Chang-Jae Chun
2025Learning Phonetic Context-Dependent Viseme for Enhancing Speech-Driven 3D Facial Animation.Hyung Kyu Kim, Hak Gu Kim
2025Towards Human-like Multimodal Conversational Agent by Generating Engaging Speech.Taesoo Kim, Yongsik Jo, Hyunmin Song, Taehwan Kim
701725 of 28,540← PreviousNext →

Comparable venues

Other A*/A conferences filed under the same field of research.