IEEE Automatic Speech Recognition and Understanding Workshop
ASRU
C
CORE rank
CORE rank (raw)
C
Fields of research
Artificial Intelligence
Papers indexed
1,323
2007–2025
Papers per year
2007209 peak2025
Most published authors
ASRU papers
1,323 records sourced from DBLP. Search titles, filter by year, sort by recency.
| Year | Title | Authors |
|---|---|---|
| 2025 | AdaBit-TasNet: Speech Separation with Inference Adaptable Precision. | Mohamed Elminshawi, Srikanth Raj Chetupalli, Emanul A. P. Habets |
| 2025 | Acoustic Phonetic Temporal Speech Representation. | Yunbin Deng |
| 2025 | KAN-AST: Kolmogorov-Arnold Network based Audio Spectrogram Transformer for Audio Classification. | Phuong Tuan Dat, Tran Huy Dat |
| 2025 | Efficient Speech Watermarking for Speech Synthesis via Progressive Knowledge Distillation. | Yang Cui, Peter Pan, Lei He, Sheng Zhao |
| 2025 | Sinba: Singing-To-Accompaniment Generation With Pitch Guidance Via Mamba-Based Language Model. | Jianwei Cui, Shihao Chen, Yu Gu, Jie Zhang, Liping Chen, Na Li, Chengxing Li, Shan Yang, Li-Rong Dai |
| 2025 | Layer-wise Analysis for Quality of Multilingual Synthesized Speech. | Erica Cooper, Takuma Okamoto, Yamato Ohtani, Tomoki Toda, Hisashi Kawai |
| 2025 | Flow-SLM: Joint Learning of Linguistic and Acoustic Information for Spoken Language Modeling. | Ju-Chieh Chou, Jiawei Zhou, Karen Livescu |
| 2025 | A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data. | Cheng-Kang Chou, Chan-Jan Hsu, Ho-Lam Chung, Liang-Hsuan Tseng, Hsi-Chun Cheng, Yu-Kuan Fu, Kuan-Po Huang, Hung-Yi Lee |
| 2025 | CAVIARES: Corpus for Audio-Visual Expressive Voice Agent. | Jinsheng Chen, Yuki Saito, Dong Yang, Naoko Tanji, Hironori Doi, Byeongseon Park, Yuma Shirahata, Kentaro Tachibana, Hiroshi Saruwatari |
| 2025 | RE-LLM: Refining Empathetic Speech-LLM Responses by Integrating Emotion Nuance. | Jing-Han Chen, Bo-Hao Su, Ya-Tse Wu, Chi-Chun Lee |
| 2025 | Time-Frequency-Based Attention Cache Memory Model for Real-Time Speech Separation. | Guo Chen, Kai Li, Runxuan Yang, Xiaolin Hu |
| 2025 | DiffRhythm+: Controllable and Flexible Full-Length Song Generation with Preference Optimization. | Huakang Chen, Yuepeng Jiang, Guobin Ma, Chunbo Hao, Shuai Wang, Jixun Yao, Ziqian Ning, Meng Meng, Jian Luan, Lei Xie |
| 2025 | Confidence-Based Self-Training for EMG-to-Speech: Leveraging Synthetic EMG for Robust Modeling. | Xiaodan Chen, Xiaoxue Gao, Mathias Quoy, Alexandre Pitti, Nancy F. Chen |
| 2025 | Qieemo: Multimodal Emotion Recognition Based on the ASR Backbone. | Jinming Chen, Jingyi Fang, Yuanzhong Zheng, Yaoxuan Wang, Haojun Fei |
| 2025 | LLM-Based Dictation Detection from Doctor-Patient Conversations. | Siyuan Chen, Mojtaba Kadkhodaie Elyaderani, Jing Su, Susanne Burger, Thomas Schaaf |
| 2025 | MEAN-RIR: Multi-Modal Environment-Aware Network for Robust Room Impulse Response Estimation. | Jiajian Chen, Jiakang Chen, Hang Chen, Qing Wang, Yu Gao, Jun Du |
| 2025 | From Simulation to Strategy: Automating Personalized Interaction Planning for Conversational Agents. | Wen-Yu Chang, Tzu-Hung Huang, Chih-Ho Chen, Yun-Nung Chen |
| 2025 | USAD: Universal Speech and Audio Representation via Distillation. | Heng-Jui Chang, Saurabhchand Bhati, James R. Glass, Alexander H. Liu |
| 2025 | Emphasis Sensitivity in Speech Representations. | Shaun Cassini, Thomas Hain, Anton Ragni |
| 2025 | Open Full-duplex Voice Agent with Speech-to-Speech Language Model. | Edresson Casanova, Chen Chen, Kevin Hu, Ankita Pasad, Elena Rastorgueva, Seelan Lakshmi Narasimhan, Slyne Deng, Ehsan Hosseini-Asl, Piotr Zelasko, Valentin Mendelev, Subhankar Ghosh, Yifan Peng, Zhehuai Chen, Jason Li, Jagadeesh Balam, Vitaly Lavrukhin, Boris Ginsburg |
| 2025 | CAMES: A Comprehensive Automatic Speech Recognition Benchmark for European Portuguese. | Carlos Carvalho, Francisco Teixeira, Catarina Botelho, Anna Pompili, Rubn Solera-Urea, Srgio Paulo, Mariana Julio, Thomas Rolland, John Mendona, Diogo A. P. Nunes, Isabel Trancoso, Alberto Abad |
| 2025 | Adaptive Audio-Visual Speech Recognition via Matryoshka-Based Multimodal LLMs. | Umberto Cappellazzo, Minsu Kim, Stavros Petridis |
| 2025 | GenVC: Self-Supervised Zero-Shot Voice Conversion. | Zexin Cai, Henry Li Xinyuan, Ashi Garg, Leibny Paola Garca-Perera, Kevin Duh, Sanjeev Khudanpur, Matthew Wiesner, Nicholas Andrews |
| 2025 | mSTEB: Massively Multilingual Evaluation of LLMs on Speech and Text Tasks. | Luel Hagos Beyene, Vivek Verma, Min Ma, Jesujoba O. Alabi, Fabian David Schmidt, Joyce Nakatumba-Nabende, David Ifeoluwa Adelani |
| 2025 | Text-Guided Speech Representations for Language Acquisition Assessment. | Ilja Baumann, Dominik Wagner, Philipp Seeberger, Korbinian Riedhammer, Tobias Bocklet |
176–200 of 1,323← PreviousNext →
Comparable venues
Other A*/A conferences filed under the same field of research.
- A*AAAINational Conference of the American Association for Artificial Intelligence
- A*ICRAIEEE International Conference on Robotics and Automation
- AInterspeechInterspeech (combined EuroSpeech and ICSLP in 2000)
- AIROSIEEE/RSJ International Conference on Intelligent Robots and Systems
- A*ACLAssociation for Computational Linguistics
- A*IJCAIInternational Joint Conference on Artificial Intelligence
- A*EMNLPEmpirical Methods in Natural Language Processing
- AGECCOGenetic and Evolutionary Computations