Skip to content

IEEE Automatic Speech Recognition and Understanding Workshop

ASRU

C

CORE rank

CORE rank (raw)

C

Fields of research

Artificial Intelligence

Papers indexed

1,323

2007–2025

Papers per year

2007209 peak2025

ASRU papers

1,323 records sourced from DBLP. Search titles, filter by year, sort by recency.

YearTitleAuthors
2021A Study of Transducer Based End-to-End ASR with ESPnet: Architecture, Auxiliary Loss and Decoding Strategies.Florian Boyer, Yusuke Shinohara, Takaaki Ishii, Hirofumi Inaguma, Shinji Watanabe
2021Target Language Extraction at Multilingual Cocktail Parties.Marvin Borsdorf, Haizhou Li, Tanja Schultz
2021Remember the Context! ASR Slot Error Correction Through Memorization.Dhanush Bekal, Ashish Shenoy, Monica Sunkara, Sravan Bodapati, Katrin Kirchhoff
2021Detecting Emotion Carriers by Combining Acoustic and Lexical Representations.Sebastian P. Bayerl, Aniruddha Tammewar, Korbinian Riedhammer, Giuseppe Riccardi
2021Improving Reverberant Speech Separation with Synthetic Room Impulse Responses.Rohith Aralikatti, Anton Ratnarajah, Zhenyu Tang, Dinesh Manocha
2021On-Device Neural Speech Synthesis.Sivanand Achanta, Albert Antony, Ladan Golipour, Jiangchuan Li, Tuomo Raitio, Ramya Rasipuram, Francesco Rossi, Jennifer Shi, Jaimin Upadhyay, David Winarsky, Hepeng Zhang
2019An Investigation into the Effectiveness of Enhancement in ASR Training and Test for Chime-5 Dinner Party Transcription.Catalin Zorila, Christoph Bddeker, Rama Doddipatla, Reinhold Haeb-Umbach
2019Controlling Emotion Strength with Relative Attribute for End-to-End Speech Synthesis.Xiaolian Zhu, Shan Yang, Geng Yang, Lei Xie
2019CNN with Phonetic Attention for Text-Independent Speaker Verification.Tianyan Zhou, Yong Zhao, Jinyu Li, Yifan Gong, Jian Wu
2019A Modularized Neural Network with Language-Specific Output Layers for Cross-Lingual Voice Conversion.Yi Zhou, Xiaohai Tian, Emre Yilmaz, Rohan Kumar Das, Haizhou Li
2019End-to-End Overlapped Speech Detection and Speaker Counting with Raw Waveform.Wangyou Zhang, Man Sun, Lan Wang, Yanmin Qian
2019A Comparison of Transformer and LSTM Encoder Decoder Models for ASR.Albert Zeyer, Parnia Bahar, Kazuki Irie, Ralf Schlter, Hermann Ney
2019A Multi Purpose and Large Scale Speech Corpus in Persian and English for Speaker and Speech Recognition: The Deepmine Database.Hossein Zeinali, Luks Burget, Jan Honza Cernock
2019End-to-End Code-Switching ASR for Low-Resourced Language Pairs.Xianghu Yue, Grandee Lee, Emre Yilmaz, Fang Deng, Haizhou Li
2019Verifying Deep Keyword Spotting Detection with Acoustic Word Embeddings.Yougen Yuan, Zhiqiang Lv, Shen Huang, Lei Xie
2019Advances in Online Audio-Visual Meeting Transcription.Takuya Yoshioka, Yan Huang, Aviv Hurvitz, Li Jiang, Sharon Koubi, Eyal Krupka, Ido Leichter, Changliang Liu, Partha Parthasarathy, Alon Vinnikov, Lingfeng Wu, Igor Abramovski, Xiong Xiao, Wayne Xiong, Huaming Wang, Zhenghao Wang, Jun Zhang, Yong Zhao, Tianyan Zhou, Cem Aksoylar, Zhuo Chen, Moshe David, Dimitrios Dimitriadis, Yifan Gong, Ilya Gurvich, Xuedong Huang
2019Improving Mandarin End-to-End Speech Synthesis by Self-Attention and Learnable Gaussian Bias.Fengyu Yang, Shan Yang, Pengcheng Zhu, Pengju Yan, Lei Xie
2019Time-Domain Speaker Extraction Network.Chenglin Xu, Wei Rao, Eng Siong Chng, Haizhou Li
2019Improving Speech Enhancement with Phonetic Embedding Features.Bo Wu, Meng Yu, Lianwu Chen, Mingjie Jin, Dan Su, Dong Yu
2019Time Domain Audio Visual Speech Separation.Jian Wu, Yong Xu, Shi-Xiong Zhang, Lianwu Chen, Meng Yu, Lei Xie, Dong Yu
2019Joint Learning of Word and Label Embeddings for Sequence Labelling in Spoken Language Understanding.Jiewen Wu, Luis Fernando D'Haro, Nancy F. Chen, Pavitra Krishnaswamy, Rafael E. Banchs
2019Learning Between Different Teacher and Student Models in ASR.Jeremy Heng Meng Wong, Mark J. F. Gales, Yu Wang
2019Zero-Shot Pronunciation Lexicons for Cross-Language Acoustic Model Transfer.Matthew Wiesner, Oliver Adams, David Yarowsky, Jan Trmal, Sanjeev Khudanpur
2019Speech Reveals Future Risk of Developing Dementia: Predictive Dementia Screening from Biographic Interviews.Jochen Weiner, Claudia Frankenberg, Johannes Schrder, Tanja Schultz
2019Leveraging Language ID in Multilingual End-to-End Speech Recognition.Austin Waters, Neeraj Gaur, Parisa Haghani, Pedro J. Moreno, Zhongdi Qu
551575 of 1,323← PreviousNext →

Comparable venues

Other A*/A conferences filed under the same field of research.