| 2024 | EACL | Zero-Shot End-to-End Spoken Language Understanding via Cross-Modal Selective Self-Training. | Jianfeng He, Julian Salazar, Kaisheng Yao, Haoqi Li, Jason Cai |
| 2023 | EMNLP | Enhancing Abstractiveness of Summarization Models through Calibrated Distillation. | Hwanjun Song, Igor Shalyminov, Hang Su, Siffi Singh, Kaisheng Yao, Saab Mansour |
| 2022 | ECCV | Switch-BERT: Learning to Model Multimodal Interactions by Switching Attention and Input. | Qingpei Guo, Kaisheng Yao, Wei Chu |
| 2021 | AAAI | Interpretable NLG for Task-oriented Dialogue Systems with Heterogeneous Rendering Machines. | Yangming Li, Kaisheng Yao |
| 2021 | ACL | Rewriter-Evaluator Architecture for Neural Machine Translation. | Yangming Li, Kaisheng Yao |
| 2021 | Interspeech | AntVoice Neural Speaker Embedding System for FFSVC 2020. | Zhiming Wang, Furong Xu, Kaisheng Yao, Yuan Cheng, Tao Xiong, Huijia Zhu |
| 2021 | NAACL | Neural Sequence Segmentation as Determining the Leftmost Segments. | Yangming Li, Lemao Liu, Kaisheng Yao |
| 2020 | AAAI | Span-Based Neural Buffer: Towards Efficient and Effective Utilization of Long-Distance Context for Neural Sequence Models. | Yangming Li, Kaisheng Yao, Libo Qin, Shuang Peng, Yijia Liu, Xiaolong Li |
| 2020 | ACL | Handling Rare Entities for Neural Sequence Labeling. | Yangming Li, Han Li, Kaisheng Yao, Xiaolong Li |
| 2020 | ACL | Slot-consistent NLG for Task-oriented Dialogue Systems with Iterative Rectification Network. | Yangming Li, Kaisheng Yao, Libo Qin, Wanxiang Che, Xiaolong Li, Ting Liu |
| 2020 | ICASSP | Multi-Resolution Multi-Head Attention in Deep Speaker Embedding. | Zhiming Wang, Kaisheng Yao, Xiaolong Li, Shuo Fang |
| 2019 | ASRU | Joint Optimization of Classification and Clustering for Deep Speaker Embedding. | Zhiming Wang, Kaisheng Yao, Shuo Fang, Xiaolong Li |
| 2019 | ICNC | Tag2Vec: Tag Embedding for Top-N Recommendation. | Ming He, Kaisheng Yao, Peng Yang, Yuan Yao |
| 2018 | WSDM | Robust Transfer Learning for Cross-domain Collaborative Filtering Using Multiple Rating Patterns Approximation. | Ming He, Jiuling Zhang, Peng Yang, Kaisheng Yao |
| 2016 | ICASSP | Highway long short-term memory RNNS for distant speech recognition. | Yu Zhang, Guoguo Chen, Dong Yu, Kaisheng Yao, Sanjeev Khudanpur, James R. Glass |
| 2016 | NAACL | Incorporating Structural Alignment Biases into an Attentional Neural Translation Model. | Trevor Cohn, Cong Duy Vu Hoang, Ekaterina Vymolova, Kaisheng Yao, Chris Dyer, Gholamreza Haffari |
| 2016 | NAACL | Recurrent Support Vector Machines For Slot Tagging In Spoken Language Understanding. | Yangyang Shi, Kaisheng Yao, Hu Chen, Dong Yu, Yi-Cheng Pan, Mei-Yuh Hwang |
| 2016 | NAACL | Deep LSTM based Feature Mapping for Query Classification. | Yangyang Shi, Kaisheng Yao, Le Tian, Daxin Jiang |
| 2015 | ASRU | Semi-supervised slot tagging in spoken language understanding using recurrent transductive support vector machines. | Yangyang Shi, Kaisheng Yao, Hu Chen, Yi-Cheng Pan, Mei-Yuh Hwang |
| 2015 | ICASSP | Estimating confidence scores on ASR results using recurrent neural networks. | Kaustubh Kalgaonkar, Chaojun Liu, Yifan Gong, Kaisheng Yao |
| 2015 | ICASSP | Feedback-based handwriting recognition from inertial sensor data for wearable devices. | Yujia Li, Kaisheng Yao, Geoffrey Zweig |
| 2015 | ICASSP | A factorization network based method for multi-lingual domain classification. | Yangyang Shi, Yi-Cheng Pan, Mei-Yuh Hwang, Kaisheng Yao, Hu Chen, Yuanhang Zou, Baolin Peng |
| 2015 | ICASSP | Contextual spoken language understanding using recurrent neural networks. | Yangyang Shi, Kaisheng Yao, Hu Chen, Yi-Cheng Pan, Mei-Yuh Hwang, Baolin Peng |
| 2015 | ICASSP | Deep neural support vector machines for speech recognition. | Shi-Xiong Zhang, Chaojun Liu, Kaisheng Yao, Yifan Gong |
| 2015 | Interspeech | Intermediate-layer DNN adaptation for offline and session-based iterative speaker adaptation. | Kshitiz Kumar, Chaojun Liu, Kaisheng Yao, Yifan Gong |
| 2015 | Interspeech | Sequence-to-sequence neural net models for grapheme-to-phoneme conversion. | Kaisheng Yao, Geoffrey Zweig |
| 2014 | ICASSP | Recurrent conditional random field for language understanding. | Kaisheng Yao, Baolin Peng, Geoffrey Zweig, Dong Yu, Xiaolong Li, Feng Gao |
| 2014 | Interspeech | An introduction to computational networks and the computational network toolkit (invited talk). | Dong Yu, Adam Eversole, Michael L. Seltzer, Kaisheng Yao, Brian Guenter, Oleksii Kuchaiev, Frank Seide, Huaming Wang, Jasha Droppo, Zhiheng Huang, Geoffrey Zweig, Christopher J. Rossbach, Jon Currey |
| 2013 | ICASSP | Recent advances in deep learning for speech research at Microsoft. | Li Deng, Jinyu Li, Jui-Ting Huang, Kaisheng Yao, Dong Yu, Frank Seide, Michael L. Seltzer, Geoffrey Zweig, Xiaodong He, Jason D. Williams, Yifan Gong, Alex Acero |
| 2013 | ICASSP | KL-divergence regularized deep neural network adaptation for improved large vocabulary speech recognition. | Dong Yu, Kaisheng Yao, Hang Su, Gang Li, Frank Seide |
| 2013 | Interspeech | Speed up of recurrent neural network language models with sentence independent subsampling stochastic gradient descent. | Yangyang Shi, Mei-Yuh Hwang, Kaisheng Yao, Martha A. Larson |
| 2013 | Interspeech | Recurrent neural networks for language understanding. | Kaisheng Yao, Geoffrey Zweig, Mei-Yuh Hwang, Yangyang Shi, Dong Yu |
| 2012 | Interspeech | A Feature Space Transformation Method for Personalization using Generalized I-Vector Clustering. | Kaisheng Yao, Yifan Gong, Chaojun Liu |
| 2011 | ICASSP | A basis method for robust estimation of constrained MLLR. | Daniel Povey, Kaisheng Yao |
| 2007 | ICASSP | An Approach to Low Footprint Pronunciation Models for Embedded Speaker Independent Name Recognition. | Kaisheng Yao, Lorin Netsch |
| 2006 | ICASSP | Speaker-Independent Name Recognition Using Improved Compensation and Acoustic Modeling Methods for Mobile Applications. | Kaisheng Yao, Lorin Netsch, Vishu Viswanathan |
| 2004 | ICASSP | Speech enhancement by perceptual filter with sequential noise parameter estimation. | Te-Won Lee, Kaisheng Yao |
| 2004 | Interspeech | Emotion verification for emotion detection and unknown emotion rejection. | Hoon-Young Cho, Kaisheng Yao, Te-Won Lee |
| 2003 | Interspeech | Speech recognition with a generative factor analyzed hidden Markov model. | Kaisheng Yao, Kuldip K. Paliwal, Te-Won Lee |
| 2003 | Interspeech | Model based noisy speech recognition with environment parameters estimated by noise adaptive speech recognition with prior. | Kaisheng Yao, Kuldip K. Paliwal, Satoshi Nakamura |
| 2003 | Interspeech | A speech processing front-end with eigenspace normalization for robust speech recognition in noisy automobile environments. | Kaisheng Yao, Erik M. Visser, Oh-Wook Kwon, Te-Won Lee |
| 2002 | ICASSP | Noise adaptive speech recognition in time-varying noise based on sequential kullback proximal algorithm. | Kaisheng Yao, Kuldip K. Paliwal, Satoshi Nakamura |
| 2002 | Interspeech | Noise adaptive speech recognition with acoustic models trained from noisy speech evaluated on Aurora-2 database. | Kaisheng Yao, Kuldip K. Paliwal, Satoshi Nakamura |
| 2002 | Interspeech | Evaluation of a noise adaptive speech recognition system on the Aurora 3 database. | Kaisheng Yao, Donglai Zhu, Satoshi Nakamura |
| 2001 | Interspeech | Feature extraction and model-based noise compensation for noisy speech recognition evaluated on AURORA 2 task. | Kaisheng Yao, Jingdong Chen, Kuldip K. Paliwal, Satoshi Nakamura |
| 2001 | Interspeech | Sequential noise compensation by a sequential kullback proximal algorithm. | Kaisheng Yao, Kuldip K. Paliwal, Satoshi Nakamura |
| 2000 | ICASSP | Soft GPD for minimum classification error rate training. | Bertram E. Shi, Kaisheng Yao, Zhigang Cao |
| 2000 | ICASSP | Residual noise compensation for robust speech recognition in nonstationary noise. | Kaisheng Yao, Bertram E. Shi, Pascale Fung, Zhigang Cao |
| 2000 | Interspeech | Residual noise compensation by a sequential EM algorithm for robust speech recognition in nonstationary noise. | Kaisheng Yao, Bertram E. Shi, Satoshi Nakamura, Zhigang Cao |
| 1999 | Interspeech | Liftered forward masking procedure for robust digits recognition. | Kaisheng Yao, Bertram E. Shi, Pascale Fung, Zhigang Cao |