| 2022 | EMNLP | Speaker Overlap-aware Neural Diarization for Multi-party Meeting Analysis. | Zhihao Du, Shiliang Zhang, Siqi Zheng, Zhi-Jie Yan |
| 2015 | ICDAR | A context-sensitive-chunk BPTT approach to training deep LSTM/BLSTM recurrent neural networks for offline handwriting recognition. | Kai Chen, Zhi-Jie Yan, Qiang Huo |
| 2015 | Interspeech | Training deep bidirectional LSTM acoustic model for LVCSR by a context-sensitive-chunk BPTT approach. | Kai Chen, Zhi-Jie Yan, Qiang Huo |
| 2013 | ICASSP | Tied-state based discriminative training of context-expanded region-dependent feature transforms for LVCSR. | Zhi-Jie Yan, Qiang Huo, Jian Xu, Yu Zhang |
| 2013 | Interspeech | A scalable approach to using DNN-derived features in GMM-HMM based acoustic modeling for LVCSR. | Zhi-Jie Yan, Qiang Huo, Jian Xu |
| 2012 | ICASSP | A study of discriminative feature extraction for i-vector based acoustic sniffing in IVN acoustic model training. | Yu Zhang, Jian Xu, Zhi-Jie Yan, Qiang Huo |
| 2011 | ICASSP | Speaker characterization using spectral subband energy ratio based on Harmonic plus Noise Model. | Yanhua Long, Zhi-Jie Yan, Frank K. Soong, Li-Rong Dai, Wu Guo |
| 2011 | ICASSP | A study of an irrelevant variability normalization based discriminative training approach for LVCSR. | Yu Zhang, Jian Xu, Zhi-Jie Yan, Qiang Huo |
| 2011 | Interspeech | Improvements in Speaker Characterization Using Spectral Subband Energy Based on Harmonic plus Noise Model. | Yanhua Long, Zhi-Jie Yan, Frank K. Soong, Li-Rong Dai, Wu Guo |
| 2011 | Interspeech | An i-vector Based Approach to Acoustic Sniffing for Irrelevant Variability Normalization Based Acoustic Model Training and Speech Recognition. | Jian Xu, Yu Zhang, Zhi-Jie Yan, Qiang Huo |
| 2011 | Interspeech | An i-vector Based Approach to Training Data Clustering for Improved Speech Recognition. | Yu Zhang, Jian Xu, Zhi-Jie Yan, Qiang Huo |
| 2010 | ICASSP | RIch-context Unit Selection (RUS) approach to high quality TTS. | Zhi-Jie Yan, Yao Qian, Frank K. Soong |
| 2010 | ICASSP | Cross-validation based decision tree clustering for HMM-based TTS. | Yu Zhang, Zhi-Jie Yan, Frank K. Soong |
| 2010 | Interspeech | A perceptual study of acceleration parameters in HMM-based TTS. | Yining Chen, Zhi-Jie Yan, Frank K. Soong |
| 2010 | Interspeech | An HMM trajectory tiling (HTT) approach to high quality TTS. | Yao Qian, Zhi-Jie Yan, Yi-Jian Wu, Frank K. Soong, Xin Zhuang, Shengyi Kong |
| 2009 | ICASSP | A trust region based optimization for maximum mutual information estimation of HMMS in speech recognition. | Zhi-Jie Yan, Cong Liu, Yu Hu, Hui Jiang |
| 2009 | Interspeech | Rich context modeling for high quality HMM-based TTS. | Zhi-Jie Yan, Yao Qian, Frank K. Soong |
| 2008 | ICASSP | Minimum word classification error training of HMMS for automatic speech recognition. | Zhi-Jie Yan, Bo Zhu, Yu Hu, Ren-Hua Wang |
| 2008 | Interspeech | Soft margin estimation with various separation levels for LVCSR. | Jinyu Li, Zhi-Jie Yan, Chin-Hui Lee, Ren-Hua Wang |
| 2007 | ASRU | A study on soft margin estimation for LVCSR. | Jinyu Li, Zhi-Jie Yan, Chin-Hui Lee, Ren-Hua Wang |
| 2007 | ICASSP | Word Graph Based Feature Enhancement for Noisy Speech Recognition. | Zhi-Jie Yan, Frank K. Soong, Ren-Hua Wang |