| 2018 | ICASSP | Emphatic Speech Generation with Conditioned Input Layer and Bidirectional LSTMS for Expressive Speech Synthesis. | Runnan Li, Zhiyong Wu, Yuchen Huang, Jia Jia, Helen Meng, Lianhong Cai |
| 2018 | ICASSP | Applying Multitask Learning to Acoustic-Phonemic Model for Mispronunciation Detection and Diagnosis in L2 English Speech. | Shaoguang Mao, Zhiyong Wu, Runnan Li, Xu Li, Helen Meng, Lianhong Cai |
| 2018 | Interspeech | Emotion Recognition from Variable-Length Speech Segments Using Deep Learning on Spectrograms. | Xi Ma, Zhiyong Wu, Jia Jia, Mingxing Xu, Helen Meng, Lianhong Cai |
| 2018 | Interspeech | Siamese Recurrent Auto-Encoder Representation for Query-by-Example Spoken Term Detection. | Ziwei Zhu, Zhiyong Wu, Runnan Li, Helen Meng, Lianhong Cai |
| 2017 | ICASSP | Multi-task learning of structured output layer bidirectional LSTMS for speech synthesis. | Runnan Li, Zhiyong Wu, Xunying Liu, Helen M. Meng, Lianhong Cai |
| 2017 | ICASSP | Learning cross-lingual knowledge with multilingual BLSTM for emphasis detection with limited training data. | Yishuang Ning, Zhiyong Wu, Runnan Li, Jia Jia, Mingxing Xu, Helen M. Meng, Lianhong Cai |
| 2017 | ICASSP | A systematic approach to compute perceptual distribution of monosyllables. | Yu-Hao Wu, Jia Jia, Feng Lu, Lianhong Cai |
| 2017 | Interspeech | Multi-Task Learning for Prosodic Structure Generation Using BLSTM RNN with Structured Output Layer. | Yuchen Huang, Zhiyong Wu, Runnan Li, Helen Meng, Lianhong Cai |
| 2017 | Interspeech | Spectro-Temporal Modelling with Time-Frequency LSTM and Structured Output Layer for Voice Conversion. | Runnan Li, Zhiyong Wu, Yishuang Ning, Lifa Sun, Helen Meng, Lianhong Cai |
| 2017 | Interspeech | Speech Emotion Recognition with Emotion-Pair Based Framework Considering Emotion Distribution Information in Dimensional Emotion Space. | Xi Ma, Zhiyong Wu, Jia Jia, Mingxing Xu, Helen Meng, Lianhong Cai |
| 2016 | ICASSP | Low level descriptors based DBLSTM bottleneck feature for speech driven talking avatar. | Xinyu Lan, Xu Li, Yishuang Ning, Zhiyong Wu, Helen Meng, Jia Jia, Lianhong Cai |
| 2016 | ICASSP | A deep bidirectional long short-term memory based multi-scale approach for music dynamic emotion prediction. | Xinxing Li, Haishu Xianyu, Jiashen Tian, Wenxiao Chen, Fanhang Meng, Mingxing Xu, Lianhong Cai |
| 2016 | ICASSP | Question detection from acoustic features using recurrent neural network with gated recurrent unit. | Yaodong Tang, Yuchen Huang, Zhiyong Wu, Helen Meng, Mingxing Xu, Lianhong Cai |
| 2016 | ICASSP | SVR based double-scale regression for dynamic emotion prediction in music. | Haishu Xianyu, Xinxing Li, Wenxiao Chen, Fanhang Meng, Jiashen Tian, Mingxing Xu, Lianhong Cai |
| 2016 | ICASSP | Learning cross-lingual information with multilingual BLSTM for speech synthesis of low-resource languages. | Quanjie Yu, Peng Liu, Zhiyong Wu, Shiyin Kang, Helen Meng, Lianhong Cai |
| 2016 | Interspeech | Phoneme Embedding and its Application to Speech Driven Talking Avatar Synthesis. | Xu Li, Zhiyong Wu, Helen M. Meng, Jia Jia, Xiaoyan Lou, Lianhong Cai |
| 2016 | Interspeech | Expressive Speech Driven Talking Avatar Synthesis with DBLSTM Using Limited Amount of Emotional Bimodal Data. | Xu Li, Zhiyong Wu, Helen M. Meng, Jia Jia, Xiaoyan Lou, Lianhong Cai |
| 2016 | Interspeech | Combining CNN and BLSTM to Extract Textual and Acoustic Features for Recognizing Stances in Mandarin Ideological Debate Competition. | Linchuan Li, Zhiyong Wu, Mingxing Xu, Helen M. Meng, Lianhong Cai |
| 2016 | Interspeech | Analysis on Gated Recurrent Unit Based Question Detection Approach. | Yaodong Tang, Zhiyong Wu, Helen M. Meng, Mingxing Xu, Lianhong Cai |
| 2015 | ACII | Understanding speaking styles of internet speech data with LSTM and low-resource training. | Xixin Wu, Zhiyong Wu, Yishuang Ning, Jia Jia, Lianhong Cai, Helen M. Meng |
| 2015 | ICASSP | A deep recurrent approach for acoustic-to-articulatory inversion. | Peng Liu, Quanjie Yu, Zhiyong Wu, Shiyin Kang, Helen M. Meng, Lianhong Cai |
| 2015 | ICASSP | HMM-based emphatic speech synthesis for corrective feedback in computer-aided pronunciation training. | Yishuang Ning, Zhiyong Wu, Jia Jia, Fanbo Meng, Helen M. Meng, Lianhong Cai |
| 2015 | ICMI | MPHA: A Personal Hearing Doctor Based on Mobile Devices. | Yu-Hao Wu, Jia Jia, Wai-Kim Leung, Yejun Liu, Lianhong Cai |
| 2015 | Interspeech | Using tilt for automatic emphasis detection with Bayesian networks. | Yishuang Ning, Zhiyong Wu, Xiaoyan Lou, Helen M. Meng, Jia Jia, Lianhong Cai |
| 2014 | ICASSP | Learning dynamic features with neural networks for phoneme recognition. | Xin Zheng, Zhiyong Wu, Helen Meng, Lianhong Cai |
| 2014 | ICASSP | Contrastive auto-encoder for phoneme recognition. | Xin Zheng, Zhiyong Wu, Helen Meng, Lianhong Cai |
| 2014 | IJCNN | Improved keyword spotting system by optimizing posterior confidence measure vector using feed-forward neural network. | Yuchen Liu, Mingxing Xu, Lianhong Cai |
| 2014 | Interspeech | Using conditional random fields to predict focus word pair in spontaneous spoken English. | Xiao Zang, Zhiyong Wu, Helen M. Meng, Jia Jia, Lianhong Cai |
| 2014 | MMM | Learning to Infer Public Emotions from Large-Scale Networked Voice Data. | Zhu Ren, Jia Jia, Lianhong Cai, Kuo Zhang, Jie Tang |
| 2013 | ICASSP | Investigation of tandem deep belief network approach for phoneme recognition. | Xin Zheng, Zhiyong Wu, Binbin Shen, Helen M. Meng, Lianhong Cai |
| 2013 | ICIP | Interpretable aesthetic features for affective image classification. | Xiaohui Wang, Jia Jia, Jiaming Yin, Lianhong Cai |
| 2013 | ICNC | SNR estimation for clipped audio based on amplitude distribution. | Xiaoqing Liu, Jia Jia, Lianhong Cai |
| 2012 | Interspeech | Hierarchical English Emphatic Speech Synthesis Based on HMM with Limited Training Data. | Fanbo Meng, Zhiyong Wu, Helen M. Meng, Jia Jia, Lianhong Cai |
| 2011 | Interspeech | Combining Active and Semi-Supervised Learning for Homograph Disambiguation in Mandarin Text-to-Speech Synthesis. | Binbin Shen, Zhiyong Wu, Yongxin Wang, Lianhong Cai |
| 2010 | ICIP | Facial expression synthesis based on motion patterns learned from face database. | Jia Jia, Shen Zhang, Lianhong Cai |
| 2010 | ICNC | Emotional talking agent: System and evaluation. | Shen Zhang, Jia Jia, Yingjin Xu, Lianhong Cai |
| 2010 | ICPR | Comparison of Syllable/Phone HMM Based Mandarin TTS. | Quansheng Duan, Shiyin Kang, Zhiyong Wu, Lianhong Cai, Zhiwei Shuang, Yong Qin |
| 2010 | Interspeech | HMM based TTS for mixed language text. | Zhiwei Shuang, Shiyin Kang, Yong Qin, Li-Rong Dai, Lianhong Cai |
| 2009 | ICASSP | Cultural style based music classification of audio signals. | Yuxiang Liu, Qiaoliang Xiang, Ye Wang, Lianhong Cai |
| 2009 | Interspeech | Voiced/unvoiced decision algorithm for HMM-based speech synthesis. | Shiyin Kang, Zhiwei Shuang, Quansheng Duan, Yong Qin, Lianhong Cai |
| 2009 | Interspeech | Syllable HMM based Mandarin TTS and comparison with concatenative TTS. | Zhiwei Shuang, Shiyin Kang, Qin Shi, Yong Qin, Lianhong Cai |
| 2008 | ICNC | Entering Tone Recognition in a Support Vector Machine Approach. | Xiangcheng Wang, Ying Liu, Lianhong Cai |
| 2007 | ACII | Affect Related Acoustic Features of Speech and Their Modification. | Dandan Cui, Fanbo Meng, Lianhong Cai, Liuyi Sun |
| 2007 | ACII | Facial Expression Synthesis Using PAD Emotional Parameters for a Chinese Expressive Avatar. | Shen Zhang, Zhiyong Wu, Helen M. Meng, Lianhong Cai |
| 2007 | ICASSP | Script Design Based on Decision Tree with Context Vector and Acoustic Distance for Mandarin TTS. | Dandan Cui, Denzhi Huang, Yuan Dong, Lianhong Cai, Haila Wang |
| 2007 | ICASSP | Head Movement Synthesis Based on Semantic and Prosodic Features for a Chinese Expressive Avatar. | Shen Zhang, Zhiyong Wu, Helen M. Meng, Lianhong Cai |
| 2007 | Interspeech | Hierarchical non-uniform unit selection based on prosodic structure. | Jun Xu, Dezhi Huang, Yongxin Wang, Yuan Dong, Lianhong Cai, Haila Wang |
| 2006 | Interspeech | Real-time synthesis of Chinese visual speech and facial expressions using MPEG-4 FAP features in a three-dimensional avatar. | Zhiyong Wu, Shen Zhang, Lianhong Cai, Helen M. Meng |
| 2006 | Interspeech | Modeling the acoustic correlates of expressive elements in text genres for expressive text-to-speech synthesis. | Hongwu Yang, Helen M. Meng, Lianhong Cai |
| 2005 | ICASSP | Prosody Analysis and Modeling for Emotional Speech Synthesis. | Dan-Ning Jiang, Wei Zhang, Liqin Shen, Lianhong Cai |
| 2005 | Interspeech | Grapheme-to-phoneme conversion based on TBL algorithm in Mandarin TTS system. | Min Zheng, Qin Shi, Wei Zhang, Lianhong Cai |
| 2004 | ICPR | Face Pose Estimation and its Application in Video Shot Selection. | Zhiguang Yang, Haizhou Ai, Bo Wu, Shihong Lao, Lianhong Cai |
| 2003 | SMC | An adaptive system for online document filtering. | Liang Ma, Qunxiu Chen, Lianhong Cai |
| 2002 | Interspeech | Clustering and feature learning based F0 prediction for Chinese speech synthesis. | Jianhua Tao, Lianhong Cai |
| 2002 | Interspeech | Prosodic phrasing with inductive learning. | Sheng Zhao, Jianhua Tao, Lianhong Cai |
| 2000 | Interspeech | The design and application of a speech database for Chinese TTS system. | Muhua Lv, Lianhong Cai |
| 2000 | Interspeech | Research on dynamic characters of Chinese pitch contours. | Zhiyong Wu, Lianhong Cai, Tongchun Zhou |