| 2024 | ICASSP | A Study on Combining Non-Parallel and Parallel Methodologies for Mandarin-English Cross-Lingual Voice Conversion. | Chang Huai You, Minghui Dong |
| 2019 | ASRU | On the Study of Generative Adversarial Networks for Cross-Lingual Voice Conversion. | Berrak Sisman, Mingyang Zhang, Minghui Dong, Haizhou Li |
| 2019 | ICASSP | Implementing Prosodic Phrasing in Chinese End-to-end Speech Synthesis. | Yanfeng Lu, Minghui Dong, Ying Chen |
| 2017 | Interspeech | Multimodal Prediction of Affective Dimensions via Fusing Multiple Regression Techniques. | Dong-Yan Huang, Wan Ding, Mingyu Xu, Huaiping Ming, Minghui Dong, Xinguo Yu, Haizhou Li |
| 2016 | ICASSP | Combining multiple kernel models for automatic intelligibility detection of pathological speech. | Dong-Yan Huang, Minghui Dong, Haizhou Li |
| 2016 | ICASSP | Exemplar-based sparse representation of timbre and prosody for voice conversion. | Huaiping Ming, Dong-Yan Huang, Lei Xie, Shaofei Zhang, Minghui Dong, Haizhou Li |
| 2016 | ICASSP | A full training framework of cross-stream dependence modelling for HMM-based singing voice synthesis. | Xin Wang, Minghui Dong, Zhen-Hua Ling |
| 2016 | ICMI | Audio and face video emotion recognition in the wild using deep neural networks and small datasets. | Wan Ding, Mingyu Xu, Dong-Yan Huang, Weisi Lin, Minghui Dong, Xinguo Yu, Haizhou Li |
| 2016 | Interspeech | SERAPHIM: A Wavetable Synthesis System with 3D Lip Animation for Real-Time Speech and Singing Applications on Mobile Platforms. | Paul Yaozhu Chan, Minghui Dong, Grace Xue Hui Ho, Haizhou Li |
| 2016 | Interspeech | SERAPHIM Live! - Singing Synthesis for the Performer, the Composer, and the 3D Game Developer. | Paul Yaozhu Chan, Minghui Dong, Grace Xue Hui Ho, Haizhou Li |
| 2016 | Interspeech | Deep Bidirectional LSTM Modeling of Timbre and Prosody for Emotional Voice Conversion. | Huaiping Ming, Dong-Yan Huang, Lei Xie, Jie Wu, Minghui Dong, Haizhou Li |
| 2015 | ACII | Fundamental frequency modeling using wavelets for emotional voice conversion. | Huaiping Ming, Dong-Yan Huang, Minghui Dong, Haizhou Li, Lei Xie, Shaofei Zhang |
| 2015 | ICASSP | Sparse representation for frequency warping based voice conversion. | Xiaohai Tian, Zhizheng Wu, Siu Wa Lee, Nguyen Quy Hy, Engsiong Chng, Minghui Dong |
| 2015 | Interspeech | A real-time variable-q non-stationary Gabor transform for pitch shifting. | Dong-Yan Huang, Minghui Dong, Haizhou Li |
| 2015 | Interspeech | An alternating optimization approach for phase retrieval. | Huaiping Ming, Dong-Yan Huang, Lei Xie, Haizhou Li, Minghui Dong |
| 2015 | Interspeech | System fusion for high-performance voice conversion. | Xiaohai Tian, Zhizheng Wu, Siu Wa Lee, Nguyen Quy Hy, Minghui Dong, Engsiong Chng |
| 2015 | Interspeech | Regularized non-negative matrix factorization using alternating direction method of multipliers and its application to source separation. | Shaofei Zhang, Dong-Yan Huang, Lei Xie, Engsiong Chng, Haizhou Li, Minghui Dong |
| 2014 | ICASSP | Intelligibility detection of pathological speech using asymmetric sparse kernel partial least squares classifier. | Dong-Yan Huang, Minghui Dong, Haizhou Li |
| 2014 | Interspeech | I | Minghui Dong, Siu Wa Lee, Haizhou Li, Paul Y. Chan, Xuejian Peng, Jochen Walter Ehnes, Dong-Yan Huang |
| 2014 | Interspeech | A comparative study of spectral transformation techniques for singing voice synthesis. | Siu Wa Lee, Zhizheng Wu, Minghui Dong, Xiaohai Tian, Haizhou Li |
| 2012 | ICASSP | Template-based personalized singing voice synthesis. | Ling Cen, Minghui Dong, Paul Y. Chan |
| 2012 | ICASSP | Generalized F0 modelling with absolute and relative pitch features for singing voice synthesis. | Siu Wa Lee, Shen Ting Ang, Minghui Dong, Haizhou Li |
| 2011 | ACII | Speech Emotion Recognition System Based on L1 Regularized Linear Regression and Decision Fusion. | Ling Cen, Zhu Liang Yu, Minghui Dong |
| 2011 | Interspeech | Singing Voice Synthesis: Singer-Dependent Vibrato Modeling and Coherent Processing of Spectral Envelope. | Siu Wa Lee, Minghui Dong |
| 2010 | Interspeech | Phonetic segmentation of singing voice using MIDI and parallel speech. | Minghui Dong, Paul Y. Chan, Ling Cen, Haizhou Li, Jason Teo, Ping Jen Kua |
| 2009 | Interspeech | Unit selection based speech synthesis for poor channel condition. | Ling Cen, Minghui Dong, Paul Y. Chan, Haizhou Li |
| 2008 | Interspeech | Multi-speaker meeting audio segmentation. | Tin Lay Nwe, Minghui Dong, Swe Zin Kalayar Khine, Haizhou Li |
| 2007 | ACL | Semantic Transliteration of Personal Names. | Haizhou Li, Khe Chai Sim, Jin-Shea Kuo, Minghui Dong |
| 2006 | Interspeech | Evaluating prosody of Mandarin speech for language learning. | Minghui Dong, Haizhou Li, Tin Lay Nwe |
| 2006 | Interspeech | Analysis and detection of speech under sleep deprivation. | Tin Lay Nwe, Haizhou Li, Minghui Dong |
| 2005 | Interspeech | A probabilistic approach to prosodic word prediction for Mandarin Chinese TTS. | Minghui Dong, Kim-Teng Lua, Haizhou Li |
| 2004 | IJCNLP | Selecting Prosody Parameters for Unit Selection Based Chinese TTS. | Minghui Dong, Kim-Teng Lua, Jun Xu |
| 2003 | Interspeech | On unit analysis for Cantonese corpus-based TTS. | Jun Xu, Thomas Choy, Minghui Dong, Cuntai Guan, Haizhou Li |
| 2002 | Interspeech | Pitch contour model for Chinese text-to-speech using CART and statistical model. | Minghui Dong, Kim-Teng Lua |
| 2002 | Interspeech | Automatic prosodic break labeling for Mandarin Chinese speech data. | Minghui Dong, Kim-Teng Lua |
| 2000 | Interspeech | Using prosody database in Chinese speech synthesis. | Minghui Dong, Kim-Teng Lua |