| 2025 | ICASSP | Self-supervised Prosody Learning at Phoneme-level with Momentum Contrast for Speech Synthesis. | Zhaoci Liu, Ya-Jun Hu, Liping Chen, Zhen-Hua Ling |
| 2025 | ICASSP | Anchored Monotonic Alignment and Representation Substitution for Rare Spontaneous Behaviors in Spontaneous Speech Synthesis. | Ning-Qian Wu, Ya-Jun Hu, Liping Chen, Zhen-Hua Ling |
| 2023 | Interspeech | Speech Synthesis with Self-Supervisedly Learnt Prosodic Representations. | Zhaoci Liu, Zhen-Hua Ling, Ya-Jun Hu, Jia Pan, Jin-Wei Wang, Yun-Di Wu |
| 2022 | ICASSP | Improving Recognition-Synthesis Based any-to-one Voice Conversion with Cyclic Training. | Yan-Nian Chen, Li-Juan Liu, Ya-Jun Hu, Yuan Jiang, Zhen-Hua Ling |
| 2022 | ICASSP | Neural Grapheme-To-Phoneme Conversion with Pre-Trained Grapheme Models. | Lu Dong, Zhiqiang Guo, Chao-Hong Tan, Ya-Jun Hu, Yuan Jiang, Zhen-Hua Ling |
| 2017 | ASRU | The USTC system for blizzard machine learning challenge 2017-ES2. | Ya-Jun Hu, Li-Juan Liu, Chuang Ding, Zhen-Hua Ling, Li-Rong Dai |
| 2017 | ASRU | The iFLYTEK system for blizzard machine learning challenge 2017-ES1. | Li-Juan Liu, Chuang Ding, Ya-Jun Hu, Zhen-Hua Ling, Yuan Jiang, Ming Zhou, Si Wei |
| 2017 | ICASSP | Extracting structural spectral features using what-where auto-encoders for statistical parametric speech synthesis. | Ya-Jun Hu, Zhen-Hua Ling, Li-Rong Dai |
| 2016 | ICASSP | Deep belief network-based post-filtering for statistical parametric speech synthesis. | Ya-Jun Hu, Zhen-Hua Ling, Li-Rong Dai |
| 2016 | ICASSP | Modeling spectral envelopes using deep conditional restricted Boltzmann machines for statistical parametric speech synthesis. | Xiang Yin, Zhen-Hua Ling, Ya-Jun Hu, Li-Rong Dai |