| 2025 | ICLR | RFWave: Multi-band Rectified Flow for Audio Waveform Reconstruction. | Peng Liu, Dongyang Dai, Zhiyong Wu |
| 2024 | IJCNN | NRAdapt: Noise-Robust Adaptive Text to Speech Using Untranscribed Data. | Ming Cheng, Shun Lei, Dongyang Dai, Zhiyong Wu, Dading Chong |
| 2024 | Interspeech | Multi-modal Adversarial Training for Zero-Shot Voice Cloning. | John Janiczek, Dading Chong, Dongyang Dai, Arlo Faria, Chao Wang, Tao Wang, Yuzong Liu |
| 2022 | ICASSP | Cloning One's Voice Using Very Limited Data in the Wild. | Dongyang Dai, Yuanzhe Chen, Li Chen, Ming Tu, Lu Liu, Rui Xia, Qiao Tian, Yuping Wang, Yuxuan Wang |
| 2021 | ICASSP | Emotion Controllable Speech Synthesis Using Emotion-Unlabeled Dataset with the Assistance of Cross-Domain Speech Emotion Recognition. | Xiong Cai, Dongyang Dai, Zhiyong Wu, Xiang Li, Jingbei Li, Helen Meng |
| 2019 | ICASSP | Learning Discriminative Features from Spectrograms Using Center Loss for Speech Emotion Recognition. | Dongyang Dai, Zhiyong Wu, Runnan Li, Xixin Wu, Jia Jia, Helen Meng |
| 2019 | ICASSP | Speech Emotion Recognition Using Capsule Networks. | Xixin Wu, Songxiang Liu, Yuewen Cao, Xu Li, Jianwei Yu, Dongyang Dai, Xi Ma, Shoukang Hu, Zhiyong Wu, Xunying Liu, Helen Meng |
| 2019 | Interspeech | Disambiguation of Chinese Polyphones in an End-to-End Framework with Semantic Features Extracted by Pre-Trained BERT. | Dongyang Dai, Zhiyong Wu, Shiyin Kang, Xixin Wu, Jia Jia, Dan Su, Dong Yu, Helen Meng |
| 2019 | Interspeech | One-Shot Voice Conversion with Global Speaker Embeddings. | Hui Lu, Zhiyong Wu, Dongyang Dai, Runnan Li, Shiyin Kang, Jia Jia, Helen Meng |