| 2025 | ACL | PodAgent: A Comprehensive Framework for Podcast Generation. | Yujia Xiao, Lei He, Haohan Guo, Fenglong Xie, Tan Lee |
| 2025 | ICASSP | ZSVC: Zero-shot Style Voice Conversion with Disentangled Latent Diffusion Models and Adversarial Training. | Xinfa Zhu, Lei He, Yujia Xiao, Xi Wang, Xu Tan, Sheng Zhao, Lei Xie |
| 2023 | Interspeech | ContextSpeech: Expressive and Efficient Text-to-Speech for Paragraph Reading. | Yujia Xiao, Shaofei Zhang, Xi Wang, Xu Tan, Lei He, Sheng Zhao, Frank K. Soong, Tan Lee |
| 2022 | ICASSP | Improving Fastspeech TTS with Efficient Self-Attention and Compact Feed-Forward Network. | Yujia Xiao, Xi Wang, Lei He, Frank K. Soong |
| 2022 | ICASSP | Prosodyspeech: Towards Advanced Prosody Model for Neural Text-to-Speech. | Yuanhao Yi, Lei He, Shifeng Pan, Xi Wang, Yujia Xiao |
| 2020 | ICASSP | Improving Prosody with Linguistic and Bert Derived Features in Multi-Speaker Based Mandarin Chinese Neural TTS. | Yujia Xiao, Lei He, Huaiping Ming, Frank K. Soong |
| 2018 | Interspeech | Paired Phone-Posteriors Approach to ESL Pronunciation Quality Assessment. | Yujia Xiao, Frank K. Soong, Wenping Hu |
| 2017 | Interspeech | Proficiency Assessment of ESL Learner's Sentence Prosody with TTS Synthesized Voice as Reference. | Yujia Xiao, Frank K. Soong |