UMETTS: A Unified Framework for Emotional Text-to-Speech Synthesis with Multimodal Prompts.
Xiang Li, Zhi-Qi Cheng, Jun-Yan He, Junyao Chen, Xiaomao Fan, Xiaojiang Peng, Alexander G. Hauptmann
Browse the full ICASSP paper archive.
Xiang Li, Zhi-Qi Cheng, Jun-Yan He, Junyao Chen, Xiaomao Fan, Xiaojiang Peng, Alexander G. Hauptmann
Browse the full ICASSP paper archive.