Prosody-TTS: Improving Prosody with Masked Autoencoder and Conditional Diffusion Model For Expressive Text-to-Speech.
Rongjie Huang, Chunlei Zhang, Yi Ren, Zhou Zhao, Dong Yu
Browse the full ACL paper archive.
Rongjie Huang, Chunlei Zhang, Yi Ren, Zhou Zhao, Dong Yu
Browse the full ACL paper archive.