Skip to content

Mellotron: Multispeaker Expressive Voice Synthesis by Conditioning on Rhythm, Pitch and Global Style Tokens.

Rafael Valle, Jason Li, Ryan Prenger, Bryan Catanzaro

Year2020
ProceedingsICASSP

Browse the full ICASSP paper archive.