Skip to content

ZET-Speech: Zero-shot adaptive Emotion-controllable Text-to-Speech Synthesis with Diffusion and Style-based Models.

Minki Kang, Wooseok Han, Sung Ju Hwang, Eunho Yang

Year2023
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.