Skip to content

High Fidelity Text-to-Speech Via Discrete Tokens Using Token Transducer and Group Masked Language Model.

Joun Yeop Lee, Myeonghun Jeong, Minchan Kim, Ji-Hyun Lee, Hoon-Young Cho, Nam Soo Kim

Year2024
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.