High Fidelity Text-to-Speech Via Discrete Tokens Using Token Transducer and Group Masked Language Model.
Joun Yeop Lee, Myeonghun Jeong, Minchan Kim, Ji-Hyun Lee, Hoon-Young Cho, Nam Soo Kim
Browse the full Interspeech paper archive.
Joun Yeop Lee, Myeonghun Jeong, Minchan Kim, Ji-Hyun Lee, Hoon-Young Cho, Nam Soo Kim
Browse the full Interspeech paper archive.