Transduce and Speak: Neural Transducer for Text-To-Speech with Semantic Token Prediction.
Minchan Kim, Myeonghun Jeong, Byoung Jin Choi, Dongjune Lee, Nam Soo Kim
Browse the full ASRU paper archive.
Minchan Kim, Myeonghun Jeong, Byoung Jin Choi, Dongjune Lee, Nam Soo Kim
Browse the full ASRU paper archive.