LiveSpeech: Low-Latency Zero-shot Text-to-Speech via Autoregressive Modeling of Audio Discrete Codes.
Trung Dang, David Aponte, Dung N. Tran, Kazuhito Koishida
Browse the full Interspeech paper archive.
Trung Dang, David Aponte, Dung N. Tran, Kazuhito Koishida
Browse the full Interspeech paper archive.