DiTTo-TTS: Diffusion Transformers for Scalable Text-to-Speech without Domain-Specific Factors.
Keon Lee, Dong Won Kim, Jaehyeon Kim, Seungjun Chung, Jaewoong Cho
Browse the full ICLR paper archive.
Keon Lee, Dong Won Kim, Jaehyeon Kim, Seungjun Chung, Jaewoong Cho
Browse the full ICLR paper archive.