ConsistencyTTA: Accelerating Diffusion-Based Text-to-Audio Generation with Consistency Distillation.
Yatong Bai, Trung Dang, Dung N. Tran, Kazuhito Koishida, Somayeh Sojoudi
Browse the full Interspeech paper archive.
Yatong Bai, Trung Dang, Dung N. Tran, Kazuhito Koishida, Somayeh Sojoudi
Browse the full Interspeech paper archive.