Skip to content

Audiobox TTA-RAG: Improving Zero-Shot and Few-Shot Text-To-Audio with Retrieval-Augmented Generation.

Mu Yang, Bowen Shi, Matthew Le, Wei-Ning Hsu, Andros Tjandra

Year2025
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.