Skip to content

SA-RAS: Speaker-Aware Style Retrieval Augmented Generation for Expressive Zero-Shot Text-to-Speech Synthesis.

Xueru Li, Jingyuan Xing, Xiaofen Xing, Zhipeng Li, Xiangmin Xu

Year2025
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.