Skip to content

SSE: A Speaking Style Extractor Based on Fine-Grained Contrastive Learning between Speech and Descriptive Text.

Zixing Zhang, Yimeng Wu, Zhongren Dong, Wulong Xiang, Shengfan Shen, Bjrn W. Schuller

Year2025
ProceedingsICASSP

Browse the full ICASSP paper archive.