SSE: A Speaking Style Extractor Based on Fine-Grained Contrastive Learning between Speech and Descriptive Text.
Zixing Zhang, Yimeng Wu, Zhongren Dong, Wulong Xiang, Shengfan Shen, Bjrn W. Schuller
Browse the full ICASSP paper archive.
Zixing Zhang, Yimeng Wu, Zhongren Dong, Wulong Xiang, Shengfan Shen, Bjrn W. Schuller
Browse the full ICASSP paper archive.