Multimodal Fusion of Music Theory-Inspired and Self-Supervised Representations for Improved Emotion Recognition.
Xiaohan Shi, Xingfeng Li, Tomoki Toda
Browse the full Interspeech paper archive.
Xiaohan Shi, Xingfeng Li, Tomoki Toda
Browse the full Interspeech paper archive.