Recurrent multi-head attention fusion network for combining audio and text for speech emotion recognition.
Chung Soo Ahn, L. L. Chamara Kasun, Sunil Sivadas, Jagath C. Rajapakse
Browse the full Interspeech paper archive.
Chung Soo Ahn, L. L. Chamara Kasun, Sunil Sivadas, Jagath C. Rajapakse
Browse the full Interspeech paper archive.