Multimodal Speaker Segmentation and Diarization Using Lexical and Acoustic Cues via Sequence to Sequence Neural Networks.
Tae Jin Park, Panayiotis G. Georgiou
Browse the full Interspeech paper archive.
Tae Jin Park, Panayiotis G. Georgiou
Browse the full Interspeech paper archive.