Seeing Voices and Hearing Voices: Learning Discriminative Embeddings Using Cross-Modal Self-Supervision.
Soo-Whan Chung, Hong-Goo Kang, Joon Son Chung
Browse the full Interspeech paper archive.
Soo-Whan Chung, Hong-Goo Kang, Joon Son Chung
Browse the full Interspeech paper archive.