Skip to content

Deep multimodal semantic embeddings for speech and images.

David F. Harwath, James R. Glass

VenueCASRU
Year2015
ProceedingsASRU

Browse the full ASRU paper archive.