Automatic metadata generation and video editing based on speech and image recognition for medical education contents.
Satoshi Tamura, Koji Hashimoto, Jiong Zhu, Satoru Hayamizu, Hirotsugu Asai, Hideki Tanahashi, Makoto Kanagawa
Browse the full Interspeech paper archive.