Two-layered audio-visual integration in voice activity detection and automatic speech recognition for robots.
Takami Yoshida, Kazuhiro Nakadai
Browse the full Interspeech paper archive.
Takami Yoshida, Kazuhiro Nakadai
Browse the full Interspeech paper archive.