Skip to content

A Transformer-Based Audio Captioning Model with Keyword Estimation.

Yuma Koizumi, Ryo Masumura, Kyosuke Nishida, Masahiro Yasuda, Shoichiro Saito

Year2020
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.