Skip to content

Show and Speak: Directly Synthesize Spoken Description of Images.

Xinsheng Wang, Siyuan Feng, Jihua Zhu, Mark Hasegawa-Johnson, Odette Scharenborg

Year2021
ProceedingsICASSP

Browse the full ICASSP paper archive.