Sound of Vision: Audio Generation from Visual Text Embedding through Training Domain Discriminator.
Jaewon Kim, Won-Gook Choi, Seyun Ahn, Joon-Hyuk Chang
Browse the full Interspeech paper archive.
Jaewon Kim, Won-Gook Choi, Seyun Ahn, Joon-Hyuk Chang
Browse the full Interspeech paper archive.