Skip to content

Text-Free Image-to-Speech Synthesis Using Learned Segmental Units.

Wei-Ning Hsu, David Harwath, Tyler Miller, Christopher Song, James R. Glass

VenueA*ACL
Year2021
ProceedingsACL/IJCNLP (1)

Browse the full ACL paper archive.