Skip to content

ImmersiveTTS: Environment-Aware Text-to-Speech with Multimodal Diffusion Transformer and Domain-Specific Representation Alignment.

Jun-Hak Yun, Seung-Bin Kim, Seong-Whan Lee

VenueA*ACL
Year2026
ProceedingsACL (1)

Browse the full ACL paper archive.