Skip to content

TokenSplit: Using Discrete Speech Representations for Direct, Refined, and Transcript-Conditioned Speech Separation and Recognition.

Hakan Erdogan, Scott Wisdom, Xuankai Chang, Zaln Borsos, Marco Tagliasacchi, Neil Zeghidour, John R. Hershey

Year2023
ProceedingsINTERSPEECH

Browse the full Interspeech paper archive.