Skip to content

Shivam Mehta

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

13

Venues

6

Active years

2022–2025

Best venue rank

A*

Where they publish

Papers

13 indexed papers, newest first.

YearVenueTitleAuthors
2025HRITake a Look, it's in a Book, a Reading Robot.Paige Tutts, Shivam Mehta, Zachary Syvenky, Bermet Burkanova, Mohammed Hfsafsti, Yue Wang, H. Henny Yeung, Gustav Eje Henter, Jean-Julien Aucouturier, Angelica Lim
2025ICASSPMake Some Noise: Towards LLM audio reasoning and generation using sound tokens.Shivam Mehta, Nebojsa Jojic, Hannes Gamper
2025InterspeechSawtArabi: A Benchmark Corpus for Arabic TTS. Standard, Dialectal and Code-Switching.Vasista Sai Lodagala, Lamya Alkanhal, Daniel Izham, Shivam Mehta, Shammur Absar Chowdhury, Aqeelah Makki, Hamdy S. Hussein, Gustav Eje Henter, Ahmed Ali
2025RO-MANEmojiVoice: Towards long-term controllable expressivity in robot speech.Paige Tutts, Shivam Mehta, Zachary Syvenky, Bermet Burkanova, Gustav Eje Henter, Angelica Lim
2024CVPRFake it to make it: Using synthetic data to remedy the data shortage in joint multi-modal speech-and-gesture synthesis.Shivam Mehta, Anna Deichler, Jim O'Regan, Birger Moll, Jonas Beskow, Gustav Eje Henter, Simon Alexanderson
2024ICASSPUnified Speech and Gesture Synthesis Using Flow Matching.Shivam Mehta, Ruibo Tu, Simon Alexanderson, Jonas Beskow, va Szkely, Gustav Eje Henter
2024ICASSPMatcha-TTS: A Fast TTS Architecture with Conditional Flow Matching.Shivam Mehta, Ruibo Tu, Jonas Beskow, va Szkely, Gustav Eje Henter
2024InterspeechShould you use a probabilistic duration model in TTS? Probably! Especially for spontaneous speech.Shivam Mehta, Harm Lameris, Rajiv Punmiya, Jonas Beskow, va Szkely, Gustav Eje Henter
2024InterspeechBeyond graphemes and phonemes: continuous phonological features in neural text-to-speech synthesis.Christina Tnnander, Shivam Mehta, Jonas Beskow, Jens Edlund
2023ICASSPProsody-Controllable Spontaneous TTS with Neural HMMS.Harm Lameris, Shivam Mehta, Gustav Eje Henter, Joakim Gustafson, va Szkely
2023ICMIDiffusion-Based Co-Speech Gesture Generation Using Joint Text and Audio Representation.Anna Deichler, Shivam Mehta, Simon Alexanderson, Jonas Beskow
2023InterspeechOverFlow: Putting flows on top of neural transducers for better TTS.Shivam Mehta, Ambika Kirkland, Harm Lameris, Jonas Beskow, va Szkely, Gustav Eje Henter
2022ICASSPNeural HMMS Are All You Need (For High-Quality Attention-Free TTS).Shivam Mehta, va Szkely, Jonas Beskow, Gustav Eje Henter