Shivam Mehta
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
13
Venues
6
Active years
2022–2025
Best venue rank
A*
Where they publish
Papers
13 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | HRI | Take a Look, it's in a Book, a Reading Robot. | Paige Tutts, Shivam Mehta, Zachary Syvenky, Bermet Burkanova, Mohammed Hfsafsti, Yue Wang, H. Henny Yeung, Gustav Eje Henter, Jean-Julien Aucouturier, Angelica Lim |
| 2025 | ICASSP | Make Some Noise: Towards LLM audio reasoning and generation using sound tokens. | Shivam Mehta, Nebojsa Jojic, Hannes Gamper |
| 2025 | Interspeech | SawtArabi: A Benchmark Corpus for Arabic TTS. Standard, Dialectal and Code-Switching. | Vasista Sai Lodagala, Lamya Alkanhal, Daniel Izham, Shivam Mehta, Shammur Absar Chowdhury, Aqeelah Makki, Hamdy S. Hussein, Gustav Eje Henter, Ahmed Ali |
| 2025 | RO-MAN | EmojiVoice: Towards long-term controllable expressivity in robot speech. | Paige Tutts, Shivam Mehta, Zachary Syvenky, Bermet Burkanova, Gustav Eje Henter, Angelica Lim |
| 2024 | CVPR | Fake it to make it: Using synthetic data to remedy the data shortage in joint multi-modal speech-and-gesture synthesis. | Shivam Mehta, Anna Deichler, Jim O'Regan, Birger Moll, Jonas Beskow, Gustav Eje Henter, Simon Alexanderson |
| 2024 | ICASSP | Unified Speech and Gesture Synthesis Using Flow Matching. | Shivam Mehta, Ruibo Tu, Simon Alexanderson, Jonas Beskow, va Szkely, Gustav Eje Henter |
| 2024 | ICASSP | Matcha-TTS: A Fast TTS Architecture with Conditional Flow Matching. | Shivam Mehta, Ruibo Tu, Jonas Beskow, va Szkely, Gustav Eje Henter |
| 2024 | Interspeech | Should you use a probabilistic duration model in TTS? Probably! Especially for spontaneous speech. | Shivam Mehta, Harm Lameris, Rajiv Punmiya, Jonas Beskow, va Szkely, Gustav Eje Henter |
| 2024 | Interspeech | Beyond graphemes and phonemes: continuous phonological features in neural text-to-speech synthesis. | Christina Tnnander, Shivam Mehta, Jonas Beskow, Jens Edlund |
| 2023 | ICASSP | Prosody-Controllable Spontaneous TTS with Neural HMMS. | Harm Lameris, Shivam Mehta, Gustav Eje Henter, Joakim Gustafson, va Szkely |
| 2023 | ICMI | Diffusion-Based Co-Speech Gesture Generation Using Joint Text and Audio Representation. | Anna Deichler, Shivam Mehta, Simon Alexanderson, Jonas Beskow |
| 2023 | Interspeech | OverFlow: Putting flows on top of neural transducers for better TTS. | Shivam Mehta, Ambika Kirkland, Harm Lameris, Jonas Beskow, va Szkely, Gustav Eje Henter |
| 2022 | ICASSP | Neural HMMS Are All You Need (For High-Quality Attention-Free TTS). | Shivam Mehta, va Szkely, Jonas Beskow, Gustav Eje Henter |