Rohan Badlani
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
10
Venues
5
Active years
2018–2025
Best venue rank
A*
Where they publish
Papers
10 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICLR | Fugatto 1: Foundational Generative Audio Transformer Opus 1. | Rafael Valle, Rohan Badlani, Zhifeng Kong, Sang-gil Lee, Arushi Goel, Sungwon Kim, Joo Felipe Santos, Shuqi Dai, Siddharth Gururani, Aya Aljafari, Alexander H. Liu, Kevin J. Shih, Ryan Prenger, Wei Ping, Chao-Han Huck Yang, Bryan Catanzaro |
| 2024 | ICASSP | Scaling Nvidia's Multi-Speaker Multi-Lingual TTS Systems With Zero-Shot TTS to Indic Languages. | Akshit Arora, Rohan Badlani, Sungwon Kim, Rafael Valle, Bryan Catanzaro |
| 2024 | ICML | Audio Flamingo: A Novel Audio Language Model with Few-Shot Learning and Dialogue Abilities. | Zhifeng Kong, Arushi Goel, Rohan Badlani, Wei Ping, Rafael Valle, Bryan Catanzaro |
| 2023 | ICASSP | Vani: Very-Lightweight Accent-Controllable TTS for Native And Non-Native Speakers With Identity Preservation. | Rohan Badlani, Akshit Arora, Subhankar Ghosh, Rafael Valle, Kevin J. Shih, Joo Felipe Santos, Boris Ginsburg, Bryan Catanzaro |
| 2023 | ICASSP | High-Acoustic Fidelity Text To Speech Synthesis With Fine-Grained Control Of Speech Attributes. | Rafael Valle, Joo Felipe Santos, Kevin J. Shih, Rohan Badlani, Bryan Catanzaro |
| 2023 | Interspeech | RAD-MMM: Multilingual Multiaccented Multispeaker Text To Speech. | Rohan Badlani, Rafael Valle, Kevin J. Shih, Joo Felipe Santos, Siddharth Gururani, Bryan Catanzaro |
| 2022 | ICASSP | One TTS Alignment to Rule Them All. | Rohan Badlani, Adrian Lancucki, Kevin J. Shih, Rafael Valle, Wei Ping, Bryan Catanzaro |
| 2018 | DSAA | Pattern-Based Automatic Parallelization of Representative-Based Clustering Algorithms. | Saiyedul Islam, Sundar Balasubramaniam, Shruti Gupta, Shikhar Brajesh, Rohan Badlani, Nitin Labhishetty, Abhinav Baid, Poonam Goyal, Navneet Goyal |
| 2018 | ICASSP | Framework for Evaluation of Sound Event Detection in Web Videos. | Rohan Badlani, Ankit Shah, Benjamin Elizalde, Anurag Kumar, Bhiksha Raj |
| 2018 | ICASSP | Content-Based Representations of Audio Using Siamese Neural Networks. | Pranay Manocha, Rohan Badlani, Anurag Kumar, Ankit Shah, Benjamin Elizalde, Bhiksha Raj |