Skip to content

Rohan Badlani

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

10

Venues

5

Active years

2018–2025

Best venue rank

A*

Where they publish

Papers

10 indexed papers, newest first.

YearVenueTitleAuthors
2025ICLRFugatto 1: Foundational Generative Audio Transformer Opus 1.Rafael Valle, Rohan Badlani, Zhifeng Kong, Sang-gil Lee, Arushi Goel, Sungwon Kim, Joo Felipe Santos, Shuqi Dai, Siddharth Gururani, Aya Aljafari, Alexander H. Liu, Kevin J. Shih, Ryan Prenger, Wei Ping, Chao-Han Huck Yang, Bryan Catanzaro
2024ICASSPScaling Nvidia's Multi-Speaker Multi-Lingual TTS Systems With Zero-Shot TTS to Indic Languages.Akshit Arora, Rohan Badlani, Sungwon Kim, Rafael Valle, Bryan Catanzaro
2024ICMLAudio Flamingo: A Novel Audio Language Model with Few-Shot Learning and Dialogue Abilities.Zhifeng Kong, Arushi Goel, Rohan Badlani, Wei Ping, Rafael Valle, Bryan Catanzaro
2023ICASSPVani: Very-Lightweight Accent-Controllable TTS for Native And Non-Native Speakers With Identity Preservation.Rohan Badlani, Akshit Arora, Subhankar Ghosh, Rafael Valle, Kevin J. Shih, Joo Felipe Santos, Boris Ginsburg, Bryan Catanzaro
2023ICASSPHigh-Acoustic Fidelity Text To Speech Synthesis With Fine-Grained Control Of Speech Attributes.Rafael Valle, Joo Felipe Santos, Kevin J. Shih, Rohan Badlani, Bryan Catanzaro
2023InterspeechRAD-MMM: Multilingual Multiaccented Multispeaker Text To Speech.Rohan Badlani, Rafael Valle, Kevin J. Shih, Joo Felipe Santos, Siddharth Gururani, Bryan Catanzaro
2022ICASSPOne TTS Alignment to Rule Them All.Rohan Badlani, Adrian Lancucki, Kevin J. Shih, Rafael Valle, Wei Ping, Bryan Catanzaro
2018DSAAPattern-Based Automatic Parallelization of Representative-Based Clustering Algorithms.Saiyedul Islam, Sundar Balasubramaniam, Shruti Gupta, Shikhar Brajesh, Rohan Badlani, Nitin Labhishetty, Abhinav Baid, Poonam Goyal, Navneet Goyal
2018ICASSPFramework for Evaluation of Sound Event Detection in Web Videos.Rohan Badlani, Ankit Shah, Benjamin Elizalde, Anurag Kumar, Bhiksha Raj
2018ICASSPContent-Based Representations of Audio Using Siamese Neural Networks.Pranay Manocha, Rohan Badlani, Anurag Kumar, Ankit Shah, Benjamin Elizalde, Bhiksha Raj