Skip to content

Kevin J. Shih

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

18

Venues

8

Active years

2013–2025

Best venue rank

A*

Where they publish

Papers

18 indexed papers, newest first.

YearVenueTitleAuthors
2025CVPREnhancing Virtual Try-On with Synthetic Pairs and Error-Aware Noise Scheduling.Nannan Li, Kevin J. Shih, Bryan A. Plummer
2025ICLRFugatto 1: Foundational Generative Audio Transformer Opus 1.Rafael Valle, Rohan Badlani, Zhifeng Kong, Sang-gil Lee, Arushi Goel, Sungwon Kim, Joo Felipe Santos, Shuqi Dai, Siddharth Gururani, Aya Aljafari, Alexander H. Liu, Kevin J. Shih, Ryan Prenger, Wei Ping, Chao-Han Huck Yang, Bryan Catanzaro
2023ICASSPVani: Very-Lightweight Accent-Controllable TTS for Native And Non-Native Speakers With Identity Preservation.Rohan Badlani, Akshit Arora, Subhankar Ghosh, Rafael Valle, Kevin J. Shih, Joo Felipe Santos, Boris Ginsburg, Bryan Catanzaro
2023ICASSPHigh-Acoustic Fidelity Text To Speech Synthesis With Fine-Grained Control Of Speech Attributes.Rafael Valle, Joo Felipe Santos, Kevin J. Shih, Rohan Badlani, Bryan Catanzaro
2023ICCVCollecting The Puzzle Pieces: Disentangled Self-Driven Human Pose Transfer by Permuting Textures.Nannan Li, Kevin J. Shih, Bryan A. Plummer
2023InterspeechRAD-MMM: Multilingual Multiaccented Multispeaker Text To Speech.Rohan Badlani, Rafael Valle, Kevin J. Shih, Joo Felipe Santos, Siddharth Gururani, Bryan Catanzaro
2022ICASSPOne TTS Alignment to Rule Them All.Rohan Badlani, Adrian Lancucki, Kevin J. Shih, Rafael Valle, Wei Ping, Bryan Catanzaro
2021ICLRFlowtron: an Autoregressive Flow-based Generative Network for Text-to-Speech Synthesis.Rafael Valle, Kevin J. Shih, Ryan Prenger, Bryan Catanzaro
2019CVPRImproving Semantic Segmentation via Video Propagation and Label Relaxation.Yi Zhu, Karan Sapra, Fitsum A. Reda, Kevin J. Shih, Shawn D. Newsam, Andrew Tao, Bryan Catanzaro
2019CVPRGraphical Contrastive Losses for Scene Graph Parsing.Ji Zhang, Kevin J. Shih, Ahmed Elgammal, Andrew Tao, Bryan Catanzaro
2019ICCVUnsupervised Video Interpolation Using Cycle Consistency.Fitsum A. Reda, Deqing Sun, Aysegul Dundar, Mohammad Shoeybi, Guilin Liu, Kevin J. Shih, Andrew Tao, Jan Kautz, Bryan Catanzaro
2018AAAILearning Interpretable Spatial Operations in a Rich 3D Blocks World.Yonatan Bisk, Kevin J. Shih, Yejin Choi, Daniel Marcu
2018ECCVImage Inpainting for Irregular Holes Using Partial Convolutions.Guilin Liu, Fitsum A. Reda, Kevin J. Shih, Ting-Chun Wang, Andrew Tao, Bryan Catanzaro
2018ECCVSDC-Net: Video Prediction Using Spatially-Displaced Convolution.Fitsum A. Reda, Guilin Liu, Kevin J. Shih, Robert Kirby, Jon Barker, David Tarjan, Andrew Tao, Bryan Catanzaro
2017ICCVAligned Image-Word Representations Improve Inductive Transfer Across Vision-Language Tasks.Tanmay Gupta, Kevin J. Shih, Saurabh Singh, Derek Hoiem
2016CVPRWhere to Look: Focus Regions for Visual Question Answering.Kevin J. Shih, Saurabh Singh, Derek Hoiem
2015BMVCPart Localization using Multi-Proposal Consensus for Fine-Grained Categorization.Kevin J. Shih, Arun Mallya, Saurabh Singh, Derek Hoiem
2013CVPRLearning Collections of Part Models for Object Recognition.Ian Endres, Kevin J. Shih, Johnston Jiaa, Derek Hoiem