Kevin J. Shih
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
18
Venues
8
Active years
2013–2025
Best venue rank
A*
Where they publish
Papers
18 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | CVPR | Enhancing Virtual Try-On with Synthetic Pairs and Error-Aware Noise Scheduling. | Nannan Li, Kevin J. Shih, Bryan A. Plummer |
| 2025 | ICLR | Fugatto 1: Foundational Generative Audio Transformer Opus 1. | Rafael Valle, Rohan Badlani, Zhifeng Kong, Sang-gil Lee, Arushi Goel, Sungwon Kim, Joo Felipe Santos, Shuqi Dai, Siddharth Gururani, Aya Aljafari, Alexander H. Liu, Kevin J. Shih, Ryan Prenger, Wei Ping, Chao-Han Huck Yang, Bryan Catanzaro |
| 2023 | ICASSP | Vani: Very-Lightweight Accent-Controllable TTS for Native And Non-Native Speakers With Identity Preservation. | Rohan Badlani, Akshit Arora, Subhankar Ghosh, Rafael Valle, Kevin J. Shih, Joo Felipe Santos, Boris Ginsburg, Bryan Catanzaro |
| 2023 | ICASSP | High-Acoustic Fidelity Text To Speech Synthesis With Fine-Grained Control Of Speech Attributes. | Rafael Valle, Joo Felipe Santos, Kevin J. Shih, Rohan Badlani, Bryan Catanzaro |
| 2023 | ICCV | Collecting The Puzzle Pieces: Disentangled Self-Driven Human Pose Transfer by Permuting Textures. | Nannan Li, Kevin J. Shih, Bryan A. Plummer |
| 2023 | Interspeech | RAD-MMM: Multilingual Multiaccented Multispeaker Text To Speech. | Rohan Badlani, Rafael Valle, Kevin J. Shih, Joo Felipe Santos, Siddharth Gururani, Bryan Catanzaro |
| 2022 | ICASSP | One TTS Alignment to Rule Them All. | Rohan Badlani, Adrian Lancucki, Kevin J. Shih, Rafael Valle, Wei Ping, Bryan Catanzaro |
| 2021 | ICLR | Flowtron: an Autoregressive Flow-based Generative Network for Text-to-Speech Synthesis. | Rafael Valle, Kevin J. Shih, Ryan Prenger, Bryan Catanzaro |
| 2019 | CVPR | Improving Semantic Segmentation via Video Propagation and Label Relaxation. | Yi Zhu, Karan Sapra, Fitsum A. Reda, Kevin J. Shih, Shawn D. Newsam, Andrew Tao, Bryan Catanzaro |
| 2019 | CVPR | Graphical Contrastive Losses for Scene Graph Parsing. | Ji Zhang, Kevin J. Shih, Ahmed Elgammal, Andrew Tao, Bryan Catanzaro |
| 2019 | ICCV | Unsupervised Video Interpolation Using Cycle Consistency. | Fitsum A. Reda, Deqing Sun, Aysegul Dundar, Mohammad Shoeybi, Guilin Liu, Kevin J. Shih, Andrew Tao, Jan Kautz, Bryan Catanzaro |
| 2018 | AAAI | Learning Interpretable Spatial Operations in a Rich 3D Blocks World. | Yonatan Bisk, Kevin J. Shih, Yejin Choi, Daniel Marcu |
| 2018 | ECCV | Image Inpainting for Irregular Holes Using Partial Convolutions. | Guilin Liu, Fitsum A. Reda, Kevin J. Shih, Ting-Chun Wang, Andrew Tao, Bryan Catanzaro |
| 2018 | ECCV | SDC-Net: Video Prediction Using Spatially-Displaced Convolution. | Fitsum A. Reda, Guilin Liu, Kevin J. Shih, Robert Kirby, Jon Barker, David Tarjan, Andrew Tao, Bryan Catanzaro |
| 2017 | ICCV | Aligned Image-Word Representations Improve Inductive Transfer Across Vision-Language Tasks. | Tanmay Gupta, Kevin J. Shih, Saurabh Singh, Derek Hoiem |
| 2016 | CVPR | Where to Look: Focus Regions for Visual Question Answering. | Kevin J. Shih, Saurabh Singh, Derek Hoiem |
| 2015 | BMVC | Part Localization using Multi-Proposal Consensus for Fine-Grained Categorization. | Kevin J. Shih, Arun Mallya, Saurabh Singh, Derek Hoiem |
| 2013 | CVPR | Learning Collections of Part Models for Object Recognition. | Ian Endres, Kevin J. Shih, Johnston Jiaa, Derek Hoiem |