Rohit Girdhar
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
31
Venues
7
Active years
2014–2025
Best venue rank
A*
Where they publish
Papers
31 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | CVPR | MotiF: Making Text Count in Image Animation with Motion Focal Loss. | Shijie Wang, Samaneh Azadi, Rohit Girdhar, Saketh Rambhatla, Chen Sun, Xi Yin |
| 2025 | ICML | LLMs can see and hear without any training. | Kumar Ashutosh, Yossi Gandelsman, Xinlei Chen, Ishan Misra, Rohit Girdhar |
| 2024 | CVPR | InstanceDiffusion: Instance-Level Control for Image Generation. | Xudong Wang, Trevor Darrell, Sai Saketh Rambhatla, Rohit Girdhar, Ishan Misra |
| 2024 | CVPR | VideoCutLER: Surprisingly Simple Unsupervised Video Instance Segmentation. | Xudong Wang, Ishan Misra, Ziyun Zeng, Rohit Girdhar, Trevor Darrell |
| 2024 | CVPR | SoundingActions: Learning How Actions Sound from Narrated Egocentric Videos. | Changan Chen, Kumar Ashutosh, Rohit Girdhar, David Harwath, Kristen Grauman |
| 2024 | CVPR | Generating Illustrated Instructions. | Sachit Menon, Ishan Misra, Rohit Girdhar |
| 2024 | ECCV | Factorizing Text-to-Video Generation by Explicit Image Conditioning. | Rohit Girdhar, Mannat Singh, Andrew Brown, Quentin Duval, Samaneh Azadi, Sai Saketh Rambhatla, Akbar Shah, Xi Yin, Devi Parikh, Ishan Misra |
| 2023 | CVPR | Learning Video Representations from Large Language Models. | Yue Zhao, Ishan Misra, Philipp Krhenbhl, Rohit Girdhar |
| 2023 | CVPR | Cut and Learn for Unsupervised Object Detection and Instance Segmentation. | Xudong Wang, Rohit Girdhar, Stella X. Yu, Ishan Misra |
| 2023 | CVPR | HierVL: Learning Hierarchical Video-Language Embeddings. | Kumar Ashutosh, Rohit Girdhar, Lorenzo Torresani, Kristen Grauman |
| 2023 | CVPR | ImageBind One Embedding Space to Bind Them All. | Rohit Girdhar, Alaaeldin El-Nouby, Zhuang Liu, Mannat Singh, Kalyan Vasudev Alwala, Armand Joulin, Ishan Misra |
| 2023 | CVPR | OmniMAE: Single Model Masked Pretraining on Images and Videos. | Rohit Girdhar, Alaaeldin El-Nouby, Mannat Singh, Kalyan Vasudev Alwala, Armand Joulin, Ishan Misra |
| 2023 | ICCV | The effectiveness of MAE pre-pretraining for billion-scale pretraining. | Mannat Singh, Quentin Duval, Kalyan Vasudev Alwala, Haoqi Fan, Vaibhav Aggarwal, Aaron Adcock, Armand Joulin, Piotr Dollr, Christoph Feichtenhofer, Ross B. Girshick, Rohit Girdhar, Ishan Misra |
| 2022 | CVPR | Masked-attention Mask Transformer for Universal Image Segmentation. | Bowen Cheng, Ishan Misra, Alexander G. Schwing, Alexander Kirillov, Rohit Girdhar |
| 2022 | CVPR | Omnivore: A Single Model for Many Visual Modalities. | Rohit Girdhar, Mannat Singh, Nikhila Ravi, Laurens van der Maaten, Armand Joulin, Ishan Misra |
| 2022 | CVPR | Ego4D: Around the World in 3, 000 Hours of Egocentric Video. | Kristen Grauman, Andrew Westbury, Eugene Byrne, Zachary Chavis, Antonino Furnari, Rohit Girdhar, Jackson Hamburger, Hao Jiang, Miao Liu, Xingyu Liu, Miguel Martin, Tushar Nagarajan, Ilija Radosavovic, Santhosh Kumar Ramakrishnan, Fiona Ryan, Jayant Sharma, Michael Wray, Mengmeng Xu, Eric Zhongcong Xu, Chen Zhao, Siddhant Bansal, Dhruv Batra, Vincent Cartillier, Sean Crane, Tien Do, Morrie Doulaty, Akshay Erapalli, Christoph Feichtenhofer, Adriano Fragomeni, Qichen Fu, Abrham Gebreselasie, Cristina Gonzlez, James Hillis, Xuhua Huang, Yifei Huang, Wenqi Jia, Weslie Khoo, Jchym Kolr, Satwik Kottur, Anurag Kumar, Federico Landini, Chao Li, Yanghao Li, Zhenqiang Li, Karttikeya Mangalam, Raghava Modhugu, Jonathan Munro, Tullie Murrell, Takumi Nishiyasu, Will Price, Paola Ruiz Puentes, Merey Ramazanova, Leda Sari, Kiran K. Somasundaram, Audrey Southerland, Yusuke Sugano, Ruijie Tao, Minh Vo, Yuchen Wang, Xindi Wu, Takuma Yagi, Ziwei Zhao, Yunyi Zhu, Pablo Arbelez, David Crandall, Dima Damen, Giovanni Maria Farinella, Christian Fuegen, Bernard Ghanem, Vamsi Krishna Ithapu, C. V. Jawahar, Hanbyul Joo, Kris Kitani, Haizhou Li, Richard A. Newcombe, Aude Oliva, Hyun Soo Park, James M. Rehg, Yoichi Sato, Jianbo Shi, Mike Zheng Shou, Antonio Torralba, Lorenzo Torresani, Mingfei Yan, Jitendra Malik |
| 2022 | ECCV | Detecting Twenty-Thousand Classes Using Image-Level Supervision. | Xingyi Zhou, Rohit Girdhar, Armand Joulin, Philipp Krhenbhl, Ishan Misra |
| 2021 | CVPR | 3D Spatial Recognition Without Spatially Labeled 3D. | Zhongzheng Ren, Ishan Misra, Alexander G. Schwing, Rohit Girdhar |
| 2021 | ICCV | Anticipative Video Transformer. | Rohit Girdhar, Kristen Grauman |
| 2021 | ICCV | An End-to-End Transformer Model for 3D Object Detection. | Ishan Misra, Rohit Girdhar, Armand Joulin |
| 2021 | ICCV | Self-Supervised Pretraining of 3D Features on any Point-Cloud. | Zaiwei Zhang, Rohit Girdhar, Armand Joulin, Ishan Misra |
| 2020 | ICLR | CATER: A diagnostic dataset for Compositional Actions & TEmporal Reasoning. | Rohit Girdhar, Deva Ramanan |
| 2020 | ICLR | MetaPix: Few-Shot Video Retargeting. | Jessica Lee, Deva Ramanan, Rohit Girdhar |
| 2019 | CVPR | Video Action Transformer Network. | Rohit Girdhar, Joo Carreira, Carl Doersch, Andrew Zisserman |
| 2019 | ICCV | DistInit: Learning Video Representations Without a Single Labeled Video. | Rohit Girdhar, Du Tran, Lorenzo Torresani, Deva Ramanan |
| 2018 | CVPR | Detect-and-Track: Efficient Pose Estimation in Videos. | Rohit Girdhar, Georgia Gkioxari, Lorenzo Torresani, Manohar Paluri, Du Tran |
| 2017 | CVPR | Binge Watching: Scaling Affordance Learning from Sitcoms. | Xiaolong Wang, Rohit Girdhar, Abhinav Gupta |
| 2017 | CVPR | ActionVLAD: Learning Spatio-Temporal Aggregation for Action Classification. | Rohit Girdhar, Deva Ramanan, Abhinav Gupta, Josef Sivic, Bryan C. Russell |
| 2016 | ECCV | Learning a Predictable and Generative Vector Representation for Objects. | Rohit Girdhar, David F. Fouhey, Mikel Rodriguez, Abhinav Gupta |
| 2016 | WACV | Cutting through the clutter: Task-relevant features for image matching. | Rohit Girdhar, David F. Fouhey, Kris M. Kitani, Abhinav Gupta, Martial Hebert |
| 2014 | ACCV | Optimizing Storage Intensive Vision Applications to Device Capacity. | Rohit Girdhar, Jayaguru Panda, C. V. Jawahar |