| 2026 | WACV | RobustGait: Robustness Analysis for Appearance Based Gait Recognition. | Reeshoon Sayera, Akash Kumar, Sirshapan Mitra, Prudvi Kamtam, Yogesh S. Rawat |
| 2025 | CVPR | A Large-Scale Analysis on Contextual Self-Supervised Video Representation Learning. | Akash Kumar, Ashlesha Kumar, Vibhav Vineet, Yogesh S. Rawat |
| 2025 | CVPR | Understanding Depth and Height Perception in Large Visual-Language Models. | Shehreen Azad, Yash Jain, Rishit Garg, Vibhav Vineet, Yogesh S. Rawat |
| 2025 | CVPR | STPro: Spatial and Temporal Progressive Learning for Weakly Supervised Spatio-Temporal Grounding. | Aaryan Garg, Akash Kumar, Yogesh S. Rawat |
| 2025 | CVPR | DIFFER: Disentangling Identity Features via Semantic Cues for Clothes-Changing Person Re-ID. | Xin Liang, Yogesh S. Rawat |
| 2025 | ICCV | Punching Bag vs. Punching Person: Motion Transferability in Videos. | Raiyaan Abdullah, Jared Claypoole, Michael Cogswell, Ajay Divakaran, Yogesh S. Rawat |
| 2025 | ICCV | GaitCrafter: Diffusion Model for Biometric Preserving Gait Synthesis. | Sirshapan Mitra, Yogesh S. Rawat |
| 2025 | ICCV | Colors See Colors Ignore: Clothes Changing ReID with Color Disentanglement. | Priyank Pathak, Yogesh S. Rawat |
| 2025 | ICCV | OmViD: Omni-Supervised Active Learning for Video Action Detection. | Aayush Jung Rana, Akash Kumar, Vibhav Vineet, Yogesh S. Rawat |
| 2025 | ICLR | Contextual Self-paced Learning for Weakly Supervised Spatio-Temporal Video Grounding. | Akash Kumar, Zsolt Kira, Yogesh S. Rawat |
| 2025 | ICLR | Lr0.Fm: low-Resolution Zero-Shot Classification Benchmark for Foundation Models. | Priyank Pathak, Shyam Marjit, Shruti Vyas, Yogesh S. Rawat |
| 2024 | CVPR | Probing Conceptual Understanding of Large Visual-Language Models. | Madeline Schiappa, Raiyaan Abdullah, Shehreen Azad, Jared Claypoole, Michael Cogswell, Ajay Divakaran, Yogesh S. Rawat |
| 2024 | CVPR | Robustness Analysis on Foundational Segmentation Models. | Madeline Chantry Schiappa, Shehreen Azad, Sachidanand VS, Yunhao Ge, Ondrej Miksik, Yogesh S. Rawat, Vibhav Vineet |
| 2024 | EMNLP | Navigating Hallucinations for Reasoning of Unintentional Activities. | Shresth Grover, Vibhav Vineet, Yogesh S. Rawat |
| 2023 | CVPR | Hybrid Active Learning via Deep Clustering for Video Action Detection. | Aayush Jung Rana, Yogesh S. Rawat |
| 2023 | CVPR | A Large-Scale Robustness Analysis of Video Action Recognition Models. | Madeline Chantry Schiappa, Naman Biyani, Prudvi Kamtam, Shruti Vyas, Hamid Palangi, Vibhav Vineet, Yogesh S. Rawat |
| 2021 | AAAI | SSA2D: Single Shot Actor-Action Detection in Videos (Student Abstract). | Aayush Jung Rana, Yogesh S. Rawat |
| 2021 | BMVC | LARNet: Latent Action Representation for Human Action Synthesis. | Naman Biyani, Aayush Jung Rana, Shruti Vyas, Yogesh S. Rawat |
| 2021 | CVPR | PLM: Partial Label Masking for Imbalanced Multi-Label Classification. | Kevin Duarte, Yogesh S. Rawat, Mubarak Shah |
| 2021 | CVPR | Modeling Multi-Label Action Dependencies for Temporal Action Localization. | Praveen Tirupattur, Kevin Duarte, Yogesh S. Rawat, Mubarak Shah |
| 2021 | ICIP | Novel View Video Prediction using a Dual Representation. | Sarah Shiraz, Krishna Regmi, Shruti Vyas, Yogesh S. Rawat, Mubarak Shah |
| 2021 | ICIP | Unsupervised Discriminative Embedding For Sub-Action Learning in Complex Activities. | Sirnam Swetha, Hilde Kuehne, Yogesh S. Rawat, Mubarak Shah |
| 2021 | ICLR | In Defense of Pseudo-Labeling: An Uncertainty-Aware Pseudo-label Selection Framework for Semi-Supervised Learning. | Mamshad Nayeem Rizve, Kevin Duarte, Yogesh S. Rawat, Mubarak Shah |
| 2021 | WACV | We don't Need Thousand Proposals: Single Shot Actor-Action Detection in Videos. | Aayush Jung Rana, Yogesh S. Rawat |
| 2020 | CVPR | Visual-Textual Capsule Routing for Text-Based Video Segmentation. | Bruce McIntosh, Kevin Duarte, Yogesh S. Rawat, Mubarak Shah |
| 2020 | ECCV | A Recurrent Transformer Network for Novel View Action Synthesis. | Kara Marie Schatz, Erik Quintanilla, Shruti Vyas, Yogesh S. Rawat |
| 2020 | ECCV | Multi-view Action Recognition Using Cross-View Video Prediction. | Shruti Vyas, Yogesh S. Rawat, Mubarak Shah |
| 2020 | ICPR | TinyVIRAT: Low-resolution Video Action Recognition. | Ugur Demir, Yogesh S. Rawat, Mubarak Shah |
| 2020 | ICPR | Gabriella: An Online System for Real-Time Activity Detection in Untrimmed Security Videos. | Mamshad Nayeem Rizve, Ugur Demir, Praveen Tirupattur, Aayush Jung Rana, Kevin Duarte, Ishan R. Dave, Yogesh S. Rawat, Mubarak Shah |
| 2019 | ICCV | CapsuleVOS: Semi-Supervised Video Object Segmentation Using Capsule Routing. | Kevin Duarte, Yogesh S. Rawat, Mubarak Shah |