| 2024 | CVPR | Large Language Models are Good Prompt Learners for Low-Shot Image Classification. | Zhaoheng Zheng, Jingmin Wei, Xuefeng Hu, Haidong Zhu, Ram Nevatia |
| 2024 | CVPR | SEAS: ShapE-Aligned Supervision for Person Re-Identification. | Haidong Zhu, Pranav Budhwant, Zhaoheng Zheng, Ram Nevatia |
| 2024 | ECCV | CaesarNeRF: Calibrated Semantic Representation for Few-Shot Generalizable Neural Rendering. | Haidong Zhu, Tianyu Ding, Tianyi Chen, Ilya Zharkov, Ram Nevatia, Luming Liang |
| 2024 | WACV | ReCLIP: Refine Contrastive Language Image Pre-Training with Source Free Domain Adaptation. | Xuefeng Hu, Ke Zhang, Lu Xia, Albert Chen, Jiajia Luo, Yuyin Sun, Ken Wang, Nan Qiao, Xiao Zeng, Min Sun, Cheng-Hao Kuo, Ram Nevatia |
| 2024 | WACV | Efficient Feature Distillation for Zero-shot Annotation Object Detection. | Zhuoming Liu, Xuefeng Hu, Ram Nevatia |
| 2024 | WACV | Leveraging Task-Specific Pre-Training to Reason across Images and Videos. | Arka Sadhu, Ram Nevatia |
| 2024 | WACV | CAILA: Concept-Aware Intra-Layer Adapters for Compositional Zero-Shot Learning. | Zhaoheng Zheng, Haidong Zhu, Ram Nevatia |
| 2024 | WACV | ShARc: Shape and Appearance Recognition for Person Identification In-the-wild. | Haidong Zhu, Wanrong Zheng, Zhaoheng Zheng, Ram Nevatia |
| 2023 | CVPR | CAT-NeRF: Constancy-Aware Tx | Haidong Zhu, Zhaoheng Zheng, Wanrong Zheng, Ram Nevatia |
| 2023 | ICRA | Multimodal Neural Radiance Field. | Haidong Zhu, Yuyin Sun, Chi Liu, Lu Xia, Jiajia Luo, Nan Qiao, Ram Nevatia, Cheng-Hao Kuo |
| 2023 | WACV | PatchZero: Defending against Adversarial Patch Attacks by Detecting and Zeroing the Patch. | Ke Xu, Yao Xiao, Zhaoheng Zheng, Kaijie Cai, Ram Nevatia |
| 2023 | WACV | Gait Recognition Using 3-D Human Body Shape Inference. | Haidong Zhu, Zhaoheng Zheng, Ram Nevatia |
| 2022 | ICASSP | Self-Supervised Learning for Sentiment Analysis via Image-Text Matching. | Haidong Zhu, Zhaoheng Zheng, Mohammad Soleymani, Ram Nevatia |
| 2022 | ICPR | Improving Weakly Supervised Scene Graph Parsing through Object Grounding. | Yizhou Zhang, Zhaoheng Zheng, Ram Nevatia, Yan Liu |
| 2022 | ICPR | OPEN: Order-preserving Pointcloud Encoder Decoder Network for Body Shape Refinement. | Haidong Zhu, Ye Yuan, Yiheng Zhu, Xiao Yang, Ram Nevatia |
| 2022 | ICPR | Temporal Shift and Attention Modules for Graphical Skeleton Action Recognition. | Haidong Zhu, Zhaoheng Zheng, Ram Nevatia |
| 2021 | CVPR | SimPLE: Similar Pseudo Label Exploitation for Semi-Supervised Classification. | Zijian Hu, Zhengyu Yang, Xuefeng Hu, Ram Nevatia |
| 2021 | CVPR | Visual Semantic Role Labeling for Video Understanding. | Arka Sadhu, Tanmay Gupta, Mark Yatskar, Ram Nevatia, Aniruddha Kembhavi |
| 2021 | ICIP | Improving Object Detection And Attribute Recognition By Feature Entanglement Reduction. | Zhaoheng Zheng, Arka Sadhu, Ram Nevatia |
| 2021 | NAACL | Video Question Answering with Phrases via Semantic Roles. | Arka Sadhu, Kan Chen, Ram Nevatia |
| 2021 | WACV | Utilizing Every Image Object for Semi-supervised Phrase Grounding. | Haidong Zhu, Arka Sadhu, Zhaoheng Zheng, Ram Nevatia |
| 2020 | CVPR | CPARR: Category-based Proposal Analysis for Referring Relationships. | Chuanzi He, Haidong Zhu, Jiyang Gao, Kan Chen, Ram Nevatia |
| 2020 | CVPR | Video Object Grounding Using Semantic Roles in Language Description. | Arka Sadhu, Kan Chen, Ram Nevatia |
| 2020 | ECCV | Curriculum DeepSDF. | Yueqi Duan, Haidong Zhu, He Wang, Li Yi, Ram Nevatia, Leonidas J. Guibas |
| 2020 | ECCV | SPAN: Spatial Pyramid Attention Network for Image Manipulation Localization. | Xuefeng Hu, Zhihan Zhang, Zhenye Jiang, Syomantak Chaudhuri, Zhenheng Yang, Ram Nevatia |
| 2020 | EMNLP | Visually Grounded Continual Learning of Compositional Phrases. | Xisen Jin, Junyi Du, Arka Sadhu, Ram Nevatia, Xiang Ren |
| 2019 | CVPR | Activity Driven Weakly Supervised Object Detection. | Zhenheng Yang, Dhruv Mahajan, Deepti Ghadiyaram, Ram Nevatia, Vignesh Ramanathan |
| 2019 | ICCV | NOTE-RCNN: NOise Tolerant Ensemble RCNN for Semi-Supervised Object Detection. | Jiyang Gao, Jiang Wang, Shengyang Dai, Li-Jia Li, Ram Nevatia |
| 2019 | ICCV | Zero-Shot Grounding of Objects From Natural Language Queries. | Arka Sadhu, Kan Chen, Ram Nevatia |
| 2019 | WACV | MAC: Mining Activity Concepts for Language-Based Temporal Localization. | Runzhou Ge, Jiyang Gao, Kan Chen, Ram Nevatia |
| 2018 | AAAI | SPOT Poachers in Action: Augmenting Conservation Drones With Automatic Detection in Near Real Time. | Elizabeth Bondi, Fei Fang, Mark Hamilton, Debarun Kar, Donnabell Dmello, Jongmoo Choi, Robert Hannaford, Arvind Iyer, Lucas Joppa, Milind Tambe, Ram Nevatia |
| 2018 | ACCV | PIRC Net: Using Proposal Indexing, Relationships and Context for Phrase Grounding. | Rama Kovvuri, Ram Nevatia |
| 2018 | CVPR | Knowledge Aided Consistency for Weakly Supervised Phrase Grounding. | Kan Chen, Jiyang Gao, Ram Nevatia |
| 2018 | CVPR | Motion-Appearance Co-Memory Networks for Video Question Answering. | Jiyang Gao, Runzhou Ge, Kan Chen, Ram Nevatia |
| 2018 | CVPR | LEGO: Learning Edge With Geometry All at Once by Watching Videos. | Zhenheng Yang, Peng Wang, Yang Wang, Wei Xu, Ram Nevatia |
| 2018 | ECCV | Visually Indicated Sound Generation by Perceptually Optimized Classification. | Kan Chen, Chuanxi Zhang, Chen Fang, Zhaowen Wang, Trung Bui, Ram Nevatia |
| 2018 | ECCV | CTAP: Complementary Temporal Action Proposal Generation. | Jiyang Gao, Kan Chen, Ram Nevatia |
| 2018 | ECCV | Every Pixel Counts: Unsupervised Geometry Learning with Holistic 3D Motion Understanding. | Zhenheng Yang, Peng Wang, Yang Wang, Wei Xu, Ram Nevatia |
| 2017 | AAAI | DECK: Discovering Event Composition Knowledge from Web Images for Zero-Shot Event Detection and Recounting in Videos. | Chuang Gan, Chen Sun, Ram Nevatia |
| 2017 | BMVC | Cascaded Boundary Regression for Temporal Action Detection. | Jiyang Gao, Zhenheng Yang, Ram Nevatia |
| 2017 | BMVC | RED: Reinforced Encoder-Decoder Networks for Action Anticipation. | Jiyang Gao, Zhenheng Yang, Ram Nevatia |
| 2017 | BMVC | Spatio-Temporal Action Detection with Cascade Proposal and Location Anticipation. | Zhenheng Yang, Jiyang Gao, Ram Nevatia |
| 2017 | CVPR | AMC: Attention Guided Multi-modal Correlation Learning for Image Search. | Kan Chen, Trung Bui, Chen Fang, Zhaowen Wang, Ram Nevatia |
| 2017 | ICCV | Query-Guided Regression Network with Context Policy for Phrase Grounding. | Kan Chen, Rama Kovvuri, Ram Nevatia |
| 2017 | ICCV | TALL: Temporal Activity Localization via Language Query. | Jiyang Gao, Chen Sun, Zhenheng Yang, Ram Nevatia |
| 2017 | ICCV | TURN TAP: Temporal Unit Regression Network for Temporal Action Proposals. | Jiyang Gao, Zhenheng Yang, Chen Sun, Kan Chen, Ram Nevatia |
| 2016 | ACCV | Image Set Classification via Template Triplets and Context-Aware Similarity Embedding. | Feng-Ju Chang, Ram Nevatia |
| 2016 | ACCV | Learning Action Concept Trees and Semantic Alignment Networks from Image-Description Data. | Jiyang Gao, Ram Nevatia |
| 2016 | CVPR | ProNet: Learning to Propose Object-Specific Boxes for Cascaded Neural Networks. | Chen Sun, Manohar Paluri, Ronan Collobert, Ram Nevatia, Lubomir D. Bourdev |
| 2016 | ICPR | Exploring deep learning based solutions in fine grained activity recognition in the wild. | Song Cao, Ram Nevatia |
| 2016 | ICPR | Segment-based models for event detection and recounting. | Rama Kovvuri, Ram Nevatia, Cees G. M. Snoek |
| 2016 | WACV | Face recognition using deep multi-pose representations. | Wael Abd-Almageed, Yue Wu, Stephen Rawls, Shai Harel, Tal Hassner, Iacopo Masi, Jongmoo Choi, Jatuporn Toy Leksut, Jungyeon Kim, Prem Natarajan, Ram Nevatia, Grard G. Medioni |
| 2016 | WACV | Tag-based video retrieval by embedding semantic content in a continuous word space. | Arnav Agharwal, Rama Kovvuri, Ram Nevatia, Cees G. M. Snoek |
| 2016 | WACV | Abstraction hierarchy and self annotation update for fine grained activity recognition. | Song Cao, Kan Chen, Ram Nevatia |
| 2016 | WACV | Activity recognition and prediction with pose based discriminative patch model. | Song Cao, Kan Chen, Ram Nevatia |
| 2015 | ICCV | Automatic Concept Discovery from Parallel Text and Visual Corpora. | Chen Sun, Chuang Gan, Ram Nevatia |
| 2015 | WACV | Forecasting Human Pose and Motion with Multibody Dynamic Model. | Song Cao, Ram Nevatia |
| 2015 | WACV | A Robust Adaptive Classifier for Detector Adaptation in a Video. | Pramod Sharma, Ram Nevatia |
| 2015 | WACV | Beyond Pedestrians: A Hybrid Approach of Tracking Multiple Articulating Humans. | Weijun Wang, Ram Nevatia, Bo Yang |
| 2014 | ACCV | Multi-state Discriminative Video Segment Selection for Complex Event Classification. | Prithviraj Banerjee, Ram Nevatia |
| 2014 | ECCV | Semantic Aware Video Transcription Using Random Forest Classifiers. | Chen Sun, Ram Nevatia |
| 2013 | AVSS | Conditional Bayesian networks for action detection. | Furqan M. Khan, Sung Chun Lee, Ram Nevatia |
| 2013 | CVPR | Efficient Detector Adaptation for Object Detection in a Video. | Pramod Sharma, Ram Nevatia |
| 2013 | ICCV | ACTIVE: Activity Concept Transitions in Video Event Classification. | Chen Sun, Ram Nevatia |
| 2013 | WACV | Large-scale web video event classification by use of Fisher Vectors. | Chen Sun, Ram Nevatia |
| 2012 | CVPR | Unsupervised incremental learning for improved object detection in a video. | Pramod Sharma, Chang Huang, Ram Nevatia |
| 2012 | CVPR | Multi-target tracking by online learning of non-linear motion patterns and robust appearance models. | Bo Yang, Ram Nevatia |
| 2012 | CVPR | An online learned CRF model for multi-target tracking. | Bo Yang, Ram Nevatia |
| 2012 | ECCV | Online Learned Discriminative Part-Based Appearance Models for Multi-human Tracking. | Bo Yang, Ram Nevatia |
| 2012 | ICPR | Robust multi-pose face tracking by multi-stage tracklet association. | Markus Roth, Martin Buml, Ram Nevatia, Rainer Stiefelhagen |
| 2012 | ICPR | Efficient incremental learning of boosted classifiers for object detection. | Pramod Sharma, Chang Huang, Ram Nevatia |
| 2012 | WACV | Simultaneous inference of activity, pose and object. | Furqan M. Khan, Vivek Kumar Singh, Ram Nevatia |
| 2012 | WACV | A systems level approach to perimeter protection. | Peter H. Tu, Ting Yu, Dashan Gao, Ram Nevatia, Sung Chun Lee, Hale Kim, Phill-Kyu Rhee, Joong-Hwan Baek |
| 2011 | AVSS | Learning neighborhood cooccurrence statistics of sparse features for human activity recognition. | Prithviraj Banerjee, Ram Nevatia |
| 2011 | AVSS | AVSS 2011 demo session: A systems level approach to perimeter protection. | Peter H. Tu, Ting Yu, Dashan Gao, Ram Nevatia, Sung Chun Lee, Hale Kim, Phill-Kyu Rhee, Joong-Hwan Baek |
| 2011 | CVPR | How does person identity recognition help multi-person tracking? | Cheng-Hao Kuo, Ram Nevatia |
| 2011 | CVPR | Learning affinities and dependencies for multi-target tracking using a CRF model. | Bo Yang, Chang Huang, Ram Nevatia |
| 2011 | ICCV | Action recognition in cluttered dynamic scenes using Pose-Specific Part Models. | Vivek Kumar Singh, Ram Nevatia |
| 2011 | WACV | Vehicle detection from low quality aerial LIDAR data. | Bo Yang, Pramod Sharma, Ram Nevatia |
| 2010 | AVSS | Dynamics Based Trajectory Segmentation for UAV videos. | Prithviraj Banerjee, Ram Nevatia |
| 2010 | CVPR | Multi-target tracking by on-line learned discriminative appearance models. | Cheng-Hao Kuo, Chang Huang, Ram Nevatia |
| 2010 | CVPR | Learning 3D action models from a few 2D videos for view invariant action recognition. | Pradeep Natarajan, Vivek Kumar Singh, Ram Nevatia |
| 2010 | CVPR | Multiple pose context trees for estimating human pose in object context. | Vivek Kumar Singh, Furqan Muhammad Khan, Ram Nevatia |
| 2010 | ECCV | Inter-camera Association of Multi-target Tracks by On-Line Learned Appearance Affinity Models. | Cheng-Hao Kuo, Chang Huang, Ram Nevatia |
| 2010 | ECCV | Efficient Inference with Multiple Heterogeneous Part Detectors for Human Pose Estimation. | Vivek Kumar Singh, Ram Nevatia, Chang Huang |
| 2009 | CVPR | Learning to associate: HybridBoosted multi-target tracker for crowded scene. | Yuan Li, Chang Huang, Ram Nevatia |
| 2009 | WACV | Extensive articulated human detection by voting Cluster Boosted Tree. | Bo Yang, Chang Huang, Ram Nevatia |
| 2008 | CVPR | View and scale invariant action recognition using multiview shape-flow models. | Pradeep Natarajan, Ram Nevatia |
| 2008 | CVPR | Optimizing discrimination-efficiency tradeoff in integrating heterogeneous local features for object detection. | Bo Wu, Ram Nevatia |
| 2008 | CVPR | Segmentation of multiple, partially occluded objects by grouping, merging, assigning part detection responses. | Bo Wu, Ram Nevatia, Yuan Li |
| 2008 | ECCV | Key Object Driven Multi-category Object Recognition, Localization and Tracking Using Spatio-temporal Context. | Yuan Li, Ram Nevatia |
| 2008 | ICPR | Human detection by searching in 3d space using camera and scene knowledge. | Yuan Li, Bo Wu, Ram Nevatia |
| 2007 | CVPR | Simultaneous Object Detection and Segmentation by Boosting Local Shape Feature based Classifier. | Bo Wu, Ram Nevatia |
| 2007 | CVPR | Improving Part based Object Detection by Unsupervised, Online Boosting. | Bo Wu, Ram Nevatia |
| 2007 | CVPR | Pedestrian Detection in Infrared Images based on Local Shape Features. | Li Zhang, Bo Wu, Ram Nevatia |
| 2006 | CVPR | Tracking of Multiple, Partially Occluded Humans based on Static Body Part Detection. | Bo Wu, Ram Nevatia |
| 2006 | CVPR | Tracking of Multiple Humans in Meetings. | Bo Wu, Ram Nevatia |
| 2004 | CVPR | An Ontology for Video Event Representation. | Ram Nevatia, Jerry R. Hobbs, Bob Bolles |
| 2003 | CVPR | Hierarchical Language-based Representation of Events in Video Streams. | Ram Nevatia, Tao Zhao, Somboon Hongeng |
| 1994 | ICPR | Parallel processing for spatial grouping and matching. | Ram Nevatia, Craig C. Reinhart |
| 1979 | IJCAI | Describing Natural Textures. | Ram Nevatia, Keith E. Price, Felicia M. Vilnrotter |