| 2026 | ACL | Zero-Shot Multimodal Retrieval with Multi-Scale Contextual Representations. | Sourajit Saha, Tejas Gokhale |
| 2025 | EMNLP | Side Effects of Erasing Concepts from Diffusion Models. | Shaswati Saha, Sourajit Saha, Manas Gaur, Tejas Gokhale |
| 2025 | ICLR | VOILA: Evaluation of MLLMs For Perceptual Understanding and Analogical Reasoning. | Nilay Yilmaz, Maitreya Patel, Yiran Lawrence Luo, Tejas Gokhale, Chitta Baral, Suren Jayasuriya, Yezhou Yang |
| 2025 | WACV | Improving Shift Invariance in Convolutional Neural Networks with Translation Invariant Polyphase Sampling. | Sourajit Saha, Tejas Gokhale |
| 2024 | AAAI | Towards Robust Visual Understanding: from Recognition to Reasoning. | Tejas Gokhale |
| 2024 | AAAI | ConceptBed: Evaluating Concept Learning Abilities of Text-to-Image Diffusion Models. | Maitreya Patel, Tejas Gokhale, Chitta Baral, Yezhou Yang |
| 2024 | CVPR | On the Robustness of Language Guidance for Low-Level Vision Tasks: Findings from Depth Estimation. | Agneet Chatterjee, Tejas Gokhale, Chitta Baral, Yezhou Yang |
| 2024 | CVPR | Grounding Stylistic Domain Generalization with Quantitative Domain Shift Measures and Synthetic Scene Images. | Yiran Luo, Joshua Feinglass, Tejas Gokhale, Kuan-Cheng Lee, Chitta Baral, Yezhou Yang |
| 2024 | ECCV | REVISION: Rendering Tools Enable Spatial Fidelity in Vision-Language Models. | Agneet Chatterjee, Yiran Luo, Tejas Gokhale, Yezhou Yang, Chitta Baral |
| 2024 | ECCV | Getting it Right: Improving Spatial Consistency in Text-to-Image Models. | Agneet Chatterjee, Gabriela Ben Melech Stan, Estelle Aflalo, Sayak Paul, Dhruba Ghosh, Tejas Gokhale, Ludwig Schmidt, Hannaneh Hajishirzi, Vasudev Lal, Chitta Baral, Yezhou Yang |
| 2023 | ACL | End-to-end Knowledge Retrieval with Multi-modal Queries. | Man Luo, Zhiyuan Fang, Tejas Gokhale, Yezhou Yang, Chitta Baral |
| 2023 | ICCV | Adversarial Bayesian Augmentation for Single-Source Domain Generalization. | Sheng Cheng, Tejas Gokhale, Yezhou Yang |
| 2023 | WACV | Improving Diversity with Adversarially Learned Transformations for Domain Generalization. | Tejas Gokhale, Rushil Anirudh, Jayaraman J. Thiagarajan, Bhavya Kailkhura, Chitta Baral, Yezhou Yang |
| 2022 | AAAI | Improving Biomedical Information Retrieval with Neural Retrievers. | Man Luo, Arindam Mitra, Tejas Gokhale, Chitta Baral |
| 2022 | ACL | Semantically Distributed Robust Optimization for Vision-and-Language Inference. | Tejas Gokhale, Abhishek Chaudhary, Pratyay Banerjee, Chitta Baral, Yezhou Yang |
| 2022 | ACL | Generalized but not Robust? Comparing the Effects of Data Modification Methods on Out-of-Domain Generalization and Adversarial Robustness. | Tejas Gokhale, Swaroop Mishra, Man Luo, Bhavdeep Singh Sachdeva, Chitta Baral |
| 2022 | ACL | To Find Waldo You Need Contextual Cues: Debiasing Who's Waldo. | Yiran Luo, Pratyay Banerjee, Tejas Gokhale, Yezhou Yang, Chitta Baral |
| 2022 | ACL | Unsupervised Natural Language Inference Using PHL Triplet Generation. | Neeraj Varshney, Pratyay Banerjee, Tejas Gokhale, Chitta Baral |
| 2022 | EMNLP | CRIPP-VQA: Counterfactual Reasoning about Implicit Physical Properties via Video Question Answering. | Maitreya Patel, Tejas Gokhale, Chitta Baral, Yezhou Yang |
| 2021 | AAAI | Attribute-Guided Adversarial Training for Robustness to Natural Perturbations. | Tejas Gokhale, Rushil Anirudh, Bhavya Kailkhura, Jayaraman J. Thiagarajan, Chitta Baral, Yezhou Yang |
| 2021 | ACL | WeaQA: Weak Supervision via Captions for Visual Question Answering. | Pratyay Banerjee, Tejas Gokhale, Yezhou Yang, Chitta Baral |
| 2021 | ICCV | Weakly Supervised Relative Spatial Reasoning for Visual Question Answering. | Pratyay Banerjee, Tejas Gokhale, Yezhou Yang, Chitta Baral |
| 2021 | NAACL | Self-Supervised Test-Time Learning for Reading Comprehension. | Pratyay Banerjee, Tejas Gokhale, Chitta Baral |
| 2020 | ECCV | VQA-LOL: Visual Question Answering Under the Lens of Logic. | Tejas Gokhale, Pratyay Banerjee, Chitta Baral, Yezhou Yang |
| 2020 | EMNLP | Video2Commonsense: Generating Commonsense Descriptions to Enrich Video Captioning. | Zhiyuan Fang, Tejas Gokhale, Pratyay Banerjee, Chitta Baral, Yezhou Yang |
| 2020 | EMNLP | MUTANT: A Training Paradigm for Out-of-Distribution Generalization in Visual Question Answering. | Tejas Gokhale, Pratyay Banerjee, Chitta Baral, Yezhou Yang |
| 2019 | CVPR | Cooking With Blocks : A Recipe for Visual Reasoning on Image-Pairs. | Tejas Gokhale, Shailaja Sampat, Zhiyuan Fang, Yezhou Yang, Chitta Baral |
| 2019 | IJCAI | Vision beyond Pixels: Visual Reasoning via Blocksworld Abstractions. | Tejas Gokhale |