| 2026 | CHI | CLARIS: Clear and Intelligible Speech from Whispered and Dysarthric Voices. | Neil Shah, Yash Sonkar, Shirish Subhash Karande, Vineet Gandhi |
| 2025 | CVPR | TIDE: Training Locally Interpretable Domain Generalization Models Enables Test-time Correction. | Aishwarya Agarwal, Srikrishna Karanam, Vineet Gandhi |
| 2025 | CVPR | VELOCITI: Benchmarking Video-Language Compositional Reasoning with Strict Entailment. | Darshana Saravanan, Varun Gupta, Darshan Singh S, Zeeshan Khan, Vineet Gandhi, Makarand Tapaswi |
| 2025 | CVPR | Pseudo-labelling meets Label Smoothing for Noisy Partial Label Learning. | Darshana Saravanan, Naresh Manwani, Vineet Gandhi |
| 2025 | CVPR | Investigating Mechanisms for In-Context Vision Language Binding. | Darshana Saravanan, Makarand Tapaswi, Vineet Gandhi |
| 2025 | ICASSP | Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues. | Rohit Girmaji, Siddharth Jain, Bhav Beri, Sarthak Bansal, Vineet Gandhi |
| 2025 | ICASSP | Prompt-to-Correct: Automated Test-Time Pronunciation Correction with Voice Prompts. | Ayan Kashyap, Neil Kumar Shah, Vineet Gandhi |
| 2025 | ICASSP | Advancing NAM-to-Speech Conversion with Novel Methods and the MultiNAM Dataset. | Neil Kumar Shah, Shirish S. Karande, Vineet Gandhi |
| 2025 | ICASSP | MRI2Speech: Speech Synthesis from Articulatory Movements Recorded by Real-time MRI. | Neil Kumar Shah, Ayan Kashyap, Shirish S. Karande, Vineet Gandhi |
| 2025 | Interspeech | NAM-to-Speech Conversion with Multitask-Enhanced Autoregressive Models. | Neil Shah, Shirish Karande, Vineet Gandhi |
| 2025 | IUI | EditIQ: Automated Cinematic Editing of Static Wide-Angle Videos via Dialogue Interpretation and Saliency Cues. | Rohit Girmaji, Bhav Beri, Ramanathan Subramanian, Vineet Gandhi |
| 2025 | NAACL | IdentifyMe: A Challenging Long-Context Mention Resolution Benchmark for LLMs. | Kawshik Manikantan, Makarand Tapaswi, Vineet Gandhi, Shubham Toshniwal |
| 2024 | EACL | ParrotTTS: Text-to-speech synthesis exploiting disentangled self-supervised representations. | Neil Kumar Shah, Saiteja Kosgi, Vishal Tambrahalli, Neha Sahipjohn, Anil Nelakanti, Vineet Gandhi |
| 2024 | EMNLP | Major Entity Identification: A Generalizable Alternative to Coreference Resolution. | Kawshik Sundar, Shubham Toshniwal, Makarand Tapaswi, Vineet Gandhi |
| 2024 | Interspeech | Towards Improving NAM-to-Speech Synthesis Intelligibility using Self-Supervised Speech Models. | Neil Kumar Shah, Shirish S. Karande, Vineet Gandhi |
| 2024 | WACV | Real Time GAZED: Online Shot Selection and Editing of Virtual Cameras from Wide-Angle Monocular Video Recordings. | Sudheer Achary, Rohit Girmaji, Adhiraj Anil Deshmukh, Vineet Gandhi |
| 2023 | ICRA | Ground then Navigate: Language-guided Navigation in Dynamic Scenes. | Kanishk Jain, Varun Chhangani, Amogh Tiwari, K. Madhava Krishna, Vineet Gandhi |
| 2023 | WACV | Bringing Generalization to Deep Multi-View Pedestrian Detection. | Jeet Vora, Swetanjal Dutta, Kanishk Jain, Shyamgopal Karthik, Vineet Gandhi |
| 2023 | RO-MAN | Instance-Level Semantic Maps for Vision Language Navigation. | Laksh Nanwani, Anmol Agarwal, Kanishk Jain, Raghav Prabhakar, Aaron Monis, Aditya Mathur, Krishna Murthy Jatavallabhula, A. H. Abdul Hafez, Vineet Gandhi, K. Madhava Krishna |
| 2022 | ACL | Comprehensive Multi-Modal Interactions for Referring Image Segmentation. | Kanishk Jain, Vineet Gandhi |
| 2022 | ICMI | Does Audio help in deep Audio-Visual Saliency prediction models? | Ritvik Agrawal, Shreyank Jyoti, Rohit Girmaji, Sarath Sivaprasad, Vineet Gandhi |
| 2022 | NAACL | Empathic Machines: Using Intermediate Features as Levers to Emulate Emotions in Text-To-Speech Systems. | Saiteja Kosgi, Sarath Sivaprasad, Niranjan Pedanekar, Anil Nelakanti, Vineet Gandhi |
| 2021 | ICLR | No Cost Likelihood Manipulation at Test Time for Making Better Mistakes in Deep Networks. | Shyamgopal Karthik, Ameya Prabhu, Puneet K. Dokania, Vineet Gandhi |
| 2021 | Interspeech | Emotional Prosody Control for Speech Generation. | Sarath Sivaprasad, Saiteja Kosgi, Vineet Gandhi |
| 2021 | IROS | ViNet: Pushing the limits of Visual Modality for Audio-Visual Saliency Prediction. | Samyak Jain, Pradeep Yarlagadda, Shreyank Jyoti, Shyamgopal Karthik, Ramanathan Subramanian, Vineet Gandhi |
| 2021 | IROS | Grounding Linguistic Commands to Navigable Regions. | Nivedita Rufus, Kanishk Jain, Unni Krishnan R. Nair, Vineet Gandhi, K. Madhava Krishna |
| 2020 | CHI | GAZED- Gaze-guided Cinematic Editing of Wide-Angle Monocular Video Recordings. | K. L. Bhanu Moorthy, Moneish Kumar, Ramanathan Subramanian, Vineet Gandhi |
| 2020 | ECCV | Cosine Meets Softmax: A Tough-to-beat Baseline for Visual Grounding. | Nivedita Rufus, Unni Krishnan R. Nair, K. Madhava Krishna, Vineet Gandhi |
| 2020 | IROS | Tidying Deep Saliency Prediction Architectures. | Navyasri Reddy, Samyak Jain, Pradeep Yarlagadda, Vineet Gandhi |
| 2020 | IROS | LiDAR guided Small obstacle Segmentation. | Aasheesh Singh, Aditya Kamireddypalli, Vineet Gandhi, K. Madhava Krishna |
| 2020 | WACV | Exploring 3 R's of Long-term Tracking: Re-detection, Recovery and Reliability. | Shyamgopal Karthik, Abhinav Moudgil, Vineet Gandhi |
| 2019 | ICASSP | Nose, Eyes and Ears: Head Pose Estimation by Locating Facial Keypoints. | Aryaman Gupta, Kalpit C. Thakkar, Vineet Gandhi, P. J. Narayanan |
| 2019 | IJCAI | Learning Unsupervised Visual Grounding Through Semantic Self-Supervision. | Syed Ashar Javed, Shreyas Saxena, Vineet Gandhi |
| 2019 | IROS | Talk to the Vehicle: Language Conditioned Autonomous Navigation of Self Driving Cars. | Sriram N. N., Tirth Maniar, Jayaganesh Kalyanasundaram, Vineet Gandhi, Brojeshwar Bhowmick, K. Madhava Krishna |
| 2018 | ACCV | Long-Term Visual Object Tracking Benchmark. | Abhinav Moudgil, Vineet Gandhi |
| 2018 | ICASSP | Document Quality Estimation Using Spatial Frequency Response. | Pranjal Kumar Rai, Sajal Maheshwari, Vineet Gandhi |
| 2018 | ICASSP | An Iterative Approach for Shadow Removal in Document Images. | Vatsal Shah, Vineet Gandhi |
| 2018 | ICRA | MergeNet: A Deep Net Architecture for Small Obstacle Discovery. | Krishnam Gupta, Syed Ashar Javed, Vineet Gandhi, K. Madhava Krishna |
| 2018 | WACV | Automated Top View Registration of Broadcast Football Videos. | Rahul Anand Sharma, Bharath Bhat, Vineet Gandhi, C. V. Jawahar |
| 2017 | ICDAR | Beyond OCRs for Document Blur Estimation. | Pranjal Kumar Rai, Sajal Maheshwari, Ishit Mehta, Parikshit Sakurikar, Vineet Gandhi |
| 2013 | CVPR | Detecting and Naming Actors in Movies Using Generative Appearance Models. | Vineet Gandhi, Rmi Ronfard |
| 2012 | ICRA | High-resolution depth maps based on TOF-stereo fusion. | Vineet Gandhi, Jan Cech, Radu Horaud |