Skip to content

Vineet Gandhi

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

42

Venues

19

Active years

2012–2026

Best venue rank

A*

Where they publish

Papers

42 indexed papers, newest first.

YearVenueTitleAuthors
2026CHICLARIS: Clear and Intelligible Speech from Whispered and Dysarthric Voices.Neil Shah, Yash Sonkar, Shirish Subhash Karande, Vineet Gandhi
2025CVPRTIDE: Training Locally Interpretable Domain Generalization Models Enables Test-time Correction.Aishwarya Agarwal, Srikrishna Karanam, Vineet Gandhi
2025CVPRVELOCITI: Benchmarking Video-Language Compositional Reasoning with Strict Entailment.Darshana Saravanan, Varun Gupta, Darshan Singh S, Zeeshan Khan, Vineet Gandhi, Makarand Tapaswi
2025CVPRPseudo-labelling meets Label Smoothing for Noisy Partial Label Learning.Darshana Saravanan, Naresh Manwani, Vineet Gandhi
2025CVPRInvestigating Mechanisms for In-Context Vision Language Binding.Darshana Saravanan, Makarand Tapaswi, Vineet Gandhi
2025ICASSPMinimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues.Rohit Girmaji, Siddharth Jain, Bhav Beri, Sarthak Bansal, Vineet Gandhi
2025ICASSPPrompt-to-Correct: Automated Test-Time Pronunciation Correction with Voice Prompts.Ayan Kashyap, Neil Kumar Shah, Vineet Gandhi
2025ICASSPAdvancing NAM-to-Speech Conversion with Novel Methods and the MultiNAM Dataset.Neil Kumar Shah, Shirish S. Karande, Vineet Gandhi
2025ICASSPMRI2Speech: Speech Synthesis from Articulatory Movements Recorded by Real-time MRI.Neil Kumar Shah, Ayan Kashyap, Shirish S. Karande, Vineet Gandhi
2025InterspeechNAM-to-Speech Conversion with Multitask-Enhanced Autoregressive Models.Neil Shah, Shirish Karande, Vineet Gandhi
2025IUIEditIQ: Automated Cinematic Editing of Static Wide-Angle Videos via Dialogue Interpretation and Saliency Cues.Rohit Girmaji, Bhav Beri, Ramanathan Subramanian, Vineet Gandhi
2025NAACLIdentifyMe: A Challenging Long-Context Mention Resolution Benchmark for LLMs.Kawshik Manikantan, Makarand Tapaswi, Vineet Gandhi, Shubham Toshniwal
2024EACLParrotTTS: Text-to-speech synthesis exploiting disentangled self-supervised representations.Neil Kumar Shah, Saiteja Kosgi, Vishal Tambrahalli, Neha Sahipjohn, Anil Nelakanti, Vineet Gandhi
2024EMNLPMajor Entity Identification: A Generalizable Alternative to Coreference Resolution.Kawshik Sundar, Shubham Toshniwal, Makarand Tapaswi, Vineet Gandhi
2024InterspeechTowards Improving NAM-to-Speech Synthesis Intelligibility using Self-Supervised Speech Models.Neil Kumar Shah, Shirish S. Karande, Vineet Gandhi
2024WACVReal Time GAZED: Online Shot Selection and Editing of Virtual Cameras from Wide-Angle Monocular Video Recordings.Sudheer Achary, Rohit Girmaji, Adhiraj Anil Deshmukh, Vineet Gandhi
2023ICRAGround then Navigate: Language-guided Navigation in Dynamic Scenes.Kanishk Jain, Varun Chhangani, Amogh Tiwari, K. Madhava Krishna, Vineet Gandhi
2023WACVBringing Generalization to Deep Multi-View Pedestrian Detection.Jeet Vora, Swetanjal Dutta, Kanishk Jain, Shyamgopal Karthik, Vineet Gandhi
2023RO-MANInstance-Level Semantic Maps for Vision Language Navigation.Laksh Nanwani, Anmol Agarwal, Kanishk Jain, Raghav Prabhakar, Aaron Monis, Aditya Mathur, Krishna Murthy Jatavallabhula, A. H. Abdul Hafez, Vineet Gandhi, K. Madhava Krishna
2022ACLComprehensive Multi-Modal Interactions for Referring Image Segmentation.Kanishk Jain, Vineet Gandhi
2022ICMIDoes Audio help in deep Audio-Visual Saliency prediction models?Ritvik Agrawal, Shreyank Jyoti, Rohit Girmaji, Sarath Sivaprasad, Vineet Gandhi
2022NAACLEmpathic Machines: Using Intermediate Features as Levers to Emulate Emotions in Text-To-Speech Systems.Saiteja Kosgi, Sarath Sivaprasad, Niranjan Pedanekar, Anil Nelakanti, Vineet Gandhi
2021ICLRNo Cost Likelihood Manipulation at Test Time for Making Better Mistakes in Deep Networks.Shyamgopal Karthik, Ameya Prabhu, Puneet K. Dokania, Vineet Gandhi
2021InterspeechEmotional Prosody Control for Speech Generation.Sarath Sivaprasad, Saiteja Kosgi, Vineet Gandhi
2021IROSViNet: Pushing the limits of Visual Modality for Audio-Visual Saliency Prediction.Samyak Jain, Pradeep Yarlagadda, Shreyank Jyoti, Shyamgopal Karthik, Ramanathan Subramanian, Vineet Gandhi
2021IROSGrounding Linguistic Commands to Navigable Regions.Nivedita Rufus, Kanishk Jain, Unni Krishnan R. Nair, Vineet Gandhi, K. Madhava Krishna
2020CHIGAZED- Gaze-guided Cinematic Editing of Wide-Angle Monocular Video Recordings.K. L. Bhanu Moorthy, Moneish Kumar, Ramanathan Subramanian, Vineet Gandhi
2020ECCVCosine Meets Softmax: A Tough-to-beat Baseline for Visual Grounding.Nivedita Rufus, Unni Krishnan R. Nair, K. Madhava Krishna, Vineet Gandhi
2020IROSTidying Deep Saliency Prediction Architectures.Navyasri Reddy, Samyak Jain, Pradeep Yarlagadda, Vineet Gandhi
2020IROSLiDAR guided Small obstacle Segmentation.Aasheesh Singh, Aditya Kamireddypalli, Vineet Gandhi, K. Madhava Krishna
2020WACVExploring 3 R's of Long-term Tracking: Re-detection, Recovery and Reliability.Shyamgopal Karthik, Abhinav Moudgil, Vineet Gandhi
2019ICASSPNose, Eyes and Ears: Head Pose Estimation by Locating Facial Keypoints.Aryaman Gupta, Kalpit C. Thakkar, Vineet Gandhi, P. J. Narayanan
2019IJCAILearning Unsupervised Visual Grounding Through Semantic Self-Supervision.Syed Ashar Javed, Shreyas Saxena, Vineet Gandhi
2019IROSTalk to the Vehicle: Language Conditioned Autonomous Navigation of Self Driving Cars.Sriram N. N., Tirth Maniar, Jayaganesh Kalyanasundaram, Vineet Gandhi, Brojeshwar Bhowmick, K. Madhava Krishna
2018ACCVLong-Term Visual Object Tracking Benchmark.Abhinav Moudgil, Vineet Gandhi
2018ICASSPDocument Quality Estimation Using Spatial Frequency Response.Pranjal Kumar Rai, Sajal Maheshwari, Vineet Gandhi
2018ICASSPAn Iterative Approach for Shadow Removal in Document Images.Vatsal Shah, Vineet Gandhi
2018ICRAMergeNet: A Deep Net Architecture for Small Obstacle Discovery.Krishnam Gupta, Syed Ashar Javed, Vineet Gandhi, K. Madhava Krishna
2018WACVAutomated Top View Registration of Broadcast Football Videos.Rahul Anand Sharma, Bharath Bhat, Vineet Gandhi, C. V. Jawahar
2017ICDARBeyond OCRs for Document Blur Estimation.Pranjal Kumar Rai, Sajal Maheshwari, Ishit Mehta, Parikshit Sakurikar, Vineet Gandhi
2013CVPRDetecting and Naming Actors in Movies Using Generative Appearance Models.Vineet Gandhi, Rmi Ronfard
2012ICRAHigh-resolution depth maps based on TOF-stereo fusion.Vineet Gandhi, Jan Cech, Radu Horaud