Skip to content

Makarand Tapaswi

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

40

Venues

14

Active years

2012–2026

Best venue rank

A*

Where they publish

Papers

40 indexed papers, newest first.

YearVenueTitleAuthors
2026WACVSTRinGS: Selective Text Refinement in Gaussian Splatting.Abhinav Raundhal, Gaurav Behera, P. J. Narayanan, Ravi Kiran Sarvadevabhatla, Makarand Tapaswi
2025CVPRVELOCITI: Benchmarking Video-Language Compositional Reasoning with Strict Entailment.Darshana Saravanan, Varun Gupta, Darshan Singh S, Zeeshan Khan, Vineet Gandhi, Makarand Tapaswi
2025CVPRInvestigating Mechanisms for In-Context Vision Language Binding.Darshana Saravanan, Makarand Tapaswi, Vineet Gandhi
2025EMNLPWhat You See is What You Ask: Evaluating Audio Descriptions.Divy Kala, Eshika Khandelwal, Makarand Tapaswi
2025ICASSPThe Sound of Water: Inferring Physical Properties from Pouring Liquids.Piyush Bagad, Makarand Tapaswi, Cees G. M. Snoek, Andrew Zisserman
2025NAACLIdentifyMe: A Challenging Long-Context Mention Resolution Benchmark for LLMs.Kawshik Manikantan, Makarand Tapaswi, Vineet Gandhi, Shubham Toshniwal
2025WACVSeeing Eye to AI: Comparing Human Gaze and Model Attention in Video Memorability.Prajneya Kumar, Eshika Khandelwal, Makarand Tapaswi, Vishnu Sreekumar
2024CVPRNurtureNet: A Multi-task Video-based Approach for Newborn Anthropometry.Yash Khandelwal, Mayur Arvind, Sriram Kumar, Ashish Gupta, Sachin Kumar Danisetty, Piyush Bagad, Anish Madan, Mayank Lunayach, Aditya Annavajjala, Abhishek Maiti, Sansiddh Jain, Aman Dalmia, Namrata Deka, Jerome White, Jigar Doshi, Angjoo Kanazawa, Rahul Panicker, Alpan Raval, Srinivas Rana, Makarand Tapaswi
2024CVPRMICap: A Unified Model for Identity-Aware Movie Descriptions.Haran Raajesh, Naveen Reddy Desanur, Zeeshan Khan, Makarand Tapaswi
2024CVPR"Previously on..." from Recaps to Story Summarization.Aditya Kumar Singh, Dhruv Srivastava, Makarand Tapaswi
2024EMNLPMajor Entity Identification: A Generalizable Alternative to Coreference Resolution.Kawshik Sundar, Shubham Toshniwal, Makarand Tapaswi, Vineet Gandhi
2023CVPRTest of Time: Instilling Video-Language Models with a Sense of Time.Piyush Bagad, Makarand Tapaswi, Cees G. M. Snoek
2023CVPRHow You Feelin'? Learning Emotions and Mental States in Movie Scenes.Dhruv Srivastava, Aditya Kumar Singh, Makarand Tapaswi
2023WWWGrapeQA: GRaph Augmentation and Pruning to Enhance Question-Answering.Dhaval Taunk, Lakshya Khanna, Siri Venkata Pavan Kumar Kandru, Vasudeva Varma, Charu Sharma, Makarand Tapaswi
2023WACVUnsupervised Audio-Visual Lecture Segmentation.Darshan Singh S, Anchit Gupta, C. V. Jawahar, Makarand Tapaswi
2022CoRLInstruction-driven history-aware policies for robotic manipulations.Pierre-Louis Guhur, Shizhe Chen, Ricardo Garcia, Makarand Tapaswi, Ivan Laptev, Cordelia Schmid
2022CVPRThink Global, Act Local: Dual-scale Graph Transformer for Vision-and-Language Navigation.Shizhe Chen, Pierre-Louis Guhur, Makarand Tapaswi, Cordelia Schmid, Ivan Laptev
2022ECCVLearning from Unlabeled 3D Environments for Vision-and-Language Navigation.Shizhe Chen, Pierre-Louis Guhur, Makarand Tapaswi, Cordelia Schmid, Ivan Laptev
2022IROSLearning Object Manipulation Skills from Video via Approximate Differentiable Physics.Vladimr Petrk, Mohammad Nomaan Qureshi, Josef Sivic, Makarand Tapaswi
2021ICCVAirbert: In-domain Pretraining for Vision-and-Language Navigation.Pierre-Louis Guhur, Makarand Tapaswi, Shizhe Chen, Ivan Laptev, Cordelia Schmid
2020CoRLLearning Object Manipulation Skills via Approximate State Estimation from Real Videos.Vladimr Petrk, Makarand Tapaswi, Ivan Laptev, Josef Sivic
2020CVPRLearning Interactions and Relationships Between Movie Characters.Anna Kukleva, Makarand Tapaswi, Ivan Laptev
2019ICCVHowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clips.Antoine Miech, Dimitri Zhukov, Jean-Baptiste Alayrac, Makarand Tapaswi, Ivan Laptev, Josef Sivic
2019ICCVVideo Face Clustering With Unknown Number of Clusters.Makarand Tapaswi, Marc T. Law, Sanja Fidler
2019ICLRVisual Reasoning by Progressive Module Networks.Seung Wook Kim, Makarand Tapaswi, Sanja Fidler
2018CVPRMovieGraphs: Towards Understanding Human-Centric Situations From Videos.Paul Vicol, Makarand Tapaswi, Llus Castrejn, Sanja Fidler
2018CVPRNow You Shake Me: Towards Automatic 4D Cinema.Yuhao Zhou, Makarand Tapaswi, Sanja Fidler
2017ICCVSituation Recognition with Graph Neural Networks.Ruiyu Li, Makarand Tapaswi, Renjie Liao, Jiaya Jia, Raquel Urtasun, Sanja Fidler
2016CVPRRecovering the Missing Link: Predicting Class-Attribute Associations for Unsupervised Zero-Shot Learning.Ziad Al-Halah, Makarand Tapaswi, Rainer Stiefelhagen
2016CVPRMovieQA: Understanding Stories in Movies through Question-Answering.Makarand Tapaswi, Yukun Zhu, Rainer Stiefelhagen, Antonio Torralba, Raquel Urtasun, Sanja Fidler
2016WACVNaming TV characters by watching and analyzing dialogs.Monica-Laura Haurilet, Makarand Tapaswi, Ziad Al-Halah, Rainer Stiefelhagen
2015CVPRBook2Movie: Aligning video scenes with book chapters.Makarand Tapaswi, Martin Buml, Rainer Stiefelhagen
2014AVSSA time pooled track kernel for person identification.Martin Buml, Makarand Tapaswi, Rainer Stiefelhagen
2014CVPRStoryGraphs: Visualizing Character Interactions as a Timeline.Makarand Tapaswi, Martin Buml, Rainer Stiefelhagen
2014ICIPCleaning up after a face tracker: False positive removal.Makarand Tapaswi, Cemal Cagn Corez, Martin Buml, Hazim Kemal Ekenel, Rainer Stiefelhagen
2013CVPRSemi-supervised Learning with Constraints for Person Identification in Multimedia Data.Martin Buml, Makarand Tapaswi, Rainer Stiefelhagen
2013InterspeechQCompere @ REPERE 2013.Herv Bredin, Johann Poignant, Guillaume Fortier, Makarand Tapaswi, Viet Bac Le, Anindya Roy, Claude Barras, Sophie Rosset, Achintya Kumar Sarkar, Qian Yang, Hua Gao, Alexis Mignon, Jakob Verbeek, Laurent Besacier, Georges Qunot, Hazim Kemal Ekenel, Rainer Stiefelhagen
2012AVSSContextual Constraints for Person Retrieval in Camera Networks.Martin Buml, Makarand Tapaswi, Arne Schumann, Rainer Stiefelhagen
2012CVPR"Knock! Knock! Who is it?" probabilistic person identification in TV-series.Makarand Tapaswi, Martin Buml, Rainer Stiefelhagen
2012ECCVFusion of Speech, Faces and Text for Person Identification in TV Broadcast.Herv Bredin, Johann Poignant, Makarand Tapaswi, Guillaume Fortier, Viet Bac Le, Thibault Napolon, Hua Gao, Claude Barras, Sophie Rosset, Laurent Besacier, Jakob Verbeek, Georges Qunot, Frdric Jurie, Hazim Kemal Ekenel