Ivan Laptev
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
102
Venues
17
Active years
1998–2026
Best venue rank
A*
Where they publish
Papers
102 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | SIGGRAPH | ActCam: Zero-Shot Joint Camera and 3D Motion Control for Video Generation. | Omar El Khalifi, Thomas Rossi, Oscar Fossey, Thibault Fouque, Ulysse Mizrahi, Philip Torr, Ivan Laptev, Fabio Pizzati, Baptiste Bellot-Gurlet |
| 2025 | ACL | LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs. | Omkar Thawakar, Dinura Dissanayake, Ketan Pravin More, Ritesh Thawkar, Ahmed Heakl, Noor Ahsan, Yuhao Li, Mohammed Zumri, Jean Lahoud, Rao Muhammad Anwer, Hisham Cholakkal, Ivan Laptev, Mubarak Shah, Fahad Shahbaz Khan, Salman H. Khan |
| 2025 | CVPR | RoomTour3D: Geometry-Aware Video-Instruction Tuning for Embodied Navigation. | Mingfei Han, Liang Ma, Kamila Zhumakhanova, Ekaterina Radionova, Jingyi Zhang, Xiaojun Chang, Xiaodan Liang, Ivan Laptev |
| 2025 | CVPR | ShowHowTo: Generating Scene-Conditioned Step-by-Step Visual Instructions. | Toms Soucek, Prajwal Gatti, Michael Wray, Ivan Laptev, Dima Damen, Josef Sivic |
| 2025 | CVPR | All Languages Matter: Evaluating LMMs on Culturally Diverse 100 Languages. | Ashmal Vayani, Dinura Dissanayake, Hasindri Watawana, Noor Ahsan, Nevasini Sasikumar, Omkar Thawakar, Henok Biadglign Ademtew, Yahya Hmaiti, Amandeep Kumar, Kartik Kuckreja, Mykola Maslych, Wafa Al Ghallabi, Mihail Minkov Mihaylov, Chao Qin, Abdelrahman M. Shaker, Mike Zhang, Mahardika Krisna Ihsani, Amiel Gian Esplana, Monil Gokani, Shachar Mirkin, Harsh Singh, Ashay Srivastava, Endre Hamerlik, Fathinah Asma Izzati, Fadillah Adamsyah Maani, Sebastian Cavada, Jenny Chim, Rohit Gupta, Sanjay Manjunath, Kamila Zhumakhanova, Feno Heriniaina Rabevohitra, Azril Hafizi Amirudin, Muhammad Ridzuan, Daniya Najiha Abdul Kareem, Ketan Pravin More, Kunyang Li, Pramesh Shakya, Muhammad Saad, Amirpouya Ghasemaghaei, Amirbek Djanibekov, Dilshod Azizov, Branislava Jankovic, Naman Bhatia, Alvaro Cabrera, Johan S. Obando-Ceron, Olympiah Otieno, Fabian Farestam, Muztoba Rabbani, Sanoojan Baliah, Santosh Sanjeev, Abduragim Shtanchaev, Maheen Fatima, Thao Nguyen, Amrin Kareem, Toluwani Aremu, Nathan Augusto Zacarias Xavier, Amit Bhatkal, Hawau Olamide Toyin, Aman Chadha, Hisham Cholakkal, Rao Muhammad Anwer, Michael Felsberg, Jorma Laaksonen, Thamar Solorio, Monojit Choudhury, Ivan Laptev, Mubarak Shah, Salman H. Khan, Fahad Shahbaz Khan |
| 2025 | EMNLP | A Culturally-diverse Multilingual Multimodal Video Benchmark & Model. | Bhuiyan Sanjid Shafique, Ashmal Vayani, Muhammad Maaz, Hanoona Abdul Rasheed, Dinura Dissanayake, Mohammed Irfan Kurpath, Yahya Hmaiti, Go Inoue, Jean Lahoud, Md. Safirur Rashid, Shadid Intisar Quasem, Maheen Fatima, Franco Vidal, Mykola Maslych, Ketan Pravin More, Sanoojan Baliah, Hasindri Watawana, Yuhao Li, Fabian Farestam, Leon Schaller, Roman Tymtsiv, Simon Weber, Hisham Cholakkal, Ivan Laptev, Shin'ichi Satoh, Michael Felsberg, Mubarak Shah, Salman H. Khan, Fahad Shahbaz Khan |
| 2025 | ICCV | ScanEdit: Hierarchically-Guided Functional 3D Scan Editing. | Mohamed El Amine Boudjoghra, Ivan Laptev, Angela Dai |
| 2025 | ICRA | ViViDex: Learning Vision-Based Dexterous Manipulation from Human Videos. | Zerui Chen, Shizhe Chen, Etienne Arlaud, Ivan Laptev, Cordelia Schmid |
| 2025 | IROS | DriveLMM-o1: A Step-by-Step Reasoning Dataset and Large Multimodal Model for Driving Scenario Understanding. | Ayesha Ishaq, Jean Lahoud, Ketan More, Omkar Thawakar, Ritesh Thawkar, Dinura Dissanayake, Noor Ahsan, Yuhao Li, Fahad Shahbaz Khan, Hisham Cholakkal, Ivan Laptev, Rao Muhammad Anwer, Salman H. Khan |
| 2025 | IROS | MALMM: Multi-Agent Large Language Models for Zero-Shot Robotic Manipulation. | Harsh Singh, Rocktim Jyoti Das, Mingfei Han, Preslav Nakov, Ivan Laptev |
| 2024 | CVPR | PairDETR : Joint Detection and Association of Human Bodies and Faces. | Ammar Ali, Georgii Gaikov, Denis Rybalchenko, Alexander Chigorin, Ivan Laptev, Sergey Zagoruyko |
| 2024 | CVPR | SUGAR : Pre-training 3D Visual Representations for Robotics. | Shizhe Chen, Ricardo Garcia, Ivan Laptev, Cordelia Schmid |
| 2024 | CVPR | GenHowTo: Learning to Generate Actions and State Transformations from Instructional Videos. | Toms Soucek, Dima Damen, Michael Wray, Ivan Laptev, Josef Sivic |
| 2023 | ACL | Tackling Ambiguity with Images: Improved Multimodal Machine Translation and Contrastive Evaluation. | Matthieu Futeral, Cordelia Schmid, Ivan Laptev, Benot Sagot, Rachel Bawden |
| 2023 | CoRL | PolarNet: 3D Point Clouds for Language-Guided Robotic Manipulation. | Shizhe Chen, Ricardo Garcia, Cordelia Schmid, Ivan Laptev |
| 2023 | CVPR | gSDF: Geometry-Driven Signed Distance Functions for 3D Hand-Object Reconstruction. | Zerui Chen, Shizhe Chen, Cordelia Schmid, Ivan Laptev |
| 2023 | CVPR | Vid2Seq: Large-Scale Pretraining of a Visual Language Model for Dense Video Captioning. | Antoine Yang, Arsha Nagrani, Paul Hongsuck Seo, Antoine Miech, Jordi Pont-Tuset, Ivan Laptev, Josef Sivic, Cordelia Schmid |
| 2023 | ICRA | Learning Video-Conditioned Policies for Unseen Manipulation Tasks. | Elliot Chane-Sane, Cordelia Schmid, Ivan Laptev |
| 2023 | ICRA | Enforcing the consensus between Trajectory Optimization and Policy Learning for precise robot control. | Quentin Le Lidec, Wilson Jallet, Ivan Laptev, Cordelia Schmid, Justin Carpentier |
| 2023 | IROS | Object Goal Navigation with Recursive Implicit Maps. | Shizhe Chen, Thomas Chabal, Ivan Laptev, Cordelia Schmid |
| 2023 | IROS | Robust Visual Sim-to-Real Transfer for Robotic Manipulation. | Ricardo Garcia, Robin Strudel, Shizhe Chen, Etienne Arlaud, Ivan Laptev, Cordelia Schmid |
| 2022 | CoRL | Instruction-driven history-aware policies for robotic manipulations. | Pierre-Louis Guhur, Shizhe Chen, Ricardo Garcia, Makarand Tapaswi, Ivan Laptev, Cordelia Schmid |
| 2022 | CVPR | Think Global, Act Local: Dual-scale Graph Transformer for Vision-and-Language Navigation. | Shizhe Chen, Pierre-Louis Guhur, Makarand Tapaswi, Cordelia Schmid, Ivan Laptev |
| 2022 | CVPR | Look for the Change: Learning Object States and State-Modifying Actions from Untrimmed Web Videos. | Toms Soucek, Jean-Baptiste Alayrac, Antoine Miech, Ivan Laptev, Josef Sivic |
| 2022 | CVPR | TubeDETR: Spatio-Temporal Video Grounding with Transformers. | Antoine Yang, Antoine Miech, Josef Sivic, Ivan Laptev, Cordelia Schmid |
| 2022 | ECCV | Learning from Unlabeled 3D Environments for Vision-and-Language Navigation. | Shizhe Chen, Pierre-Louis Guhur, Makarand Tapaswi, Cordelia Schmid, Ivan Laptev |
| 2022 | ECCV | AlignSDF: Pose-Aligned Signed Distance Fields for Hand-Object Reconstruction. | Zerui Chen, Yana Hasson, Cordelia Schmid, Ivan Laptev |
| 2021 | CVPR | Thinking Fast and Slow: Efficient Text-to-Visual Retrieval With Transformers. | Antoine Miech, Jean-Baptiste Alayrac, Ivan Laptev, Josef Sivic, Andrew Zisserman |
| 2021 | ICCV | Airbert: In-domain Pretraining for Vision-and-Language Navigation. | Pierre-Louis Guhur, Makarand Tapaswi, Shizhe Chen, Ivan Laptev, Cordelia Schmid |
| 2021 | ICCV | Segmenter: Transformer for Semantic Segmentation. | Robin Strudel, Ricardo Garcia, Ivan Laptev, Cordelia Schmid |
| 2021 | ICCV | Just Ask: Learning to Answer Questions from Millions of Narrated Videos. | Antoine Yang, Antoine Miech, Josef Sivic, Ivan Laptev, Cordelia Schmid |
| 2021 | ICML | Goal-Conditioned Reinforcement Learning with Imagined Subgoals. | Elliot Chane-Sane, Cordelia Schmid, Ivan Laptev |
| 2020 | CoRL | Learning Object Manipulation Skills via Approximate State Estimation from Real Videos. | Vladimr Petrk, Makarand Tapaswi, Ivan Laptev, Josef Sivic |
| 2020 | CoRL | Learning Obstacle Representations for Neural Motion Planning. | Robin Strudel, Ricardo Garcia, Justin Carpentier, Jean-Paul Laumond, Ivan Laptev, Cordelia Schmid |
| 2020 | CVPR | Action Modifiers: Learning From Adverbs in Instructional Videos. | Hazel Doughty, Ivan Laptev, Walterio W. Mayol-Cuevas, Dima Damen |
| 2020 | CVPR | Leveraging Photometric Consistency Over Time for Sparsely Supervised Hand-Object Reconstruction. | Yana Hasson, Bugra Tekin, Federica Bogo, Ivan Laptev, Marc Pollefeys, Cordelia Schmid |
| 2020 | CVPR | Learning Interactions and Relationships Between Movie Characters. | Anna Kukleva, Makarand Tapaswi, Ivan Laptev |
| 2020 | CVPR | End-to-End Learning of Visual Representations From Uncurated Instructional Videos. | Antoine Miech, Jean-Baptiste Alayrac, Lucas Smaira, Ivan Laptev, Josef Sivic, Andrew Zisserman |
| 2020 | ECCV | Learning Actionness via Long-Range Temporal Order Verification. | Dimitri Zhukov, Jean-Baptiste Alayrac, Ivan Laptev, Josef Sivic |
| 2020 | IROS | Learning visual policies for building 3D shape categories. | Alexander Pashevich, Igor Kalevatykh, Ivan Laptev, Cordelia Schmid |
| 2020 | ICRA | Learning to combine primitive skills: A step towards versatile robotic manipulation §. | Robin Strudel, Alexander Pashevich, Igor Kalevatykh, Ivan Laptev, Josef Sivic, Cordelia Schmid |
| 2019 | CVPR | Learning Joint Reconstruction of Hands and Manipulated Objects. | Yana Hasson, Gl Varol, Dimitrios Tzionas, Igor Kalevatykh, Michael J. Black, Ivan Laptev, Cordelia Schmid |
| 2019 | CVPR | Deep Metric Learning Beyond Binary Supervision. | Sungyeon Kim, Minkyo Seo, Ivan Laptev, Minsu Cho, Suha Kwak |
| 2019 | CVPR | Estimating 3D Motion and Forces of Person-Object Interactions From Monocular Video. | Zongmian Li, Jir Sedlr, Justin Carpentier, Ivan Laptev, Nicolas Mansard, Josef Sivic |
| 2019 | CVPR | Leveraging the Present to Anticipate the Future in Videos. | Antoine Miech, Ivan Laptev, Josef Sivic, Heng Wang, Lorenzo Torresani, Du Tran |
| 2019 | CVPR | Cross-Task Weakly Supervised Learning From Instructional Videos. | Dimitri Zhukov, Jean-Baptiste Alayrac, Ramazan Gokberk Cinbis, David F. Fouhey, Ivan Laptev, Josef Sivic |
| 2019 | ICCV | HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clips. | Antoine Miech, Dimitri Zhukov, Jean-Baptiste Alayrac, Makarand Tapaswi, Ivan Laptev, Josef Sivic |
| 2019 | ICCV | Detecting Unseen Visual Relations Using Analogies. | Julia Peyre, Josef Sivic, Ivan Laptev, Cordelia Schmid |
| 2019 | ICMV | Margin based knowledge distillation for mobile face recognition. | Dmitry Nekhaev, Sergey Milyaev, Ivan Laptev |
| 2019 | IROS | Learning to Augment Synthetic Images for Sim2Real Policy Transfer. | Alexander Pashevich, Robin Strudel, Igor Kalevatykh, Ivan Laptev, Cordelia Schmid |
| 2018 | ECCV | MobileFace: 3D Face Reconstruction with Efficient CNN Regression. | Nikolai Chinaev, Alexander Chigorin, Ivan Laptev |
| 2018 | ECCV | BodyNet: Volumetric Inference of 3D Human Body Shapes. | Gl Varol, Duygu Ceylan, Bryan C. Russell, Jimei Yang, Ersin Yumer, Ivan Laptev, Cordelia Schmid |
| 2017 | CVPR | Learning from Synthetic Humans. | Gl Varol, Javier Romero, Xavier Martin, Naureen Mahmood, Michael J. Black, Ivan Laptev, Cordelia Schmid |
| 2017 | ICCV | Joint Discovery of Object States and Manipulation Actions. | Jean-Baptiste Alayrac, Josef Sivic, Ivan Laptev, Simon Lacoste-Julien |
| 2017 | ICCV | Learning from Video and Text via Large-Scale Discriminative Clustering. | Antoine Miech, Jean-Baptiste Alayrac, Piotr Bojanowski, Ivan Laptev, Josef Sivic |
| 2017 | ICCV | Weakly-Supervised Learning of Visual Relations. | Julia Peyre, Ivan Laptev, Cordelia Schmid, Josef Sivic |
| 2016 | CVPR | Unsupervised Learning from Narrated Instruction Videos. | Jean-Baptiste Alayrac, Piotr Bojanowski, Nishant Agrawal, Josef Sivic, Ivan Laptev, Simon Lacoste-Julien |
| 2016 | CVPR | Thin-Slicing for Pose: Learning to Understand Pose without Explicit Pose Estimation. | Suha Kwak, Minsu Cho, Ivan Laptev |
| 2016 | CVPR | Instance-Level Video Segmentation from Object Tracks. | Guillaume Seguin, Piotr Bojanowski, Rmi Lajugie, Ivan Laptev |
| 2016 | ECCV | ContextLocNet: Context-Aware Deep Network Models for Weakly Supervised Localization. | Vadim Kantorov, Maxime Oquab, Minsu Cho, Ivan Laptev |
| 2016 | ECCV | Hollywood in Homes: Crowdsourcing Data Collection for Activity Understanding. | Gunnar A. Sigurdsson, Gl Varol, Xiaolong Wang, Ali Farhadi, Ivan Laptev, Abhinav Gupta |
| 2016 | HCOMP | Much Ado About Time: Exhaustive Annotation of Temporal Data. | Gunnar A. Sigurdsson, Olga Russakovsky, Ali Farhadi, Ivan Laptev, Abhinav Gupta |
| 2015 | CVPR | On pairwise costs for network flow multi-object tracking. | Visesh Chari, Simon Lacoste-Julien, Ivan Laptev, Josef Sivic |
| 2015 | CVPR | Is object localization for free? - Weakly-supervised learning with convolutional neural networks. | Maxime Oquab, Lon Bottou, Ivan Laptev, Josef Sivic |
| 2015 | ICCV | Weakly-Supervised Alignment of Video with Text. | Piotr Bojanowski, Rmi Lajugie, Edouard Grave, Francis R. Bach, Ivan Laptev, Jean Ponce, Cordelia Schmid |
| 2015 | ICCV | P-CNN: Pose-Based CNN Features for Action Recognition. | Guilhem Chron, Ivan Laptev, Cordelia Schmid |
| 2015 | ICCV | Unsupervised Object Discovery and Tracking in Video Collections. | Suha Kwak, Minsu Cho, Ivan Laptev, Jean Ponce, Cordelia Schmid |
| 2015 | ICCV | Context-Aware CNNs for Person Head Detection. | Tuan-Hung Vu, Anton Osokin, Ivan Laptev |
| 2014 | CVPR | Efficient Feature Extraction, Encoding, and Classification for Action Recognition. | Vadim Kantorov, Ivan Laptev |
| 2014 | CVPR | Learning and Transferring Mid-level Image Representations Using Convolutional Neural Networks. | Maxime Oquab, Lon Bottou, Ivan Laptev, Josef Sivic |
| 2014 | ECCV | Weakly Supervised Action Labeling in Videos under Ordering Constraints. | Piotr Bojanowski, Rmi Lajugie, Francis R. Bach, Ivan Laptev, Jean Ponce, Cordelia Schmid, Josef Sivic |
| 2014 | ECCV | Predicting Actions from Static Scenes. | Tuan-Hung Vu, Catherine Olsson, Ivan Laptev, Aude Oliva, Josef Sivic |
| 2013 | ICCV | Pose Estimation and Segmentation of People in 3D Movies. | Karteek Alahari, Guillaume Seguin, Josef Sivic, Ivan Laptev |
| 2013 | ICCV | Finding Actors and Actions in Movies. | Piotr Bojanowski, Francis R. Bach, Ivan Laptev, Jean Ponce, Cordelia Schmid, Josef Sivic |
| 2012 | ECCV | Object Detection Using Strongly-Supervised Deformable Part Models. | Hossein Azizpour, Ivan Laptev |
| 2012 | ECCV | Scene Semantics from Long-Term Observation of People. | Vincent Delaitre, David F. Fouhey, Ivan Laptev, Josef Sivic, Abhinav Gupta, Alexei A. Efros |
| 2012 | ECCV | People Watching: Human Actions as a Cue for Single View Geometry. | David F. Fouhey, Vincent Delaitre, Abhinav Gupta, Alexei A. Efros, Ivan Laptev, Josef Sivic |
| 2012 | ICIP | Actlets: A novel local representation for human action recognition in video. | Muhammad Muneeb Ullah, Ivan Laptev |
| 2011 | CVPR | Track to the future: Spatio-temporal video segmentation with long-range motion cues. | Jos Lezama, Karteek Alahari, Josef Sivic, Ivan Laptev |
| 2011 | ICCV | Density-aware person detection and tracking in crowds. | Mikel Rodriguez, Ivan Laptev, Josef Sivic, Jean-Yves Audibert |
| 2011 | ICCV | Data-driven crowd analysis in videos. | Mikel Rodriguez, Josef Sivic, Ivan Laptev, Jean-Yves Audibert |
| 2011 | ICIP | Joint pose estimation and action recognition in image graphs. | Kumar Raja, Ivan Laptev, Patrick Prez, Lionel Oisel |
| 2010 | BMVC | Recognizing human actions in still images: a study of bag-of-features and part-based representations. | Vincent Delaitre, Ivan Laptev, Josef Sivic |
| 2010 | BMVC | Improving bag-of-features action recognition with non-local cues. | Muhammad Muneeb Ullah, Sobhan Naderi Parizi, Ivan Laptev |
| 2010 | ECCV | Semi-supervised Learning of Facial Attributes in Video. | Neva Cherniavsky, Ivan Laptev, Josef Sivic, Andrew Zisserman |
| 2010 | ICPR | Recognizing Human Action in the Wild. | Ivan Laptev |
| 2009 | BMVC | Multi-view Synchronization of Human Actions and Dynamic Scenes. | Emilie Dexter, Patrick Prez, Ivan Laptev |
| 2009 | BMVC | Evaluation of Local Spatio-temporal Features for Action Recognition. | Heng Wang, Muhammad Muneeb Ullah, Alexander Klser, Ivan Laptev, Cordelia Schmid |
| 2009 | CVPR | Actions in context. | Marcin Marszalek, Ivan Laptev, Cordelia Schmid |
| 2009 | DICTA | Modeling Image Context Using Object Centered Grid. | Sobhan Naderi Parizi, Ivan Laptev, Alireza Tavakoli Targhi |
| 2009 | ICCV | Automatic annotation of human actions in video. | Olivier Duchenne, Ivan Laptev, Josef Sivic, Francis R. Bach, Jean Ponce |
| 2008 | CVPR | Learning realistic human actions from movies. | Ivan Laptev, Marcin Marszalek, Cordelia Schmid, Benjamin Rozenfeld |
| 2008 | ECCV | Cross-View Action Recognition from Temporal Self-similarities. | Imran N. Junejo, Emilie Dexter, Ivan Laptev, Patrick Prez |
| 2007 | ICCV | Retrieving actions in movies. | Ivan Laptev, Patrick Prez |
| 2006 | BMVC | Improvements of Object Detection Using Boosted Histograms. | Ivan Laptev |
| 2005 | ICCV | Periodic Motion Detection and Segmentation via Approximate Sequence Alignment. | Ivan Laptev, Serge J. Belongie, Patrick Prez, Josh Wills |
| 2004 | ICPR | Velocity Adaptation of Space-Time Interest Points. | Ivan Laptev, Tony Lindeberg |
| 2004 | ICPR | Galilean-Diagonalized Spatio-Temporal Interest Operators. | Tony Lindeberg, Amir Akbarzadeh, Ivan Laptev |
| 2004 | ICPR | Recognizing Human Actions: A Local SVM Approach. | Christian Schldt, Ivan Laptev, Barbara Caputo |
| 2003 | ICCV | Space-time Interest Points. | Ivan Laptev, Tony Lindeberg |
| 1998 | ECCV | Multi-scale and Snakes for Automatic Road Extraction. | Helmut Mayer, Ivan Laptev, Albert Baumgartner |
| 1998 | RoboCup | Agilo RoboCuppers: RoboCup Team Description. | Michael Klupsch, Maximilian Lckenhaus, Christoph Zierl, Ivan Laptev, Thorsten Bandlow, Marc Grimme, Ignaz Kellerer, Fabian Schwarzer |