Skip to content

Josef Sivic

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

93

Venues

13

Active years

2003–2026

Best venue rank

A*

Where they publish

Papers

93 indexed papers, newest first.

YearVenueTitleAuthors
2026SIGGRAPHAutoregressive Modeling of Film with Applications in Video Montage.Marcelo Sandoval-Castaeda, Fabian Caba Heilbron, Shiry Ginosar, Bryan C. Russell, Josef Sivic, Alexei A. Efros, Gregory Shakhnarovich
2025CVPRImproving Personalized Search with Regularized Low-Rank Parameter Updates.Fiona Ryan, Josef Sivic, Fabian Caba Heilbron, Judy Hoffman, James M. Rehg, Bryan C. Russell
2025CVPRShowHowTo: Generating Scene-Conditioned Step-by-Step Visual Instructions.Toms Soucek, Prajwal Gatti, Michael Wray, Ivan Laptev, Dima Damen, Josef Sivic
2025ICCVDiscovering Divergent Representations Between Text-To-Image Models.Lisa Dunlap, Joseph E. Gonzalez, Trevor Darrell, Fabian Caba Heilbron, Josef Sivic, Bryan C. Russell
2025ICCVLarge-Scale Pre-Training for Grounded Video Caption Generation.Evangelos Kazakos, Cordelia Schmid, Josef Sivic
2025ICCVResidualViT for Efficient Temporally Dense Video Encoding.Mattia Soldan, Fabian Caba Heilbron, Bernard Ghanem, Josef Sivic, Bryan C. Russell
2025ICLRLearning to engineer protein flexibility.Petr Kouba, Joan Planas-Iglesias, Jir Damborsk, Jir Sedlr, Stanislav Mazurenko, Josef Sivic
2025ICLR6D Object Pose Tracking in Internet Videos for Robotic Manipulation.Georgy Ponimatkin, Martin Cfka, Toms Soucek, Mdric Fourmy, Yann Labb, Vladimr Petrk, Josef Sivic
2025SIGGRAPHEditDuet: A Multi-Agent System for Video Non-Linear Editing.Marcelo Sandoval-Castaeda, Bryan C. Russell, Josef Sivic, Gregory Shakhnarovich, Fabian Caba Heilbron
2024ACCVNewMove: Customizing Text-to-Video Models with Novel Motions.Joanna Materzynska, Josef Sivic, Eli Shechtman, Antonio Torralba, Richard Zhang, Bryan C. Russell
2024CVPRGenHowTo: Learning to Generate Actions and State Transformations from Instructional Videos.Toms Soucek, Dima Damen, Michael Wray, Ivan Laptev, Josef Sivic
2024ICLRLearning to design protein-protein interactions with enhanced generalization.Anton Bushuiev, Roman Bushuiev, Petr Kouba, Anatolii Filkin, Marketa Gabrielova, Michal Gabriel, Jir Sedlr, Toms Pluskal, Jir Damborsk, Stanislav Mazurenko, Josef Sivic
2023CVPRLanguage-Guided Music Recommendation for Video via Prompt Analogies.Daniel McKee, Justin Salamon, Josef Sivic, Bryan C. Russell
2023CVPRVid2Seq: Large-Scale Pretraining of a Visual Language Model for Dense Video Captioning.Antoine Yang, Arsha Nagrani, Paul Hongsuck Seo, Antoine Miech, Jordi Pont-Tuset, Ivan Laptev, Josef Sivic, Cordelia Schmid
2023CVPRMeta-Personalizing Vision-Language Models to Find Named Instances in Video.Chun-Hsiao Yeh, Bryan C. Russell, Josef Sivic, Fabian Caba Heilbron, Simon Jenni
2023ICRADifferentiable Collision Detection: a Randomized Smoothing Approach.Louis Montaut, Quentin Le Lidec, Antoine Bambade, Vladimr Petrk, Josef Sivic, Justin Carpentier
2023ICRAMulti-Contact Task and Motion Planning Guided by Video Demonstration.Kateryna Zorina, David Kovr, Florent Lamiraux, Nicolas Mansard, Justin Carpentier, Josef Sivic, Vladimr Petrk
2022CoRLMegaPose: 6D Pose Estimation of Novel Objects via Render & Compare.Yann Labb, Lucas Manuelli, Arsalan Mousavian, Stephen Tyree, Stan Birchfield, Jonathan Tremblay, Justin Carpentier, Mathieu Aubry, Dieter Fox, Josef Sivic
2022CVPRFocal Length and Object Pose Estimation via Render and Compare.Georgy Ponimatkin, Yann Labb, Bryan C. Russell, Mathieu Aubry, Josef Sivic
2022CVPRLook for the Change: Learning Object States and State-Modifying Actions from Untrimmed Web Videos.Toms Soucek, Jean-Baptiste Alayrac, Antoine Miech, Ivan Laptev, Josef Sivic
2022CVPRTubeDETR: Spatio-Temporal Video Grounding with Transformers.Antoine Yang, Antoine Miech, Josef Sivic, Ivan Laptev, Cordelia Schmid
2022ECCVDrive&Segment: Unsupervised Semantic Segmentation of Urban Scenes via Cross-Modal Distillation.Antonn Vobeck, David Hurych, Oriane Simoni, Spyros Gidaris, Andrei Bursuc, Patrick Prez, Josef Sivic
2022IROSLearning Object Manipulation Skills from Video via Approximate Differentiable Physics.Vladimr Petrk, Mohammad Nomaan Qureshi, Josef Sivic, Makarand Tapaswi
2021AAAIArtificial Dummies for Urban Dataset Augmentation.Antonn Vobeck, David Hurych, Michal Uricr, Patrick Prez, Josef Sivic
2021CVPRSingle-View Robot Pose and Joint Angle Estimation via Render & Compare.Yann Labb, Justin Carpentier, Mathieu Aubry, Josef Sivic
2021CVPRThinking Fast and Slow: Efficient Text-to-Visual Retrieval With Transformers.Antoine Miech, Jean-Baptiste Alayrac, Ivan Laptev, Josef Sivic, Andrew Zisserman
2021ICCVWeakly Supervised Human-Object Interaction Detection in Video via Contrastive Spatiotemporal Regions.Shuang Li, Yilun Du, Antonio Torralba, Josef Sivic, Bryan C. Russell
2021ICCVJust Ask: Learning to Answer Questions from Millions of Narrated Videos.Antoine Yang, Antoine Miech, Josef Sivic, Ivan Laptev, Cordelia Schmid
2020CoRLLearning Object Manipulation Skills via Approximate State Estimation from Real Videos.Vladimr Petrk, Makarand Tapaswi, Ivan Laptev, Josef Sivic
2020CVPREnd-to-End Learning of Visual Representations From Uncurated Instructional Videos.Antoine Miech, Jean-Baptiste Alayrac, Lucas Smaira, Ivan Laptev, Josef Sivic, Andrew Zisserman
2020ECCVCosyPose: Consistent Multi-view Multi-object 6D Pose Estimation.Yann Labb, Justin Carpentier, Mathieu Aubry, Josef Sivic
2020ECCVEfficient Neighbourhood Consensus Networks via Submanifold Sparse Convolutions.Ignacio Rocco, Relja Arandjelovic, Josef Sivic
2020ECCVLearning Actionness via Long-Range Temporal Order Verification.Dimitri Zhukov, Jean-Baptiste Alayrac, Ivan Laptev, Josef Sivic
2020ICRALearning to combine primitive skills: A step towards versatile robotic manipulation §.Robin Strudel, Alexander Pashevich, Igor Kalevatykh, Ivan Laptev, Josef Sivic, Cordelia Schmid
2019CVPRD2-Net: A Trainable CNN for Joint Description and Detection of Local Features.Mihai Dusmanu, Ignacio Rocco, Toms Pajdla, Marc Pollefeys, Josef Sivic, Akihiko Torii, Torsten Sattler
2019CVPREstimating 3D Motion and Forces of Person-Object Interactions From Monocular Video.Zongmian Li, Jir Sedlr, Justin Carpentier, Ivan Laptev, Nicolas Mansard, Josef Sivic
2019CVPRLeveraging the Present to Anticipate the Future in Videos.Antoine Miech, Ivan Laptev, Josef Sivic, Heng Wang, Lorenzo Torresani, Du Tran
2019CVPRCross-Task Weakly Supervised Learning From Instructional Videos.Dimitri Zhukov, Jean-Baptiste Alayrac, Ramazan Gokberk Cinbis, David F. Fouhey, Ivan Laptev, Josef Sivic
2019ICCVHowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clips.Antoine Miech, Dimitri Zhukov, Jean-Baptiste Alayrac, Makarand Tapaswi, Ivan Laptev, Josef Sivic
2019ICCVDetecting Unseen Visual Relations Using Analogies.Julia Peyre, Josef Sivic, Ivan Laptev, Cordelia Schmid
2019ICCVIs This the Right Place? Geometric-Semantic Pose Verification for Indoor Visual Localization.Hajime Taira, Ignacio Rocco, Jir Sedlr, Masatoshi Okutomi, Josef Sivic, Toms Pajdla, Torsten Sattler, Akihiko Torii
2018CVPREnd-to-End Weakly-Supervised Semantic Alignment.Ignacio Rocco, Relja Arandjelovic, Josef Sivic
2018CVPRBenchmarking 6DOF Outdoor Visual Localization in Changing Conditions.Torsten Sattler, Will Maddern, Carl Toft, Akihiko Torii, Lars Hammarstrand, Erik Stenborg, Daniel Safari, Masatoshi Okutomi, Marc Pollefeys, Josef Sivic, Fredrik Kahl, Toms Pajdla
2018CVPRInLoc: Indoor Visual Localization With Dense Matching and View Synthesis.Hajime Taira, Masatoshi Okutomi, Torsten Sattler, Mircea Cimpoi, Marc Pollefeys, Josef Sivic, Toms Pajdla, Akihiko Torii
2018EMNLPLocalizing Moments in Video with Temporal Language.Lisa Anne Hendricks, Oliver Wang, Eli Shechtman, Josef Sivic, Trevor Darrell, Bryan C. Russell
2017CVPRActionVLAD: Learning Spatio-Temporal Aggregation for Action Classification.Rohit Girdhar, Deva Ramanan, Abhinav Gupta, Josef Sivic, Bryan C. Russell
2017CVPRConvolutional Neural Network Architecture for Geometric Matching.Ignacio Rocco, Relja Arandjelovic, Josef Sivic
2017CVPRAre Large-Scale 3D Models Really Necessary for Accurate Visual Localization?Torsten Sattler, Akihiko Torii, Josef Sivic, Marc Pollefeys, Hajime Taira, Masatoshi Okutomi, Toms Pajdla
2017ICCVJoint Discovery of Object States and Manipulation Actions.Jean-Baptiste Alayrac, Josef Sivic, Ivan Laptev, Simon Lacoste-Julien
2017ICCVLocalizing Moments in Video with Natural Language.Lisa Anne Hendricks, Oliver Wang, Eli Shechtman, Josef Sivic, Trevor Darrell, Bryan C. Russell
2017ICCVLearning from Video and Text via Large-Scale Discriminative Clustering.Antoine Miech, Jean-Baptiste Alayrac, Piotr Bojanowski, Ivan Laptev, Josef Sivic
2017ICCVWeakly-Supervised Learning of Visual Relations.Julia Peyre, Ivan Laptev, Cordelia Schmid, Josef Sivic
2016CVPRUnsupervised Learning from Narrated Instruction Videos.Jean-Baptiste Alayrac, Piotr Bojanowski, Nishant Agrawal, Josef Sivic, Ivan Laptev, Simon Lacoste-Julien
2016CVPRNetVLAD: CNN Architecture for Weakly Supervised Place Recognition.Relja Arandjelovic, Petr Gront, Akihiko Torii, Toms Pajdla, Josef Sivic
2015CVPROn pairwise costs for network flow multi-object tracking.Visesh Chari, Simon Lacoste-Julien, Ivan Laptev, Josef Sivic
2015CVPRIs object localization for free? - Weakly-supervised learning with convolutional neural networks.Maxime Oquab, Lon Bottou, Ivan Laptev, Josef Sivic
2015CVPR24/7 place recognition by view synthesis.Akihiko Torii, Relja Arandjelovic, Josef Sivic, Masatoshi Okutomi, Toms Pajdla
2015ICCPLinking Past to Present: Discovering Style in Two Centuries of Architecture.Stefan Lee, Nicolas Maisonneuve, David J. Crandall, Alexei A. Efros, Josef Sivic
2014CVPRSeeing 3D Chairs: Exemplar Part-Based 2D-3D Alignment Using a Large Dataset of CAD Models.Mathieu Aubry, Daniel Maturana, Alexei A. Efros, Bryan C. Russell, Josef Sivic
2014CVPRLearning and Transferring Mid-level Image Representations Using Convolutional Neural Networks.Maxime Oquab, Lon Bottou, Ivan Laptev, Josef Sivic
2014ECCVWeakly Supervised Action Labeling in Videos under Ordering Constraints.Piotr Bojanowski, Rmi Lajugie, Francis R. Bach, Ivan Laptev, Jean Ponce, Cordelia Schmid, Josef Sivic
2014ECCVPredicting Actions from Static Scenes.Tuan-Hung Vu, Catherine Olsson, Ivan Laptev, Aude Oliva, Josef Sivic
2013CVPRLearning and Calibrating Per-Location Classifiers for Visual Place Recognition.Petr Gront, Guillaume Obozinski, Josef Sivic, Toms Pajdla
2013CVPRVisual Place Recognition with Repetitive Structures.Akihiko Torii, Josef Sivic, Toms Pajdla, Masatoshi Okutomi
2013ICCVPose Estimation and Segmentation of People in 3D Movies.Karteek Alahari, Guillaume Seguin, Josef Sivic, Ivan Laptev
2013ICCVFinding Actors and Actions in Movies.Piotr Bojanowski, Francis R. Bach, Ivan Laptev, Jean Ponce, Cordelia Schmid, Josef Sivic
2012ECCVScene Semantics from Long-Term Observation of People.Vincent Delaitre, David F. Fouhey, Ivan Laptev, Josef Sivic, Abhinav Gupta, Alexei A. Efros
2012ECCVPeople Watching: Human Actions as a Cue for Single View Geometry.David F. Fouhey, Vincent Delaitre, Abhinav Gupta, Alexei A. Efros, Ivan Laptev, Josef Sivic
2011CVPRTrack to the future: Spatio-temporal video segmentation with long-range motion cues.Jos Lezama, Karteek Alahari, Josef Sivic, Ivan Laptev
2011ICCVDensity-aware person detection and tracking in crowds.Mikel Rodriguez, Ivan Laptev, Josef Sivic, Jean-Yves Audibert
2011ICCVData-driven crowd analysis in videos.Mikel Rodriguez, Josef Sivic, Ivan Laptev, Jean-Yves Audibert
2010BMVCRecognizing human actions in still images: a study of bag-of-features and part-based representations.Vincent Delaitre, Ivan Laptev, Josef Sivic
2010CVPRNon-uniform deblurring for shaken images.Oliver Whyte, Josef Sivic, Andrew Zisserman, Jean Ponce
2010ECCVSemi-supervised Learning of Facial Attributes in Video.Neva Cherniavsky, Ivan Laptev, Josef Sivic, Andrew Zisserman
2010ECCVAvoiding Confusing Features in Place Recognition.Jan Knopp, Josef Sivic, Toms Pajdla
2010ECCVDescriptor Learning for Efficient Retrieval.James Philbin, Michael Isard, Josef Sivic, Andrew Zisserman
2009BMVCGet Out of my Picture! Internet-based Inpainting.Oliver Whyte, Josef Sivic, Andrew Zisserman
2009CVPR"Who are you?" - Learning person specific classifiers from video.Josef Sivic, Mark Everingham, Andrew Zisserman
2009ICCVAutomatic annotation of human actions in video.Olivier Duchenne, Ivan Laptev, Josef Sivic, Francis R. Bach, Jean Ponce
2008BMVCGeometric LDA: A Generative Model for Particular Object Discovery.James Philbin, Josef Sivic, Andrew Zisserman
2008CVPRLost in quantization: Improving particular object retrieval in large scale image databases.James Philbin, Ondrej Chum, Michael Isard, Josef Sivic, Andrew Zisserman
2008CVPRCreating and exploring a large photorealistic virtual space.Josef Sivic, Biliana Kaneva, Antonio Torralba, Shai Avidan, William T. Freeman
2008CVPRUnsupervised discovery of visual object class hierarchies.Josef Sivic, Bryan C. Russell, Andrew Zisserman, William T. Freeman, Alexei A. Efros
2008ECCVSIFT Flow: Dense Correspondence across Different Scenes.Ce Liu, Jenny Yuen, Antonio Torralba, Josef Sivic, William T. Freeman
2007CVPRObject retrieval with large vocabularies and fast spatial matching.James Philbin, Ondrej Chum, Michael Isard, Josef Sivic, Andrew Zisserman
2007ICCVTotal Recall: Automatic Query Expansion with a Generative Feature Model for Object Retrieval.Ondrej Chum, James Philbin, Josef Sivic, Michael Isard, Andrew Zisserman
2006BMVCHello! My name is... Buffy'' -- Automatic Naming of Characters in TV Video.Mark Everingham, Josef Sivic, Andrew Zisserman
2006BMVCFinding People in Repeated Shots of the Same Scene.Josef Sivic, C. Lawrence Zitnick, Richard Szeliski
2006CVPRUsing Multiple Segmentations to Discover Objects and their Extent in Image Collections.Bryan C. Russell, William T. Freeman, Alexei A. Efros, Josef Sivic, Andrew Zisserman
2005ICCVDiscovering Objects and their Localization in Images.Josef Sivic, Bryan C. Russell, Alexei A. Efros, Andrew Zisserman, William T. Freeman
2004CVPRVideo Data Mining Using Configurations of Viewpoint Invariant Regions.Josef Sivic, Andrew Zisserman
2004ECCVObject Level Grouping for Video Shots.Josef Sivic, Frederik Schaffalitzky, Andrew Zisserman
2003ICCVVideo Google: A Text Retrieval Approach to Object Matching in Videos.Josef Sivic, Andrew Zisserman