Skip to content

Anna Rohrbach

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

46

Venues

13

Active years

2015–2026

Best venue rank

A*

Where they publish

Papers

46 indexed papers, newest first.

YearVenueTitleAuthors
2026ACLVeriTaS: The First Dynamic Benchmark for Multimodal Automated Fact-Checking.Mark Rothermel, Marcus Kornmann, Marcus Rohrbach, Anna Rohrbach
2025CVPRV^2Dial: Unification of Video and Visual Dialog via Multimodal Experts.Adnen Abdessaied, Anna Rohrbach, Marcus Rohrbach, Andreas Bulling
2025ICCVChrono: A Simple Blueprint for Representing Time in MLLMs.Boris Meinardus, Hector G. Rodriguez, Anil Batra, Anna Rohrbach, Marcus Rohrbach
2025ICMLDEFAME: Dynamic Evidence-based FAct-checking with Multimodal Experts.Tobias Braun, Mark Rothermel, Marcus Rohrbach, Anna Rohrbach
2024WACVShape-Guided Diffusion with Inside-Outside Attention.Dong Huk Park, Grace Luo, Clayton Toste, Samaneh Azadi, Xihui Liu, Maka Karalashvili, Anna Rohrbach, Trevor Darrell
2024WACVSimple Token-Level Confidence Improves Caption Correctness.Suzanne Petryk, Spencer Whitehead, Joseph E. Gonzalez, Trevor Darrell, Anna Rohrbach, Marcus Rohrbach
2023CVPRMammalNet: A Large-Scale Video Benchmark for Mammal Recognition and Behavior Understanding.Jun Chen, Ming Hu, Darren J. Coker, Michael L. Berumen, Blair R. Costelloe, Sara Beery, Anna Rohrbach, Mohamed Elhoseiny
2023ICLRUsing Language to Extend to Unseen Domains.Lisa Dunlap, Clara Mohri, Devin Guillory, Han Zhang, Trevor Darrell, Joseph E. Gonzalez, Aditi Raghunathan, Anna Rohrbach
2023WACVWatch Those Words: Video Falsification Detection Using Word-Conditioned Facial Motion.Shruti Agarwal, Liwen Hu, Evonne Ng, Trevor Darrell, Hao Li, Anna Rohrbach
2023WACVMore Control for Free! Image Synthesis with Semantic Diffusion Guidance.Xihui Liu, Dong Huk Park, Samaneh Azadi, Gong Zhang, Arman Chopikyan, Yuxiao Hu, Humphrey Shi, Anna Rohrbach, Trevor Darrell
2022ACLReCLIP: A Strong Zero-Shot Baseline for Referring Expression Comprehension.Sanjay Subramanian, William Merrill, Trevor Darrell, Matt Gardner, Sameer Singh, Anna Rohrbach
2022CVPRDETReg: Unsupervised Pretraining with Region Priors for Object Detection.Amir Bar, Xin Wang, Vadim Kantorov, Colorado J. Reed, Roei Herzig, Gal Chechik, Anna Rohrbach, Trevor Darrell, Amir Globerson
2022CVPRObject-Region Video Transformers.Roei Herzig, Elad Ben-Avraham, Karttikeya Mangalam, Amir Bar, Gal Chechik, Anna Rohrbach, Trevor Darrell, Amir Globerson
2022CVPROn Guiding Visual Attention with Language Specification.Suzanne Petryk, Lisa Dunlap, Keyan Nasseri, Joseph Gonzalez, Trevor Darrell, Anna Rohrbach
2022ECCVThe Abduction of Sherlock Holmes: A Dataset for Visual Abductive Reasoning.Jack Hessel, Jena D. Hwang, Jae Sung Park, Rowan Zellers, Chandra Bhagavatula, Anna Rohrbach, Kate Saenko, Yejin Choi
2022ECCVTL;DW? Summarizing Instructional Videos with Task Relevance and Cross-Modal Saliency.Medhini Narasimhan, Arsha Nagrani, Chen Sun, Michael Rubinstein, Trevor Darrell, Anna Rohrbach, Cordelia Schmid
2022ECCVReliable Visual Question Answering: Abstain Rather Than Answer Incorrectly.Spencer Whitehead, Suzanne Petryk, Vedaad Shakib, Joseph Gonzalez, Trevor Darrell, Anna Rohrbach, Marcus Rohrbach
2022EMNLPG3: Geolocation via Guidebook Grounding.Grace Luo, Giscard Biamby, Trevor Darrell, Daniel Fried, Anna Rohrbach
2022EMNLPFocus! Relevant and Sufficient Context Selection for News Image Captioning.Mingyang Zhou, Grace Luo, Anna Rohrbach, Zhou Yu
2022ICLRHow Much Can CLIP Benefit Vision-and-Language Tasks?Sheng Shen, Liunian Harold Li, Hao Tan, Mohit Bansal, Anna Rohrbach, Kai-Wei Chang, Zhewei Yao, Kurt Keutzer
2022NAACLTwitter-COMMs: Detecting Climate, COVID, and Military Multimodal Misinformation.Giscard Biamby, Grace Luo, Trevor Darrell, Anna Rohrbach
2022NAACLExposing the Limits of Video-Text Models through Contrast Sets.Jae Sung Park, Sheng Shen, Ali Farhadi, Trevor Darrell, Yejin Choi, Anna Rohrbach
2021EMNLPNewsCLIPpings: Automatic Generation of Out-of-Context Multimodal Media.Grace Luo, Trevor Darrell, Anna Rohrbach
2021ICMLCompositional Video Synthesis with Action Graphs.Amir Bar, Roei Herzig, Xiaolong Wang, Anna Rohrbach, Gal Chechik, Trevor Darrell, Amir Globerson
2020CVPRAdvisable Learning for Self-Driving Vehicles by Internalizing Observation-to-Action Rules.Jinkyu Kim, Suhong Moon, Anna Rohrbach, Trevor Darrell, John F. Canny
2020ECCVIdentity-Aware Multi-sentence Video Description.Jae Sung Park, Trevor Darrell, Anna Rohrbach
2019ACLAre You Looking? Grounding to Multiple Modalities in Vision-and-Language Navigation.Ronghang Hu, Daniel Fried, Anna Rohrbach, Dan Klein, Trevor Darrell, Kate Saenko
2019CVPRAdversarial Inference for Multi-Sentence Video Description.Jae Sung Park, Marcus Rohrbach, Trevor Darrell, Anna Rohrbach
2019CVPRAdversarial Inference for Multi-Sentence Video Description.Jae Sung Park, Marcus Rohrbach, Trevor Darrell, Anna Rohrbach
2019ICCVLanguage-Conditioned Graph Networks for Relational Reasoning.Ronghang Hu, Anna Rohrbach, Trevor Darrell, Kate Saenko
2019ICCVRobust Change Captioning.Dong Huk Park, Trevor Darrell, Anna Rohrbach
2018ACCVVideo Object Segmentation with Language Referring Expressions.Anna Khoreva, Anna Rohrbach, Bernt Schiele
2018CVPRMultimodal Explanations: Justifying Decisions and Pointing to the Evidence.Dong Huk Park, Lisa Anne Hendricks, Zeynep Akata, Anna Rohrbach, Bernt Schiele, Trevor Darrell, Marcus Rohrbach
2018CVPRFooling Vision and Language Models Despite Localization and Attention Mechanism.Xiaojun Xu, Xinyun Chen, Chang Liu, Anna Rohrbach, Trevor Darrell, Dawn Song
2018ECCVWomen Also Snowboard: Overcoming Bias in Captioning Models.Lisa Anne Hendricks, Kaylee Burns, Kate Saenko, Trevor Darrell, Anna Rohrbach
2018ECCVVideo Object Segmentation with Referring Expressions.Anna Khoreva, Anna Rohrbach, Bernt Schiele
2018ECCVTextual Explanations for Self-Driving Vehicles.Jinkyu Kim, Anna Rohrbach, Trevor Darrell, John F. Canny, Zeynep Akata
2018EMNLPObject Hallucination in Image Captioning.Anna Rohrbach, Lisa Anne Hendricks, Kaylee Burns, Trevor Darrell, Kate Saenko
2018LRECA vision-grounded dataset for predicting typical locations for verbs.Nelson Mukuze, Anna Rohrbach, Vera Demberg, Bernt Schiele
2017CoRLGradient-free Policy Architecture Search and Adaptation.Sayna Ebrahimi, Anna Rohrbach, Trevor Darrell
2017CVPRA Dataset and Exploration of Models for Understanding Video Data through Fill-in-the-Blank Question-Answering.Tegan Maharaj, Nicolas Ballas, Anna Rohrbach, Aaron C. Courville, Christopher Joseph Pal
2017CVPRGenerating Descriptions with Grounded and Co-referenced People.Anna Rohrbach, Marcus Rohrbach, Siyu Tang, Seong Joon Oh, Bernt Schiele
2016AAAICommonsense in Parts: Mining Part-Whole Relations from the Web and Image Tags.Niket Tandon, Charles Hariman, Jacopo Urbani, Anna Rohrbach, Marcus Rohrbach, Gerhard Weikum
2016ECCVGrounding of Textual Phrases in Images by Reconstruction.Anna Rohrbach, Marcus Rohrbach, Ronghang Hu, Trevor Darrell, Bernt Schiele
2016EMNLPMultimodal Compact Bilinear Pooling for Visual Question Answering and Visual Grounding.Akira Fukui, Dong Huk Park, Daylen Yang, Anna Rohrbach, Trevor Darrell, Marcus Rohrbach
2015CVPRA dataset for Movie Description.Anna Rohrbach, Marcus Rohrbach, Niket Tandon, Bernt Schiele