Skip to content

Zsolt Kira

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

75

Venues

15

Active years

2004–2026

Best venue rank

A*

Where they publish

Papers

75 indexed papers, newest first.

YearVenueTitleAuthors
2026WACVGrounding Descriptions in Images informs Zero-Shot Visual Recognition.Shaunak Halbe, Junjiao Tian, K. J. Joseph, James Seale Smith, Katherine Stevo, Vineeth N. Balasubramanian, Zsolt Kira
2025CVPRFRAMES-VQA: Benchmarking Fine-Tuning Robustness across Multi-Modal Shifts in Visual Question Answering.Chengyue Huang, Brisa Maneechotesuwan, Shivang Chopra, Zsolt Kira
2025CVPRWhen Domain Generalization meets Generalized Category Discovery: An Adaptive Task-Arithmetic Driven Approach.Vaibhav Rathore, Shubhranil B, Saikat Dutta, Sarthak Mehrotra, Zsolt Kira, Biplab Banerjee
2025CVPRFrom Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons.Andrew Szot, Bogdan Mazoure, Omar Attia, Aleksei Timofeev, Harsh Agrawal, R. Devon Hjelm, Zhe Gan, Zsolt Kira, Alexander Toshev
2025ICCVEmbodiedSplat: Personalized Real-To-Sim-To-Real Navigation with Gaussian Splats From a Mobile Device.Gunjan Chhablani, Xiaomeng Ye, Muhammad Zubair Irshad, Zsolt Kira
2025ICLRContextual Self-paced Learning for Weakly Supervised Spatio-Temporal Video Grounding.Akash Kumar, Zsolt Kira, Yogesh S. Rawat
2025ICLRDirectional Gradient Projection for Robust Fine-Tuning of Foundation Models.Chengyue Huang, Junjiao Tian, Brisa Maneechotesuwan, Shivang Chopra, Zsolt Kira
2025IJCAIRenderBender: A Survey on Adversarial Attacks Using Differentiable Rendering.Matthew Hull, Haoran Wang, Matthew Lau, Alec Helbling, Mansi Phute, Chao Zhang, Zsolt Kira, Willian T. Lunardi, Martin Andreoni, Wenke Lee, Duen Horng Chau
2024CVPRGOAT-Bench: A Benchmark for Multi-Modal Lifelong Navigation.Mukul Khanna, Ram Ramrakhya, Gunjan Chhablani, Sriram Yenamandra, Thophile Gervet, Matthew Chang, Zsolt Kira, Devendra Singh Chaplot, Dhruv Batra, Roozbeh Mottaghi
2024CVPRSeeing the Unseen: Visual Common Sense for Semantic Placement.Ram Ramrakhya, Aniruddha Kembhavi, Dhruv Batra, Zsolt Kira, Kuo-Hao Zeng, Luca Weihs
2024CVPRContinual Diffusion with STAMINA: STack-And-Mask INcremental Adapters.James Seale Smith, Yen-Chang Hsu, Zsolt Kira, Yilin Shen, Hongxia Jin
2024CVPRAdaptive Memory Replay for Continual Learning.James Seale Smith, Lazar Valkov, Shaunak Halbe, Vyshnavi Gutta, Rogrio Feris, Zsolt Kira, Leonid Karlinsky
2024CVPRDiffuse, Attend, and Segment: Unsupervised Zero-Shot Segmentation using Stable Diffusion.Junjiao Tian, Lavisha Aggarwal, Andrea Colaco, Zsolt Kira, Mar Gonzlez-Franco
2024ECCVReinforcement Learning via Auxiliary Task Distillation.Abhinav Narayan Harish, Larry Heck, Josiah P. Hanna, Zsolt Kira, Andrew Szot
2024ECCVNeRF-MAE: Masked AutoEncoders for Self-supervised 3D Representation Learning for Neural Radiance Fields.Muhammad Zubair Irshad, Sergey Zakharov, Vitor Guizilini, Adrien Gaidon, Zsolt Kira, Rares Ambrus
2024ICLRHabitat 3.0: A Co-Habitat for Humans, Avatars, and Robots.Xavier Puig, Eric Undersander, Andrew Szot, Mikael Dallaire Cote, Tsung-Yen Yang, Ruslan Partsey, Ruta Desai, Alexander Clegg, Michal Hlavac, So Yeon Min, Vladimir Vondrus, Thophile Gervet, Vincent-Pierre Berges, John M. Turner, Oleksandr Maksymets, Zsolt Kira, Mrinal Kalakrishnan, Jitendra Malik, Devendra Singh Chaplot, Unnat Jain, Dhruv Batra, Akshara Rai, Roozbeh Mottaghi
2024ICRAN-QR: Natural Quick Response Codes for Multi-Robot Instance Correspondence.Nathaniel Moore Glaser, Rajashree Ravi, Zsolt Kira
2024ICRAFSD: Fast Self-Supervised Single RGB-D to Categorical 3D Objects.Mayank Lunayach, Sergey Zakharov, Dian Chen, Rares Ambrus, Zsolt Kira, Muhammad Zubair Irshad
2024WACVLatentDR: Improving Model Generalization Through Sample-Aware Latent Degradation and Restoration.Ran Liu, Sahil Khose, Jingyun Xiao, Lakshmi Sathidevi, Keerthan Ramnath, Zsolt Kira, Eva L. Dyer
2024WACVMissing Modality Robustness in Semi-Supervised Multi-Modal Semantic Segmentation.Harsh Maheshwari, Yen-Cheng Liu, Zsolt Kira
2023CoRLHomeRobot: Open-Vocabulary Mobile Manipulation.Sriram Yenamandra, Arun Ramachandran, Karmesh Yadav, Austin S. Wang, Mukul Khanna, Thophile Gervet, Tsung-Yen Yang, Vidhi Jain, Alexander Clegg, John M. Turner, Zsolt Kira, Manolis Savva, Angel X. Chang, Devendra Singh Chaplot, Dhruv Batra, Roozbeh Mottaghi, Yonatan Bisk, Chris Paxton
2023CVPRHAAV: Hierarchical Aggregation of Augmented Views for Image Captioning.Chia-Wen Kuo, Zsolt Kira
2023CVPRConStruct-VL: Data-Free Continual Structured VL Concepts Learning.James Seale Smith, Paola Cascante-Bonilla, Assaf Arbelle, Donghyun Kim, Rameswar Panda, David D. Cox, Diyi Yang, Zsolt Kira, Rogrio Feris, Leonid Karlinsky
2023CVPRCODA-Prompt: COntinual Decomposed Attention-Based Prompting for Rehearsal-Free Continual Learning.James Seale Smith, Leonid Karlinsky, Vyshnavi Gutta, Paola Cascante-Bonilla, Donghyun Kim, Assaf Arbelle, Rameswar Panda, Rogrio Feris, Zsolt Kira
2023CVPRA Closer Look at Rehearsal-Free Continual Learning.James Seale Smith, Junjiao Tian, Shaunak Halbe, Yen-Chang Hsu, Zsolt Kira
2023CVPRTrainable Projected Gradient Method for Robust Fine-Tuning.Junjiao Tian, Xiaoliang Dai, Chih-Yao Ma, Zecheng He, Yen-Cheng Liu, Zsolt Kira
2023ICCVNeO 360: Neural Fields for Sparse View Synthesis of Outdoor Scenes.Muhammad Zubair Irshad, Sergey Zakharov, Katherine Liu, Vitor Guizilini, Thomas Kollar, Adrien Gaidon, Zsolt Kira, Rares Ambrus
2023ICLRBC-IRL: Learning Generalizable Reward Functions from Demonstrations.Andrew Szot, Amy Zhang, Dhruv Batra, Zsolt Kira, Franziska Meier
2023ICMLAdaptive Coordination in Social Embodied Rearrangement.Andrew Szot, Unnat Jain, Dhruv Batra, Zsolt Kira, Ruta Desai, Akshara Rai
2023IJCNNConstraintMatch for Semi-constrained Clustering.Jann Goschenhofer, Bernd Bischl, Zsolt Kira
2023ICRACommunication-Critical Planning via Multi-Agent Trajectory Exchange.Nathaniel Moore Glaser, Zsolt Kira
2023WACVStructure-Encoding Auxiliary Tasks for Improved Visual Representation in Vision-and-Language Navigation.Chia-Wen Kuo, Chih-Yao Ma, Judy Hoffman, Zsolt Kira
2022CVPRBeyond a Pre-Trained Object Detector: Cross-Modal Textual and Visual Context for Image Captioning.Chia-Wen Kuo, Zsolt Kira
2022CVPRUnbiased Teacher v2: Semi-supervised Object Detection for Anchor-free and Anchor-based Detectors.Yen-Cheng Liu, Chih-Yao Ma, Zsolt Kira
2022ECCVShAPO: Implicit Representations for Multi-object Shape, Appearance, and Pose Optimization.Muhammad Zubair Irshad, Sergey Zakharov, Rares Ambrus, Thomas Kollar, Zsolt Kira, Adrien Gaidon
2022ECCVOpen-Set Semi-Supervised Object Detection.Yen-Cheng Liu, Chih-Yao Ma, Xiaoliang Dai, Junjiao Tian, Peter Vajda, Zijian He, Zsolt Kira
2022ICRACenterSnap: Single-Shot Multi-Object 3D Shape Reconstruction and Categorical 6D Pose and Size Estimation.Muhammad Zubair Irshad, Thomas Kollar, Michael Laskey, Kevin Stone, Zsolt Kira
2022ICRAStriking the Right Balance: Recall Loss for Semantic Segmentation.Junjiao Tian, Niluthpol Chowdhury Mithun, Zachary Seymour, Han-Pang Chiu, Zsolt Kira
2021ICCVAlways Be Dreaming: A New Approach for Data-Free Class-Incremental Learning.James Seale Smith, Yen-Chang Hsu, Jonathan C. Balloch, Yilin Shen, Hongxia Jin, Zsolt Kira
2021ICLRUnbiased Teacher for Semi-Supervised Object Detection.Yen-Cheng Liu, Chih-Yao Ma, Zijian He, Chia-Wen Kuo, Kan Chen, Peizhao Zhang, Bichen Wu, Zsolt Kira, Peter Vajda
2021ICRAHierarchical Cross-Modal Agent for Robotics Vision-and-Language Navigation.Muhammad Zubair Irshad, Chih-Yao Ma, Zsolt Kira
2021IJCNNMemory-Efficient Semi-Supervised Continual Learning: The World is its Own Replay Buffer.James Seale Smith, Jonathan C. Balloch, Yen-Chang Hsu, Zsolt Kira
2021IROSOvercoming Obstructions via Bandwidth-Limited Multi-Agent Spatial Handshaking.Nathaniel Glaser, Yen-Cheng Liu, Junjiao Tian, Zsolt Kira
2020AAAIPath Ranking with Attention to Type Hierarchies.Weiyu Liu, Angel Andres Daruna, Zsolt Kira, Sonia Chernova
2020CVPRAction Segmentation With Joint Self-Supervised Temporal Domain Adaptation.Min-Hung Chen, Baopu Li, Yingze Bao, Ghassan AlRegib, Zsolt Kira
2020CVPRGeneralized ODIN: Detecting Out-of-Distribution Image Without Learning From Out-of-Distribution Data.Yen-Chang Hsu, Yilin Shen, Hongxia Jin, Zsolt Kira
2020CVPRWhen2com: Multi-Agent Perception via Communication Graph Grouping.Yen-Cheng Liu, Junjiao Tian, Nathaniel Glaser, Zsolt Kira
2020ECCVFeatMatch: Feature-Based Augmentation for Semi-supervised Learning.Chia-Wen Kuo, Chih-Yao Ma, Jia-Bin Huang, Zsolt Kira
2020ECCVLearning to Generate Grounded Visual Captions Without Localization Supervision.Chih-Yao Ma, Yannis Kalantidis, Ghassan AlRegib, Peter Vajda, Marcus Rohrbach, Zsolt Kira
2020ICRAWho2com: Collaborative Perception via Learnable Handshake Communication.Yen-Cheng Liu, Junjiao Tian, Chih-Yao Ma, Nathan Glaser, Chia-Wen Kuo, Zsolt Kira
2020ICRAUNO: Uncertainty-aware Noisy-Or Multimodal Fusion for Unanticipated Input Degradation.Junjiao Tian, Wesley Cheung, Nathaniel Glaser, Yen-Cheng Liu, Zsolt Kira
2019CVPRThe Regretful Agent: Heuristic-Aided Navigation Through Progress Estimation.Chih-Yao Ma, Zuxuan Wu, Ghassan AlRegib, Caiming Xiong, Zsolt Kira
2019ICCVTemporal Attentive Alignment for Large-Scale Video Domain Adaptation.Min-Hung Chen, Zsolt Kira, Ghassan Alregib, Jaekwon Yoo, Ruxin Chen, Jian Zheng
2019ICLRA Closer Look at Few-shot Classification.Wei-Yu Chen, Yen-Cheng Liu, Zsolt Kira, Yu-Chiang Frank Wang, Jia-Bin Huang
2019ICLRMulti-class classification without multi-class labels.Yen-Chang Hsu, Zhaoyang Lv, Joel Schlosser, Phillip Odom, Zsolt Kira
2019ICLRSelf-Monitoring Navigation Agent via Auxiliary Progress Estimation.Chih-Yao Ma, Jiasen Lu, Zuxuan Wu, Ghassan AlRegib, Zsolt Kira, Richard Socher, Caiming Xiong
2019ICRARoboCSE: Robot Common Sense Embedding.Angel Andres Daruna, Weiyu Liu, Zsolt Kira, Sonia Chernova
2019WACVData-Efficient Graph Embedding Learning for PCB Component Detection.Chia-Wen Kuo, Jacob Ashmore, David Huggins, Zsolt Kira
2018CVPRAttend and Interact: Higher-Order Object Interactions for Video Understanding.Chih-Yao Ma, Asim Kadav, Iain Melvin, Zsolt Kira, Ghassan AlRegib, Hans Peter Graf
2018ICLRLearning to cluster in order to transfer across domains and tasks.Yen-Chang Hsu, Zhaoyang Lv, Zsolt Kira
2018IJCNNLearning to Cluster for Proposal-Free Instance Segmentation.Yen-Chang Hsu, Zheng Xu, Zsolt Kira, Jiawei Huang
2016ECCVA Continuous Optimization Approach for Efficient and Accurate Scene Flow.Zhaoyang Lv, Chris Beall, Pablo F. Alcantarilla, Fuxin Li, Zsolt Kira, Frank Dellaert
2016ICRAFusing LIDAR and images for pedestrian detection using convolutional neural networks.Joel Schlosser, Christopher K. Chow, Zsolt Kira
2015ICRAAn evaluation of features for classifier transfer during target handoff across aerial and ground robots.Zsolt Kira
2014BMVCMining Structure Fragments for Smart Bundle Adjustment.Luca Carlone, Pablo Fernndez Alcantarilla, Han-Pang Chiu, Zsolt Kira, Frank Dellaert
2014ICRAEliminating conditionally independent sets in factor graphs: A unifying perspective based on smart factors.Luca Carlone, Zsolt Kira, Chris Beall, Vadim Indelman, Frank Dellaert
2014IROSTransfer of sparse coding representations and object classifiers across heterogeneous robots.Zsolt Kira
2012ICASSPUnsupervised topic modeling for leader detection in spoken discourse.Raia Hadsell, Zsolt Kira, Wen Wang, Kristin Precoda
2012ICASSPDetecting leadership and cohesion in spoken interactions.Wen Wang, Kristin Precoda, Raia Hadsell, Zsolt Kira, Colleen Richey, Gabriel Jiva
2012IROSLong-Range Pedestrian Detection using stereo and a cascade of convolutional network classifiers.Zsolt Kira, Raia Hadsell, Garbis Salgian, Supun Samarasekera
2009FlAIRSMapping Grounded Object Properties across Perceptually Heterogeneous Embodiments.Zsolt Kira
2009IROSTransferring embodied concepts between perceptually heterogeneous robots.Zsolt Kira
2007IROSModeling cross-sensory and sensorimotor correlations to detect and localize faults in mobile robots.Zsolt Kira
2006IROSContinuous and Embedded Learning for Multi-Agent Systems.Zsolt Kira, Alan C. Schultz
2004IROSForgetting bad behavior: memory for case-based navigation.Zsolt Kira, Ronald C. Arkin