| 2025 | Hijacking Vision-and-Language Navigation Agents with Adversarial Environmental Attacks. | Zijiao Yang, Xiangxi Shi, Eric Slyman, Stefan Lee |
| 2025 | MegaFusion: Extend Diffusion Models towards Higher-resolution Image Generation without Further Tuning. | Haoning Wu, Shaocheng Shen, Qiang Hu, Xiaoyun Zhang, Ya Zhang, Yanfeng Wang |
| 2025 | Adaptive Deviation Learning for Visual Anomaly Detection with Data Contamination. | Anindya Sundar Das, Guansong Pang, Monowar Bhuyan |
| 2025 | Active Learning for Vision-Language Models. | Bardia Safaei, Vishal M. Patel |
| 2025 | Breaking the Frame: Visual Place Recognition by Overlap Prediction. | Tong Wei, Philipp Lindenberger, Jir Matas, Daniel Barath |
| 2025 | DreaMo: Articulated 3D Reconstruction from a Single Casual Video. | Tao Tu, Ming-Feng Li, Chieh Hubert Lin, Yen-Chi Cheng, Min Sun, Ming-Hsuan Yang |
| 2025 | Event-Guided Low-Light Video Semantic Segmentation. | Zhen Yao, Mooi Choo Chuah |
| 2025 | Loose Social-Interaction Recognition in Real-World Therapy Scenarios. | Abid Ali, Rui Dai, Ashish Marisetty, Guillaume Astruc, Monique Thonnat, Jean-Marc Odobez, Susanne Thmmler, Franois Brmond |
| 2025 | DiffPAD: Denoising Diffusion-Based Adversarial Patch Decontamination. | Jia Fu, Xiao Zhang, Sepideh Pashami, Fatemeh Rahimian, Anders Holst |
| 2025 | Effective Scene Graph Generation by Statistical Relation Distillation. | Thanh-Son Nguyen, Hong Yang, Basura Fernando |
| 2025 | Feasibility of Federated Learning from Client Databases with Different Brain Diseases and MRI Modalities. | Felix Wagner, Wentian Xu, Pramit Saha, Ziyun Liang, Daniel Whitehouse, David K. Menon, Virginia F. J. Newcombe, Natalie Voets, J. Alison Noble, Konstantinos Kamnitsas |
| 2025 | A Rapid Test for Accuracy and Bias of Face Recognition Technology. | Manuel Knott, Ignacio Serna, Ethan Mann, Pietro Perona |
| 2025 | Generalizable Single-View Object Pose Estimation by Two-Side Generating and Matching. | Yujing Sun, Caiyi Sun, Yuan Liu, Yuexin Ma, Siu Ming Yiu |
| 2025 | Enhancing Novel Object Detection via Cooperative Foundational Models. | Rohit K. Bharadwaj, Muzammal Naseer, Salman Khan, Fahad Shahbaz Khan |
| 2025 | BioNet and NeFF: Crop Biomass Prediction from Point Clouds to Drone Imagery. | Xuesong Li, Zeeshan Hayder, Ali Zia, Connor Cassidy, Shiming Liu, Warwick Stiller, Eric A. Stone, Warren Conaty, Lars Petersson, Vivien Rolland |
| 2025 | Guardian of the Ensembles: Introducing Pairwise Adversarially Robust Loss for Resisting Adversarial Attacks in DNN Ensembles. | Shubhi Shukla, Subhadeep Dalui, Manaar Alam, Shubhajit Datta, Arijit Mondal, Debdeep Mukhopadhyay, Partha Pratim Chakrabarti |
| 2025 | Flatness Improves Backbone Generalisation in Few-Shot Classification. | Rui Li, Martin Trapp, Marcus Klasson, Arno Solin |
| 2024 | Learning to Recognize Occluded and Small Objects with Partial Inputs. | Hasib Zunair, A. Ben Hamza |
| 2024 | RDIR: Capturing Temporally-Invariant Representations of Multiple Objects in Videos. | Piotr Zielinski, Tomasz Kajdanowicz |
| 2024 | Multimodal Deep Learning for Remote Stress Estimation Using CCT-LSTM. | Sayyedjavad Ziaratnia, Tipporn Laohakangvalvit, Midori Sugaya, Peeraya Sripian |
| 2024 | ShARc: Shape and Appearance Recognition for Person Identification In-the-wild. | Haidong Zhu, Wanrong Zheng, Zhaoheng Zheng, Ram Nevatia |
| 2024 | TinyWT: A Large-Scale Wind Turbine Dataset of Satellite Images for Tiny Object Detection. | Mingye Zhu, Zhicheng Yang, Hang Zhou, Chen Du, Andy J. Y. Wong, Yibing Wei, Zhuo Deng, Mei Han, Jui-Hsin Lai |
| 2024 | CATS: Combined Activation and Temporal Suppression for Efficient Network Inference. | Zeqi Zhu, Arash Pourtaherian, Luc Waeijen, Ibrahim Batuhan Akkaya, Egor Bondarev, Orlando Moreira |
| 2024 | FELGA: Unsupervised Fragment Embedding for Fine-Grained Cross-Modal Association. | Yaoxin Zhuo, Baoxin Li |
| 2024 | Consistent Multimodal Generation via A Unified GAN Framework. | Zhen Zhu, Yijun Li, Weijie Lyu, Krishna Kumar Singh, Zhixin Shu, Sren Pirk, Derek Hoiem |