| 2025 | ExpertAF: Expert Actionable Feedback from Video. | Kumar Ashutosh, Tushar Nagarajan, Georgios Pavlakos, Kris Kitani, Kristen Grauman |
| 2025 | Seeing like a Cephalopod: Colour Vision with a Monochrome Event Camera. | Sami Arja, Nimrod Kruger, Alexandre Marcireau, Nicholas Owen Ralph, Saeed Afshar, Gregory Cohen |
| 2025 | Learned Lightweight Smartphone ISP with Unpaired Data. | Andrei Arhire, Radu Timofte |
| 2025 | LLMPi: Optimizing LLMs for High-Throughput on Raspberry Pi. | Mahsa Ardakani, Jinendra Malekar, Ramtin Zand |
| 2025 | CAV-MAE Sync: Improving Contrastive Audio-Visual Mask Autoencoders via Fine-Grained Alignment. | Edson Araujo, Andrew Rouditchenko, Yuan Gong, Saurabhchand Bhati, Samuel Thomas, Brian Kingsbury, Leonid Karlinsky, Rogrio Feris, James R. Glass, Hilde Kuehne |
| 2025 | Making Every Event Count: Balancing Data Efficiency and Accuracy in Event Camera Subsampling. | Hesam Araghi, Jan van Gemert, Nergis Tomen |
| 2025 | Decoupling Identity Confounders for Enhanced Facial Expression Recognition: An Information-Theoretic Approach. | Mohd Aquib, Nishchal K. Verma, M. Jaleel Akhtar |
| 2025 | Open-World Amodal Appearance Completion. | Jiayang Ao, Yanbei Jiang, Qiuhong Ke, Krista A. Ehinger |
| 2025 | CryptoFace: End-to-End Encrypted Face Recognition. | Wei Ao, Vishnu Naresh Boddeti |
| 2025 | CheXwhatsApp: A Dataset for Exploring Challenges in the Diagnosis of Chest X-rays through Mobile Devices. | Mariamma Antony, Rajiv Porana, Sahil M. Lathiya, Siva Teja Kakileti, Chiranjib Bhattacharyya |
| 2025 | Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention. | Wenbin An, Feng Tian, Sicong Leng, Jiahao Nie, Haonan Lin, Qianying Wang, Ping Chen, Xiaoqin Zhang, Shijian Lu |
| 2025 | Generalized Few-shot 3D Point Cloud Segmentation with Vision-Language Model. | Zhaochong An, Guolei Sun, Yun Liu, Runjia Li, Junlin Han, Ender Konukoglu, Serge J. Belongie |
| 2025 | Cross-View Completion Models are Zero-shot Correspondence Estimators. | Honggyu An, Jin Hyeon Kim, Seonghoon Park, Jaewoo Jung, Jisang Han, Sunghwan Hong, Seungryong Kim |
| 2025 | NExNet Seg: Neuron Expansion Network for Medical Image Segmentation. | Abel A. Reyes Angulo, Sidike Paheding |
| 2025 | Detecting Localized Deepfake Manipulations Using Action Unit-Guided Video Representations. | Tharun Anand, Siva Sankar, Pravin Nair |
| 2025 | FlexiDiT: Your Diffusion Transformer Can Easily Generate High-Quality Samples with Less Compute. | Sotiris Anagnostidis, Gregor Bachmann, Yeongmin Kim, Jonas Kohler, Markos Georgopoulos, Artsiom Sanakoyeu, Yuming Du, Albert Pumarola, Ali K. Thabet, Edgar Schnfeld |
| 2025 | Sample- and Parameter-Efficient Auto-Regressive Image Models. | Elad Amrani, Leonid Karlinsky, Alex M. Bronstein |
| 2025 | Defending Against Frequency-Based Attacks with Diffusion Models. | Fatemeh Amerehi, Patrick Healy |
| 2025 | s2p-hd: Gpu-Accelerated Binocular Stereo Pipeline for Large-Scale Same-Date Stereo. | Tristan Amadei, Enric Meinhardt-Llopis, Carlo de Franchis, Jrmy Anger, Thibaud Ehret, Gabriele Facciolo |
| 2025 | Generative Multiview Relighting for 3D Reconstruction under Extreme Illumination Variation. | Hadi Alzayer, Philipp Henzler, Jonathan T. Barron, Jia-Bin Huang, Pratul P. Srinivasan, Dor Verbin |
| 2025 | CLIP-SLA: Parameter-Efficient CLIP Adaptation for Continuous Sign Language Recognition. | Sarah N. Alyami, Hamzah Luqman |
| 2025 | Read My Ears! Horse Ear Movement Detection for Equine Affective State Assessment. | Joo Alves, Pia Haubro Andersen, Rikke Gade |
| 2025 | DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models. | Saeed Ranjbar Alvar, Gursimran Singh, Mohammad Akbari, Yong Zhang |
| 2025 | Enhanced Multi-View Pedestrian Detection Using Probabilistic Occupancy Volume. | Reef Alturki, Adrian Hilton, Jean-Yves Guillemaut |
| 2025 | CE-NPBG: Connectivity Enhanced Neural Point-Based Graphics for Novel View Synthesis in Autonomous Driving Scenes. | Mohammad Altillawi, Fengyi Shen, Liudi Yang, Sai Manoj Prakhya, Ziyuan Liu |