| 2025 | Generating Long-Take Videos via Effective Keyframes and Guidance. | Hsin-Ping Huang, Yu-Chuan Su, Ming-Hsuan Yang |
| 2025 | VILLS: Video-Image Learning to Learn Semantics for Person Re-Identification. | Siyuan Huang, Ram Prabhakar, Yuxiang Guo, Rama Chellappa, Cheng Peng |
| 2025 | Who Brings the Frisbee: Probing Hidden Hallucination Factors in Large Vision-Language Model via Causality Analysis. | Po-Hsuan Huang, Jeng-Lin Li, Chin-Po Chen, Ming-Ching Chang, Wei-Chao Chen |
| 2025 | Infant Action Generative Modeling. | Xiaofei Huang, Elaheh Hatamimajoumerd, Amal Mathew, Sarah Ostadabbas |
| 2025 | Dual-Schedule Inversion: Training- and Tuning-Free Inversion for Real Image Editing. | Jiancheng Huang, Yi Huang, Jianzhuang Liu, Donghao Zhou, Yifan Liu, Shifeng Chen |
| 2025 | Spatio-Temporal Context Prompting for Zero-Shot Action Detection. | Wei-Jhe Huang, Min-Hung Chen, Shang-Hong Lai |
| 2025 | VADet: Multi-Frame LiDAR 3D Object Detection Using Variable Aggregation. | Chengjie Huang, Vahdat Abdelzad, Sean Sedwards, Krzysztof Czarnecki |
| 2025 | Recoverable Anonymization for Pose Estimation: A Privacy-Enhancing Approach. | Wenjun Huang, Yang Ni, Arghavan Rezvani, Sungheon Jeong, Hanning Chen, Yezi Liu, Fei Wen, Mohsen Imani |
| 2025 | Covariance-Based Space Regularization for Few-Shot Class Incremental Learning. | Yijie Hu, Guanyu Yang, Zhaorui Tan, Xiaowei Huang, Kaizhu Huang, Qiufeng Wang |
| 2025 | Shadow Removal Refinement via Material-Consistent Shadow Edges. | Shilin Hu, Hieu Le, ShahRukh Athar, Sagnik Das, Dimitris Samaras |
| 2025 | ProMM-RS: Exploring Probabilistic Learning for Multi-Modal Remote Sensing Image Representations. | Nicolas Houdr, Diego Marcos, Dino Ienco, Laurent Wendling, Camille Kurtz, Sylvain Lobry |
| 2025 | SUM: Saliency Unification Through Mamba for Visual Attention Modeling. | Alireza Hosseini, Amirhossein Kazerouni, Saeed Akhavan, Michael Brudno, Babak Taati |
| 2025 | GeoPos: A Minimal Positional Encoding for Enhanced Fine-Grained Details in Image Synthesis Using Convolutional Neural Networks. | Mehran Hosseini, Peyman Hosseini |
| 2025 | Invariant Shape Representation Learning for Image Classification. | Tonmoy Hossain, Jing Ma, Jundong Li, Miaomiao Zhang |
| 2025 | Dense Depth from Event Focal Stack. | Kenta Horikawa, Mariko Isogawa, Hideo Saito, Shohei Mori |
| 2025 | Shape-Biased Texture Agnostic Representations for Improved Textureless and Metallic Object Detection and 6D Pose Estimation. | Peter Hnig, Stefan Thalhammer, Jean-Baptiste Weibel, Matthias Hirschmanner, Markus Vincze |
| 2025 | D2FP: Learning Implicit Prior for Human Parsing. | Junyoung Hong, Hyeri Yang, Ye Ju Kim, Haerim Kim, Shinwoong Kim, Euna Shim, Kyungjae Lee |
| 2025 | Difficulty, Diversity, and Plausibility: Dynamic Data-Free Quantization. | Cheeun Hong, Sungyong Baik, Junghun Oh, Kyoung Mu Lee |
| 2025 | Semantic Clustering of Image Retrieval Databases used for Visual Localization. | Henry Hlzemann, Torsten Fiolka |
| 2025 | NeRFs are Mirror Detectors: Using Structural Similarity for Multi-View Mirror Scene Reconstruction with 3D Surface Primitives. | Leif Van Holland, Michael Weinmann, Jan U. Mller, Patrick Stotko, Reinhard Klein |
| 2025 | Joint Co-Speech Gesture and Expressive Talking Face Generation Using Diffusion with Adapters. | Steven Hogue, Chenxu Zhang, Yapeng Tian, Xiaohu Guo |
| 2025 | On Neural BRDFs: A Thorough Comparison of State-of-the-Art Approaches. | Florian Hofherr, Bjoern Haefner, Daniel Cremers |
| 2025 | F2FLDM: Latent Diffusion Models with Histopathology Pre-Trained Embeddings for Unpaired Frozen Section to FFPE Translation. | Man Minh Ho, Shikha Dubey, Yosep Chong, Beatrice Knudsen, Tolga Tasdizen |
| 2025 | Re-identifying People in Video via Learned Temporal Attention and Multi-modal Foundation Models. | Cole Hill, Florence Yellin, Krishna Regmi, Dawei Du, Scott McCloskey |
| 2025 | SEER-ZSL: Semantic Encoder-Enhanced Representations for Generalized Zero-Shot Learning. | William Heyden, Habib Ullah, Muhammad Salman Siddiqui, Fadi Al Machot |