| 2024 | Instance-Dependent Noise Refinement in Segment Anything Model for Weakly Supervised Object Detection. | Fariborz Taherkhani, Ehsan Kazemi |
| 2024 | HAHA: Highly Articulated Gaussian Human Avatars with Textured Mesh Prior. | David Svitov, Pietro Morerio, Lourdes Agapito, Alessio Del Bue |
| 2024 | GTA: Global Tracklet Association for Multi-object Tracking in Sports. | Jiacheng Sun, Hsiang-Wei Huang, Cheng-Yen Yang, Zhongyu Jiang, Jenq-Neng Hwang |
| 2024 | Seeing Through Expert's Eyes: Leveraging Radiologist Eye Gaze and Speech Report with Graph Neural Networks for Chest X-Ray Image Classification. | Jamalia Sultana, Ruwen Qin, Zhaozheng Yin |
| 2024 | M-RAT: a Multi-grained Retrieval Augmentation Transformer for Image Captioning. | Jiayan Song, Renjie Pan, Jun Zhou, Hua Yang |
| 2024 | Capture Concept Through Comparison: Vision-and-Language Representation Learning with Intrinsic Information Mining. | Yun-Zhu Song, Yi-Syuan Chen, Tzu-Ling Lin, Bei Liu, Jianlong Fu, Hong-Han Shuai |
| 2024 | LoGDesc: Local Geometric Features Aggregation for Robust Point Cloud Registration. | Karim Slimani, Brahim Tamadazte, Catherine Achard |
| 2024 | Every Shot Counts: Using Exemplars for Repetition Counting in Videos. | Saptarshi Sinha, Alexandros Stergiou, Dima Damen |
| 2024 | CCNDF: Curvature Constrained Neural Distance Fields from 3D LiDAR Sequences. | Akshit Singh, Karan Bhakuni, Rajendra Nagar |
| 2024 | RayEmb: Arbitrary Landmark Detection in X-Ray Images Using Ray Embedding Subspace. | Pragyan Shrestha, Chun Xie, Yuichi Yoshii, Itaru Kitahara |
| 2024 | EDAF: Early Detection of Atrial Fibrillation from Post-stroke Brain MRI. | Mohammad Javad Shokri, Nandakishor Desai, Aravinda S. Rao, Angelos Sharobeam, Bernard Yan, Marimuthu Palaniswami |
| 2024 | VIFA: An Efficient Visible and Infrared Image Fusion Architecture for Multi-task Applications via Continual Learning. | Jiaxing Shi, Ao Ren, Wei Zhuang, Yang Hua, ZhiYong Qin, Zhenyu Wang, Yang Song, Yujuan Tan, Duo Liu |
| 2024 | Learning Dual Hierarchical Representation for 3D Surface Reconstruction. | Jiyoon Shin, Youngwook Kim, Sangwoo Hong, Jungwoo Lee |
| 2024 | Feature Estimation of Global Language Processing in EEG Using Attention Maps. | Dai Shimizu, Ko Watanabe, Andreas Dengel |
| 2024 | Do They Share the Same Tail? Learning Individual Compositional Attribute Prototype for Generalized Zero-Shot Learning. | Yuyan Shi, Chenyi Jiang, Run Shi, Haofeng Zhang |
| 2024 | EmoTalker: Audio Driven Emotion Aware Talking Head Generation. | Xiaoqian Shen, Faizan Farooq Khan, Mohamed Elhoseiny |
| 2024 | Decoupled DETR for Few-Shot Object Detection. | Zeyu Shangguan, Lian Huai, Tong Liu, Yuyu Liu, Xingqun Jiang |
| 2024 | QR-DETR: Query Routing for Detection Transformer. | Tharsan Senthivel, Ngoc-Son Vu |
| 2024 | MGNiceNet: Unified Monocular Geometric Scene Understanding. | Markus Schn, Michael Buchholz, Klaus Dietmayer |
| 2024 | WARMOS: Enhancing Weather-Affected Referred Moving Object Segmentation. | Prafulla Saxena, Dinesh Kumar Tyagi, Santosh Kumar Vipparthi, Subrahmanyam Murala |
| 2024 | Contrastive Learning Using Synthetic Images Generated from Real Images. | Tenta Sasaya, Shintaro Yamamoto, Takashi Ida, Takahiro Takimoto |
| 2024 | Estimating Soil Organic Carbon from Multispectral Images Using Physics-Informed Neural Networks. | James Sargeant, Shyh Wei Teng, M. Manzur Murshed, Manoranjan Paul, David Brennan |
| 2024 | A Multi-phase Multi-graph Approach for Focal Liver Lesion Classification on CT Scans. | Tran Bao Sam, Ta Duc Huy, Cong Tuyen Dao, Thanh Tin Lam, Van Ha Tang, Steven Q. H. Truong |
| 2024 | Redefining Normal: A Novel Object-Level Approach for Multi-object Novelty Detection. | Mohammadreza Salehi, Nikolaos Apostolikas, Efstratios Gavves, Cees G. M. Snoek, Yuki M. Asano |
| 2024 | Spatial Clustering and Machine Learning for Crime Prediction: A Case Study on Women Safety in Bhopal. | Yamini Sahu, Vaibhav Kumar |