| 2024 | Improving Depth Gradient Continuity in Transformers: A Comparative Study on Monocular Depth Estimation with CNN. | Jiawei Yao, Tong Wu, Xiaofeng Zhang |
| 2024 | Balancing Calibration and Performance: Stochastic Depth in Segmentation BNNs. | Linghong Yao, Denis Hadjivelichkov, Andromachi Maria Delfaki, Yuanchang Liu, Brooks Paige, Dimitrios Kanoulas |
| 2024 | APTPose: Anatomy-aware Pre-Training for 3D Human Pose Estimation. | Qing-Wen Yang, Kai-Wen Duan, Ting-Yi Lu, Kevin Lin, Cheng-Yen Yang, Lijuan Wang, Jenq-Neng Hwang, Shang-Hong Lai |
| 2024 | PlainMamba: Improving Non-Hierarchical Mamba in Visual Recognition. | Chenhongyi Yang, Zehui Chen, Miguel Espinosa, Linus Ericsson, Zhenyu Wang, Jiaming Liu, Elliot J. Crowley |
| 2024 | Impact of Surface Reflections in Maritime Obstacle Detection. | Samed Yalin, Hazim Kemal Ekenel |
| 2024 | ICAF-4: An Integrated Framework of Category-level Articulated Object Perception and Manipulation for Embodied Intelligence. | Wenbo Xu, Li Zhang, Qiankun Li, Qi Wu, Lin Yuanbo Wu, Liu Liu |
| 2024 | Lightweight Human Pose Estimation with Enhanced Knowledge Review. | Hao Xu, Shengye Yan, Wei Zheng |
| 2024 | FFR-UNet: Feature Filter-Refinement UNet for Medical Image Segmentation. | Weixin Xu |
| 2024 | Open-World Semi-Supervised Learning under Compound Distribution Shifts. | Shijia Xu, Lin Zhao, Jialiang Tang, Guangyu Li, Chen Gong |
| 2024 | PT43D: A Probabilistic Transformer for Generating 3D Shapes from Single Highly-Ambiguous RGB Images. | Yiheng Xiong, Angela Dai |
| 2024 | Are Sparse Neural Networks Better Hard Sample Learners? | Qiao Xiao, Boqian Wu, Lu Yin, Christopher Neil Gadzinski, Tianjin Huang, Mykola Pechenizkiy, Decebal Constantin Mocanu |
| 2024 | IRFusionFormer: Enhancing Pavement Crack Segmentation with RGB-T Fusion and Topological-Based Loss. | Ruiqiang Xiao, Xiaohu Chen |
| 2024 | InSpaceType: Dataset and Benchmark for Reconsidering Cross-Space Type Performance in Indoor Monocular Depth. | Cho-Ying Wu, Quankai Gao, Chin-Cheng Hsu, Te-Lin Wu, Jing-Wen Chen, Ulrich Neumann |
| 2024 | Out-Of-Distribution Detection for Audio-visual Generalized Zero-Shot Learning: A General Framework. | Liuyuan Wen |
| 2024 | Interactive Image Segmentation with Temporal Information Augmented. | Qiaoqiao Wei, Hui Zhang, Jun-Hai Yong |
| 2024 | Self-Supervised Real-World Denoising by Jointly Learning Visible and Invisible Noise. | Shaoyu Wang, Changze Zhou, Bolin Song, Yiyang Wang |
| 2024 | Gaussian Splatting in Mirrors: Reflection-aware Rendering via Virtual Camera Optimization. | Zihan Wang, Shuzhe Wang, Matias Turkulainen, Junyuan Fang, Juho Kannala |
| 2024 | Multi-modal Crowd Counting via Modal Emulation. | Chenhao Wang, Xiaopeng Hong, Zhiheng Ma, Yupeng Wei, Yabin Wang, Xiaopeng Fan |
| 2024 | Learning Object Placement via Convolution Scoring Attention. | Yibin Wang, Yuchao Feng, Jianwei Zheng |
| 2024 | Weakly-supervised Localization of Manipulated Image Regions Using Multi-resolution Learned Features. | Ziyong Wang, Charith Abhayaratne |
| 2024 | Syn-to-Real Unsupervised Domain Adaptation for Indoor 3D Object Detection. | Yunsong Wang, Na Zhao, Gim Hee Lee |
| 2024 | Sign Stitching: A Novel Approach to Sign Language Production. | Harry Walsh, Ben Saunders, Richard Bowden |
| 2024 | MixMask: Revisiting Masking Strategy for Siamese ConvNets. | Kirill Vishniakov, Eric P. Xing, Zhiqiang Shen |
| 2024 | MCDS-VSS: Moving Camera Dynamic Scene Video Semantic Segmentation by Filtering with Self-Supervised Geometry and Motion. | Angel Villar-Corrales, Moritz Austermann, Sven Behnke |
| 2024 | DiffusedWrinkles: A Diffusion-Based Model for Data-Driven Garment Animation. | Raquel Vidaurre, Elena Garces, Dan Casas |