| 2025 | RampNet: A Two-Stage Pipeline for Bootstrapping Curb Ramp Detection in Streetscape Images from Open Government Metadata. | John S. O'Meara, Jared Hwang, Zeyu Wang, Michael Saugstad, Jon E. Froehlich |
| 2025 | Certifiably Optimal Anisotropic Rotation Averaging. | Carl Olsson, Yaroslava Lochman, Johan Malmport, Christopher Zach |
| 2025 | Probing the Representational Power of Sparse Autoencoders in Vision Models. | Matthew L. Olson, Musashi Hinck, Neale Ratzlaff, Changbai Li, Phillip Howard, Vasudev Lal, Shao-Yen Tseng |
| 2025 | WCCA-AK: A Multimodal Dataset of Andr Kim's Fashion Legacy for AI-Driven Cultural Heritage Research. | SeongYeon Oh, Soyoung Lee, Hyeon Seong Jeong, Sangwoo Jo, JinYoung Kim, Yeonseo Choi, Young Joon Yoo, Taehoon Kim |
| 2025 | Generative Modeling of Shape-Dependent Self-Contact Human Poses. | Takehiko Ohkawa, Jihyun Lee, Shunsuke Saito, Jason M. Saragih, Fabian Prada, Yichen Xu, Shoou-I Yu, Ryosuke Furuta, Yoichi Sato, Takaaki Shiratori |
| 2025 | LaViPlan: Language-Guided Visual Path Planning with RLVR. | Hayeon Oh |
| 2025 | Visual Surface Wave Elastography: Revealing Subsurface Physical Properties via Visible Surface Waves. | Alexander C. Ogren, Berthy T. Feng, Jihoon Ahn, Katherine L. Bouman, Chiara Daraio |
| 2025 | Improving U-Net Confidence on TEM Image Data with L2-Regularization, Transfer Learning, and Deep Fine-Tuning. | Aiden Ochoa, Xinyuan Xu, Xing Wang |
| 2025 | Hierarchical 3D Scene Graphs Construction Outdoors. | Jon Nyffeler, Federico Tombari, Daniel Barath |
| 2025 | Breast Cancer Detection with Topological Deep Learning. | Brighton Nuwagira, Adrian Rodriguez, Qiwei Li, Baris Coskunuzer |
| 2025 | Assessing the Quality of Soccer Shots from Single-Camera Video with Vision-Language Models and Motion Features. | Filip Noworolnik, Joanna Jaworek-Korjakowska |
| 2025 | From Binary to Semantic: Utilizing Large-Scale Binary Occupancy Data for 3D Semantic Occupancy Prediction. | Chihiro Noguchi, Takaki Yamamoto |
| 2025 | Text Image Generation for Low-Resource Languages with Dual Translation Learning. | Chihiro Noguchi, Shun Fukuda, Shoichiro Mihara, Masao Yamanaka |
| 2025 | What Makes for Text to 360-Degree Panorama Generation with Stable Diffusion? | Jinhong Ni, Chang-Bin Zhang, Qiang Zhang, Jing Zhang |
| 2025 | WonderTurbo: Generating Interactive 3D World in 0.72 Seconds. | Chaojun Ni, Xiaofeng Wang, Zheng Zhu, Weijie Wang, Haoyun Li, Guosheng Zhao, Jie Li, Wenkang Qin, Guan Huang, Wenjun Mei |
| 2025 | Face Video Steganography for Privacy-Protection Automatic Depression Assessment. | Xinyi Ni, Zijian Wu, Lu Liu, Siyang Song |
| 2025 | Beyond Appearance: Geometric Cues for Robust Video Instance Segmentation. | Quanzhu Niu, Yikang Zhou, Shihao Chen, Tao Zhang, Shunping Ji |
| 2025 | ChatReID: Open-Ended Interactive Person Retrieval via Hierarchical Progressive Tuning for Vision Language Models. | Ke Niu, Haiyang Yu, Mengyang Zhao, Teng Fu, Siyang Yi, Wei Lu, Bin Li, Xuelin Qian, Xiangyang Xue |
| 2025 | Enhancing Adversarial Transferability by Balancing Exploration and Exploitation with Gradient-Guided Sampling. | Zenghao Niu, Weicheng Xie, Siyang Song, Zitong Yu, Feng Liu, Linlin Shen |
| 2025 | The Inter-Intra Modal Measure: A Predictive Lens on Fine-Tuning Outcomes in Vision-Language Models. | Laura Niss, Kevin Vogt-Lowell, Theodoros Tsiligkaridis |
| 2025 | Planar Affine Rectification from Local Change of Scale and Orientation. | Yuval Nissan, Marc Pollefeys, Daniel Barath |
| 2025 | Enhancing Spatial Reasoning in Multimodal Large Language Models Through Reasoning-Based Segmentation. | Zhenhua Ning, Zhuotao Tian, Shaoshuai Shi, Guangming Lu, Daojing He, Wenjie Pei, Li Jiang |
| 2025 | Color Matching Using Hypernetwork-Based Kolmogorov-Arnold Networks. | Artem V. Nikonorov, Georgy Perevozchikov, Andrei Korepanov, Nancy Mehta, Mahmoud Afifi, Egor Ershov, Radu Timofte |
| 2025 | Dual Orthogonal Guidance for Robust Diffusion-Based Handwritten Text Generation. | Konstantina Nikolaidou, George Retsinas, Giorgos Sfikas, Silvia Cascianelli, Rita Cucchiara, Marcus Liwicki |
| 2025 | HIVE: A Hyperbolic Interactive Visualization Explorer for Representation Learning. | Thijmen Nijdam, Derck W. E. Prinzhorn, Jurgen de Heus, Thomas Brouwer |