IEEE Conference on Computer Vision and Pattern Recognition
CVPR
A*
CORE rank
CORE rank (raw)
A*
Acceptance rate
23.6% (2024)
Fields of research
Computer Vision and Multimedia Computation
Papers indexed
31,646
1988–2025
Papers per year
19883,543 peak2025
Most published authors
CVPR papers
31,646 records sourced from DBLP. Search titles, filter by year, sort by recency.
| Year | Title | Authors |
|---|---|---|
| 2025 | The Devil is in Temporal Token: High Quality Video Reasoning Segmentation. | Sitong Gong, Yunzhi Zhuge, Lu Zhang, Zongxin Yang, Pingping Zhang, Huchuan Lu |
| 2025 | Rectification-specific Supervision and Constrained Estimator for Online Stereo Rectification. | Rui Gong, Kim-Hui Yap, Weide Liu, Xulei Yang, Jun Cheng |
| 2025 | Monocular and Generalizable Gaussian Talking Head Animation. | Shengjie Gong, Haojie Li, Jiapeng Tang, Dongming Hu, Shuangping Huang, Hao Chen, Tianshui Chen, Zhuoman Liu |
| 2025 | The Art of Deception: Color Visual Illusions and Diffusion Models. | Alexandra Gomez-Villa, Kai Wang, C. Alejandro Prraga, Bartlomiej Twardowski, Jesus Malo, Javier Vazquez-Corral, Joost van de Weijer |
| 2025 | From Broadcast to Minimap: Achieving State-of-the-Art SoccerNet Game State Reconstruction. | Vladimir Golovkin, Nikolay Nemtsev, Vasyl Shandyba, Oleg Udin, Nikita Kasatkin, Pavel Kononov, Anton Afanasiev, Sergey Ulasen, Andrei Boiarov |
| 2025 | Learning to Drive from a World Model. | Mitchell Goff, Greg Hogan, George Hotz, Armand du Parc Locmaria, Kacper Raczy, Harald Schfer, Adeeb Shihadeh, Weixing Zhang, Yassine Yousfi |
| 2025 | Seeing More with Less: Human-like Representations in Vision Models. | Andrey Gizdov, Shimon Ullman, Daniel Harari |
| 2025 | Probabilistic Online Event Downsampling. | Andreu Girbau-Xalabarder, Jun Nagata, Shinichi Sumiyoshi |
| 2025 | Vit4V: a Video Classification Method for the Detection of Varroa Destructor from Honeybees. | Luca Giovannesi, Paolo Russo, Roberto Beraldi |
| 2025 | CYFLOD: Cyclic Filtering and Loss Damping for Alleviating Noisy Labels in Fine-grained Visual Classification. | Nauman Ullah Gilal, Khaled A. Al-Thelaya, Fahad Majeed, Zhihe Lu, Sabri Boughorbel, Jens Schneider, Marco Agus |
| 2025 | End-to-End Implicit Neural Representations for Classification. | Alexander Gielisse, Jan van Gemert |
| 2025 | VRAG: Retrieval-Augmented Video Question Answering for Long-Form Videos. | Bao Tran Gia, Khiem Le, Tien Do, Tien-Dung Mai, Thanh Duc Ngo, Duy-Dinh Le, Shin'ichi Satoh |
| 2025 | Towards Faster and More Compact Foundation Models for Molecular Property Prediction. | Yasir Ghunaim, Andrs Villa, Gergo Ignacz, Gyorgy Szekely, Motasem Alfarra, Bernard Ghanem |
| 2025 | Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment. | Soumya Suvra Ghosal, Souradip Chakraborty, Vaibhav Singh, Tianrui Guan, Mengdi Wang, Ahmad Beirami, Furong Huang, Alvaro Velasquez, Dinesh Manocha, Amrit Singh Bedi |
| 2025 | CASP: Compression of Large Multimodal Models Based on Attention Sparsity. | Mohsen Gholami, Mohammad Akbari, Kevin Cannons, Yong Zhang |
| 2025 | CompGS: Unleashing 2D Compositionality for Compositional Text-to-3D via Dynamically Optimizing 3D Gaussians. | Chongjian Ge, Chenfeng Xu, Yuanfeng Ji, Chensheng Peng, Masayoshi Tomizuka, Ping Luo, Mingyu Ding, Varun Jampani, Wei Zhan |
| 2025 | AutoPresent: Designing Structured Visuals from Scratch. | Jiaxin Ge, Zora Zhiruo Wang, Xuhui Zhou, Yi-Hao Peng, Sanjay Subramanian, Qinyue Tan, Maarten Sap, Alane Suhr, Daniel Fried, Graham Neubig, Trevor Darrell |
| 2025 | Arc2Avatar: Generating Expressive 3D Avatars from a Single Image via ID Guidance. | Dimitrios Gerogiannis, Foivos Paraperas Papantoniou, Rolandos Alexandros Potamias, Alexandros Lattas, Stefanos Zafeiriou |
| 2025 | FSboard: Over 3 Million Characters of ASL Fingerspelling Collected via Smartphones. | Manfred Georg, Garrett Tanzer, Esha Uboweja, Saad Hassan, Maximus Shengelia, Sam S. Sepah, Sean Forbes, Thad Starner |
| 2025 | The Illusion of Unlearning: The Unstable Nature of Machine Unlearning in Text-to-Image Diffusion Models. | Naveen George, Karthik Nandan Dasaraju, Rutheesh Reddy Chittepu, Konda Reddy Mopuri |
| 2025 | LongVALE: Vision-Audio-Language-Event Benchmark Towards Time-Aware Omni-Modal Perception of Long Videos. | Tiantian Geng, Jinrui Zhang, Qingni Wang, Teng Wang, Jinming Duan, Feng Zheng |
| 2025 | HORP: Human-Object Relation Priors Guided HOI Detection. | Pei Geng, Jian Yang, Shanshan Zhang |
| 2025 | Motion Prompting: Controlling Video Generation with Motion Trajectories. | Daniel Geng, Charles Herrmann, Junhwa Hur, Forrester Cole, Serena Zhang, Tobias Pfaff, Tatiana Lopez-Guevara, Yusuf Aytar, Michael Rubinstein, Chen Sun, Oliver Wang, Andrew Owens, Deqing Sun |
| 2025 | Divot: Diffusion Powers Video Tokenizer for Comprehension and Generation. | Yuying Ge, Yizhuo Li, Yixiao Ge, Ying Shan |
| 2025 | Pureformer: Transformer-Based Image Denoising. | Arnim Gautam, Aditi Pawar, Aishwarya Joshi, Satya Narayan Tazi, Sachin Chaudhary, Praful Hambarde, Akshay Dudhane, Santosh Kumar Vipparthi, Subrahmanyam Murala |
2,601–2,625 of 31,646← PreviousNext →
Comparable venues
Other A*/A conferences filed under the same field of research.