| 2025 | MATATA: Weakly Supervised End-to-End MAthematical Tool-Augmented Reasoning for Tabular Applications. | Vishnou Vinayagame, Gregory Senay, Luis Mart |
| 2025 | Archival Faces: Detection of Faces in Digitized Historical Documents. | Marek Vasko, Adam Herout, Michal Hradis |
| 2025 | Unchecked and Overlooked: Addressing the Checkbox Blind Spot in Large Language Models with CheckboxQA. | Michal Turski, Mateusz Chilinski, Lukasz Borchmann |
| 2025 | PALM-LAY: A Multi-script Cross-Regional Dataset for Layout Analysis of Palm Leaf Manuscripts. | Nimol Thuon, Jun Du, Panhapin Theang, Ratana Thuon |
| 2025 | QUEST: Quality-Aware Semi-supervised Table Extraction for Business Documents. | Eliott Thomas, Mickal Coustaty, Aurlie Joseph, Gaspar Deloin, Elodie Carel, Vincent Poulain D'Andecy, Jean-Marc Ogier |
| 2025 | SFDLA: Source-Free Document Layout Analysis. | Sebastian Tewes, Yufan Chen, Omar Moured, Jiaming Zhang, Rainer Stiefelhagen |
| 2025 | Dual Downsample Vision Transformer for Handwritten Text Recognition. | Yew Lee Tan, Ernest Yu Kai Chew, Jung-Jae Kim, Adams Wai-Kin Kong |
| 2025 | Ar-Q-Former: Historical Newspaper Article Separation Based on Multimodal Transformer Structure. | Wenjun Sun, Nancy Girdhar, Tran Thi Hong Hanh, Carlos-Emiliano Gonzlez-Gallardo, Mickal Coustaty, Antoine Doucet |
| 2025 | Few-Shot Segmentation of Historical Maps via Linear Probing of Vision Foundation Models. | Rafael Sterzinger, Marco Peer, Robert Sablatnig |
| 2025 | UniLayDet: Simple Multi-dataset Document Layout Analysis. | Prasidh Srikumar, Ajoy Mondal, C. V. Jawahar |
| 2025 | AttentionLeak: What Does Human Attention Reveal About Information Visualisation? | Malte Snnichsen, Mayar Elfares, Yao Wang, Ralf Ksters, Alina Roitberg, Andreas Bulling |
| 2025 | Optimizing Thai-English Spoken Question Answering Interaction for Open Environments with Limited Resources. | Sattaya Singkul, Atthakorn Petchsod, Panya Sunantasaengtong, Theerat Sakdejayont, Tawunrat Chalothorn |
| 2025 | From Conversations to Insights: A Multimodal Approach to Discussion Summarization. | Punit Kumar Singh, Nishant Kumar, Hrushik Mehta, Sriparna Saha |
| 2025 | Classifying the Unknown: In-Context Learning for Open-Vocabulary Text and Symbol Recognition. | Tom Simon, William Mocar, Pierrick Tranouez, Clment Chatelain, Thierry Paquet |
| 2025 | DevInSight: Weaving Path Development Into Online Signature Verification. | Yilin Shi, Lei Jiang, Hao Ni, Lianwen Jin |
| 2025 | SemiTabDETR: End-to-End Semi-supervised Table Detection with Transformer-Based Enhanced Query Approach. | Tahira Shehzadi, Didier Stricker, Muhammad Zeshan Afzal |
| 2025 | Evaluating Popular Scene Text Detection and Recognition Methods on Tombstones. | Mathias Seuret, Oliver Traub, Ning Guo, Florian Kordon, Thomas Gorges, Vincent Christlein |
| 2025 | The Return of Structural Handwritten Mathematical Expression Recognition. | Jakob Seitz, Tobias Lengfeld, Radu Timofte |
| 2025 | A New Multimodal Cross-Domain Network for Classification of Challenging Scene Images. | Shashwat Sarkar, Kunal Purkayastha, Shivakumara Palaiahnakote, Umapada Pal, Muhammad Hammad Saleem, Palash Ghosal |
| 2025 | DP-DocLDM: Differentially Private Document Image Generation Using Latent Diffusion Models. | Saifullah Saifullah, Stefan Agne, Andreas Dengel, Sheraz Ahmed |
| 2025 | Evaluating Compliance with Visualization Guidelines in Diagrams for Scientific Publications Using Large Vision Language Models. | Johannes Rckert, Louise Bloch, Christoph M. Friedrich |
| 2025 | DocForgeNet: Dual Cross-Stream Fusion Network for Robust Forgery Detection in Scanned Documents. | Nauman Riaz, Stefan Agne, Andreas Dengel, Sheraz Ahmed |
| 2025 | KuiSCIMA V2.0: Improved Baselines, Calibration, and Cross-Notation Generalization for Historical Chinese Music Notations in Jiang Kui's Baishidaoren Gequ. | Tristan Repolusk, Eduardo E. Veas |
| 2025 | DocAnnot - Accelerating the Creation of Key Information Extraction Datasets with GenAI-Powered Auto-annotation. | Siddartha Reddy, Harikrishnan P. M., Goutham Vignesh, Varun V, Vishal Vaddina |
| 2025 | Interpretable Writer Recognition via Vectors of Locally Aggregated Characters. | Tim Raven, Vincent Christlein, Gernot A. Fink |