| 2024 | DocTabQA: Answering Questions from Long Documents Using Tables. | Haochen Wang, Kai Hu, Haoyu Dong, Liangcai Gao |
| 2024 | One-Shot Transformer-Based Framework for Visually-Rich Document Understanding. | Huynh The Vu, Van Pham Hoai, Jeff Yang |
| 2024 | ViT-ED: Transformer Network for Image Similarity Measurement. | Manh-Tu Vu, Marie Beurton-Aimar |
| 2024 | Clustering Running Titles to Understand the Printing of Early Modern Books. | Nikolai Vogler, Kartik Goyal, Samuel V. Lemley, D. J. Schuldt, Christopher N. Warren, Max G'Sell, Taylor Berg-Kirkpatrick |
| 2024 | An Interpretable Deep Learning Approach for Morphological Script Type Analysis. | Malamatenia Vlachou-Efstathiou, Ioannis Siglidis, Dominique Stutzmann, Mathieu Aubry |
| 2024 | Comics Datasets Framework: Mix of Comics Datasets for Detection Benchmarking. | Emanuele Vivoli, Irene Campaioli, Mariateresa Nardoni, Niccol Biondi, Marco Bertini, Dimosthenis Karatzas |
| 2024 | Multimodal Transformer for Comics Text-Cloze. | Emanuele Vivoli, Joan Lafuente Baeza, Ernest Valveny Llobet, Dimosthenis Karatzas |
| 2024 | Enhancing Recognition of Historical Musical Pieces with Synthetic and Composed Images. | Manuel Villarreal, Joan-Andreu Snchez |
| 2024 | Reading Order Independent Metrics for Information Extraction in Handwritten Documents. | David Villanova-Aparisi, Solne Tarride, Carlos D. Martnez-Hinarejos, Vernica Romero, Christopher Kermorvant, Moiss Pastor-i-Gadea |
| 2024 | Zipf Curves and Basic Text Analytics from Untranscribed Manuscript Images. | Enrique Vidal, Alejandro H. Toselli |
| 2024 | Detecting and Deciphering Damaged Medieval Armenian Inscriptions Using YOLO and Vision Transformers. | Chahan Vidal-Gorne, Alinor Decours-Perez |
| 2024 | Image-to-Image Translation Approach for Page Layout Analysis and Artificial Generation of Historical Manuscripts. | Chahan Vidal-Gorne, Jean-Baptiste Camps |
| 2024 | Robust Handwritten Signature Representation with Continual Learning of Synthetic Data over Predefined Real Feature Space. | Talles Brito Viana, Victor L. F. Souza, Adriano L. I. Oliveira, Rafael M. O. Cruz, Robert Sabourin |
| 2024 | Text Line Segmentation on Ancient Egyptian Papyri: Layout Analysis with Object Detection Networks and Connected Components. | Stephan M. Unter |
| 2024 | Content-Based Similarity for Automatic Scoring of Handwritten Descriptive Answers. | Nghia Thanh Truong, Hung Tuan Nguyen, Nam Tuan Ly, Toshihiko Horie, Masaki Nakagawa |
| 2024 | GDP: Generic Document Pretraining to Improve Document Understanding. | Akkshita Trivedi, Akarsh Upadhyay, Rudrabha Mukhopadhyay, Santanu Chaudhury |
| 2024 | SketchGPT: Autoregressive Modeling for Sketch Generation and Recognition. | Adarsh Tiwari, Sanket Biswas, Josep Llads |
| 2024 | Privacy-Aware Document Visual Question Answering. | Rubn Tito, Khanh Nguyen, Marlon Tobaben, Raouf Kerkouche, Mohamed Ali Souibgui, Kangsoo Jung, Joonas Jlk, Vincent Poulain D'Andecy, Aurlie Joseph, Lei Kang, Ernest Valveny, Antti Honkela, Mario Fritz, Dimosthenis Karatzas |
| 2024 | Ablation Study of a Multimodal Gat Network on Perfect Synthetic and Real-world Data to Investigate the Influence of Language Models in Invoice Recognition. | Lukas Walter Thie |
| 2024 | Improving Automatic Text Recognition with Language Models in the PyLaia Open-Source Library. | Solne Tarride, Yoann Schneider, Marie Generali-Lince, Mlodie Boillet, Bastien Abadie, Christopher Kermorvant |
| 2024 | Revisiting N-Gram Models: Their Impact in Modern Neural Networks for Handwritten Text Recognition. | Solne Tarride, Christopher Kermorvant |
| 2024 | Recurrent Few-Shot Model for Document Verification. | Maxime Talarmain, Carlos Boned Riera, Sanket Biswas, Oriol Ramos Terrades |
| 2024 | Global-SEG: Text Semantic Segmentation Based on Global Semantic Pair Relations. | Wenjun Sun, Tran Thi Hong Hanh, Carlos-Emiliano Gonzlez-Gallardo, Mickal Coustaty, Antoine Doucet |
| 2024 | Drawing the Line: Deep Segmentation for Extracting Art from Ancient Etruscan Mirrors. | Rafael Sterzinger, Simon Brenner, Robert Sablatnig |
| 2024 | ComicBERT: A Transformer Model and Pre-training Strategy for Contextual Understanding in Comics. | Grkan Soykan, Deniz Yuret, Tevfik Metin Sezgin |