Skip to content

DIVE-Doc: Downscaling Foundational Image Visual Encoder into Hierarchical Architecture for DocVQA.

Rayane Bencharef, Abderrahmane Rahiche, Mohamed Cheriet

VenueA*ICCV
Year2025
ProceedingsICCVW

Browse the full ICCV paper archive.