Skip to content

Rethinking Agentic and End-to-End Large Multimodal Models for Vision Tasks.

Yixin Wang, Xinyu Wang

Year2025
ProceedingsDICTA

Browse the full DICTA paper archive.