DocThinker: Explainable Multimodal Large Language Models with Rule-Based Reinforcement Learning for Document Understanding.
Wenwen Yu, Zhibo Yang, Yuliang Liu, Xiang Bai
Browse the full ICCV paper archive.
Wenwen Yu, Zhibo Yang, Yuliang Liu, Xiang Bai
Browse the full ICCV paper archive.