DocKylin: A Large Multimodal Model for Visual Document Understanding with Efficient Visual Slimming.
Jiaxin Zhang, Wentao Yang, Songxuan Lai, Zecheng Xie, Lianwen Jin
Browse the full AAAI paper archive.
Jiaxin Zhang, Wentao Yang, Songxuan Lai, Zecheng Xie, Lianwen Jin
Browse the full AAAI paper archive.