Attention-Driven GUI Grounding: Leveraging Pretrained Multimodal Large Language Models Without Fine-Tuning.
Hai-Ming Xu, Qi Chen, Lei Wang, Lingqiao Liu
Browse the full AAAI paper archive.
Hai-Ming Xu, Qi Chen, Lei Wang, Lingqiao Liu
Browse the full AAAI paper archive.