Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models.
Ankit Yadav, Lingqiao Liu, Yuankai Qi
Browse the full DICTA paper archive.
Ankit Yadav, Lingqiao Liu, Yuankai Qi
Browse the full DICTA paper archive.