Keypoint-Integrated Instruction-Following Data Generation for Enhanced Human Pose and Action Understanding in Multimodal Models.
Dewen Zhang, Wangpeng An, Hayaru Shouno
Browse the full ACIVS paper archive.
Dewen Zhang, Wangpeng An, Hayaru Shouno
Browse the full ACIVS paper archive.