SpatialVLM: Endowing Vision-Language Models with Spatial Reasoning Capabilities.
Boyuan Chen, Zhuo Xu, Sean Kirmani, Brian Ichter, Dorsa Sadigh, Leonidas J. Guibas, Fei Xia
Browse the full CVPR paper archive.
Boyuan Chen, Zhuo Xu, Sean Kirmani, Brian Ichter, Dorsa Sadigh, Leonidas J. Guibas, Fei Xia
Browse the full CVPR paper archive.