Compositional Video Understanding with Spatiotemporal Structure-based Transformers.
Hoyeoung Yun, Jinwoo Ahn, Minseo Kim, Eun-Sol Kim
Browse the full CVPR paper archive.
Hoyeoung Yun, Jinwoo Ahn, Minseo Kim, Eun-Sol Kim
Browse the full CVPR paper archive.