Skip to content

Compositional Video Understanding with Spatiotemporal Structure-based Transformers.

Hoyeoung Yun, Jinwoo Ahn, Minseo Kim, Eun-Sol Kim

VenueA*CVPR
Year2024
ProceedingsCVPR

Browse the full CVPR paper archive.