What to look at and where: Semantic and Spatial Refined Transformer for detecting human-object interactions.
A. S. M. Iftekhar, Hao Chen, Kaustav Kundu, Xinyu Li, Joseph Tighe, Davide Modolo
Browse the full CVPR paper archive.
A. S. M. Iftekhar, Hao Chen, Kaustav Kundu, Xinyu Li, Joseph Tighe, Davide Modolo
Browse the full CVPR paper archive.