Skip to content

Leveraging Gaze and Set-of-Mark in VLLMs for Human-Object Interaction Anticipation from Egocentric Videos.

Daniele Materia, Francesco Ragusa, Giovanni Maria Farinella

VenueBICPR
Year2026
ProceedingsICPR (8)

Browse the full ICPR paper archive.