Leveraging Gaze and Set-of-Mark in VLLMs for Human-Object Interaction Anticipation from Egocentric Videos.
Daniele Materia, Francesco Ragusa, Giovanni Maria Farinella
Browse the full ICPR paper archive.
Daniele Materia, Francesco Ragusa, Giovanni Maria Farinella
Browse the full ICPR paper archive.