Skip to content

Learning Video Context as Interleaved Multimodal Sequences.

Kevin Qinghong Lin, Pengchuan Zhang, Difei Gao, Xide Xia, Joya Chen, Ziteng Gao, Jinheng Xie, Xuhong Xiao, Mike Zheng Shou

VenueA*ECCV
Year2024
ProceedingsECCV (49)

Browse the full ECCV paper archive.