UniAVLM: Unified Large Audio-Visual Language Models for Comprehensive Video Understanding.
Lecheng Yan, Chenyang Lyu, Wenxi Li, Younes Samih, Shaochen Jiang
Browse the full PRICAI paper archive.
Lecheng Yan, Chenyang Lyu, Wenxi Li, Younes Samih, Shaochen Jiang
Browse the full PRICAI paper archive.