OmAgent: A Multi-modal Agent Framework for Complex Video Understanding with Task Divide-and-Conquer.
Lu Zhang, Tiancheng Zhao, Heting Ying, Yibo Ma, Kyusong Lee
Browse the full EMNLP paper archive.
Lu Zhang, Tiancheng Zhao, Heting Ying, Yibo Ma, Kyusong Lee
Browse the full EMNLP paper archive.