Skip to content

Shoubin Yu

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

12

Venues

6

Active years

2024–2026

Best venue rank

A*

Where they publish

Papers

12 indexed papers, newest first.

YearVenueTitleAuthors
2026LAKA Novel Approach to Evaluating the Effectiveness of Large Language Models for Multimodal Analysis of Embodied Learning in Classrooms.Joyce Horn Fonteles, Nithin Sivakumaran, Clayton Cohn, Austin Coursey, Shoubin Yu, Elias Stengel-Eskin, Ashwin T. S., Mohit Bansal, Gautam Biswas
2025CVPRMotion-Grounded Video Reasoning: Understanding and Perceiving Motion at Pixel Level.Andong Deng, Tongjia Chen, Shoubin Yu, Taojiannan Yang, Lincoln Spencer, Yapeng Tian, Ajmal Saeed Mian, Mohit Bansal, Chen Chen
2025CVPRVideoTree: Adaptive Tree-based Video Representation for LLM Reasoning on Long Videos.Ziyang Wang, Shoubin Yu, Elias Stengel-Eskin, Jaehong Yoon, Feng Cheng, Gedas Bertasius, Mohit Bansal
2025EMNLPVideo-RTS: Rethinking Reinforcement Learning and Test-Time Scaling for Efficient and Enhanced Video Reasoning.Ziyang Wang, Jaehong Yoon, Shoubin Yu, Md Mohaiminul Islam, Gedas Bertasius, Mohit Bansal
2025EMNLPRACCooN: Versatile Instructional Video Editing with Auto-Generated Narratives.Jaehong Yoon, Shoubin Yu, Mohit Bansal
2025EMNLPMEXA: Towards General Multimodal Reasoning with Dynamic Multi-Expert Aggregation.Shoubin Yu, Yue Zhang, Ziyang Wang, Jaehong Yoon, Mohit Bansal
2025ICCVVEGGIE: Instructional Editing and Reasoning Video Concepts with Grounded Generation.Shoubin Yu, Difan Liu, Ziqiao Ma, Yicong Hong, Yang Zhou, Hao Tan, Joyce Chai, Mohit Bansal
2025ICLRBootstrapping Language-Guided Navigation Learning with Self-Refining Data Flywheel.Zun Wang, Jialu Li, Yicong Hong, Songze Li, Kunchang Li, Shoubin Yu, Yi Wang, Yu Qiao, Yali Wang, Mohit Bansal, Limin Wang
2025ICLRSAFREE: Training-Free and Adaptive Guard for Safe Text-to-Image And Video Generation.Jaehong Yoon, Shoubin Yu, Vaidehi Patil, Huaxiu Yao, Mohit Bansal
2025ICLRCREMA: Generalizable and Efficient Video-Language Reasoning via Multimodal Modular Fusion.Shoubin Yu, Jaehong Yoon, Mohit Bansal
2025ICMIA Multimodal Classroom Video Question-Answering Framework for Automated Understanding of Collaborative Learning.Nithin Sivakumaran, Chia-Yu Yang, Abhay Zala, Shoubin Yu, Daeun Hong, Xiaotian Zou, Elias Stengel-Eskin, Dan Carpenter, Wookhee Min, Cindy E. Hmelo-Silver, Jonathan P. Rowe, James C. Lester, Mohit Bansal
2024EMNLPA Simple LLM Framework for Long-Range Video Question-Answering.Ce Zhang, Taixi Lu, Md Mohaiminul Islam, Ziyang Wang, Shoubin Yu, Mohit Bansal, Gedas Bertasius