| 2026 | ACIIDS | FOCUS +: Enhancing Sustained Attention Through Comparative Evaluation of Digital and Physical Interventions. | Trang Nguyen, Mckenna Kirk, Giang Ho, Thuy Duong Ta, Du Tran, Chi Thanh Vi |
| 2025 | CVPR | SEAL: Semantic Attention Learning for Long Video Representation. | Lan Wang, Yujia Chen, Du Tran, Vishnu Naresh Boddeti, Wen-Sheng Chu |
| 2024 | CVPR | Open-world Instance Segmentation: Top-down Learning with Bottom-up Supervision. | Tarun Kalluri, Weiyao Wang, Heng Wang, Manmohan Chandraker, Lorenzo Torresani, Du Tran |
| 2023 | CVPR | Relational Space-Time Query in Long-Form Videos. | Xitong Yang, Fu-Jen Chu, Matt Feiszli, Raghav Goyal, Lorenzo Torresani, Du Tran |
| 2023 | WACV | FLAVR: Flow-Agnostic Video Representations for Fast Frame Interpolation. | Tarun Kalluri, Deepak Pathak, Manmohan Chandraker, Du Tran |
| 2022 | CVPR | Open-World Instance Segmentation: Exploiting Pseudo Ground Truth From Learned Pairwise Affinity. | Weiyao Wang, Matt Feiszli, Heng Wang, Jitendra Malik, Du Tran |
| 2022 | CVPR | Long-Short Temporal Contrastive Learning of Video Transformers. | Jue Wang, Gedas Bertasius, Du Tran, Lorenzo Torresani |
| 2021 | ICCV | Unidentified Video Objects: A Benchmark for Dense, Open-World Segmentation. | Weiyao Wang, Matt Feiszli, Heng Wang, Du Tran |
| 2020 | AAAI | FASTER Recurrent Networks for Efficient Video Classification. | Linchao Zhu, Du Tran, Laura Sevilla-Lara, Yi Yang, Matt Feiszli, Heng Wang |
| 2020 | CVPR | What Makes Training Multi-Modal Classification Networks Hard? | Weiyao Wang, Du Tran, Matt Feiszli |
| 2020 | CVPR | Video Modeling With Correlation Networks. | Heng Wang, Du Tran, Lorenzo Torresani, Matt Feiszli |
| 2019 | CVPR | Large-Scale Weakly-Supervised Pre-Training for Video Action Recognition. | Deepti Ghadiyaram, Du Tran, Dhruv Mahajan |
| 2019 | CVPR | Leveraging the Present to Anticipate the Future in Videos. | Antoine Miech, Ivan Laptev, Josef Sivic, Heng Wang, Lorenzo Torresani, Du Tran |
| 2019 | ICCV | DistInit: Learning Video Representations Without a Single Labeled Video. | Rohit Girdhar, Du Tran, Lorenzo Torresani, Deva Ramanan |
| 2019 | ICCV | SCSampler: Sampling Salient Clips From Video for Efficient Action Recognition. | Bruno Korbar, Du Tran, Lorenzo Torresani |
| 2019 | ICCV | Video Classification With Channel-Separated Convolutional Networks. | Du Tran, Heng Wang, Matt Feiszli, Lorenzo Torresani |
| 2018 | CVPR | Detect-and-Track: Efficient Pose Estimation in Videos. | Rohit Girdhar, Georgia Gkioxari, Lorenzo Torresani, Manohar Paluri, Du Tran |
| 2018 | CVPR | A Closer Look at Spatiotemporal Convolutions for Action Recognition. | Du Tran, Heng Wang, Lorenzo Torresani, Jamie Ray, Yann LeCun, Manohar Paluri |
| 2018 | ECCV | Scenes-Objects-Actions: A Multi-task, Multi-label Video Dataset. | Jamie Ray, Heng Wang, Du Tran, Yufei Wang, Matt Feiszli, Lorenzo Torresani, Manohar Paluri |
| 2016 | CVPR | Deep End2End Voxel2Voxel Prediction. | Du Tran, Lubomir D. Bourdev, Rob Fergus, Lorenzo Torresani, Manohar Paluri |
| 2015 | ICCV | Learning Spatiotemporal Features with 3D Convolutional Networks. | Du Tran, Lubomir D. Bourdev, Rob Fergus, Lorenzo Torresani, Manohar Paluri |
| 2011 | CVPR | Optimal spatio-temporal path discovery for video event detection. | Du Tran, Junsong Yuan |
| 2008 | ECCV | Human Activity Recognition with Metric Learning. | Du Tran, Alexander Sorokin |