| 2020 | CVPR | Large Scale Video Representation Learning via Relational Graph Clustering. | Hyodong Lee, Joonseok Lee, Joe Yue-Hei Ng, Paul Natsev |
| 2019 | WACV | TAN: Temporal Aggregation Network for Dense Multi-Label Action Recognition. | Xiyang Dai, Bharat Singh, Joe Yue-Hei Ng, Larry S. Davis |
| 2018 | WACV | ActionFlowNet: Learning Motion Representation for Action Recognition. | Joe Yue-Hei Ng, Jonghyun Choi, Jan Neumann, Larry S. Davis |
| 2018 | WACV | Temporal Difference Networks for Video Action Recognition. | Joe Yue-Hei Ng, Larry S. Davis |
| 2017 | CVPR | FASON: First and Second Order Information Fusion Network for Texture Recognition. | Xiyang Dai, Joe Yue-Hei Ng, Larry S. Davis |
| 2017 | CVPR | Generating Holistic 3D Scene Abstractions for Text-Based Image Retrieval. | Ang Li, Jin Sun, Joe Yue-Hei Ng, Ruichi Yu, Vlad I. Morariu, Larry S. Davis |
| 2015 | CVPR | Beyond short snippets: Deep networks for video classification. | Joe Yue-Hei Ng, Matthew J. Hausknecht, Sudheendra Vijayanarasimhan, Oriol Vinyals, Rajat Monga, George Toderici |
| 2015 | CVPR | Exploiting local features from deep networks for image retrieval. | Joe Yue-Hei Ng, Fan Yang, Larry S. Davis |