| 2025 | CVPR | MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis. | Ho Kei Cheng, Masato Ishii, Akio Hayakawa, Takashi Shibuya, Alexander G. Schwing, Yuki Mitsufuji |
| 2025 | ICCV | The Curse of Conditions: Analyzing and Improving Optimal Transport for Conditional Flow-Based Generation. | Ho Kei Cheng, Alexander Gerhard Schwing |
| 2024 | CVPR | Putting the Object Back into Video Object Segmentation. | Ho Kei Cheng, Seoung Wug Oh, Brian L. Price, Joon-Young Lee, Alexander G. Schwing |
| 2023 | ICCV | Tracking Anything with Decoupled Video Segmentation. | Ho Kei Cheng, Seoung Wug Oh, Brian L. Price, Alexander G. Schwing, Joon-Young Lee |
| 2022 | ECCV | XMem: Long-Term Video Object Segmentation with an Atkinson-Shiffrin Memory Model. | Ho Kei Cheng, Alexander G. Schwing |
| 2021 | CVPR | Modular Interactive Video Object Segmentation: Interaction-to-Mask, Propagation and Difference-Aware Fusion. | Ho Kei Cheng, Yu-Wing Tai, Chi-Keung Tang |
| 2020 | CVPR | CascadePSP: Toward Class-Agnostic and Very High-Resolution Segmentation via Global and Local Refinement. | Ho Kei Cheng, Jihoon Chung, Yu-Wing Tai, Chi-Keung Tang |