| 2025 | CVPR | InterMimic: Towards Universal Whole-Body Control for Physics-Based Human-Object Interactions. | Sirui Xu, Hung Yu Ling, Yu-Xiong Wang, Liang-Yan Gui |
| 2025 | CVPR | InterAct: Advancing Large-Scale Versatile 3D Human-Object Interaction Generation. | Sirui Xu, Dongting Li, Yucheng Zhang, Xiyan Xu, Qi Long, Ziyin Wang, Yunzhi Lu, Shuchang Dong, Hezi Jiang, Akshat Gupta, Yu-Xiong Wang, Liang-Yan Gui |
| 2025 | ICLR | DICE: End-to-end Deformation Capture of Hand-Face Interactions from a Single Image. | Qingxuan Wu, Zhiyang Dou, Sirui Xu, Soshi Shimada, Chen Wang, Zhengming Yu, Yuan Liu, Cheng Lin, Zeyu Cao, Taku Komura, Vladislav Golyanik, Christian Theobalt, Wenping Wang, Lingjie Liu |
| 2023 | ICCV | InterDiff: Generating 3D Human-Object Interactions with Physics-Informed Diffusion. | Sirui Xu, Zhengyuan Li, Yu-Xiong Wang, Liang-Yan Gui |
| 2023 | ICLR | Stochastic Multi-Person 3D Motion Forecasting. | Sirui Xu, Yu-Xiong Wang, Liangyan Gui |
| 2022 | ECCV | Diverse Human Motion Prediction Guided by Multi-level Spatial-Temporal Anchors. | Sirui Xu, Yu-Xiong Wang, Liang-Yan Gui |
| 2019 | ICASSP | Spatial and Channel Attention Based Convolutional Neural Networks for Modeling Noisy Speech. | Sirui Xu, Eric Fosler-Lussier |
| 2018 | ICASSP | Application of Progressive Neural Networks for Multi-Stream Wfst Combination in One-Pass Decoding. | Sirui Xu, Eric Fosler-Lussier |
| 2016 | Interspeech | A WFST Framework for Single-Pass Multi-Stream Decoding. | Sirui Xu, Eric Fosler-Lussier |