| 2026 | ACL | MoCa: Modality-aware Continual Pre-training Makes Better Bidirectional Multimodal Embeddings. | Haonan Chen, Hong Liu, Yuping Luo, Liang Wang, Nan Yang, Furu Wei, Zhicheng Dou |
| 2022 | IROS | Towards Learning to Play Piano with Dexterous Hands and Touch. | Huazhe Xu, Yuping Luo, Shaoxiong Wang, Trevor Darrell, Roberto Calandra |
| 2021 | ICLR | Towards Resolving the Implicit Bias of Gradient Descent for Matrix Factorization: Greedy Low-Rank Learning. | Zhiyuan Li, Yuping Luo, Kaifeng Lyu |
| 2020 | ICLR | Learning Self-Correctable Policies and Value Functions from Demonstrations with Negative Sampling. | Yuping Luo, Huazhe Xu, Tengyu Ma |
| 2020 | ICML | Provable Representation Learning for Imitation Learning via Bi-level Optimization. | Sanjeev Arora, Simon S. Du, Sham M. Kakade, Yuping Luo, Nikunj Saunshi |
| 2020 | ICML | On the Expressivity of Neural Networks for Deep Reinforcement Learning. | Kefan Dong, Yuping Luo, Tianhe Yu, Chelsea Finn, Tengyu Ma |
| 2019 | ICLR | Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees. | Yuping Luo, Huazhe Xu, Yuanzhi Li, Yuandong Tian, Trevor Darrell, Tengyu Ma |
| 2017 | ICASSP | Learning online alignments with continuous rewards policy gradient. | Yuping Luo, Chung-Cheng Chiu, Navdeep Jaitly, Ilya Sutskever |