| 2026 | ACL | Efficiently Learning To Reason or Not to Reason: Root-token Policy Optimization for Adaptive Thinking. | Taehyeon Kim, Hyunsoo Lee, Youngsoo Jang, Moontae Lee |
| 2026 | EACL | IRPO: Implicit Policy Regularized Preference Optimization. | Youngsoo Jang, Yu Jin Kim, Geon-Hyeong Kim, Honglak Lee, Moontae Lee |
| 2025 | ICML | Online Pre-Training for Offline-to-Online Reinforcement Learning. | Yongjae Shin, Jeonghye Kim, Whiyoung Jung, Sunghoon Hong, Deunsol Yoon, Youngsoo Jang, Geon-Hyeong Kim, Jongseong Chae, Youngchul Sung, Kanghoon Lee, Woohyung Lim |
| 2024 | ACL | Semantic Skill Grounding for Embodied Instruction-Following in Cross-Domain Environments. | Sangwoo Shin, Seunghyun Kim, Youngsoo Jang, Moontae Lee, Honguk Woo |
| 2024 | CVPR | Show, Think, and Tell: Thought-Augmented Fine-Tuning of Large Language Models for Video Captioning. | Byoungjip Kim, Dasol Hwang, Sungjun Cho, Youngsoo Jang, Honglak Lee, Moontae Lee |
| 2024 | EMNLP | Prospector: Improving LLM Agents with Self-Asking and Trajectory Ranking. | Byoungjip Kim, Youngsoo Jang, Lajanugen Logeswaran, Geon-Hyeong Kim, Yu Jin Kim, Honglak Lee, Moontae Lee |
| 2024 | ICML | Degeneration-free Policy Optimization: RL Fine-Tuning for Language Models without Degeneration. | Youngsoo Jang, Geon-Hyeong Kim, Byoungjip Kim, Yu Jin Kim, Honglak Lee, Moontae Lee |
| 2023 | ICML | Information-Theoretic State Space Model for Multi-View Reinforcement Learning. | HyeongJoo Hwang, Seokin Seo, Youngsoo Jang, Sungyoon Kim, Geon-Hyeong Kim, Seunghoon Hong, Kee-Eung Kim |
| 2022 | ICLR | GPT-Critic: Offline Reinforcement Learning for End-to-End Task-Oriented Dialogue Systems. | Youngsoo Jang, Jongmin Lee, Kee-Eung Kim |
| 2021 | ICLR | Monte-Carlo Planning and Learning with Language Action Value Estimates. | Youngsoo Jang, Seokin Seo, Jongmin Lee, Kee-Eung Kim |
| 2020 | AAAI | Bayes-Adaptive Monte-Carlo Planning and Learning for Goal-Oriented Dialogues. | Youngsoo Jang, Jongmin Lee, Kee-Eung Kim |
| 2020 | ACL | End-to-End Neural Pipeline for Goal-Oriented Dialogue Systems using GPT-2. | DongHoon Ham, Jeong-Gwan Lee, Youngsoo Jang, Kee-Eung Kim |
| 2020 | ICML | Variational Inference for Sequential Data with Future Likelihood Estimates. | Geon-Hyeong Kim, Youngsoo Jang, Hongseok Yang, Kee-Eung Kim |
| 2019 | ACML | Trust Region Sequential Variational Inference. | Geon-hyeong Kim, Youngsoo Jang, Jongmin Lee, Wonseok Jeon, Hongseok Yang, Kee-Eung Kim |
| 2019 | EMNLP | PyOpenDial: A Python-based Domain-Independent Toolkit for Developing Spoken Dialogue Systems with Probabilistic Rules. | Youngsoo Jang, Jongmin Lee, Jaeyoung Park, Kyeng-Hun Lee, Pierre Lison, Kee-Eung Kim |
| 2017 | IJCAI | Constrained Bayesian Reinforcement Learning via Approximate Linear Programming. | Jongmin Lee, Youngsoo Jang, Pascal Poupart, Kee-Eung Kim |