| 2025 | ICLR | DiTTo-TTS: Diffusion Transformers for Scalable Text-to-Speech without Domain-Specific Factors. | Keon Lee, Dong Won Kim, Jaehyeon Kim, Seungjun Chung, Jaewoong Cho |
| 2025 | ICML | Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities. | Sreyan Ghosh, Zhifeng Kong, Sonal Kumar, S. Sakshi, Jaehyeon Kim, Wei Ping, Rafael Valle, Dinesh Manocha, Bryan Catanzaro |
| 2025 | ICML | Efficient Generative Modeling with Residual Vector Quantization-Based Tokens. | Jaehyeon Kim, Taehong Moon, Keon Lee, Jaewoong Cho |
| 2025 | ICML | How to Move Your Dragon: Text-to-Motion Synthesis for Large-Vocabulary Objects. | Wonkwang Lee, Jongwon Jeong, Taehong Moon, Hyeon-Jong Kim, Jaehyeon Kim, Gunhee Kim, Byeong-Uk Lee |
| 2024 | ICLR | CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech. | Jaehyeon Kim, Keon Lee, Seungjun Chung, Jaewoong Cho |
| 2023 | ICML | QASA: Advanced Question Answering on Scientific Articles. | Yoonjoo Lee, Kyungjae Lee, Sunghyun Park, Dasol Hwang, Jaehyeon Kim, Hong-In Lee, Moontae Lee |
| 2021 | ICML | Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech. | Jaehyeon Kim, Jungil Kong, Juhee Son |
| 2019 | ICML | FloWaveNet : A Generative Flow for Raw Audio. | Sungwon Kim, Sang-gil Lee, Jongyoon Song, Jaehyeon Kim, Sungroh Yoon |