| 2026 | ACL | Affectron: Emotional Speech Synthesis with Affective and Contextually Aligned Nonverbal Vocalizations. | Deok-Hyeon Cho, Hyung-Seok Oh, Seung-Bin Kim, Seong-Whan Lee |
| 2025 | EMNLP | FillerSpeech: Towards Human-Like Text-to-Speech Synthesis with Filler Insertion and Filler Style Control. | Seung-Bin Kim, Junhyeok Cha, Hyung-Seok Oh, Heejin Choi, Seong-Whan Lee |
| 2025 | ICASSP | JELLY: Joint Emotion Recognition and Context Reasoning with LLMs for Conversational Speech Synthesis. | Junhyeok Cha, Seung-Bin Kim, Hyung-Seok Oh, Seong-Whan Lee |
| 2025 | Interspeech | VibE-SVC: Vibrato Extraction with High-frequency F0 Contour for Singing Voice Conversion. | Joon-Seung Choi, Dong-Min Byun, Hyung-Seok Oh, Seong-Whan Lee |
| 2025 | Interspeech | DiEmo-TTS: Disentangled Emotion Representations via Self-Supervised Distillation for Cross-Speaker Emotion Transfer in Text-to-Speech. | Deok-Hyeon Cho, Hyung-Seok Oh, Seung-Bin Kim, Seong-Whan Lee |
| 2025 | Interspeech | EmoSphere-SER: Enhancing Speech Emotion Recognition Through Spherical Representation with Auxiliary Classification. | Deok-Hyeon Cho, Hyung-Seok Oh, Seung-Bin Kim, Seong-Whan Lee |
| 2024 | Interspeech | EmoSphere-TTS: Emotional Style and Intensity Modeling via Spherical Emotion Vector for Controllable Emotional Text-to-Speech. | Deok-Hyeon Cho, Hyung-Seok Oh, Seung-Bin Kim, Sang-Hoon Lee, Seong-Whan Lee |
| 2023 | Interspeech | HierVST: Hierarchical Adaptive Zero-shot Voice Style Transfer. | Sang-Hoon Lee, Ha-Yeong Choi, Hyung-Seok Oh, Seong-Whan Lee |