| 2024 | ACL | Textless Acoustic Model with Self-Supervised Distillation for Noise-Robust Expressive Speech-to-Speech Translation. | Min-Jae Hwang, Ilia Kulikov, Benjamin N. Peloquin, Hongyu Gong, Peng-Jen Chen, Ann Lee |
| 2022 | Interspeech | TTS-by-TTS 2: Data-Selective Augmentation for Neural Speech Synthesis Using Ranking Support Vector Machine with Variational Autoencoder. | Eunwoo Song, Ryuichi Yamamoto, Ohsung Kwon, Chan-Ho Song, Min-Jae Hwang, Suhyeon Oh, Hyun-Wook Yoon, Jin-Seob Kim, Jae-Min Kim |
| 2022 | Interspeech | Language Model-Based Emotion Prediction Methods for Emotional Speech Synthesis Systems. | Hyun-Wook Yoon, Ohsung Kwon, Hoyeon Lee, Ryuichi Yamamoto, Eunwoo Song, Jae-Min Kim, Min-Jae Hwang |
| 2021 | ICASSP | TTS-by-TTS: TTS-Driven Data Augmentation for Fast and High-Quality Speech Synthesis. | Min-Jae Hwang, Ryuichi Yamamoto, Eunwoo Song, Jae-Min Kim |
| 2021 | ICASSP | Parallel Waveform Synthesis Based on Generative Adversarial Networks with Voicing-Aware Conditional Discriminators. | Ryuichi Yamamoto, Eunwoo Song, Min-Jae Hwang, Jae-Min Kim |
| 2021 | Interspeech | High-Fidelity Parallel WaveGAN with Multi-Band Harmonic-Plus-Noise Model. | Min-Jae Hwang, Ryuichi Yamamoto, Eunwoo Song, Jae-Min Kim |
| 2021 | Interspeech | LiteTTS: A Lightweight Mel-Spectrogram-Free Text-to-Wave Synthesizer Based on Generative Adversarial Networks. | Huu-Kim Nguyen, Kihyuk Jeong, Seyun Um, Min-Jae Hwang, Eunwoo Song, Hong-Goo Kang |
| 2020 | ICASSP | Improving LPCNET-Based Text-to-Speech with Linear Prediction-Structured Mixture Density Network. | Min-Jae Hwang, Eunwoo Song, Ryuichi Yamamoto, Frank K. Soong, Hong-Goo Kang |
| 2020 | Interspeech | Neural Text-to-Speech with a Modeling-by-Generation Excitation Vocoder. | Eunwoo Song, Min-Jae Hwang, Ryuichi Yamamoto, Jin-Seob Kim, Ohsung Kwon, Jae-Min Kim |
| 2019 | Interspeech | Parameter Enhancement for MELP Speech Codec in Noisy Communication Environment. | Min-Jae Hwang, Hong-Goo Kang |
| 2018 | ICASSP | Modeling-By-Generation-Structured Noise Compensation Algorithm for Glottal Vocoding Speech Synthesis System. | Min-Jae Hwang, Eunwoo Song, Kyungguen Byun, Hong-Goo Kang |
| 2018 | Interspeech | A Unified Framework for the Generation of Glottal Signals in Deep Learning-based Parametric Speech Synthesis Systems. | Min-Jae Hwang, Eunwoo Song, Jin-Seob Kim, Hong-Goo Kang |