| 2025 | AAAI | LAMA-UT: Language Agnostic Multilingual ASR Through Orthography Unification and Language-Specific Transliteration. | Sangmin Lee, Woo-Jin Chung, Hong-Goo Kang |
| 2025 | EMNLP | UniCoM: A Universal Code-Switching Speech Generator. | Sangmin Lee, Woojin Chung, Seyun Um, Hong-Goo Kang |
| 2025 | ICASSP | StableQuant: Layer Adaptive Post-Training Quantization for Speech Foundation Models. | Yeona Hong, Hyewon Han, Woo-Jin Chung, Hong-Goo Kang |
| 2025 | Interspeech | Neural Spectral Band Generation for Audio Coding. | Woongjib Choi, Byeong Hyeon Kim, Hyungseob Lim, Inseon Jang, Hong-Goo Kang |
| 2025 | Interspeech | Quadruple Path Modeling with Latent Feature Transfer for Permutation-free Continuous Speech Separation. | Jihyun Kim, Doyeon Kim, Hyewon Han, Jinyoung Lee, Jonguk Yoo, Chang Woo Han, Jeongook Song, Hoon-Young Cho, Hong-Goo Kang |
| 2025 | Interspeech | Towards an Ultra-Low-Delay Neural Audio Coding with Computational Efficiency. | Byeong Hyeon Kim, Hyungseob Lim, Inseon Jang, Hong-Goo Kang |
| 2025 | Interspeech | SpeechMLC: Speech Multi-label Classification. | Miseul Kim, Seyun Um, Hyeonjin Cha, Hong-Goo Kang |
| 2024 | ICASSP | On Fine-Tuning Pre-Trained Speech Models With EMA-Target Self-Supervised Loss. | Hejung Yang, Hong-Goo Kang |
| 2024 | Interspeech | Speaker-Independent Acoustic-to-Articulatory Inversion through Multi-Channel Attention Discriminator. | Woo-Jin Chung, Hong-Goo Kang |
| 2024 | Interspeech | Speak in the Scene: Diffusion-based Acoustic Scene Transfer toward Immersive Speech Generation. | Miseul Kim, Soo-Whan Chung, Youna Ji, Hong-Goo Kang, Min-Seok Choi |
| 2024 | Interspeech | Enhanced Deep Speech Separation in Clustered Ad Hoc Distributed Microphone Environments. | Jihyun Kim, Stijn Kindt, Nilesh Madhu, Hong-Goo Kang |
| 2024 | Interspeech | PARAN: Variational Autoencoder-based End-to-End Articulation-to-Speech System for Speech Intelligibility. | Seyun Um, Doyeon Kim, Hong-Goo Kang |
| 2024 | Interspeech | UNIQUE : Unsupervised Network for Integrated Speech Quality Evaluation. | Juhwan Yoon, WooSeok Ko, Seyun Um, Sungwoong Hwang, Soojoong Hwang, Changhwan Kim, Hong-Goo Kang |
| 2023 | ICASSP | Progressive Multi-Stage Neural Audio Codec with Psychoacoustic Loss and Discriminator. | Byeong Hyeon Kim, Hyungseob Lim, Jihyun Lee, Inseon Jang, Hong-Goo Kang |
| 2023 | ICASSP | Style Modeling for Multi-Speaker Articulation-to-Speech. | Miseul Kim, Zhenyu Piao, Jihyun Lee, Hong-Goo Kang |
| 2023 | ICASSP | End-to-End Neural Audio Coding in the MDCT Domain. | Hyungseob Lim, Jihyun Lee, Byeong Hyeon Kim, Inseon Jang, Hong-Goo Kang |
| 2023 | ICASSP | HappyQuokka System for ICASSP 2023 Auditory EEG Challenge. | Zhenyu Piao, Miseul Kim, Hyungchan Yoon, Hong-Goo Kang |
| 2023 | Interspeech | MF-PAM: Accurate Pitch Estimation through Periodicity Analysis and Multi-level Feature Fusion. | Woo-Jin Chung, Doyeon Kim, Soo-Whan Chung, Hong-Goo Kang |
| 2023 | Interspeech | HD-DEMUCS: General Speech Restoration with Heterogeneous Decoders. | Doyeon Kim, Soo-Whan Chung, Hyewon Han, Youna Ji, Hong-Goo Kang |
| 2023 | Interspeech | Contrastive Learning based Deep Latent Masking for Music Source Separation. | Jihyun Kim, Hong-Goo Kang |
| 2023 | Interspeech | Feature Normalization for Fine-tuning Self-Supervised Models in Speech Enhancement. | Hejung Yang, Hong-Goo Kang |
| 2023 | Interspeech | Pruning Self-Attention for Zero-Shot Multi-Speaker Text-to-Speech. | Hyungchan Yoon, Changhwan Kim, Eunwoo Song, Hyun-Wook Yoon, Hong-Goo Kang |
| 2023 | Interspeech | Adversarial Learning of Intermediate Acoustic Feature for End-to-End Lightweight Text-to-Speech. | Hyungchan Yoon, Seyun Um, Changhwan Kim, Hong-Goo Kang |
| 2022 | ICASSP | Phase Continuity: Learning Derivatives of Phase Spectrum for Speech Enhancement. | Doyeon Kim, Hyewon Han, Hyeon-Kyeong Shin, Soo-Whan Chung, Hong-Goo Kang |
| 2022 | ICASSP | Progressive Multi-Stage Neural Audio Coding with Guided References. | Chanwoo Lee, Hyungseob Lim, Jihyun Lee, Inseon Jang, Hong-Goo Kang |
| 2022 | ICASSP | Adversarial Audio Synthesis Using a Harmonic-Percussive Discriminator. | Jihyun Lee, Hyungseob Lim, Chanwoo Lee, Inseon Jang, Hong-Goo Kang |
| 2022 | Interspeech | Light-Weight Speaker Verification with Global Context Information. | Miseul Kim, Zhenyu Piao, Seyun Um, Ran Lee, Jaemin Joh, Seungshin Lee, Hong-Goo Kang |
| 2022 | Interspeech | FluentTTS: Text-dependent Fine-grained Style Control for Multi-style TTS. | Changhwan Kim, Seyun Um, Hyungchan Yoon, Hong-Goo Kang |
| 2022 | Interspeech | Learning Audio-Text Agreement for Open-vocabulary Keyword Spotting. | Hyeon-Kyeong Shin, Hyewon Han, Doyeon Kim, Soo-Whan Chung, Hong-Goo Kang |
| 2021 | CVPR | Looking Into Your Speech: Learning Cross-Modal Affinity for Audio-Visual Speech Separation. | Jiyoung Lee, Soo-Whan Chung, Sunok Kim, Hong-Goo Kang, Kwanghoon Sohn |
| 2021 | Interspeech | LiteTTS: A Lightweight Mel-Spectrogram-Free Text-to-Wave Synthesizer Based on Generative Adversarial Networks. | Huu-Kim Nguyen, Kihyuk Jeong, Seyun Um, Min-Jae Hwang, Eunwoo Song, Hong-Goo Kang |
| 2020 | ACSSC | A Study on Conditional Features for a Flow-based Neural Vocoder. | Hyungseob Lim, Suhyeon Oh, Kyungguen Byun, Hong-Goo Kang |
| 2020 | ICASSP | Improving LPCNET-Based Text-to-Speech with Linear Prediction-Structured Mixture Density Network. | Min-Jae Hwang, Eunwoo Song, Ryuichi Yamamoto, Frank K. Soong, Hong-Goo Kang |
| 2020 | ICASSP | Emotional Speech Synthesis with Rich and Granularized Control. | Se-Yun Um, Sangshin Oh, Kyungguen Byun, Inseon Jang, Chunghyun Ahn, Hong-Goo Kang |
| 2020 | Interspeech | FaceFilter: Audio-Visual Speech Separation Using Still Images. | Soo-Whan Chung, Soyeon Choe, Joon Son Chung, Hong-Goo Kang |
| 2020 | Interspeech | Seeing Voices and Hearing Voices: Learning Discriminative Embeddings Using Cross-Modal Self-Supervision. | Soo-Whan Chung, Hong-Goo Kang, Joon Son Chung |
| 2020 | Interspeech | MIRNet: Learning Multiple Identities Representations in Overlapped Speech. | Hyewon Han, Soo-Whan Chung, Hong-Goo Kang |
| 2020 | Interspeech | A Cross-Channel Attention-Based Wave-U-Net for Multi-Channel Speech Enhancement. | Minh Tri Ho, Jinyoung Lee, Bong-Ki Lee, Dong Hoon Yi, Hong-Goo Kang |
| 2020 | Interspeech | Intra-Class Variation Reduction of Speaker Representation in Disentanglement Framework. | Yoohwan Kwon, Soo-Whan Chung, Hong-Goo Kang |
| 2020 | MMSP | Speaker-Adaptive Neural Vocoders for Parametric Speech Synthesis Systems. | Eunwoo Song, Jin-Seob Kim, Kyungguen Byun, Hong-Goo Kang |
| 2019 | ICASSP | Perfect Match: Improved Cross-modal Embeddings for Audio-visual Synchronisation. | Soo-Whan Chung, Joon Son Chung, Hong-Goo Kang |
| 2019 | ICASSP | Gradient-based Active Learning Query Strategy for End-to-end Speech Recognition. | Yang Yuan, Soo-Whan Chung, Hong-Goo Kang |
| 2019 | Interspeech | Parameter Enhancement for MELP Speech Codec in Noisy Communication Environment. | Min-Jae Hwang, Hong-Goo Kang |
| 2018 | ICASSP | Modeling-By-Generation-Structured Noise Compensation Algorithm for Glottal Vocoding Speech Synthesis System. | Min-Jae Hwang, Eunwoo Song, Kyungguen Byun, Hong-Goo Kang |
| 2018 | ICASSP | Dnn-Based Wireless Positioning in an Outdoor Environment. | Jin-Young Lee, Chahyeon Eom, Youngsu Kwak, Hong-Goo Kang, Chungyong Lee |
| 2018 | Interspeech | A Unified Framework for the Generation of Glottal Signals in Deep Learning-based Parametric Speech Synthesis Systems. | Min-Jae Hwang, Eunwoo Song, Jin-Seob Kim, Hong-Goo Kang |
| 2017 | ASRU | Perceptual quality and modeling accuracy of excitation parameters in DLSTM-based speech synthesis systems. | Eunwoo Song, Frank K. Soong, Hong-Goo Kang |
| 2016 | Interspeech | Improved Time-Frequency Trajectory Excitation Vocoder for DNN-Based Speech Synthesis. | Eunwoo Song, Frank K. Soong, Hong-Goo Kang |
| 2015 | ICASSP | Coherent channel based subband multichannel dereverberation. | JeeSok Lee, Sejin Oh, Hong-Goo Kang |
| 2015 | ICASSP | Improved time-frequency trajectory excitation modeling for a statistical parametric speech synthesis system. | Eunwoo Song, Young-Sun Joo, Hong-Goo Kang |
| 2015 | Interspeech | Systematic integration of acoustic echo canceller and noise reduction modules for voice communication systems. | Hyeonjoo Kang, JeeSok Lee, Soonho Baek, Hong-Goo Kang |
| 2015 | Interspeech | Deep neural network-based statistical parametric speech synthesis system using improved time-frequency trajectory excitation model. | Eunwoo Song, Hong-Goo Kang |
| 2014 | ICASSP | Mean normalization of power function based cepstral coefficients for robust speech recognition in noisy environment. | Soonho Baek, Hong-Goo Kang |
| 2014 | ICASSP | Detecting pathological speech using contour modeling of harmonic-to-noise ratio. | Jung-Won Lee, Samuel Kim, Hong-Goo Kang |
| 2014 | ICASSP | Factored adaptation of speaker and environment using orthogonal subspace transforms. | Hyunson Seo, Hong-Goo Kang, Michael L. Seltzer |
| 2014 | ICASSP | A maximum a Posterior-based reconstruction approach to speech bandwidth expansion in noise. | Hyunson Seo, Hong-Goo Kang, Frank K. Soong |
| 2013 | ASRU | Vector Taylor series based HMM adaptation for generalized cepstrum in noisy environment. | Soonho Baek, Hong-Goo Kang |
| 2013 | ICASSP | Enhancement of spectral clarity for HMM-based text-to-speech systems. | Young-Sun Joo, Chi-Sang Jung, Hong-Goo Kang |
| 2013 | Interspeech | A source-filter based adaptive harmonic model and its application to speech prosody modification. | JeeSok Lee, Frank K. Soong, Hong-Goo Kang |
| 2011 | ICASSP | Enhanced long-term predictor for Unified Speech and Audio Coding. | Jeongook Song, Hyen-O Oh, Hong-Goo Kang |
| 2011 | Interspeech | Classification of Fricatives Using Feature Extrapolation of Acoustic-Phonetic Features in Telephone Speech. | Jung-Won Lee, Jeung-Yoon Choi, Hong-Goo Kang |
| 2010 | ICASSP | Binaural loudness based speech reinforcement with a closed-form solution. | Ho Seon Shin, Min-Seok Choi, Taesu Kim, Hong-Goo Kang |
| 2010 | Interspeech | A variable frame length and rate algorithm based on the spectral kurtosis measure for speaker verification. | Chi-Sang Jung, Kyu Jeong Han, Hyunson Seo, Shrikanth S. Narayanan, Hong-Goo Kang |
| 2010 | MMSP | Enhancing loudspeaker-based 3D audio with room modeling. | Myung-Suk Song, Cha Zhang, Dinei A. F. Florncio, Hong-Goo Kang |
| 2009 | ICASSP | Normalized minimum-redundancy and maximum-relevancy based feature selection for speaker verification systems. | Chi-Sang Jung, Moo Young Kim, Hong-Goo Kang |
| 2008 | ICASSP | Designing a unified speech/audio codec by adopting a single channel harmonic source separation module. | Sang-Wook Shin, Chang-Heon Lee, Hyen-O Oh, Hong-Goo Kang |
| 2007 | ICASSP | A Soft-Decision Adaptation Mode Controller for an Efficient Frequency-Domain Generalized Sidelobe Canceller. | Min-Seok Choi, Chang-Hyun Baik, Young-Cheol Park, Hong-Goo Kang |
| 2007 | Interspeech | Speech quality estimation using packet loss effects in CELP-type speech coders. | Min-Ki Lee, Kyung-Tae Kim, Hong-Goo Kang, Dae Hee Youn |
| 2006 | ICASSP | On the Use of Voting Methods for Speaker Identification Based on Various Resolution Filterbanks. | Bong-Jin Lee, Sung-Wan Yoon, Hong-Goo Kang, Dae Hee Youn |
| 2006 | Interspeech | An efficient segment-based speech compression technique for hand-held TTS systems. | Chang-Heon Lee, Sung-Kyo Jung, Thomas Eriksson, Won-Suk Jun, Hong-Goo Kang |
| 2006 | Interspeech | Performance analysis of various single channel speech enhancement algorithms for automatic speech recognition. | Myung-Suk Song, Chang-Heon Lee, Hong-Goo Kang |
| 2005 | ICASSP | An improved estimation of a priori speech absence probability for speech enhancement : in perspective of speech perception. | Min-Seok Choi, Hong-Goo Kang |
| 2005 | Interspeech | A noise-robust pitch synchronous feature extraction algorithm for speaker recognition systems. | Samuel Kim, Sung-Wan Yoon, Thomas Eriksson, Hong-Goo Kang, Dae Hee Youn |
| 2004 | ICASSP | Improvement issues on transcoding algorithms: for the flexible usage to the various pairs of speech codec. | Jin-Kyu Choi, Chang-Heon Lee, Hong-Goo Kang, Young-Cheol Park, Dae Hee Youn |
| 2004 | ICASSP | A bit-rate/bandwidth scalable speech coder based on ITU-T G.723.1 standard. | Sung-Kyo Jung, Kyung-Tae Kim, Hong-Goo Kang |
| 2004 | ICASSP | A pitch synchronous feature extraction method for speaker recognition. | Samuel Kim, Thomas Eriksson, Hong-Goo Kang, Dae Hee Youn |
| 2004 | Interspeech | Theory for speaker recognition over IP. | Thomas Eriksson, Samuel Kim, Hong-Goo Kang, Chungyong Lee |
| 2004 | Interspeech | Performance analysis of transcoding algorithms in packet-loss environments. | Sung-Kyo Jung, Hong-Goo Kang, Dae Hee Youn, Chang-Heon Lee |
| 2004 | Interspeech | On the time variability of vocal tract for speaker recognition. | Samuel Kim, Thomas Eriksson, Hong-Goo Kang |
| 2004 | Interspeech | Temporal normalization techniques for transform-type speech coding and application to split-band wideband coders. | Kyung-Tae Kim, Sung-Kyo Jung, MiSuk Lee, Hong-Goo Kang, Dae Hee Youn |
| 2003 | ICASSP | A cascaded algebraic codebook structure to improve the performance of speech coder. | Sung-Kyo Jung, Kyoung-Tae Kim, Hong-Goo Kang, Dae Hee Youn |
| 2003 | ICASSP | A packet loss concealment algorithm based on time-scale modification for CELP-type speech coders. | Moon-Keun Lee, Sung-Kyo Jung, Hong-Goo Kang, Young-Cheol Park, Dae Hee Youn |
| 2003 | Interspeech | Transcoding algorithm for g.723.1 and AMR speech coders: for interoperability between voIP and mobile networks. | Sung-Wan Yoon, Jin-Kyu Choi, Hong-Goo Kang, Dae Hee Youn |
| 2002 | ICASSP | A phase generation method for speech reconstruction from spectral envelope and pitch intervals. | Hong-Goo Kang, Hong Kook Kim |
| 2001 | ICASSP | A candidate for the ITU-T 4 kbit/s speech coding standard. | Jes Thyssen, Yang Gao, Adil Benyassine, Eyal Shlomot, Carlo Murgia, Huan-yu Su, Kazunori Mano, Yusuke Hiwasaki, Hiroyuki Ehara, Kazutoshi Yasunaga, Claude Lamblin, Balzs Kvesi, Joachim Stegmann, Hong-Goo Kang |
| 2001 | Interspeech | Acoustic feature compensation based on decomposition of speech and noise for ASR in noisy environments. | Hong Kook Kim, Richard C. Rose, Hong-Goo Kang |
| 2000 | ICASSP | Low-rate quantization of spectrum parameters. | Thomas Eriksson, Hong-Goo Kang, Per Hedelin |
| 1999 | ICASSP | Pitch quantization in low bit-rate speech coding. | Thomas Eriksson, Hong-Goo Kang |
| 1999 | ICASSP | Phase adjustment in waveform interpolation. | Hong-Goo Kang, Dipanjan Sen |
| 1999 | Interspeech | Low delay analysis/synthesis schemes for joint speech enhancement and low bit rate speech coding. | Rainer Martin, Hong-Goo Kang, Richard V. Cox |
| 1998 | ICASSP | Quantization of the spectral envelope for sinusoidal coders. | Thomas Eriksson, Hong-Goo Kang, Yannis Stylianou |
| 1997 | Interspeech | A 3 channel digital CVSD bit-rate conversion system using a general purpose DSP. | Yong-Soo Choi, Hong-Goo Kang, Sung-Youn Kim, Young-Cheol Park, Dae Hee Youn |
| 1997 | Interspeech | Improved regular pulse VSELP coding of speech at low bit-rates. | Yong-Soo Choi, Hong-Goo Kang, Sang-Wook Park, Jae-Ha Yoo, Dae Hee Youn |
| 1996 | ICASSP | A fast VSELP speech coder based on mutually orthonormal regular pulse vectors. | Yong-Soo Choi, Hong-Goo Kang, Dae Hee Youn |
| 1995 | Interspeech | A low bit-rate speech coder using the perceptual properties of the human ear. | Hong-Goo Kang, Jeong Tae Seo, Il-Whan Cha, Dae Hee Youn |