| 2026 | AAAI | OnEDIT: Online Editing with Decoupled Implicit Task for Large Language Models. | Chae-Won Lee, Jae-Hong Lee, Ji-Hun Kang, Joon-Hyuk Chang |
| 2025 | ASRU | A Momentum-Based Framework with Contrastive Data Generation for Robust Sound Source Localization. | Hyun-Soo Kim, Da-Hee Yang, Joon-Hyuk Chang |
| 2025 | ASRU | Improving Noise Robust Audio-Visual Speech Recognition via Router-Gated Cross-Modal Feature Fusion. | DongHoon Lim, YoungChae Kim, Dong-Hyun Kim, Da-Hee Yang, Joon-Hyuk Chang |
| 2025 | ICASSP | Diffusion-based Target Device Style Transfer for Robust Acoustic Scene Classification. | Won-Gook Choi, Joon-Hyuk Chang |
| 2025 | ICASSP | Trainable Adaptive Score Normalization for Automatic Speaker Verification. | Jeong-Hwan Choi, Ju-Seok Seong, Ye-Rin Jeoung, Joon-Hyuk Chang |
| 2025 | ICASSP | Multimodal Emotion Recognition with Target Speaker-Based Facial Embeddings. | Serin Heo, Jehyun Kyung, Joon-Hyuk Chang |
| 2025 | ICASSP | Progressive Subband Modeling for Artifacts-free Speech Super-resolution. | Donghyun Kim, Joon-Hyuk Chang |
| 2025 | ICASSP | Few-shot Keyword-incremental Learning Using Compositional Information. | Ilseok Kim, Ju-Seok Seong, Joon-Hyuk Chang |
| 2025 | ICASSP | Quad-Net: Melspectrogram Vocoder with Convolutional Layers Restricted by the Quadrature Mirror Filter for Perfect Reconstruction. | Nam-Seok Song, Joon-Hyuk Chang |
| 2025 | Interspeech | Optimizing CLAP Reward with LLM Feedback for Semantically Aligned and Diverse Automated Audio Captioning. | Seyun Ahn, Pil Moo Byun, Won-Gook Choi, Joon-Hyuk Chang |
| 2025 | Interspeech | Temp4Cap: Temporally-aligned Automated Audio Captioning. | Ho-Young Choi, Jae-Heung Cho, Pil Moo Byun, Won-Gook Choi, Joon-Hyuk Chang |
| 2025 | Interspeech | Spatially Weighted Contrastive Learning for Robust Sound Source Localization. | Hyun-Soo Kim, Da-Hee Yang, Joon-Hyuk Chang |
| 2025 | Interspeech | Improving Generalization of End-to-End ASR through Diversity and Independence Regularization. | Ye-Eun Ko, Mun-Hak Lee, Dong-Hyun Kim, Joon-Hyuk Chang |
| 2025 | Interspeech | Enhancing Target-speaker Automatic Speech Recognition Using Multiple Speaker Embedding Extractors with Virtual Speaker Embedding. | Ju-Seok Seong, Jeong-Hwan Choi, Ye-Rin Jeoung, Ilseok Kim, Joon-Hyuk Chang |
| 2024 | ICASSP | Generalized Specaugment via Multi-Rectangle Inverse Masking For Acoustic Scene Classification. | Pil Moo Byun, Joon-Hyuk Chang |
| 2024 | ICASSP | Adversarial Learning on Compressed Posterior Space for Non-Iterative Score-based End-to-End Text-to-Speech. | Won-Gook Choi, Donghyun Seong, Joon-Hyuk Chang |
| 2024 | ICASSP | Improving Target Sound Extraction with Timestamp Knowledge Distillation. | Dail Kim, Min-Sang Baek, Yungyeo Kim, Joon-Hyuk Chang |
| 2024 | ICASSP | Class: Continual Learning Approach for Speech Super-Resolution. | Donghyun Kim, Yungyeo Kim, Joon-Hyuk Chang |
| 2024 | ICASSP | Text-Only Unsupervised Domain Adaptation for Neural Transducer-Based ASR Personalization Using Synthesized Data. | Dong-Hyun Kim, Jae-Hong Lee, Joon-Hyuk Chang |
| 2024 | ICLR | Continual Momentum Filtering on Parameter Space for Online Test-time Adaptation. | Jae-Hong Lee, Joon-Hyuk Chang |
| 2024 | ICML | Stationary Latent Weight Inference for Unreliable Observations from Online Test-Time Adaptation. | Jae-Hong Lee, Joon-Hyuk Chang |
| 2024 | Interspeech | Retrieval-Augmented Classifier Guidance for Audio Generation. | Ho-Young Choi, Won-Gook Choi, Joon-Hyuk Chang |
| 2024 | Interspeech | Efficient Speaker Embedding Extraction Using a Twofold Sliding Window Algorithm for Speaker Diarization. | Jeong-Hwan Choi, Ye-Rin Jeoung, Ilseok Kim, Joon-Hyuk Chang |
| 2024 | Interspeech | Whisper Multilingual Downstream Task Tuning Using Task Vectors. | Ji-Hun Kang, Jae-Hong Lee, Mun-Hak Lee, Joon-Hyuk Chang |
| 2024 | Interspeech | Mitigating Overfitting in Structured Pruning of ASR Models with Gradient-Guided Parameter Regularization. | Dong-Hyun Kim, Joon-Hyuk Chang |
| 2024 | Interspeech | Sound of Vision: Audio Generation from Visual Text Embedding through Training Domain Discriminator. | Jaewon Kim, Won-Gook Choi, Seyun Ahn, Joon-Hyuk Chang |
| 2024 | Interspeech | Few-Shot Keyword-Incremental Learning with Total Calibration. | Ilseok Kim, Ju-Seok Seong, Joon-Hyuk Chang |
| 2024 | Interspeech | Guided conditioning with predictive network on score-based diffusion model for speech enhancement. | Dail Kim, Da-Hee Yang, Donghyun Kim, Joon-Hyuk Chang, Jeonghwan Choi, Moa Lee, Jaemo Yang, Han-gil Moon |
| 2024 | Interspeech | Enhancing Multimodal Emotion Recognition through ASR Error Compensation and LLM Fine-Tuning. | Jehyun Kyung, Serin Heo, Joon-Hyuk Chang |
| 2024 | Interspeech | Neural ATSM: Fully Neural Network-based Adaptive Time-Scale Modification Using Sentence-Specific Dynamic Control. | Jaeuk Lee, Sohee Jang, Joon-Hyuk Chang |
| 2024 | Interspeech | Online Subloop Search via Uncertainty Quantization for Efficient Test-Time Adaptation. | Jae-Hong Lee, Sang-Eon Lee, Dong-Hyun Kim, Do-Hee Kim, Joon-Hyuk Chang |
| 2024 | Interspeech | Balanced-Wav2Vec: Enhancing Stability and Robustness of Representation Learning Through Sample Reweighting Techniques. | Mun-Hak Lee, Jae-Hong Lee, Do-Hee Kim, Ye-Eun Ko, Joon-Hyuk Chang |
| 2024 | Interspeech | H4C-TTS: Leveraging Multi-Modal Historical Context for Conversational Text-to-Speech. | Donghyun Seong, Joon-Hyuk Chang |
| 2024 | Interspeech | TSP-TTS: Text-based Style Predictor with Residual Vector Quantization for Expressive Text-to-Speech. | Donghyun Seong, Hoyoung Lee, Joon-Hyuk Chang |
| 2023 | ASRU | Extending Self-Distilled Self-Supervised Learning For Semi-Supervised Speaker Verification. | Jeong-Hwan Choi, Jehyun Kyung, Ju-Seok Seong, Ye-Rin Jeoung, Joon-Hyuk Chang |
| 2023 | ASRU | AWMC: Online Test-Time Adaptation Without Mode Collapse for Continual Adaptation. | Jae-Hong Lee, Do-Hee Kim, Joon-Hyuk Chang |
| 2023 | ASRU | Cross-Modal Learning for CTC-Based ASR: Leveraging CTC-Bertscore and Sequence-Level Training. | Mun-Hak Lee, Sang-Eon Lee, Ji-Eun Choi, Joon-Hyuk Chang |
| 2023 | ASRU | Knowledge Distillation From Offline to Streaming Transducer: Towards Accurate and Fast Streaming Model by Matching Alignments. | Ji-Hwan Mo, Jae-Jin Jeon, Mun-Hak Lee, Joon-Hyuk Chang |
| 2023 | ASRU | Towards Robust Packet Loss Concealment System With ASR-Guided Representations. | Da-Hee Yang, Joon-Hyuk Chang |
| 2023 | ICASSP | CAN2V: Can-Bus Data-Based Seq2seq Model for Vehicle Velocity Prediction. | Jae-Heung Cho, Joon-Hyuk Chang |
| 2023 | ICASSP | M-CTRL: A Continual Representation Learning Framework with Slowly Improving Past Pre-Trained Model. | Jin-Seong Choi, Jae-Hong Lee, Chae-Won Lee, Joon-Hyuk Chang |
| 2023 | ICASSP | Adaptive Time-Scale Modification for Improving Speech Intelligibility Based On Phoneme Clustering For Streaming Services. | Sohee Jang, Jiye Kim, Yeon-Ju Kim, Joon-Hyuk Chang |
| 2023 | ICASSP | Improving Transformer-Based End-to-End Speaker Diarization by Assigning Auxiliary Losses to Attention Heads. | Ye-Rin Jeoung, Joon-Young Yang, Jeong-Hwan Choi, Joon-Hyuk Chang |
| 2023 | ICASSP | Repackagingaugment: Overcoming Prediction Error Amplification in Weight-Averaged Speech Recognition Models Subject to Self-Training. | Jae-Hong Lee, Dong-Hyun Kim, Joon-Hyuk Chang |
| 2023 | ICASSP | Noise-Aware Target Extension with Self-Distillation for Robust Speech Recognition. | Ju-Seok Seong, Jeong-Hwan Choi, Jehyun Kyung, Ye-Rin Jeoung, Joon-Hyuk Chang |
| 2023 | ICASSP | Selective Film Conditioning with CTC-Based ASR Probability for Speech Enhancement. | Da-Hee Yang, Joon-Hyuk Chang |
| 2023 | Interspeech | Deeply Supervised Curriculum Learning for Deep Neural Network-based Sound Source Localization. | Min-Sang Baek, Joon-Young Yang, Joon-Hyuk Chang |
| 2023 | Interspeech | SR-SRP: Super-Resolution based SRP-PHAT for Sound Source Localization and Tracking. | Jae-Heung Cho, Joon-Hyuk Chang |
| 2023 | Interspeech | Resolution Consistency Training on Time-Frequency Domain for Semi-Supervised Sound Event Detection. | Won-Gook Choi, Joon-Hyuk Chang |
| 2023 | Interspeech | Prior-free Guided TTS: An Improved and Efficient Diffusion-based Text-Guided Speech Synthesis. | Won-Gook Choi, So-Jeong Kim, Tae-Ho Kim, Joon-Hyuk Chang |
| 2023 | Interspeech | Self-Distillation into Self-Attention Heads for Improving Transformer-based End-to-End Neural Speaker Diarization. | Ye-Rin Jeoung, Jeong-Hwan Choi, Ju-Seok Seong, Jehyun Kyung, Joon-Hyuk Chang |
| 2023 | Interspeech | Intra-ensemble: A New Method for Combining Intermediate Outputs in Transformer-based Automatic Speech Recognition. | Do-Hee Kim, Ji-Eun Choi, Joon-Hyuk Chang |
| 2023 | Interspeech | General-purpose Adversarial Training for Enhanced Automatic Speech Recognition Model Generalization. | Do-Hee Kim, Daeyeol Shim, Joon-Hyuk Chang |
| 2023 | Interspeech | Improving Joint Speech and Emotion Recognition Using Global Style Tokens. | Jehyun Kyung, Ju-Seok Seong, Jeong-Hwan Choi, Ye-Rin Jeoung, Joon-Hyuk Chang |
| 2023 | Interspeech | HAD-ANC: A Hybrid System Comprising an Adaptive Filter and Deep Neural Networks for Active Noise Control. | JungPhil Park, Jeong-Hwan Choi, Yungyeo Kim, Joon-Hyuk Chang |
| 2022 | ICASSP | Knowledge Distillation from Language Model to Acoustic Model: A Hierarchical Multi-Task Learning Approach. | Mun-Hak Lee, Joon-Hyuk Chang |
| 2022 | Interspeech | Convolutional Recurrent Neural Network with Auxiliary Stream for Robust Variable-Length Acoustic Scene Classification. | Joon-Hyuk Chang, Won-Gook Choi |
| 2022 | Interspeech | Improved CNN-Transformer using Broadcasted Residual Learning for Text-Independent Speaker Verification. | Jeong-Hwan Choi, Joon-Young Yang, Ye-Rin Jeoung, Joon-Hyuk Chang |
| 2022 | Interspeech | HYU Submission for the SASV Challenge 2022: Reforming Speaker Embeddings with Spoofing-Aware Conditioning. | Jeong-Hwan Choi, Joon-Young Yang, Ye-Rin Jeoung, Joon-Hyuk Chang |
| 2022 | Interspeech | Adversarial and Sequential Training for Cross-lingual Prosody Transfer TTS. | Min-Kyung Kim, Joon-Hyuk Chang |
| 2022 | Interspeech | W2V2-Light: A Lightweight Version of Wav2vec 2.0 for Automatic Speech Recognition. | Dong-Hyun Kim, Jae-Hong Lee, Ji-Hwan Mo, Joon-Hyuk Chang |
| 2022 | Interspeech | One-Shot Speaker Adaptation Based on Initialization by Generative Adversarial Networks for TTS. | Jaeuk Lee, Joon-Hyuk Chang |
| 2022 | Interspeech | Advanced Speaker Embedding with Predictive Variance of Gaussian Distribution for Speaker Adaptation in TTS. | Jaeuk Lee, Joon-Hyuk Chang |
| 2022 | Interspeech | Regularizing Transformer-based Acoustic Models by Penalizing Attention Weights. | Mun-Hak Lee, Joon-Hyuk Chang, Sang-Eon Lee, Ju-Seok Seong, Chanhee Park, Haeyoung Kwon |
| 2022 | Interspeech | CTRL: Continual Representation Learning to Transfer Information of Pre-trained for WAV2VEC 2.0. | Jae-Hong Lee, Chae Won Lee, Jin-Seong Choi, Joon-Hyuk Chang, Woo Kyeong Seong, Jeonghan Lee |
| 2022 | Interspeech | FiLM Conditioning with Enhanced Feature to the Transformer-based End-to-End Noisy Speech Recognition. | Da-Hee Yang, Joon-Hyuk Chang |
| 2021 | ASRU | Short-Utterance Embedding Enhancement Method Based on Time Series Forecasting Technique for Text-Independent Speaker Verification. | Jeong-Hwan Choi, Joon-Young Yang, Joon-Hyuk Chang |
| 2021 | Interspeech | Deep Neural Network Calibration for E2E Speech Recognition System. | Mun-Hak Lee, Joon-Hyuk Chang |
| 2021 | ISCAS | MIMO Noise Suppression Preserving Spatial Cues for Sound Source Localization in Mobile Robot. | Jung-Hee Kim, Jeong-Hwan Choi, Jinyoung Son, Gyeong-Su Kim, Jihwan Park, Joon-Hyuk Chang |
| 2020 | Interspeech | Attention Wave-U-Net for Acoustic Echo Cancellation. | Jung-Hee Kim, Joon-Hyuk Chang |
| 2020 | Interspeech | Virtual Acoustic Channel Expansion Based on Neural Networks for Weighted Prediction Error-Based Speech Dereverberation. | Joon-Young Yang, Joon-Hyuk Chang |
| 2019 | Interspeech | Joint Optimization of Neural Acoustic Beamforming and Dereverberation with x-Vectors for Robust Speaker Verification. | Joon-Young Yang, Joon-Hyuk Chang |
| 2018 | IROS | DNN-based Speech Recognition System dealing with Motor State as Auxiliary Information of DNN for Head Shaking Robot. | Moa Lee, Joon-Hyuk Chang |
| 2016 | ICASSP | Dual-microphone voice activity detection based on using optimally weighted maximum a posteriori probabilities. | Seng Hyun Huang, Jihwan Park, Joon-Hyuk Chang |
| 2015 | Interspeech | A statistical model-based voice activity detection using multiple DNNs and noise awareness. | Inyoung Hwang, Jaeseong Sim, Sang-Hyeon Kim, Kwang-Sub Song, Joon-Hyuk Chang |
| 2014 | Interspeech | Enhanced muting method in packet loss concealment of ITU-t g.722 using sigmoid function with on-line optimized parameters. | Bong-Ki Lee, Inyoung Hwang, Jihwan Park, Joon-Hyuk Chang |
| 2013 | Interspeech | Enhanced muting method in packet loss concealment of ITU-t g.722 employing optimized sigmoid function. | Bong-Ki Lee, Chungsoo Lim, Jihwan Park, Joon-Hyuk Chang |
| 2012 | ICASSP | Adaptive noise power estimation using spectral difference for robust speech enhancement. | Jae-Hun Choi, Sang-Kyun Kim, Joon-Hyuk Chang |
| 2012 | ICASSP | New techniques for improving the practicality of an SVM-based speech/music classifier. | Chungsoo Lim, Seong-Ro Lee, Yeonwoo Lee, Joon-Hyuk Chang |
| 2011 | Interspeech | A Soft Decision-Based Speech Enhancement Using Acoustic Noise Classification. | Jae-Hun Choi, Sang-Kyun Kim, Joon-Hyuk Chang |
| 2010 | ICASSP | A statistical model-based double-talk detection incorporating soft decision. | Yun-Sik Park, Ji-Hyun Song, Sang-Ick Kang, Woojung Lee, Joon-Hyuk Chang |
| 2010 | Interspeech | Toward detecting voice activity employing soft decision in second-order conditional MAP. | Sang-Kyun Kim, Jae-Hun Choi, Sang-Ick Kang, Ji-Hyun Song, Joon-Hyuk Chang |
| 2010 | Interspeech | On using Gaussian mixture model for double-talk detection in acoustic echo suppression. | Ji-Hyun Song, Kyu-Ho Lee, Yun-Sik Park, Sang-Ick Kang, Joon-Hyuk Chang |
| 2009 | ICASSP | Speech enhancement based on minima controlled recursive averaging incorporating conditional maximum a posteriori criterion. | Jong-Mo Kum, Yun-Sik Park, Joon-Hyuk Chang |
| 2009 | Interspeech | Enhanced minimum statistics technique incorporating soft decision for noise suppression. | Yun-Sik Park, Ji-Hyun Song, Jae-Hun Choi, Joon-Hyuk Chang |
| 2009 | Interspeech | Soft decision-based acoustic echo suppression in a frequency domain. | Yun-Sik Park, Ji-Hyun Song, Jae-Hun Choi, Joon-Hyuk Chang |
| 2008 | Interspeech | A statistical model-based voice activity detection employing minimum classification error technique. | Sang-Ick Kang, Ji-Hyun Song, Kye-Hwan Lee, Yun-Sik Park, Joon-Hyuk Chang |
| 2008 | Interspeech | Group delay function for improved gender identification. | Kye-Hwan Lee, Sang-Ick Kang, Ji-Hyun Song, Joon-Hyuk Chang |
| 2007 | Interspeech | A uniformly most powerful test for statistical model-based voice activity detection. | Keun Won Jang, Dong Kook Kim, Joon-Hyuk Chang |
| 2007 | Interspeech | Voice activity detection based on support vector machine using effective feature vectors. | Q-Haing Jo, Yun-Sik Park, Kye-Hwan Lee, Ji-Hyun Song, Joon-Hyuk Chang |
| 2006 | Interspeech | Signal modification incorporating perceptual weighting filter. | Joon-Hyuk Chang, Woohyung Lim, Nam Soo Kim |
| 2005 | ICASSP | Voice Activity Detection based on Generalized Gamma Distribution. | Jong Won Shin, Joon-Hyuk Chang, Hwan Sik Yun, Nam Soo Kim |
| 2005 | Interspeech | A new structural preprocessor for low-bit rate speech coding. | Joon-Hyuk Chang, Jong Won Shin, Seung Yeol Lee, Nam Soo Kim |
| 2004 | Interspeech | Inner product based-multiband vector quantization for wideband speech coding at 16 kbps. | Seung Yeol Lee, Nam Soo Kim, Joon-Hyuk Chang |
| 2004 | Interspeech | Speech probability distribution based on generalized gama distribution. | Jong Won Shin, Joon-Hyuk Chang, Nam Soo Kim |
| 2003 | Interspeech | Likelihood ratio test with complex laplacian model for voice activity detection. | Joon-Hyuk Chang, Jong Won Shin, Nam Soo Kim |
| 2002 | ICASSP | Generalized analysis-by-synthesis based on system identification. | Nam Soo Kim, Joon-Hyuk Chang |
| 2000 | Interspeech | Speech enhancement: new approaches to soft decision. | Joon-Hyuk Chang, Nam Soo Kim |