Skip to content

Joon-Hyuk Chang

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

98

Venues

8

Active years

2000–2026

Best venue rank

A*

Where they publish

Papers

98 indexed papers, newest first.

YearVenueTitleAuthors
2026AAAIOnEDIT: Online Editing with Decoupled Implicit Task for Large Language Models.Chae-Won Lee, Jae-Hong Lee, Ji-Hun Kang, Joon-Hyuk Chang
2025ASRUA Momentum-Based Framework with Contrastive Data Generation for Robust Sound Source Localization.Hyun-Soo Kim, Da-Hee Yang, Joon-Hyuk Chang
2025ASRUImproving Noise Robust Audio-Visual Speech Recognition via Router-Gated Cross-Modal Feature Fusion.DongHoon Lim, YoungChae Kim, Dong-Hyun Kim, Da-Hee Yang, Joon-Hyuk Chang
2025ICASSPDiffusion-based Target Device Style Transfer for Robust Acoustic Scene Classification.Won-Gook Choi, Joon-Hyuk Chang
2025ICASSPTrainable Adaptive Score Normalization for Automatic Speaker Verification.Jeong-Hwan Choi, Ju-Seok Seong, Ye-Rin Jeoung, Joon-Hyuk Chang
2025ICASSPMultimodal Emotion Recognition with Target Speaker-Based Facial Embeddings.Serin Heo, Jehyun Kyung, Joon-Hyuk Chang
2025ICASSPProgressive Subband Modeling for Artifacts-free Speech Super-resolution.Donghyun Kim, Joon-Hyuk Chang
2025ICASSPFew-shot Keyword-incremental Learning Using Compositional Information.Ilseok Kim, Ju-Seok Seong, Joon-Hyuk Chang
2025ICASSPQuad-Net: Melspectrogram Vocoder with Convolutional Layers Restricted by the Quadrature Mirror Filter for Perfect Reconstruction.Nam-Seok Song, Joon-Hyuk Chang
2025InterspeechOptimizing CLAP Reward with LLM Feedback for Semantically Aligned and Diverse Automated Audio Captioning.Seyun Ahn, Pil Moo Byun, Won-Gook Choi, Joon-Hyuk Chang
2025InterspeechTemp4Cap: Temporally-aligned Automated Audio Captioning.Ho-Young Choi, Jae-Heung Cho, Pil Moo Byun, Won-Gook Choi, Joon-Hyuk Chang
2025InterspeechSpatially Weighted Contrastive Learning for Robust Sound Source Localization.Hyun-Soo Kim, Da-Hee Yang, Joon-Hyuk Chang
2025InterspeechImproving Generalization of End-to-End ASR through Diversity and Independence Regularization.Ye-Eun Ko, Mun-Hak Lee, Dong-Hyun Kim, Joon-Hyuk Chang
2025InterspeechEnhancing Target-speaker Automatic Speech Recognition Using Multiple Speaker Embedding Extractors with Virtual Speaker Embedding.Ju-Seok Seong, Jeong-Hwan Choi, Ye-Rin Jeoung, Ilseok Kim, Joon-Hyuk Chang
2024ICASSPGeneralized Specaugment via Multi-Rectangle Inverse Masking For Acoustic Scene Classification.Pil Moo Byun, Joon-Hyuk Chang
2024ICASSPAdversarial Learning on Compressed Posterior Space for Non-Iterative Score-based End-to-End Text-to-Speech.Won-Gook Choi, Donghyun Seong, Joon-Hyuk Chang
2024ICASSPImproving Target Sound Extraction with Timestamp Knowledge Distillation.Dail Kim, Min-Sang Baek, Yungyeo Kim, Joon-Hyuk Chang
2024ICASSPClass: Continual Learning Approach for Speech Super-Resolution.Donghyun Kim, Yungyeo Kim, Joon-Hyuk Chang
2024ICASSPText-Only Unsupervised Domain Adaptation for Neural Transducer-Based ASR Personalization Using Synthesized Data.Dong-Hyun Kim, Jae-Hong Lee, Joon-Hyuk Chang
2024ICLRContinual Momentum Filtering on Parameter Space for Online Test-time Adaptation.Jae-Hong Lee, Joon-Hyuk Chang
2024ICMLStationary Latent Weight Inference for Unreliable Observations from Online Test-Time Adaptation.Jae-Hong Lee, Joon-Hyuk Chang
2024InterspeechRetrieval-Augmented Classifier Guidance for Audio Generation.Ho-Young Choi, Won-Gook Choi, Joon-Hyuk Chang
2024InterspeechEfficient Speaker Embedding Extraction Using a Twofold Sliding Window Algorithm for Speaker Diarization.Jeong-Hwan Choi, Ye-Rin Jeoung, Ilseok Kim, Joon-Hyuk Chang
2024InterspeechWhisper Multilingual Downstream Task Tuning Using Task Vectors.Ji-Hun Kang, Jae-Hong Lee, Mun-Hak Lee, Joon-Hyuk Chang
2024InterspeechMitigating Overfitting in Structured Pruning of ASR Models with Gradient-Guided Parameter Regularization.Dong-Hyun Kim, Joon-Hyuk Chang
2024InterspeechSound of Vision: Audio Generation from Visual Text Embedding through Training Domain Discriminator.Jaewon Kim, Won-Gook Choi, Seyun Ahn, Joon-Hyuk Chang
2024InterspeechFew-Shot Keyword-Incremental Learning with Total Calibration.Ilseok Kim, Ju-Seok Seong, Joon-Hyuk Chang
2024InterspeechGuided conditioning with predictive network on score-based diffusion model for speech enhancement.Dail Kim, Da-Hee Yang, Donghyun Kim, Joon-Hyuk Chang, Jeonghwan Choi, Moa Lee, Jaemo Yang, Han-gil Moon
2024InterspeechEnhancing Multimodal Emotion Recognition through ASR Error Compensation and LLM Fine-Tuning.Jehyun Kyung, Serin Heo, Joon-Hyuk Chang
2024InterspeechNeural ATSM: Fully Neural Network-based Adaptive Time-Scale Modification Using Sentence-Specific Dynamic Control.Jaeuk Lee, Sohee Jang, Joon-Hyuk Chang
2024InterspeechOnline Subloop Search via Uncertainty Quantization for Efficient Test-Time Adaptation.Jae-Hong Lee, Sang-Eon Lee, Dong-Hyun Kim, Do-Hee Kim, Joon-Hyuk Chang
2024InterspeechBalanced-Wav2Vec: Enhancing Stability and Robustness of Representation Learning Through Sample Reweighting Techniques.Mun-Hak Lee, Jae-Hong Lee, Do-Hee Kim, Ye-Eun Ko, Joon-Hyuk Chang
2024InterspeechH4C-TTS: Leveraging Multi-Modal Historical Context for Conversational Text-to-Speech.Donghyun Seong, Joon-Hyuk Chang
2024InterspeechTSP-TTS: Text-based Style Predictor with Residual Vector Quantization for Expressive Text-to-Speech.Donghyun Seong, Hoyoung Lee, Joon-Hyuk Chang
2023ASRUExtending Self-Distilled Self-Supervised Learning For Semi-Supervised Speaker Verification.Jeong-Hwan Choi, Jehyun Kyung, Ju-Seok Seong, Ye-Rin Jeoung, Joon-Hyuk Chang
2023ASRUAWMC: Online Test-Time Adaptation Without Mode Collapse for Continual Adaptation.Jae-Hong Lee, Do-Hee Kim, Joon-Hyuk Chang
2023ASRUCross-Modal Learning for CTC-Based ASR: Leveraging CTC-Bertscore and Sequence-Level Training.Mun-Hak Lee, Sang-Eon Lee, Ji-Eun Choi, Joon-Hyuk Chang
2023ASRUKnowledge Distillation From Offline to Streaming Transducer: Towards Accurate and Fast Streaming Model by Matching Alignments.Ji-Hwan Mo, Jae-Jin Jeon, Mun-Hak Lee, Joon-Hyuk Chang
2023ASRUTowards Robust Packet Loss Concealment System With ASR-Guided Representations.Da-Hee Yang, Joon-Hyuk Chang
2023ICASSPCAN2V: Can-Bus Data-Based Seq2seq Model for Vehicle Velocity Prediction.Jae-Heung Cho, Joon-Hyuk Chang
2023ICASSPM-CTRL: A Continual Representation Learning Framework with Slowly Improving Past Pre-Trained Model.Jin-Seong Choi, Jae-Hong Lee, Chae-Won Lee, Joon-Hyuk Chang
2023ICASSPAdaptive Time-Scale Modification for Improving Speech Intelligibility Based On Phoneme Clustering For Streaming Services.Sohee Jang, Jiye Kim, Yeon-Ju Kim, Joon-Hyuk Chang
2023ICASSPImproving Transformer-Based End-to-End Speaker Diarization by Assigning Auxiliary Losses to Attention Heads.Ye-Rin Jeoung, Joon-Young Yang, Jeong-Hwan Choi, Joon-Hyuk Chang
2023ICASSPRepackagingaugment: Overcoming Prediction Error Amplification in Weight-Averaged Speech Recognition Models Subject to Self-Training.Jae-Hong Lee, Dong-Hyun Kim, Joon-Hyuk Chang
2023ICASSPNoise-Aware Target Extension with Self-Distillation for Robust Speech Recognition.Ju-Seok Seong, Jeong-Hwan Choi, Jehyun Kyung, Ye-Rin Jeoung, Joon-Hyuk Chang
2023ICASSPSelective Film Conditioning with CTC-Based ASR Probability for Speech Enhancement.Da-Hee Yang, Joon-Hyuk Chang
2023InterspeechDeeply Supervised Curriculum Learning for Deep Neural Network-based Sound Source Localization.Min-Sang Baek, Joon-Young Yang, Joon-Hyuk Chang
2023InterspeechSR-SRP: Super-Resolution based SRP-PHAT for Sound Source Localization and Tracking.Jae-Heung Cho, Joon-Hyuk Chang
2023InterspeechResolution Consistency Training on Time-Frequency Domain for Semi-Supervised Sound Event Detection.Won-Gook Choi, Joon-Hyuk Chang
2023InterspeechPrior-free Guided TTS: An Improved and Efficient Diffusion-based Text-Guided Speech Synthesis.Won-Gook Choi, So-Jeong Kim, Tae-Ho Kim, Joon-Hyuk Chang
2023InterspeechSelf-Distillation into Self-Attention Heads for Improving Transformer-based End-to-End Neural Speaker Diarization.Ye-Rin Jeoung, Jeong-Hwan Choi, Ju-Seok Seong, Jehyun Kyung, Joon-Hyuk Chang
2023InterspeechIntra-ensemble: A New Method for Combining Intermediate Outputs in Transformer-based Automatic Speech Recognition.Do-Hee Kim, Ji-Eun Choi, Joon-Hyuk Chang
2023InterspeechGeneral-purpose Adversarial Training for Enhanced Automatic Speech Recognition Model Generalization.Do-Hee Kim, Daeyeol Shim, Joon-Hyuk Chang
2023InterspeechImproving Joint Speech and Emotion Recognition Using Global Style Tokens.Jehyun Kyung, Ju-Seok Seong, Jeong-Hwan Choi, Ye-Rin Jeoung, Joon-Hyuk Chang
2023InterspeechHAD-ANC: A Hybrid System Comprising an Adaptive Filter and Deep Neural Networks for Active Noise Control.JungPhil Park, Jeong-Hwan Choi, Yungyeo Kim, Joon-Hyuk Chang
2022ICASSPKnowledge Distillation from Language Model to Acoustic Model: A Hierarchical Multi-Task Learning Approach.Mun-Hak Lee, Joon-Hyuk Chang
2022InterspeechConvolutional Recurrent Neural Network with Auxiliary Stream for Robust Variable-Length Acoustic Scene Classification.Joon-Hyuk Chang, Won-Gook Choi
2022InterspeechImproved CNN-Transformer using Broadcasted Residual Learning for Text-Independent Speaker Verification.Jeong-Hwan Choi, Joon-Young Yang, Ye-Rin Jeoung, Joon-Hyuk Chang
2022InterspeechHYU Submission for the SASV Challenge 2022: Reforming Speaker Embeddings with Spoofing-Aware Conditioning.Jeong-Hwan Choi, Joon-Young Yang, Ye-Rin Jeoung, Joon-Hyuk Chang
2022InterspeechAdversarial and Sequential Training for Cross-lingual Prosody Transfer TTS.Min-Kyung Kim, Joon-Hyuk Chang
2022InterspeechW2V2-Light: A Lightweight Version of Wav2vec 2.0 for Automatic Speech Recognition.Dong-Hyun Kim, Jae-Hong Lee, Ji-Hwan Mo, Joon-Hyuk Chang
2022InterspeechOne-Shot Speaker Adaptation Based on Initialization by Generative Adversarial Networks for TTS.Jaeuk Lee, Joon-Hyuk Chang
2022InterspeechAdvanced Speaker Embedding with Predictive Variance of Gaussian Distribution for Speaker Adaptation in TTS.Jaeuk Lee, Joon-Hyuk Chang
2022InterspeechRegularizing Transformer-based Acoustic Models by Penalizing Attention Weights.Mun-Hak Lee, Joon-Hyuk Chang, Sang-Eon Lee, Ju-Seok Seong, Chanhee Park, Haeyoung Kwon
2022InterspeechCTRL: Continual Representation Learning to Transfer Information of Pre-trained for WAV2VEC 2.0.Jae-Hong Lee, Chae Won Lee, Jin-Seong Choi, Joon-Hyuk Chang, Woo Kyeong Seong, Jeonghan Lee
2022InterspeechFiLM Conditioning with Enhanced Feature to the Transformer-based End-to-End Noisy Speech Recognition.Da-Hee Yang, Joon-Hyuk Chang
2021ASRUShort-Utterance Embedding Enhancement Method Based on Time Series Forecasting Technique for Text-Independent Speaker Verification.Jeong-Hwan Choi, Joon-Young Yang, Joon-Hyuk Chang
2021InterspeechDeep Neural Network Calibration for E2E Speech Recognition System.Mun-Hak Lee, Joon-Hyuk Chang
2021ISCASMIMO Noise Suppression Preserving Spatial Cues for Sound Source Localization in Mobile Robot.Jung-Hee Kim, Jeong-Hwan Choi, Jinyoung Son, Gyeong-Su Kim, Jihwan Park, Joon-Hyuk Chang
2020InterspeechAttention Wave-U-Net for Acoustic Echo Cancellation.Jung-Hee Kim, Joon-Hyuk Chang
2020InterspeechVirtual Acoustic Channel Expansion Based on Neural Networks for Weighted Prediction Error-Based Speech Dereverberation.Joon-Young Yang, Joon-Hyuk Chang
2019InterspeechJoint Optimization of Neural Acoustic Beamforming and Dereverberation with x-Vectors for Robust Speaker Verification.Joon-Young Yang, Joon-Hyuk Chang
2018IROSDNN-based Speech Recognition System dealing with Motor State as Auxiliary Information of DNN for Head Shaking Robot.Moa Lee, Joon-Hyuk Chang
2016ICASSPDual-microphone voice activity detection based on using optimally weighted maximum a posteriori probabilities.Seng Hyun Huang, Jihwan Park, Joon-Hyuk Chang
2015InterspeechA statistical model-based voice activity detection using multiple DNNs and noise awareness.Inyoung Hwang, Jaeseong Sim, Sang-Hyeon Kim, Kwang-Sub Song, Joon-Hyuk Chang
2014InterspeechEnhanced muting method in packet loss concealment of ITU-t g.722 using sigmoid function with on-line optimized parameters.Bong-Ki Lee, Inyoung Hwang, Jihwan Park, Joon-Hyuk Chang
2013InterspeechEnhanced muting method in packet loss concealment of ITU-t g.722 employing optimized sigmoid function.Bong-Ki Lee, Chungsoo Lim, Jihwan Park, Joon-Hyuk Chang
2012ICASSPAdaptive noise power estimation using spectral difference for robust speech enhancement.Jae-Hun Choi, Sang-Kyun Kim, Joon-Hyuk Chang
2012ICASSPNew techniques for improving the practicality of an SVM-based speech/music classifier.Chungsoo Lim, Seong-Ro Lee, Yeonwoo Lee, Joon-Hyuk Chang
2011InterspeechA Soft Decision-Based Speech Enhancement Using Acoustic Noise Classification.Jae-Hun Choi, Sang-Kyun Kim, Joon-Hyuk Chang
2010ICASSPA statistical model-based double-talk detection incorporating soft decision.Yun-Sik Park, Ji-Hyun Song, Sang-Ick Kang, Woojung Lee, Joon-Hyuk Chang
2010InterspeechToward detecting voice activity employing soft decision in second-order conditional MAP.Sang-Kyun Kim, Jae-Hun Choi, Sang-Ick Kang, Ji-Hyun Song, Joon-Hyuk Chang
2010InterspeechOn using Gaussian mixture model for double-talk detection in acoustic echo suppression.Ji-Hyun Song, Kyu-Ho Lee, Yun-Sik Park, Sang-Ick Kang, Joon-Hyuk Chang
2009ICASSPSpeech enhancement based on minima controlled recursive averaging incorporating conditional maximum a posteriori criterion.Jong-Mo Kum, Yun-Sik Park, Joon-Hyuk Chang
2009InterspeechEnhanced minimum statistics technique incorporating soft decision for noise suppression.Yun-Sik Park, Ji-Hyun Song, Jae-Hun Choi, Joon-Hyuk Chang
2009InterspeechSoft decision-based acoustic echo suppression in a frequency domain.Yun-Sik Park, Ji-Hyun Song, Jae-Hun Choi, Joon-Hyuk Chang
2008InterspeechA statistical model-based voice activity detection employing minimum classification error technique.Sang-Ick Kang, Ji-Hyun Song, Kye-Hwan Lee, Yun-Sik Park, Joon-Hyuk Chang
2008InterspeechGroup delay function for improved gender identification.Kye-Hwan Lee, Sang-Ick Kang, Ji-Hyun Song, Joon-Hyuk Chang
2007InterspeechA uniformly most powerful test for statistical model-based voice activity detection.Keun Won Jang, Dong Kook Kim, Joon-Hyuk Chang
2007InterspeechVoice activity detection based on support vector machine using effective feature vectors.Q-Haing Jo, Yun-Sik Park, Kye-Hwan Lee, Ji-Hyun Song, Joon-Hyuk Chang
2006InterspeechSignal modification incorporating perceptual weighting filter.Joon-Hyuk Chang, Woohyung Lim, Nam Soo Kim
2005ICASSPVoice Activity Detection based on Generalized Gamma Distribution.Jong Won Shin, Joon-Hyuk Chang, Hwan Sik Yun, Nam Soo Kim
2005InterspeechA new structural preprocessor for low-bit rate speech coding.Joon-Hyuk Chang, Jong Won Shin, Seung Yeol Lee, Nam Soo Kim
2004InterspeechInner product based-multiband vector quantization for wideband speech coding at 16 kbps.Seung Yeol Lee, Nam Soo Kim, Joon-Hyuk Chang
2004InterspeechSpeech probability distribution based on generalized gama distribution.Jong Won Shin, Joon-Hyuk Chang, Nam Soo Kim
2003InterspeechLikelihood ratio test with complex laplacian model for voice activity detection.Joon-Hyuk Chang, Jong Won Shin, Nam Soo Kim
2002ICASSPGeneralized analysis-by-synthesis based on system identification.Nam Soo Kim, Joon-Hyuk Chang
2000InterspeechSpeech enhancement: new approaches to soft decision.Joon-Hyuk Chang, Nam Soo Kim