Skip to content

Tan Lee

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

132

Venues

7

Active years

1995–2025

Best venue rank

A*

Where they publish

Papers

132 indexed papers, newest first.

YearVenueTitleAuthors
2025ACLPodAgent: A Comprehensive Framework for Podcast Generation.Yujia Xiao, Lei He, Haohan Guo, Fenglong Xie, Tan Lee
2024ICASSPEfficient Black-Box Speaker Verification Model Adaptation With Reprogramming And Backend Learning.Jingyu Li, Tan Lee
2024ICASSPSparsely Shared Lora on Whisper for Child Speech Recognition.Wei Liu, Ying Qin, Zhiyuan Peng, Tan Lee
2024ICASSPModeling Intrapersonal and Interpersonal Influences for Automatic Estimation of Therapist Empathy in Counseling Conversation.Dehua Tao, Tan Lee, Harold Chui, Sarah Luk
2024ICASSPCreating Personalized Synthetic Voices from Articulation Impaired Speech Using Augmented Reconstruction Loss.Yusheng Tian, Jingyu Li, Tan Lee
2024InterspeechA Parameter-efficient Language Extension Framework for Multilingual ASR.Wei Liu, Jingyong Hou, Dong Yang, Muyong Cao, Tan Lee
2024InterspeechLUPET: Incorporating Hierarchical Information Path into Multilingual ASR.Wei Liu, Jingyong Hou, Dong Yang, Muyong Cao, Tan Lee
2024InterspeechLearning Representation of Therapist Empathy in Counseling Conversation Using Siamese Hierarchical Attention Network.Dehua Tao, Tan Lee, Harold Chui, Sarah Luk
2023ASRUDiffusion-Based Mel-Spectrogram Enhancement for Personalized Speech Synthesis with Found Data.Yusheng Tian, Wei Liu, Tan Lee
2023ICASSPConvolution-Based Channel-Frequency Attention for Text-Independent Speaker Verification.Jingyu Li, Yusheng Tian, Tan Lee
2023ICASSPLeveraging Phone-Level Linguistic-Acoustic Similarity For Utterance-Level Pronunciation Scoring.Wei Liu, Kaiqi Fu, Xiaohai Tian, Shuju Shi, Wei Li, Zejun Ma, Tan Lee
2023ICASSPAn ASR-Free Fluency Scoring Approach with Self-Supervised Learning.Wei Liu, Kaiqi Fu, Xiaohai Tian, Shuju Shi, Wei Li, Zejun Ma, Tan Lee
2023ICASSPCovariance Regularization for Probabilistic Linear Discriminant Analysis.Zhiyuan Peng, Mingjie Shao, Xuanji He, Xu Li, Tan Lee, Ke Ding, Guanglu Wan
2023InterspeechModel Compression for DNN-based Speaker Verification Using Weight Quantization.Jingyu Li, Wei Liu, Zhaoyang Zhang, Jiong Wang, Tan Lee
2023InterspeechCoMFLP: Correlation Measure Based Fast Search on ASR Layer Pruning.Wei Liu, Zhiyuan Peng, Tan Lee
2023InterspeechA Study on Using Duration and Formant Features in Automatic Detection of Speech Sound Disorder in Children.Si Ioi Ng, Cymie Wing-Yee Ng, Tan Lee
2023InterspeechA Study on Prosodic Entrainment in Relation to Therapist Empathy in Counseling Conversation.Dehua Tao, Tan Lee, Harold Chui, Sarah Luk
2023InterspeechCreating Personalized Synthetic Voices from Post-Glossectomy Speech with Guided Diffusion Models.Yusheng Tian, Guangyan Zhang, Tan Lee
2023InterspeechContextSpeech: Expressive and Efficient Text-to-Speech for Paragraph Reading.Yujia Xiao, Shaofei Zhang, Xi Wang, Xu Tan, Lei He, Sheng Zhao, Frank K. Soong, Tan Lee
2022ICASSPA Study on the Efficacy of Model Pre-Training In Developing Neural Text-to-Speech System.Guangyan Zhang, Yichong Leng, Daxin Tan, Ying Qin, Kaitao Song, Xu Tan, Sheng Zhao, Tan Lee
2022InterspeechDurational Patterning at Discourse Boundaries in Relation to Therapist Empathy in Psychotherapy.Jonathan Him Nok Lee, Dehua Tao, Harold Chui, Tan Lee, Sarah Luk, Nicolette Wing Tung Lee, Koonkan Fung
2022InterspeechEDITnet: A Lightweight Network for Unsupervised Domain Adaptation in Speaker Verification.Jingyu Li, Wei Liu, Tan Lee
2022InterspeechAutomatic Detection of Speech Sound Disorder in Child Speech Using Posterior-based Speaker Representations.Si Ioi Ng, Cymie Wing-Yee Ng, Jiarui Wang, Tan Lee
2022InterspeechUnifying Cosine and PLDA Back-ends for Speaker Verification.Zhiyuan Peng, Xuanji He, Ke Ding, Tan Lee, Guanglu Wan
2022InterspeechEnvironment Aware Text-to-Speech Synthesis.Daxin Tan, Guangyan Zhang, Tan Lee
2022InterspeechCharacterizing Therapist's Speaking Style in Relation to Empathy in Psychotherapy.Dehua Tao, Tan Lee, Harold Chui, Sarah Luk
2022InterspeechHierarchical Attention Network for Evaluating Therapist Empathy in Counseling Session.Dehua Tao, Tan Lee, Harold Chui, Sarah Luk
2022InterspeechTransport-Oriented Feature Aggregation for Speaker Embedding Learning.Yusheng Tian, Jingyu Li, Tan Lee
2022InterspeechMixed-Phoneme BERT: Improving BERT with Mixed Phoneme and Sup-Phoneme Representations for Text to Speech.Guangyan Zhang, Kaitao Song, Xu Tan, Daxin Tan, Yuzi Yan, Yanqing Liu, Gang Wang, Wei Zhou, Tao Qin, Tan Lee, Sheng Zhao
2021ASRUImproving Text-Independent Speaker Verification with Auxiliary Speakers Using Graph.Jingyu Li, Si Ioi Ng, Tan Lee
2021ASRUUtterance-Level Neural Confidence Measure for End-to-End Children Speech Recognition.Wei Liu, Tan Lee
2021ASRUEditSpeech: A Text Based Speech Editing System Using Partial Inference and Bidirectional Fusion.Daxin Tan, Liqun Deng, Yu Ting Yeung, Xin Jiang, Xiao Chen, Tan Lee
2021InterspeechDetection of Consonant Errors in Disordered Speech Based on Consonant-Vowel Segment Embedding.Si Ioi Ng, Cymie Wing-Yee Ng, Jingyu Li, Tan Lee
2021InterspeechPairing Weak with Strong: Twin Models for Defending Against Adversarial Attack on Speaker Verification.Zhiyuan Peng, Xu Li, Tan Lee
2021InterspeechFine-Grained Style Modeling, Transfer and Prediction in Text-to-Speech Synthesis via Phone-Level Content-Style Disentanglement.Daxin Tan, Tan Lee
2021InterspeechApplying the Information Bottleneck Principle to Prosodic Representation Learning.Guangyan Zhang, Ying Qin, Daxin Tan, Tan Lee
2020ICASSPResting-State EEG-Based Biometrics with Signals Features Extracted by Multivariate Empirical Mode Decomposition.Matthew King-Hang Ma, Tan Lee, Manson Cheuk-Man Fong, William Shi-Yuan Wang
2020ICASSPMixture Factorized Auto-Encoder for Unsupervised Hierarchical Deep Factorization of Speech Signal.Zhiyuan Peng, Siyuan Feng, Tan Lee
2020ICASSPTime-Frequency Feature Decomposition Based on Sound Duration for Acoustic Scene Classification.Yuzhong Wu, Tan Lee
2020InterspeechText-Independent Speaker Verification with Dual Attention Network.Jingyu Li, Tan Lee
2020InterspeechAdvancing Multiple Instance Learning with Attention Modeling for Categorical Speech Emotion Recognition.Shuiyang Mao, P. C. Ching, C.-C. Jay Kuo, Tan Lee
2020InterspeechEmotion Profile Refinery for Speech Emotion Classification.Shuiyang Mao, Pak-Chung Ching, Tan Lee
2020InterspeechEigenEmo: Spectral Utterance Representation Using Dynamic Mode Decomposition for Speech Emotion Classification.Shuiyang Mao, P. C. Ching, Tan Lee
2020InterspeechAutomatic Detection of Phonological Errors in Child Speech Using Siamese Recurrent Autoencoder.Si Ioi Ng, Tan Lee
2020InterspeechCUCHILD: A Large-Scale Cantonese Corpus of Child Speech for Phonology and Articulation Assessment.Si Ioi Ng, Cymie Wing-Yee Ng, Jiarui Wang, Tan Lee, Kathy Yuet-Sheung Lee, Michael Chi-Fai Tong
2020InterspeechLearning Syllable-Level Discrete Prosodic Representation for Expressive Speech Generation.Guangyan Zhang, Ying Qin, Tan Lee
2019ICASSPRevisiting Hidden Markov Models for Speech Emotion Recognition.Shuiyang Mao, Dehua Tao, Guangyan Zhang, P. C. Ching, Tan Lee
2019ICASSPAdversarial Multi-task Deep Features and Unsupervised Back-end Adaptation for Language Recognition.Zhiyuan Peng, Siyuan Feng, Tan Lee
2019ICASSPCombining Phone Posteriorgrams from Strong and Weak Recognizers for Automatic Speech Assessment of People with Aphasia.Ying Qin, Tan Lee, Anthony Pak-Hin Kong
2019ICASSPEnhancing Sound Texture in CNN-based Acoustic Scene Classification.Yuzhong Wu, Tan Lee
2019ICASSPBLHUC: Bayesian Learning of Hidden Unit Contributions for Deep Neural Network Speaker Adaptation.Xurong Xie, Xunying Liu, Tan Lee, Shoukang Hu, Lan Wang
2019InterspeechImproving Unsupervised Subword Modeling via Disentangled Speech Representation Learning and Transformation.Siyuan Feng, Tan Lee
2019InterspeechCombining Adversarial Training and Disentangled Speech Representation for Robust Zero-Resource Subword Modeling.Siyuan Feng, Tan Lee, Zhiyuan Peng
2019InterspeechDeep Learning of Segment-Level Feature Representation with Multiple Instance Learning for Utterance-Level Speech Emotion Recognition.Shuiyang Mao, P. C. Ching, Tan Lee
2019InterspeechAutomatic Assessment of Language Impairment Based on Raw ASR Output.Ying Qin, Tan Lee, Anthony Pak-Hin Kong
2019InterspeechChild Speech Disorder Detection with Siamese Recurrent Network Using Speech Attribute Features.Jiarui Wang, Ying Qin, Zhiyuan Peng, Tan Lee
2019InterspeechFast DNN Acoustic Model Speaker Adaptation by Learning Hidden Unit Contribution Features.Xurong Xie, Xunying Liu, Tan Lee, Lan Wang
2018ICASSPAutomatic Speech Assessment for Aphasic Patients Based on Syllable-Level Embedding and Supra-Segmental Duration Features.Ying Qin, Tan Lee, Anthony Pak-Hin Kong
2018ICASSPReducing Model Complexity for DNN Based Large-Scale Audio Classification.Yuzhong Wu, Tan Lee
2018InterspeechImproving Cross-Lingual Knowledge Transferability Using Multilingual TDNN-BLSTM with Language-Dependent Pre-Final Layer.Siyuan Feng, Tan Lee
2018InterspeechExploiting Speaker and Phonetic Diversity of Mismatched Language Resources for Unsupervised Subword Modeling.Siyuan Feng, Tan Lee
2018InterspeechCross-cultural (A)symmetries in Audio-visual Attitude Perception.Hansjrg Mixdorff, Albert Rilliard, Tan Lee, Matthew K. H. Ma, Angelika Hnemann
2018InterspeechAutomatic Speech Assessment for People with Aphasia Using TDNN-BLSTM with Multi-Task Learning.Ying Qin, Tan Lee, Siyuan Feng, Anthony Pak-Hin Kong
2017ICASSPPolyphonic piano note transcription with non-negative matrix factorization of differential spectrogram.Lufei Gao, Li Su, Yi-Hsuan Yang, Tan Lee
2017ICASSPShefce: A Cantonese-English bilingual speech corpus for pronunciation assessment.Raymond W. M. Ng, Alvin C. M. Kwan, Tan Lee, Thomas Hain
2017InterspeechOn the Linguistic Relevance of Speech Units Learned by Unsupervised Acoustic Modeling.Siyuan Feng, Tan Lee
2017InterspeechAcoustic Assessment of Disordered Voice with Continuous Speech Based on Utterance-Level ASR Posterior Features.Yuanyuan Liu, Tan Lee, P. C. Ching, Thomas K. T. Law, Kathy Y. S. Lee
2017InterspeechRNN-LDA Clustering for Feature Based DNN Adaptation.Xurong Xie, Xunying Liu, Tan Lee, Lan Wang
2016ICASSPAutomatic speech recognition for acoustical analysis and assessment of cantonese pathological voice and speech.Tan Lee, Yuanyuan Liu, Pei-Wen Huang, Jen-Tzung Chien, Wang-Kong Lam, Yu Ting Yeung, Thomas K. T. Law, Kathy Y. S. Lee, Anthony Pak-Hin Kong, Sam-Po Law
2016InterspeechHybrid Accelerated Optimization for Speech Recognition.Jen-Tzung Chien, Pei-Wen Huang, Tan Lee
2016InterspeechPredicting Severity of Voice Disorder from DNN-HMM Acoustic Posteriors.Tan Lee, Yuanyuan Liu, Yu Ting Yeung, Thomas K. T. Law, Kathy Y. S. Lee
2015InterspeechModeling temporal dependency for robust estimation of LP model parameters in speech enhancement.Chun Hoy Wong, Tan Lee, Yu Ting Yeung, Pak-Chung Ching
2015MMSPMulti-pitch estimation based on sparse representation with pre-screened dictionary.Lufei Gao, Tan Lee
2014InterspeechA graph-based Gaussian component clustering approach to unsupervised acoustic modeling.Haipeng Wang, Tan Lee, Cheung-Chi Leung, Bin Ma, Haizhou Li
2014InterspeechLarge-margin conditional random fields for single-microphone speech separation.Yu Ting Yeung, Tan Lee, Cheung-Chi Leung
2014ISMCorrecting Chord Classification Errors Based on Tonal Organization Information of Classical Music.Wang-Kong Lam, Tan Lee
2013ICASSPEvaluation of pitch estimation algorithms on separated speech.Feng Huang, Yu Ting Yeung, Tan Lee
2013ICASSPUsing parallel tokenizers with DTW matrix combination for low-resource spoken term detection.Haipeng Wang, Tan Lee, Cheung-Chi Leung, Bin Ma, Haizhou Li
2013ICASSPUsing dynamic conditional random field on single-microphone speech separation.Yu Ting Yeung, Tan Lee, Cheung-Chi Leung
2013InterspeechUnsupervised mining of acoustic subword units with segment-level Gaussian posteriorgrams.Haipeng Wang, Tan Lee, Cheung-Chi Leung, Bin Ma, Haizhou Li
2012ICASSPSparsity-based confidence measure for pitch estimation in noisy speech.Feng Huang, Tan Lee
2012ICASSPTransform-domain Wiener filter for speech periodicity enhancement.Feng Huang, Tan Lee, W. Bastiaan Kleijn
2012ICASSPAn acoustic segment modeling approach to query-by-example spoken term detection.Haipeng Wang, Cheung-Chi Leung, Tan Lee, Bin Ma, Haizhou Li
2012ICASSPIntegrating multiple observations for model-based single-microphone speech separation with conditional random fields.Yu Ting Yeung, Tan Lee, Cheung-Chi Leung
2012InterspeechRobust Pitch Estimation Using l1-regularized Maximum Likelihood Estimation.Feng Huang, Tan Lee
2011ICASSPScore fusion and calibration in multiple language detectors with large performance variation.Raymond W. M. Ng, Cheung-Chi Leung, Tan Lee, Bin Ma, Haizhou Li
2010ICASSPProsodic attribute model for spoken language identification.Raymond W. M. Ng, Cheung-Chi Leung, Tan Lee, Bin Ma, Haizhou Li
2010InterspeechCross-lingual speaker adaptation via Gaussian component mapping.Houwei Cao, Tan Lee, P. C. Ching
2010InterspeechPitch estimation in noisy speech based on temporal accumulation of spectrum peaks.Feng Huang, Tan Lee
2010InterspeechPerception-based automatic approximation of F0 contours in Cantonese speech.Yujia Li, Tan Lee
2010InterspeechTowards long-range prosodic attribute modeling for language recognition.Raymond W. M. Ng, Cheung-Chi Leung, Ville Hautamki, Tan Lee, Bin Ma, Haizhou Li
2010InterspeechExploitation of phase information for speaker recognition.Ning Wang, P. C. Ching, Tan Lee
2009InterspeechEffects of language mixing for automatic recognition of Cantonese-English code-mixing utterances.Houwei Cao, P. C. Ching, Tan Lee
2009InterspeechModel-based speech separation: identifying transcription using orthogonality.Siu Wa Lee, Frank K. Soong, Tan Lee
2009InterspeechExploration of vocal excitation modulation features for speaker recognition.Ning Wang, P. C. Ching, Tan Lee
2008InterspeechLanguage modeling for speech recognition of spoken Cantonese.Yu Ting Yeung, Houwei Cao, Nengheng Zheng, Tan Lee, P. C. Ching
2008InterspeechProsody for Mandarin speech recognition: a comparative study of read and spontaneous speech.Yu Ting Yeung, Yao Qian, Tan Lee, Frank K. Soong
2007InterspeechModeling tones in hakka on the basis of the command-response model.Wentao Gu, Rerrario Shui-Ching Ho, Tan Lee
2007InterspeechPerceptual equivalence of approximated Cantonese tone contours.Yujia Li, Tan Lee
2006ICASSPUse of Vocal Source Features in Speaker Segmentation.Wai Nang Chan, Tan Lee, Nengheng Zheng, Hua Ouyang
2006ICASSPFeature Extraction From Talking Mouths for Video-Based Bi-Modal Speaker Verification.Hua Ouyang, Tan Lee, Wai Nang Chan
2006ICASSPTone-Enhanced Generalized Character Posterior Probability (GCPP) for Cantonese LVCSR.Yao Qian, Frank K. Soong, Tan Lee
2006InterspeechAutomatic speech recognition of Cantonese-English code-mixing utterances.Joyce Y. C. Chan, P. C. Ching, Tan Lee, Houwei Cao
2006InterspeechImproved tone modeling for Mandarin broadcast news speech recognition.Xin Lei, Man-Hung Siu, Mei-Yuh Hwang, Mari Ostendorf, Tan Lee
2006InterspeechTowards automatic parameter extraction of command-response model for Cantonese.Raymond W. M. Ng, Tan Lee, Wentao Gu
2005ICASSPStatic and Dynamic Spectral Features: Their Noise Robustness and Optimal Weights for ASR.Chen Yang, Frank K. Soong, Tan Lee
2005InterspeechDevelopment of a Cantonese-English code-mixing speech corpus.Joyce Y. C. Chan, P. C. Ching, Tan Lee
2004InterspeechTone information as a confidence measure for improving Cantonese LVCSR.Yao Qian, Tan Lee, Frank K. Soong
2004InterspeechTime -frequency analysis of vocal source signal for speaker recognition.Nengheng Zheng, P. C. Ching, Tan Lee
2004InterspeechExplicit duration modeling for Cantonese connected-digit recognition.Yu Zhu, Tan Lee
2004ISCASNoise-robust automatic speech recognition using mainlobe-resilient time-frequency quantile-based noise estimation.Siu Wa Lee, Pak-Chung Ching, Tan Lee
2003InterspeechModeling Cantonese pronunciation variation by acoustic model refinement.Patgi Kam, Tan Lee, Frank K. Soong
2003InterspeechOverlapped di-tone modeling for tone recognition in continuous Cantonese speech.Yao Qian, Tan Lee, Yujia Li
2003ISCASAn HMM-based speech recognition IC.Wei Han, Kwok-Wai Hon, Cheong-Fat Chan, Tan Lee, Chiu-sing Choy, Kong-Pang Pun, Pak-Chung Ching
2002InterspeechUnsupervised n-best based model adaptation using model-level confidence measures.Ka-Yan Kwan, Tan Lee, Chen Yang
2002InterspeechModeling tones in continuous Cantonese speech.Tan Lee, Greg Kochanski, Chilin Shih, Yujia Li
2001InterspeechCantonese text-to-speech synthesis using sub-syllable units.Ka Man Law, Tan Lee, Wai H. Lau
2001InterspeechISIS: a learning system with combined interaction and delegation dialogs.Helen M. Meng, Shuk Fong Chan, Yee Fong Wong, Cheong Chat Chan, Yiu Wing Wong, Tien Ying Fung, Wai Ching Tsui, Ke Chen, Lan Wang, Ting-Yao Wu, Xiaolong Li, Tan Lee, Wing Nin Choi, P. C. Ching, Huisheng Chi
2000ICASSPAcoustic modeling for Chinese speech recognition: a comparative study of Mandarin and Cantonese.Sheng Gao, Tan Lee, Yiu Wing Wong, Bo Xu, Pak-Chung Ching, Taiyi Huang
2000InterspeechLexical tree decoding with a class-based language model for Chinese speech recognition.Wing Nin Choi, Yiu Wing Wong, Tan Lee, P. C. Ching
2000InterspeechIncorporating tone information into Cantonese large-vocabulary continuous speech recognition.Wai H. Lau, Tan Lee, Yiu Wing Wong, P. C. Ching
2000InterspeechUsing cross-syllable units for Cantonese speech synthesis.Ka Man Law, Tan Lee
2000InterspeechISIS: A multilingual spoken dialog system developed with CORBA and KQML agents.Helen M. Meng, Shuk Fong Chan, Yee Fong Wong, Tien Ying Fung, Wai Ching Tsui, Tin Hang Lo, Cheong Chat Chan, Ke Chen, Lan Wang, Ting-Yao Wu, Xiaolong Li, Tan Lee, Wing Nin Choi, Yiu Wing Wong, P. C. Ching, Huisheng Chi
1999ICASSPTwo-dimensional multi-resolution analysis of speech signals and its application to speech recognition.Chun-Ping Chan, Yiu Wing Wong, Tan Lee, Pak-Chung Ching
1999InterspeechMicro-prosodic control in cantonese text-to-speech synthesis.Tan Lee, Helen M. Meng, Wai H. Lau, Wai Kit Lo, P. C. Ching
1999InterspeechAcoustic modeling and language modeling for cantonese LVCSR.Yiu Wing Wong, Ka-Fai Chow, Wai H. Lau, Wai Kit Lo, Tan Lee, Pak-Chung Ching
1998InterspeechContext-dependent duration modelling for continuous speech recognition.Tan Lee, Rolf Carlson, Bjrn Granstrm
1997ICASSPDevelopment of a large vocabulary speech database for Cantonese.Pak-Chung Ching, Ka-Fai Chow, Tan Lee, Alfred Ying Pang Ng, Lai-Wan Chan
1997ICASSPA neural network based speech recognition system for isolated Cantonese syllables.Tan Lee, Pak-Chung Ching
1996InterspeechOn improving discrimination capability of an RNN based recognizer.Tan Lee, P. C. Ching
1995ICASSPRecurrent neural networks for speech modeling and speech recognition.Tan Lee, Pak-Chung Ching, Lai-Wan Chan
1995InterspeechAn RNN based speech recognition system with discriminative training.Tan Lee, P. C. Ching, Lai-Wan Chan