Skip to content

Khe Chai Sim

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

101

Venues

10

Active years

2004–2024

Best venue rank

A*

Where they publish

Papers

101 indexed papers, newest first.

YearVenueTitleAuthors
2024ICASSPImproving Speech Recognition for African American English with Audio Classification.Shefali Garg, Zhouyuan Huo, Khe Chai Sim, Suzan Schwartz, Mason Chua, Alna Aksnova, Tsendsuren Munkhdalai, Levi King, Darryl Wright, Zion Mengesha, Dongseong Hwang, Tara N. Sainath, Franoise Beaufays, Pedro Moreno Mengibar
2024ICASSPA Comparison of Parameter-Efficient ASR Domain Adaptation Methods for Universal Speech and Language Models.Khe Chai Sim, Zhouyuan Huo, Tsendsuren Munkhdalai, Nikhil Siddhartha, Adam Stooke, Zhong Meng, Bo Li, Tara N. Sainath
2024InterspeechAdaRA: Adaptive Rank Allocation of Residual Adapters for Speech Foundation Model.Zhouyuan Huo, Dongseong Hwang, Gan Song, Khe Chai Sim, Weiran Wang
2024InterspeechContextual Biasing with the Knuth-Morris-Pratt Matching Algorithm.Weiran Wang, Zelin Wu, Diamantino Caseiro, Tsendsuren Munkhdalai, Khe Chai Sim, Pat Rondon, Golan Pundak, Gan Song, Rohit Prabhavalkar, Zhong Meng, Ding Zhao, Tara Sainath, Yanzhang He, Pedro Moreno Mengibar
2024NAACLMassive End-to-end Speech Recognition Models with Time Reduction.Weiran Wang, Rohit Prabhavalkar, Haozhe Shan, Zhong Meng, Dongseong Hwang, Qiujia Li, Khe Chai Sim, Bo Li, James Qin, Xingyu Cai, Adam Stooke, Chengjian Zheng, Yanzhang He, Tara N. Sainath, Pedro Moreno Mengibar
2023ASRUContextual Spelling Correction with Large Language Models.Gan Song, Zelin Wu, Golan Pundak, Angad Chandorkar, Kandarp Joshi, Xavier Velez, Diamantino Caseiro, Ben Haynor, Weiran Wang, Nikhil Siddhartha, Pat Rondon, Khe Chai Sim
2023ICASSPResource-Efficient Transfer Learning from Speech Foundation Model Using Hierarchical Feature Fusion.Zhouyuan Huo, Khe Chai Sim, Bo Li, Dongseong Hwang, Tara N. Sainath, Trevor Strohman
2023ICASSPComparison of Soft and Hard Target RNN-T Distillation for Large-Scale ASR.Dongseong Hwang, Khe Chai Sim, Yu Zhang, Trevor Strohman
2023ICASSPEfficient Domain Adaptation for Speech Foundation Models.Bo Li, Dongseong Hwang, Zhouyuan Huo, Junwen Bai, Guru Prakash, Tara N. Sainath, Khe Chai Sim, Yu Zhang, Wei Han, Trevor Strohman, Franoise Beaufays
2023InterspeechRe-investigating the Efficient Transfer Learning of Speech Foundation Model using Feature Fusion Methods.Zhouyuan Huo, Khe Chai Sim, Dongseong Hwang, Tsendsuren Munkhdalai, Tara N. Sainath, Pedro Moreno Mengibar
2023InterspeechDual-Mode NAM: Effective Top-K Context Injection for End-to-End ASR.Zelin Wu, Tsendsuren Munkhdalai, Pat Rondon, Golan Pundak, Khe Chai Sim, Christopher Li
2022ICASSPJoint Unsupervised and Supervised Training for Multilingual ASR.Junwen Bai, Bo Li, Yu Zhang, Ankur Bapna, Nikhil Siddhartha, Khe Chai Sim, Tara N. Sainath
2022ICASSPLarge-Scale ASR Domain Adaptation Using Self- and Semi-Supervised Learning.Dongseong Hwang, Ananya Misra, Zhouyuan Huo, Nikhil Siddhartha, Shefali Garg, David Qiu, Khe Chai Sim, Trevor Strohman, Franoise Beaufays, Yanzhang He
2022ICASSPFast Contextual Adaptation with Neural Associative Memory for On-Device Personalized Speech Recognition.Tsendsuren Munkhdalai, Khe Chai Sim, Angad Chandorkar, Fan Gao, Mason Chua, Trevor Strohman, Franoise Beaufays
2022InterspeechUserLibri: A Dataset for ASR Personalization Using Only Text.Theresa Breiner, Swaroop Ramaswamy, Ehsan Variani, Shefali Garg, Rajiv Mathews, Khe Chai Sim, Kilol Gupta, Mingqing Chen, Lara McConnaughey
2022InterspeechIncremental Layer-Wise Self-Supervised Learning for Efficient Unsupervised Speech Domain Adaptation On Device.Zhouyuan Huo, Dongseong Hwang, Khe Chai Sim, Shefali Garg, Ananya Misra, Nikhil Siddhartha, Trevor Strohman, Franoise Beaufays
2022InterspeechPseudo Label Is Better Than Human Label.Dongseong Hwang, Khe Chai Sim, Zhouyuan Huo, Trevor Strohman
2022InterspeechOn-the-fly ASR Corrections with Audio Exemplars.Golan Pundak, Tsendsuren Munkhdalai, Khe Chai Sim
2021InterspeechA Comparison of Supervised and Unsupervised Pre-Training of End-to-End Models.Ananya Misra, Dongseong Hwang, Zhouyuan Huo, Shefali Garg, Nikhil Siddhartha, Arun Narayanan, Khe Chai Sim
2021InterspeechRobust Continuous On-Device Personalization for Automatic Speech Recognition.Khe Chai Sim, Angad Chandorkar, Fan Gao, Mason Chua, Tsendsuren Munkhdalai, Franoise Beaufays
2020ICASSPLow-Rank Gradient Approximation for Memory-Efficient on-Device Training of Deep Neural Network.Mary Gooneratne, Khe Chai Sim, Petr Zadrazil, Andreas Kabel, Franoise Beaufays, Giovanni Motta
2019ASRUPersonalization of End-to-End Speech Recognition on Mobile Devices for Named Entities.Khe Chai Sim, Leif Johnson, Giovanni Motta, Lillian Zhou, Franoise Beaufays, Arnaud Benard, Dhruv Guliani, Andreas Kabel, Nikhil Khare, Tamar Lucassen, Petr Zadrazil, Harry Zhang
2019ICASSPStreaming End-to-end Speech Recognition for Mobile Devices.Yanzhang He, Tara N. Sainath, Rohit Prabhavalkar, Ian McGraw, Raziel Alvarez, Ding Zhao, David Rybach, Anjuli Kannan, Yonghui Wu, Ruoming Pang, Qiao Liang, Deepti Bhatia, Yuan Shangguan, Bo Li, Golan Pundak, Khe Chai Sim, Tom Bagby, Shuo-Yiin Chang, Kanishka Rao, Alexander Gruenstein
2019ICASSPImproving CTC Using Stimulated Learning for Sequence Modeling.Jahn Heymann, Khe Chai Sim, Bo Li
2019InterspeechAn Investigation into On-Device Personalization of End-to-End Automatic Speech Recognition Models.Khe Chai Sim, Petr Zadrazil, Franoise Beaufays
2018ICASSPUnderstanding Recurrent Neural State Using Memory Signatures.Skanda Koppula, Khe Chai Sim, Kean K. Chin
2018ICASSPMulti-Dialect Speech Recognition with a Single Sequence-to-Sequence Model.Bo Li, Tara N. Sainath, Khe Chai Sim, Michiel Bacchiani, Eugene Weinstein, Patrick Nguyen, Zhifeng Chen, Yanghui Wu, Kanishka Rao
2018ICASSPlearning Effective Factorized Hidden Layer Bases Using Student-Teacher Training for LSTM Acoustic Model Adaptation.Lahiru Samarakoon, Brian Mak, Khe Chai Sim
2018InterspeechDomain Adaptation Using Factorized Hidden Layer for Robust Automatic Speech Recognition.Khe Chai Sim, Arun Narayanan, Ananya Misra, Anshuman Tripathi, Golan Pundak, Tara N. Sainath, Parisa Haghani, Bo Li, Michiel Bacchiani
2017ASRUImproving the efficiency of forward-backward algorithm using batched computation in TensorFlow.Khe Chai Sim, Arun Narayanan, Tom Bagby, Tara N. Sainath, Michiel Bacchiani
2017ICASSPAn investigation into learning effective speaker subspaces for robust unsupervised DNN adaptation.Lahiru Samarakoon, Khe Chai Sim, Brian Mak
2017InterspeechAcoustic Modeling for Google Home.Bo Li, Tara N. Sainath, Arun Narayanan, Joe Caroselli, Michiel Bacchiani, Ananya Misra, Izhak Shafran, Hasim Sak, Golan Pundak, Kean K. Chin, Khe Chai Sim, Ron J. Weiss, Kevin W. Wilson, Ehsan Variani, Chanwoo Kim, Olivier Siohan, Mitchel Weintraub, Erik McDermott, Richard Rose, Matt Shannon
2017InterspeechLearning Factorized Transforms for Unsupervised Adaptation of LSTM-RNN Acoustic Models.Lahiru Samarakoon, Brian Mak, Khe Chai Sim
2017InterspeechAn Efficient Phone N-Gram Forward-Backward Computation Using Dense Matrix Multiplication.Khe Chai Sim, Arun Narayanan
2016ICASSPJoint acoustic factor learning for robust deep neural network based automatic speech recognition.Souvik Kundu, Gautam Mantena, Yanmin Qian, Tian Tan, Marc Delcroix, Khe Chai Sim
2016ICASSPOn combining i-vectors and discriminative adaptation methods for unsupervised speaker normalization in DNN acoustic models.Lahiru Samarakoon, Khe Chai Sim
2016ICASSPSpeaker-aware training of LSTM-RNNS for acoustic modelling.Tian Tan, Yanmin Qian, Dong Yu, Souvik Kundu, Liang Lu, Khe Chai Sim, Xiong Xiao, Yu Zhang
2016ICASSPTowards implicit complexity control using variable-depth deep neural networks for automatic speech recognition.Shawn Tan, Khe Chai Sim
2016InterspeechIncorporating a Generative Front-End Layer to Deep Neural Network for Noise Robust Automatic Speech Recognition.Souvik Kundu, Khe Chai Sim, Mark J. F. Gales
2016InterspeechMicrophone Distance Adaptation Using Cluster Adaptive Training for Robust Far Field Speech Recognition.Animesh Prasad, Khe Chai Sim
2016InterspeechSubspace LHUC for Fast Adaptation of Deep Neural Network Acoustic Models.Lahiru Samarakoon, Khe Chai Sim
2016InterspeechMulti-Attribute Factorized Hidden Layer Adaptation for DNN Acoustic Models.Lahiru Samarakoon, Khe Chai Sim
2016InterspeechStimulated Deep Neural Network for Speech Recognition.Chunyang Wu, Penny Karanasou, Mark J. F. Gales, Khe Chai Sim
2015ASRULearning factorized feature transforms for speaker normalization.Lahiru Samarakoon, Khe Chai Sim
2015ASRUOn constructing and analysing an interpretable brain model for the DNN based on hidden activity patterns.Khe Chai Sim
2015ASRUImproving the interpretability of deep neural networks with stimulated learning.Shawn Tan, Khe Chai Sim, Mark J. F. Gales
2015ICASSPAn investigation of augmenting speaker representations to improve speaker normalisation for DNN-based speech recognition.Hengguan Huang, Khe Chai Sim
2014COLINGA Beam-Search Decoder for Disfluency Detection.Xuancong Wang, Hwee Tou Ng, Khe Chai Sim
2014EMNLPCombining Punctuation and Disfluency Prediction: An Empirical Study.Xuancong Wang, Khe Chai Sim, Hwee Tou Ng
2014ICASSPSecond order vector taylor series based robust speech recognition.Suliang Bu, Yanmin Qian, Khe Chai Sim, Yongbin You, Kai Yu
2014ICASSPAn ideal hidden-activation mask for deep neural networks based noise-robust speech recognition.Bo Li, Khe Chai Sim
2014ICASSPOn combining DNN and GMM with unsupervised speaker adaptation for robust automatic speech recognition.Shilin Liu, Khe Chai Sim
2014ICASSPRefinements of regression-based context-dependent modelling of deep neural networks for automatic speech recognition.Guangsen Wang, Khe Chai Sim
2014InterspeechModeling long temporal contexts for robust DNN-based speech recognition.Bo Li, Khe Chai Sim
2014InterspeechJoint adaptation and adaptive training of TVWR for robust automatic speech recognition.Shilin Liu, Khe Chai Sim
2013ASRUImproving robustness of deep neural networks via spectral masking for automatic speech recognition.Bo Li, Khe Chai Sim
2013ASRUMulti-stream temporally varying weight regression for cross-lingual speech recognition.Shilin Liu, Khe Chai Sim
2013ASRUContext-dependent modelling of deep neural network using logistic regression.Guangsen Wang, Khe Chai Sim
2013ICASSPNoise adaptive front-end normalization based on Vector Taylor Series for Deep Neural Networks in robust speech recognition.Bo Li, Khe Chai Sim
2013ICASSPApproximated Parallel Model Combination for efficient noise-robust speech recognition.Khe Chai Sim
2013InterspeechAn investigation of spectral restoration algorithms for deep neural networks based noise robust speech recognition.Bo Li, Yu Tsao, Khe Chai Sim
2013InterspeechParameter clustering for temporally varying weight regression for automatic speech recognition.Shilin Liu, Khe Chai Sim
2013InterspeechAn investigation of temporally varying weight regression for noise robust speech recognition.Shilin Liu, Khe Chai Sim
2013InterspeechIntegrating conditional random fields and joint multi-gram model with syllabic features for grapheme-to-phone conversion.Xiaoxuan Wang, Khe Chai Sim
2012ACLProbabilistic Integration of Partial Lexical Information for Noise Robust Haptic Voice Recognition.Khe Chai Sim
2012ICASSPImplicit trajectory modelling using temporally varying weight regression for automatic speech recognition.Shilin Liu, Khe Chai Sim
2012ICASSPAn investigation of tied-mixture GMM based triphone state clustering.Guangsen Wang, Khe Chai Sim
2012ICMIDesign and implementation of the note-taking style haptic voice recognition for mobile devices.Seungwhan Moon, Khe Chai Sim
2012ICMISpeak-as-you-swipe (SAYS): a multimodal interface combining speech and gesture keyboard synchronously for continuous mobile text entry.Khe Chai Sim
2012ICMIICMI'12 grand challenge: haptic voice recognition.Khe Chai Sim, Shengdong Zhao, Kai Yu, Hank Liao
2012ICMIImproving mandarin predictive text input by augmenting pinyin initials with speech and tonal information.Guangsen Wang, Bo Li, Shilin Liu, Xuancong Wang, Xiaoxuan Wang, Khe Chai Sim
2012InterspeechA Weighted Combination of Speech with Text-based Models for Arabic Diacritization.Aisha S. Azim, Xiaoxuan Wang, Khe Chai Sim
2012InterspeechA Two-stage Speaker Adaptation Approach for Subspace Gaussian Mixture Model based Nonnative Speech Recognition.Bo Li, Khe Chai Sim
2012InterspeechDynamic Conditional Random Fields for Joint Sentence Boundary and Punctuation Prediction.Xuancong Wang, Hwee Tou Ng, Khe Chai Sim
2011ASRUA Trajectory-based Parallel Model Combination with a unified static and dynamic parameter compensation for noisy speech recognition.Khe Chai Sim, Minh-Thang Luong
2011InterspeechSequential Classification Criteria for NNs in Automatic Speech Recognition.Guangsen Wang, Khe Chai Sim
2011InterspeechComparison of Smoothing Techniques for Robust Context Dependent Acoustic Modelling in Hybrid NN/HMM Systems.Guangsen Wang, Khe Chai Sim
2010ICASSPA minimum variance asynchronous Detection Error Trade-off performance analysis for multi-class detection problems.Khe Chai Sim
2010ICASSPAdaptive score fusion using Weighted Logistic Linear Regression for spoken language recognition.Khe Chai Sim, Kong-Aik Lee
2010InterspeechComparison of discriminative input and output transformations for speaker adaptation in the hybrid NN/HMM systems.Bo Li, Khe Chai Sim
2010InterspeechHidden logistic linear regression for support vector machine based phone verification.Bo Li, Khe Chai Sim
2010InterspeechProbabilistic state clustering using conditional random field for context-dependent acoustic modelling.Khe Chai Sim
2010InterspeechSemi-parametric trajectory modelling using temporally varying feature mapping for speech recognition.Khe Chai Sim, Shilin Liu
2009ASRUDiscriminative Product-of-Expert acoustic mapping for cross-lingual phone recognition.Khe Chai Sim
2009ICASSPThe I4U system in NIST 2008 speaker recognition evaluation.Haizhou Li, Bin Ma, Kong-Aik Lee, Hanwu Sun, Donglai Zhu, Khe Chai Sim, Changhuai You, Rong Tong, Ismo Krkkinen, Chien-Lin Huang, Vladimir Pervouchine, Wu Guo, Yijie Li, Li-Rong Dai, Mohaddeseh Nosratighods, Tharmarajah Thiruvaran, Julien Epps, Eliathamby Ambikairajah, Chng Eng Siong, Tanja Schultz, Qin Jin
2009InterspeechStream-based context-sensitive phone mapping for cross-lingual speech recognition.Khe Chai Sim, Haizhou Li
2008ICASSPRobust phone set mapping using decision tree clustering for cross-lingual phone recognition.Khe Chai Sim, Haizhou Li
2008InterspeechContext-sensitive probabilistic phone mapping model for cross-lingual speech recognition.Khe Chai Sim, Haizhou Li
2008PACLICNIST 2007 Language Recognition Evaluation: From the Perspective of IIR.Haizhou Li, Bin Ma, Kong-Aik Lee, Khe Chai Sim, Hanwu Sun, Rong Tong, Donglai Zhu, Changhuai You
2008SIGIRA lattice-based approach to query-by-example spoken document retrieval.Tee Kiah Chia, Khe Chai Sim, Haizhou Li, Hwee Tou Ng
2007ACLSemantic Transliteration of Personal Names.Haizhou Li, Khe Chai Sim, Jin-Shea Kuo, Minghui Dong
2007ICASSPConsensus Network Decoding for Statistical Machine Translation System Combination.Khe Chai Sim, William J. Byrne, Mark J. F. Gales, Hichem Sahbi, Philip C. Woodland
2007ICASSPImproving Speech Transcription for Mandarin-English Translation.Marcus Tomalin, Mark J. F. Gales, X. Andrew Liu, Khe Chai Sim, Rohit Sinha, Lan Wang, Philip C. Woodland, Kai Yu
2007InterspeechFusion of contrastive acoustic models for parallel phonotactic spoken language identification.Khe Chai Sim, Haizhou Li
2006ICASSPThe Cu-Htk Mandarin Broadcast News Transcription System.Rohit Sinha, Mark J. F. Gales, Do Yeong Kim, X. Andrew Liu, Khe Chai Sim, Philip C. Woodland
2005ICASSPDevelopment of the CUHTK 2004 Mandarin Conversational Telephone Speech Transcription System.Mark J. F. Gales, Bin Jia, X. Andrew Liu, Khe Chai Sim, Philip C. Woodland, Kai Yu
2005ICASSPDevelopment of the CU-HTK 2004 Broadcast News Transcription Systems.Do Yeong Kim, Ho Yin Chan, Gunnar Evermann, Mark J. F. Gales, David Mrva, Khe Chai Sim, Philip C. Woodland
2005ICASSPInvestigation of Acoustic Modeling Techniques for LVCSR Systems.Xunying Liu, Mark J. F. Gales, Khe Chai Sim, Kai Yu
2005ICASSPAdaptation of Precision Matrix Models on Large Vocabulary Continuous Speech Recognition.Khe Chai Sim, Mark J. F. Gales
2005InterspeechTemporally varying model parameters for large vocabulary continuous speech recognition.Khe Chai Sim, Mark J. F. Gales
2004ICASSPBasis superposition precision matrix modelling for large vocabulary continuous speech recognition.Khe Chai Sim, Mark J. F. Gales