Skip to content

Frank K. Soong

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

208

Venues

9

Active years

1978–2023

Best venue rank

Multiconference

Where they publish

Papers

208 indexed papers, newest first.

YearVenueTitleAuthors
2023InterspeechContextSpeech: Expressive and Efficient Text-to-Speech for Paragraph Reading.Yujia Xiao, Shaofei Zhang, Xi Wang, Xu Tan, Lei He, Sheng Zhao, Frank K. Soong, Tan Lee
2022ICASSPA Universal Ordinal Regression for Assessing Phoneme-Level Pronunciation.Shaoguang Mao, Frank K. Soong, Yan Xia, Jonathan Tien
2022ICASSPImproving Fastspeech TTS with Efficient Self-Attention and Compact Feed-Forward Network.Yujia Xiao, Xi Wang, Lei He, Frank K. Soong
2022ICASSPAn Approach to Mispronunciation Detection and Diagnosis with Acoustic, Phonetic and Linguistic (APL) Embeddings.Wenxuan Ye, Shaoguang Mao, Frank K. Soong, Wenshan Wu, Yan Xia, Jonathan Tien, Zhiyong Wu
2022InterspeechNeural Lexicon Reader: Reduce Pronunciation Errors in End-to-end TTS by Leveraging External Textual Knowledge.Mutian He, Jingzhou Yang, Lei He, Frank K. Soong
2022InterspeechA Multi-Stage Multi-Codebook VQ-VAE Approach to High-Performance Neural TTS.Haohan Guo, Feng-Long Xie, Frank K. Soong, Xixin Wu, Helen Meng
2021ICASSPSpeech Bert Embedding for Improving Prosody in Neural TTS.Liping Chen, Yan Deng, Xi Wang, Frank K. Soong, Lei He
2021ICASSPMBNET: MOS Prediction for Synthesized Speech with Mean-Bias Network.Yichong Leng, Xu Tan, Sheng Zhao, Frank K. Soong, Xiang-Yang Li, Tao Qin
2021ICASSPImproving Pronunciation Assessment Via Ordinal Regression with Anchored Reference Samples.Bin Su, Shaoguang Mao, Frank K. Soong, Yan Xia, Jonathan Tien, Zhiyong Wu
2021ICASSPA New High Quality Trajectory Tiling Based Hybrid TTS In Real Time.Feng-Long Xie, Xinhui Li, Wen-Chao Su, Li Lu, Frank K. Soong
2021InterspeechImproving Performance of Seen and Unseen Speech Style Transfer in End-to-End Neural TTS.Xiaochun An, Frank K. Soong, Lei Xie
2020ICASSPImproving LPCNET-Based Text-to-Speech with Linear Prediction-Structured Mixture Density Network.Min-Jae Hwang, Eunwoo Song, Ryuichi Yamamoto, Frank K. Soong, Hong-Goo Kang
2020ICASSPImproving Prosody with Linguistic and Bert Derived Features in Multi-Speaker Based Mandarin Chinese Neural TTS.Yujia Xiao, Lei He, Huaiping Ming, Frank K. Soong
2020ICASSPAn Improved Frame-Unit-Selection Based Voice Conversion System Without Parallel Training Data.Feng-Long Xie, Xinhui Li, Bo Liu, Yibin Zheng, Li Meng, Li Lu, Frank K. Soong
2020InterspeechAn Efficient Subband Linear Prediction for LPCNet-Based Neural Synthesis.Yang Cui, Xi Wang, Lei He, Frank K. Soong
2020InterspeechTransfer Learning for Improving Singing-Voice Detection in Polyphonic Instrumental Music.Yuanbo Hou, Frank K. Soong, Jian Luan, Shengchen Li
2019ICASSPDomain Adversarial Training for Improving Keyword Spotting Performance of ESL Speech.Jingyong Hou, Pengcheng Guo, Sining Sun, Frank K. Soong, Wenping Hu, Lei Xie
2019ICASSPNN-based Ordinal Regression for Assessing Fluency of ESL Speech.Shaoguang Mao, Zhiyong Wu, Jingshuai Jiang, Peiyun Liu, Frank K. Soong
2019ICASSPA Pitch-aware Approach to Single-channel Speech Separation.Ke Wang, Frank K. Soong, Lei Xie
2019InterspeechA New GAN-Based End-to-End TTS Training Algorithm.Haohan Guo, Frank K. Soong, Lei He, Lei Xie
2019InterspeechExploiting Syntactic Features in a Parsed Tree to Improve End-to-End TTS.Haohan Guo, Frank K. Soong, Lei He, Lei Xie
2019InterspeechForward-Backward Decoding for Regularizing End-to-End TTS.Yibin Zheng, Xi Wang, Lei He, Shifeng Pan, Frank K. Soong, Zhengqi Wen, Jianhua Tao
2018ICASSPExploring Sequential Characteristics in Speaker Bottleneck Feature for Text-Dependent Speaker Verification.Liping Chen, Yong Zhao, Shi-Xiong Zhang, Jie Li, Guoli Ye, Frank K. Soong
2018InterspeechA New Glottal Neural Vocoder for Speech Synthesis.Yang Cui, Xi Wang, Lei He, Frank K. Soong
2018InterspeechPaired Phone-Posteriors Approach to ESL Pronunciation Quality Assessment.Yujia Xiao, Frank K. Soong, Wenping Hu
2017ASRUImproving native language (L1) identifation with better VAD and TDNN trained separately on native and non-native English corpora.Yao Qian, Keelan Evanini, Patrick L. Lange, Robert A. Pugh, Rutuja Ubale, Frank K. Soong
2017ASRUPerceptual quality and modeling accuracy of excitation parameters in DLSTM-based speech synthesis systems.Eunwoo Song, Frank K. Soong, Hong-Goo Kang
2017InterspeechImproving Sub-Phone Modeling for Better Native Language Identification with Non-Native English Speech.Yao Qian, Keelan Evanini, Xinhao Wang, David Suendermann-Oeft, Robert A. Pugh, Patrick L. Lange, Hillary R. Molloy, Frank K. Soong
2017InterspeechProficiency Assessment of ESL Learner's Sentence Prosody with TTS Synthesized Voice as Reference.Yujia Xiao, Frank K. Soong
2017InterspeechDNN i-Vector Speaker Verification with Short, Text-Constrained Test Utterances.Jinghua Zhong, Wenping Hu, Frank K. Soong, Helen Meng
2016ICASSPUnsupervised speaker adaptation for DNN-based TTS synthesis.Yuchen Fan, Yao Qian, Frank K. Soong, Lei He
2016ICASSPSpeaker and language factorization in DNN-based TTS synthesis.Yuchen Fan, Yao Qian, Frank K. Soong, Lei He
2016ICASSPA KL divergence and DNN approach to cross-lingual TTS.Feng-Long Xie, Frank K. Soong, Haifeng Li
2016InterspeechImproved Time-Frequency Trajectory Excitation Vocoder for DNN-Based Speech Synthesis.Eunwoo Song, Frank K. Soong, Hong-Goo Kang
2016InterspeechA KL Divergence and DNN-Based Approach to Voice Conversion without Parallel Training Sentences.Feng-Long Xie, Frank K. Soong, Haifeng Li
2016NAACLLearning Distributed Word Representations For Bidirectional LSTM Recurrent Neural Network.Peilu Wang, Yao Qian, Frank K. Soong, Lei He, Hai Zhao
2015ICASSPMulti-speaker modeling and speaker adaptation for DNN-based TTS synthesis.Yuchen Fan, Yao Qian, Frank K. Soong, Lei He
2015ICASSPPhoto-real talking head with deep bidirectional LSTM.Bo Fan, Lijuan Wang, Frank K. Soong, Lei Xie
2015ICASSPWord embedding for recurrent neural network based TTS synthesis.Peilu Wang, Yao Qian, Frank K. Soong, Lei He, Hai Zhao
2015ICASSPAA spectral space warping approach to cross-lingual voice transformation in HMM-based TTS.Hao Wang, Frank K. Soong, Helen Meng
2015InterspeechSequence generation error (SGE) minimization based deep neural networks training for text-to-speech synthesis.Yuchen Fan, Yao Qian, Frank K. Soong, Lei He
2014ICASSPA DNN-based acoustic modeling of tonal language and its application to Mandarin pronunciation training.Wenping Hu, Yao Qian, Frank K. Soong
2014ICASSPOn the training aspects of Deep Neural Network (DNN) for parametric TTS synthesis.Yao Qian, Yuchen Fan, Wenping Hu, Frank K. Soong
2014ICASSPA maximum a Posterior-based reconstruction approach to speech bandwidth expansion in noise.Hyunson Seo, Hong-Goo Kang, Frank K. Soong
2014InterspeechTTS synthesis with bidirectional LSTM based recurrent neural networks.Yuchen Fan, Yao Qian, Feng-Long Xie, Frank K. Soong
2014InterspeechSequence error (SE) minimization training of neural network for voice conversion.Feng-Long Xie, Yao Qian, Yuchen Fan, Frank K. Soong, Haifeng Li
2014InterspeechModeling DCT parameterized F0 trajectory at intonation phrase level with DNN or decision tree.Xiang Yin, Ming Lei, Yao Qian, Frank K. Soong, Lei He, Zhen-Hua Ling, Li-Rong Dai
2013ICASSPA fast table lookup based, statistical model driven non-uniform unit selection TTS.Yao Qian, Frank K. Soong, Xiaobo Zhou, Yundi Qian, Xiaotian Zhang
2013InterspeechA new DNN-based high quality pronunciation evaluation for computer-aided language learning (CALL).Wenping Hu, Yao Qian, Frank K. Soong
2013InterspeechA source-filter based adaptive harmonic model and its application to speech prosody modification.JeeSok Lee, Frank K. Soong, Hong-Goo Kang
2013InterspeechBinocular photometric stereo acquisition and reconstruction for 3d talking head applications.Chaoyang Wang, Lijuan Wang, Yasuyuki Matsushita, Bojun Huang, Magnetro Chen, Frank K. Soong
2013InterspeechA new language independent, photo-realistic talking head driven by voice only.Xinjian Zhang, Lijuan Wang, Gang Li, Frank Seide, Frank K. Soong
2012ICASSPImproved minimum converted trajectory error training for real-time speech-to-lips conversion.Wei Han, Lijuan Wang, Frank K. Soong, Bo Yuan
2012ICASSPHigh quality lip-sync animation for 3D photo-realistic talking head.Lijuan Wang, Wei Han, Frank K. Soong
2012ICASSPModeling pitch trajectory by hierarchical HMM with minimum generation error training.Yi-Jian Wu, Frank K. Soong
2012ICASSPNoise estimation using a constrained sequential HMM IN log-spectral domain.Dongwen Ying, Xugang Lu, Junfeng Li, Yonghong Yan, Jianwu Dang, Frank K. Soong
2012InterspeechTurning a Monolingual Speaker into Multilingual for a Mixed-language TTS.Ji He, Yao Qian, Frank K. Soong, Sheng Zhao
2012InterspeechThe Use of DBN-HMMs for Mispronunciation Detection and Diagnosis in L2 English to Support Computer-Aided Pronunciation Training.Xiaojun Qian, Helen M. Meng, Frank K. Soong
2012InterspeechObjective Intelligibility Assessment of Text-to-Speech System using Template Constrained Generalized Posterior Probability.Linfang Wang, Lijuan Wang, Yan Teng, Zhe Geng, Frank K. Soong
2012InterspeechConstrained Multichannel Speech Dereverberation.Meng Yu, Frank K. Soong
2011ICASSPImproved F0 modeling and generation in voice conversion.Aki Kunikoshi, Yao Qian, Frank K. Soong, Nobuaki Minematsu
2011ICASSPSpeaker characterization using spectral subband energy ratio based on Harmonic plus Noise Model.Yanhua Long, Zhi-Jie Yan, Frank K. Soong, Li-Rong Dai, Wu Guo
2011ICASSPA frame mapping based HMM approach to cross-lingual voice transformation.Yao Qian, Ji Xu, Frank K. Soong
2011ICASSPSynthesizing visual speech trajectory with minimum generation error.Lijuan Wang, Yi-Jian Wu, Xiaodan Zhuang, Frank K. Soong
2011ICASSPA Sparse and Low-rank approach to efficient face alignment for photo-real talking head synthesis.King Keung Wu, Lijuan Wang, Frank K. Soong, Yeung Yam
2011InterspeechImprovements in Speaker Characterization Using Spectral Subband Energy Based on Harmonic plus Noise Model.Yanhua Long, Zhi-Jie Yan, Frank K. Soong, Li-Rong Dai, Wu Guo
2011InterspeechA New Phonetic Candidate Generator for Improving Search Query Efficiency.Bo Peng, Yao Qian, Frank K. Soong, Bo Zhang
2011InterspeechOn Mispronunciation Lexicon Generation Using Joint-Sequence Multigrams in Computer-Aided Pronunciation Training (CAPT).Xiaojun Qian, Helen M. Meng, Frank K. Soong
2011InterspeechText Driven 3D Photo-Realistic Talking Head.Lijuan Wang, Wei Han, Frank K. Soong, Qiang Huo
2010ICASSPRIch-context Unit Selection (RUS) approach to high quality TTS.Zhi-Jie Yan, Yao Qian, Frank K. Soong
2010ICASSPImproved modeling for F0 generation and V/U decision in HMM-based TTS.Qingqing Zhang, Frank K. Soong, Yao Qian, Zhijie Yan, Jielin Pan, Yonghong Yan
2010ICASSPCross-validation based decision tree clustering for HMM-based TTS.Yu Zhang, Zhi-Jie Yan, Frank K. Soong
2010InterspeechA perceptual study of acceleration parameters in HMM-based TTS.Yining Chen, Zhi-Jie Yan, Frank K. Soong
2010InterspeechA hierarchical F0 modeling method for HMM-based speech synthesis.Ming Lei, Yi-Jian Wu, Frank K. Soong, Zhen-Hua Ling, Li-Rong Dai
2010InterspeechDiscriminative acoustic model for improving mispronunciation detection and diagnosis in computer-aided pronunciation training (CAPT).Xiaojun Qian, Frank K. Soong, Helen M. Meng
2010InterspeechAn HMM trajectory tiling (HTT) approach to high quality TTS.Yao Qian, Zhi-Jie Yan, Yi-Jian Wu, Frank K. Soong, Xin Zhuang, Shengyi Kong
2010InterspeechSynthesizing photo-real talking head via trajectory-guided sample selection.Lijuan Wang, Xiaojun Qian, Wei Han, Frank K. Soong
2010InterspeechFormant-based frequency warping for improving speaker adaptation in HMM TTS.Xin Zhuang, Yao Qian, Frank K. Soong, Yi-Jian Wu, Bo Zhang
2010InterspeechA minimum converted trajectory error (MCTE) approach to high quality speech-to-lips conversion.Xiaodan Zhuang, Lijuan Wang, Frank K. Soong, Mark Hasegawa-Johnson
2009ICASSPImproving mispronunciation detection using machine learning.Yuqiang Chen, Chao Huang, Frank K. Soong
2009ICASSPState mapping for cross-language speaker adaptation in TTS.Yining Chen, Yang Jiao, Yao Qian, Frank K. Soong
2009ICASSPImproved prosody generation by maximizing joint likelihood of state and longer units.Yao Qian, Zhizheng Wu, Frank K. Soong
2009ICASSPAn evidence framework for Bayesian learning of continuous-density hidden Markov models.Yu Zhang, Peng Liu, Jen-Tzung Chien, Frank K. Soong
2009InterspeechModel-based speech separation: identifying transcription using orthogonality.Siu Wa Lee, Frank K. Soong, Tan Lee
2009InterspeechA minimum v/u error approach to F0 generation in HMM-based TTS.Yao Qian, Frank K. Soong, Miaomiao Wang, Zhizheng Wu
2009InterspeechAuto-checking speech transcriptions by multiple template constrained posterior.Lijuan Wang, Shenghao Qin, Frank K. Soong
2009InterspeechRich context modeling for high quality HMM-based TTS.Zhi-Jie Yan, Yao Qian, Frank K. Soong
2008ICASSPDiscriminative training for improving letter-to-sound conversion performance.Yining Chen, Peng Liu, Jia-Li You, Frank K. Soong
2008ICASSPA cross-language state mapping approach to bilingual (Mandarin-English) TTS.Hui Liang, Yao Qian, Frank K. Soong, Gongshen Liu
2008ICASSPPrefix tree based auto-completion for convenient bi-modal chinese character input.Peng Liu, Lei Ma, Frank K. Soong
2008ICASSPSymbol graph based discriminative training and rescoring for improved math symbol recognition.Zhen Xuan Luo, Yu Shi, Frank K. Soong
2008ICASSPTemplate constrained posterior for verifying phone transcriptions.Lijuan Wang, Tao Hu, Frank K. Soong
2008ICASSPImproving letter-to-sound conversion performance with automatically generated new words.Jia-Li You, Yining Chen, Frank K. Soong, Jin-Lin Wang
2008ICASSPAutomatic mispronunciation detection for Mandarin.Feng Zhang, Chao Huang, Frank K. Soong, Min Chu, Ren-Hua Wang
2008ICPRRadical based fine trajectory HMMs of online handwritten characters.Peng Liu, Lei Ma, Frank K. Soong
2008ICPRA symbol graph based handwritten math expression recognition.Yu Shi, Frank K. Soong
2008InterspeechDuration refinement by jointly optimizing state and longer unit likelihood.Boyang Gao, Yao Qian, Zhizheng Wu, Frank K. Soong
2008InterspeechMispronunciation detection for Mandarin Chinese.Chao Huang, Feng Zhang, Frank K. Soong, Min Chu
2008InterspeechAn ellipsoid constrained quadratic programming perspective to discriminative training of HMMs.Peng Liu, Frank K. Soong
2008InterspeechGenerating natural F0 trajectory with additive trees.Yao Qian, Hui Liang, Frank K. Soong
2008InterspeechGPU-accelerated Gaussian clustering for fMPE discriminative training.Yu Shi, Frank Seide, Frank K. Soong
2008InterspeechEfficient handwriting correction of speech recognition errors with template constrained posterior (TCP).Lijuan Wang, Tao Hu, Peng Liu, Frank K. Soong
2008InterspeechA real-time text to audio-visual speech synthesis system.Lijuan Wang, Xiaojun Qian, Lei Ma, Yao Qian, Yining Chen, Frank K. Soong
2008InterspeechProsody for Mandarin speech recognition: a comparative study of read and spontaneous speech.Yu Ting Yeung, Yao Qian, Tan Lee, Frank K. Soong
2007ASRUA constrained line search approach to general discriminative HMM training.Peng Liu, Cong Liu, Hui Jiang, Frank K. Soong, Ren-Hua Wang
2007HCIEnrich Web Applications with Voice Internet Persona Text-to-Speech for Anyone, Anywhere.Min Chu, Yusheng Li, Xin Zou, Frank K. Soong
2007ICASSPDivergence-Based Similarity Measure for Spoken Document Retrieval.Peng Liu, Frank K. Soong, Jian-Lai Zhou
2007ICASSPA New Minimum Divergence Approach to Discriminative Training.Jun Du, Peng Liu, Hui Jiang, Frank K. Soong, Ren-Hua Wang
2007ICASSPA Constrained Line Search Optimization for Discriminative Training in Speech Recognition.Cong Liu, Peng Liu, Hui Jiang, Frank K. Soong, Ren-Hua Wang
2007ICASSPAgreement Learning for Automatic Accent Annotation.Xinqiang Ni, Yining Chen, Min Chu, Frank K. Soong, Yong Zhao, Ping Zhang
2007ICASSPFull HMM Training for Minimizing Generation Error in Synthesis.Yi-Jian Wu, Ren-Hua Wang, Frank K. Soong
2007ICASSPA Segmentation Posterior Based Endpointing Algorithm.Yanlu Xie, Yu Shi, Frank K. Soong, Beiqian Dai
2007ICASSPWord Graph Based Feature Enhancement for Noisy Speech Recognition.Zhi-Jie Yan, Frank K. Soong, Ren-Hua Wang
2007ICASSPGeneralized Segment Posterior Probability for Automatic Mandarin Pronunciation Evaluation.Jing Zheng, Chao Huang, Min Chu, Frank K. Soong, Weiping Ye
2007ICDARA MSD-HMM Approach to Pen Trajectory Modeling for Online Handwriting Recognition.Lei Ma, Frank K. Soong, Peng Liu, Yi-Jian Wu
2007ICDARA Unified Framework for Symbol Segmentation and Recognition of Handwritten Mathematical Expressions.Yu Shi, HaiYang Li, Frank K. Soong
2007ICDARMinimum Error Discriminative Training for Radical-Based Online Chinese Handwriting Recognition.Yu Zhang, Peng Liu, Frank K. Soong
2007InterspeechModel-based speech separation with single-microphone input.Siu Wa Lee, Frank K. Soong, Pak-Chung Ching
2007InterspeechIterative unit selection with unnatural prosody detection.Dacheng Lin, Yong Zhao, Frank K. Soong, Min Chu, Jieyu Zhao
2007InterspeechAn unsupervised approach to automatic prosodic annotation.Xinqiang Ni, Yining Chen, Frank K. Soong, Min Chu, Ping Zhang
2007InterspeechRobust F0 modeling for Mandarin speech recognition in noise.Sheng Qiang, Yao Qian, Frank K. Soong, Congfu Xu
2007InterspeechContext constrained-generalized posterior probability for verifying phone transcriptions.Hua Zhang, Lijuan Wang, Frank K. Soong, Wenju Liu
2006ICASSPWeighted Likelihood Ratio (WLR) Hidden Markov Model for Noisy Speech Recognition.Chao Huang, Yingchun Huang, Frank K. Soong, Jianlai Zhou
2006ICASSPSyllable Lattice Based Re-Scoring For Speaker Verification.Minho Jin, Frank K. Soong, Chang D. Yoo
2006ICASSPAn Iterative Trajectory Regeneration Algorithm for Separating Mixed Speech Sources.Siu Wa Lee, Frank K. Soong, Pak-Chung Ching
2006ICASSPTone-Enhanced Generalized Character Posterior Probability (GCPP) for Cantonese LVCSR.Yao Qian, Frank K. Soong, Tan Lee
2006ICASSPAuto-Segmentation Based Partitioning and Clustering Approach to Robust Endpointing.Yu Shi, Frank K. Soong, Jian-Lai Zhou
2006ICASSPA Comparative Study of Discriminative Methods for Reranking LVCSR N-Best Hypotheses in Domain Adaptation and Generalization.Zhengyu Zhou, Jianfeng Gao, Frank K. Soong, Helen Meng
2006ICASSPImproved Chinese Character Input by Merging Speech and Handwriting Recognition Hypotheses.Xi Zhou, Ye Tian, Jian-Lai Zhou, Frank K. Soong, Beiqian Dai
2006ICMIWord graph based speech rcognition error correction by handwriting input.Peng Liu, Frank K. Soong
2006InterspeechMinimum divergence based discriminative training.Jun Du, Peng Liu, Frank K. Soong, Jian-Lai Zhou, Ren-Hua Wang
2006InterspeechGeneralization of the minimum classification error (MCE) training based on maximizing generalized posterior probability (GPP).Qiang Fu, Antonio Moreno-Daniel, Biing-Hwang Juang, Jian-Lai Zhou, Frank K. Soong
2006InterspeechAuto-segmentation based VAD for robust ASR.Yu Shi, Frank K. Soong, Jian-Lai Zhou
2006InterspeechA multi-space distribution (MSD) approach to speech recognition of tonal languages.Huanliang Wang, Yao Qian, Frank K. Soong, Jian-Lai Zhou, Jiqing Han
2005ICASSPOptimal Clustering and Non-Uniform Allocation of Gaussian Kernels in Scalar Dimension for HMM Compression.Xiao-Bing Li, Frank K. Soong, Tor Andr Myrvoll, Ren-Hua Wang
2005ICASSPGeneralized Posterior Probability for Minimum Error Verification of Recognized Sentences.Wai Kit Lo, Frank K. Soong
2005ICASSPStatic and Dynamic Spectral Features: Their Noise Robustness and Optimal Weights for ASR.Chen Yang, Frank K. Soong, Tan Lee
2005InterspeechHarmonic filtering for joint estimation of pitch and voiced source with single-microphone input.Siu Wa Lee, Frank K. Soong, Pak-Chung Ching
2005InterspeechBackground model based posterior probability for measuring confidence.Peng Liu, Ye Tian, Jian-Lai Zhou, Frank K. Soong
2005InterspeechPhonetic transcription verification with generalized posterior probability.Lijuan Wang, Yong Zhao, Min Chu, Frank K. Soong, Zhigang Cao
2005InterspeechRefining phoneme segmentations using speaker-adaptive context dependent boundary models.Yong Zhao, Lijuan Wang, Min Chu, Frank K. Soong, Zhigang Cao
2004COLINGA Unified Approach in Speech-to-Speech Translation: Integrating Features of Speech recognition and Machine Translation.Ruiqiang Zhang, Gen-ichiro Kikui, Hirofumi Yamamoto, Frank K. Soong, Taro Watanabe, Wai Kit Lo
2004InterspeechRobust verification of recognized words in noise.Wai Kit Lo, Frank K. Soong, Satoshi Nakamura
2004InterspeechTone information as a confidence measure for improving Cantonese LVCSR.Yao Qian, Tan Lee, Frank K. Soong
2004InterspeechOptimal acoustic and language model weights for minimizing word verification errors.Frank K. Soong, Wai Kit Lo, Satoshi Nakamura
2004InterspeechImproved spoken language translation using n-best speech recognition hypotheses.Ruiqiang Zhang, Gen-ichiro Kikui, Hirofumi Yamamoto, Frank K. Soong, Taro Watanabe, Eiichiro Sumita, Wai Kit Lo
2003ICASSPCombining neighboring filter channels to improve quantile based histogram equalization.Florian Hilger, Hermann Ney, Olivier Siohan, Frank K. Soong
2003ICASSPOptimal clustering of multivariate normal distributions using divergence and its application to HMM adaptation.Tor Andr Myrvoll, Frank K. Soong
2003InterspeechModeling Cantonese pronunciation variation by acoustic model refinement.Patgi Kam, Tan Lee, Frank K. Soong
2003InterspeechOn divergence based clustering of normal distributions and its application to HMM adaptation.Tor Andr Myrvoll, Frank K. Soong
2002ICASSPA dynamic in-search discriminative training approach for large vocabulary speech recognition.Hui Jiang, Olivier Siohan, Frank K. Soong, Chin-Hui Lee
2002ICASSPClassifier design for verification of multi-class recognition decision.Tomoko Matsui, Frank K. Soong, Biing-Hwang Juang
2002InterspeechBell labs approach to Aurora evaluation on connected digit recognition.Jingdong Chen, Dimitris Dimitriadis, Hui Jiang, Qi Li, Tor Andr Myrvoll, Olivier Siohan, Frank K. Soong
2002InterspeechRecognition of noisy speech using normalized moments.Jingdong Chen, Yiteng Huang, Qi Li, Frank K. Soong
2001ICASSPHierarchical stochastic feature matching for robust speech recognition.Hui Jiang, Frank K. Soong, Chin-Hui Lee
2001InterspeechEvaluating the Aurora connected digit recognition task - a bell labs approach.Mohamed Afify, Hui Jiang, Filipp Korkmazskiy, Chin-Hui Lee, Qi Li, Olivier Siohan, Frank K. Soong, Arun C. Surendran
2001InterspeechA data selection strategy for utterance verification in continuous speech recognition.Hui Jiang, Frank K. Soong, Chin-Hui Lee
2001InterspeechAn auditory system-based feature for robust speech recognition.Qi Li, Frank K. Soong, Olivier Siohan
2001InterspeechA real-time Japanese broadcast news closed-captioning system.Olivier Siohan, Akio Ando, Mohamed Afify, Hui Jiang, Chin-Hui Lee, Qi Li, Feng Liu, Kazuo Onoe, Frank K. Soong, Qiru Zhou
2000InterspeechA high-performance auditory feature for robust speech recognition.Qi Li, Frank K. Soong, Olivier Siohan
2000InterspeechHands-free human-machine dialogue - corpora, technology and evaluation.Frank K. Soong, Eric A. Woudenberg
1999ICASSPHidden Markov models with divergence based vector quantized variances.Jae H. Kim, Raziel Haimi-Cohen, Frank K. Soong
1999ICASSPA block least squares approach to acoustic echo cancellation.Eric A. Woudenberg, Frank K. Soong, Biing-Hwang Juang
1998InterspeechImproved utterance rejection using length dependent thresholds.Sunil K. Gupta, Frank K. Soong
1997ICASSPGeneralized mixture of HMMs for continuous speech recognition.Filipp Korkmazskiy, Biing-Hwang Juang, Frank K. Soong
1996ICASSPHigh-accuracy connected digit recognition for mobile applications.Sunil K. Gupta, Frank K. Soong, Raziel Haimi-Cohen
1996InterspeechQuantizing mixture-weights in a tied-mixture HMM.Sunil K. Gupta, Frank K. Soong, Raziel Haimi-Cohen
1995ICASSPAn orthogonal polynomial representation of speech signals and its probabilistic model for text independent speaker verification.Chi-Shi Liu, Hsiao-Chuan Wang, Frank K. Soong, Chao-Shih Huang
1995InterspeechLarge vocabulary, word-based Mandarin dictation system.Jung-Kuei Chen, Lin-Shan Lee, Frank K. Soong
1995InterspeechOptimizing baseforms for HMM-based speech recognition.Torbjrn Svendsen, Frank K. Soong, Heiko Purnhagen
1994ICASSPDiscriminative training of high performance speech recognizer using N best candidates.Jung-Kuei Chen, Frank K. Soong
1994ICASSPLarge vocabulary word recognition based on tree-trellis search.Jung-Kuei Chen, Frank K. Soong, Lin-Shan Lee
1994InterspeechCepstral channel normalization techniques for HMM-based speaker verification.Aaron E. Rosenberg, Chin-Hui Lee, Frank K. Soong
1992ICASSPContinuous probabilistic acoustic map for speaker recognition.Belle L. Tseng, Frank K. Soong, Aaron E. Rosenberg
1992InterspeechThe use of cohort normalized scores for speaker verification.Aaron E. Rosenberg, Joel DeLong, Chin-Hui Lee, Biing-Hwang Juang, Frank K. Soong
1992InterspeechContinuous mixture HMM-LR using the a* algorithm for continuous speech recognition.Kouichi Yamaguchi, Shigeki Sagayama, Kenji Kita, Frank K. Soong
1991ICASSPA tree-trellis based fast search for finding the N-best sentence hypotheses in continuous speech recognition.Frank K. Soong, Eng-Fong Huang
1990ICASSPStatistical segmentation and word modeling techniques in isolated word recognition.S. A. Euler, Biing-Hwang Juang, Chin-Hui Lee, Frank K. Soong
1990ICASSPA probabilistic acoustic map based discriminative HMM training.Eng-Fong Huang, Frank K. Soong
1990ICASSPSpeaker recognition based on source coding approaches.Biing-Hwang Juang, Frank K. Soong
1990ICASSPSub-word unit talker verification using hidden Markov models.Aaron E. Rosenberg, Chin-Hui Lee, Frank K. Soong
1990ICASSPOptimal quantization of LSP parameters using delayed decisions.Frank K. Soong, Biing-Hwang Juang
1990InterspeechExperiments in automatic talker verification using sub-word unit hidden Markov models.Aaron E. Rosenberg, Chin-Hui Lee, Frank K. Soong, Maureen A. McGee
1990InterspeechA tree-trellis based fast search for finding the n best sentence hypotheses in continuous speech recognition.Frank K. Soong, Eng-Fong Huang
1990NAACLA Tree.Trellis Based Fast Search for Finding the N Best Sentence Hypotheses in Continuous Speech Recognition.Frank K. Soong, Eng-Fong Huang
1989ICASSPWord recognition using whole word and subword models.Chin-Hui Lee, Biing-Hwang Juang, Frank K. Soong, Lawrence R. Rabiner
1989ICASSPA phonetically labeled acoustic segment (PLAS) approach to speech analysis-synthesis.Frank K. Soong
1988ICASSPA segment model based approach to speech recognition.Chin-Hui Lee, Frank K. Soong, Biing-Hwang Juang
1988ICASSPHigh performance connected digit recognition, using hidden Markov models.Lawrence R. Rabiner, Jay G. Wilpon, Frank K. Soong
1988ICASSPOptimal quantization of LSP parameters [speech coding].Frank K. Soong, Bling-Hwang Juang
1987ICASSPA training procedure for a segment-based-network approach to isolated word recognition.Frank K. Soong
1987ICASSPA frequency-weighted Itakura spectral distortion measure and its application to speech recognition in noise.Frank K. Soong, M. Mohan Sondhi
1987ICASSPOn the automatic segmentation of speech signals.Torbjrn Svendsen, Frank K. Soong
1986ICASSPEvaluation of a vector quantization talker recognition system in text independent and text dependent modes.Aaron E. Rosenberg, Frank K. Soong
1986ICASSPA high quality subband speech coder with backward adaptive predictor and optimal time-frequency bit assignment.Frank K. Soong, Richard V. Cox, Nikil S. Jayant
1986ICASSPOn the use of instantaneous and transitional spectral information in speaker recognition.Frank K. Soong, Aaron E. Rosenberg
1985ICASSPComparative study of several distortion measures for speech recognition.N. Nocerino, Frank K. Soong, Lawrence R. Rabiner, Dennis H. Klatt
1985ICASSPAn efficient vector-quantization preprocessor for speaker independent isolated word recognition.Kuk-Chin Pan, Frank K. Soong, Lawrence R. Rabiner, A. F. Bergh
1985ICASSPSubband coding of speech using backward adaptive prediction and bit allocation.Frank K. Soong, Richard V. Cox, Nikil S. Jayant
1985ICASSPA vector quantization approach to speaker recognition.Frank K. Soong, Aaron E. Rosenberg, Lawrence R. Rabiner, Biing-Hwang Juang
1984ICASSPOn the use of transient information in speech recognition.Jean-Sylvain Linard, Frank K. Soong
1984ICASSPLine spectrum pair (LSP) and speech data compression.Frank K. Soong, Biing-Hwang Juang
1982ICASSPOn the high resolution and unbiased frequency estimates of sinusoids in white noise-A new adaptive approach.Frank K. Soong, Allen M. Peterson
1982ICASSPFast least-squares (LS) in the voice echo cancellation application.Frank K. Soong, Allen M. Peterson
1981ICASSPOn the asymptotic behavior of a complex adaptive line enchancer (CALE).Frank K. Soong, S. Shankar Narayan, Allen M. Peterson
1980ICASSPFast spectral estimation of speech signal in analytic form.Frank K. Soong, Allen M. Peterson
1978ICASSPObservations on linear estimation.Leland B. Jackson, Frank K. Soong
1978ICASSPFrequency estimation by linear prediction.Leland B. Jackson, Donald W. Tufts, Frank K. Soong, Rahul M. Rao