| 2024 | ICASSP | SpeechDPR: End-To-End Spoken Passage Retrieval For Open-Domain Spoken Question Answering. | Chyi-Jiunn Lin, Guan-Ting Lin, Yung-Sung Chuang, Wei-Lun Wu, Shang-Wen Li, Abdelrahman Mohamed, Hung-Yi Lee, Lin-Shan Lee |
| 2022 | Interspeech | DUAL: Discrete Spoken Unit Adaptive Learning for Textless Spoken Question Answering. | Guan-Ting Lin, Yung-Sung Chuang, Ho-Lam Chung, Shu-Wen Yang, Hsuan-Jui Chen, Shuyan Annie Dong, Shang-Wen Li, Abdelrahman Mohamed, Hung-yi Lee, Lin-Shan Lee |
| 2021 | ICASSP | Fragmentvc: Any-To-Any Voice Conversion by End-To-End Extracting and Fusing Fine-Grained Voice Fragments with Attention. | Yist Y. Lin, Chung-Ming Chien, Jheng-Hao Lin, Hung-yi Lee, Lin-Shan Lee |
| 2021 | Interspeech | Towards Lifelong Learning of End-to-End ASR. | Heng-Jui Chang, Hung-yi Lee, Lin-Shan Lee |
| 2020 | ICASSP | Sequence-to-Sequence Automatic Speech Recognition with Word Embedding Regularization and Fused Decoding. | Alexander H. Liu, Tzu-Wei Sung, Shun-Po Chuang, Hung-yi Lee, Lin-Shan Lee |
| 2020 | ICASSP | Towards Unsupervised Speech Recognition and Synthesis with Quantized Speech Representation Learning. | Alexander H. Liu, Tao Tu, Hung-yi Lee, Lin-Shan Lee |
| 2020 | ICASSP | Interrupted and Cascaded Permutation Invariant Training for Speech Separation. | Gene-Ping Yang, Szu-Lin Wu, Yao-Wen Mao, Hung-yi Lee, Lin-Shan Lee |
| 2020 | Interspeech | SpeechBERT: An Audio-and-Text Jointly Learned Language Model for End-to-End Spoken Question Answering. | Yung-Sung Chuang, Chi-Liang Liu, Hung-yi Lee, Lin-Shan Lee |
| 2020 | Interspeech | Doing Something we Never could with Spoken Language Technologies-from early days to the era of deep learning. | Lin-Shan Lee |
| 2019 | ICASSP | Adversarial Training of End-to-end Speech Recognition Using a Criticizing Language Model. | Alexander H. Liu, Hung-yi Lee, Lin-Shan Lee |
| 2019 | ICASSP | Towards End-to-end Speech-to-text Translation with Two-pass Decoding. | Tzu-Wei Sung, Jun-You Liu, Hung-yi Lee, Lin-Shan Lee |
| 2019 | Interspeech | Completely Unsupervised Phoneme Recognition by a Generative Adversarial Network Harmonized with Iteratively Refined Hidden Markov Models. | Kuan-Yu Chen, Che-Ping Tsai, Da-Rong Liu, Hung-yi Lee, Lin-Shan Lee |
| 2019 | Interspeech | Improved Speech Separation with Time-and-Frequency Cross-Domain Joint Embedding and Clustering. | Gene-Ping Yang, Chao-I Tuan, Hung-yi Lee, Lin-Shan Lee |
| 2018 | ICASSP | Scalable Sentiment for Sequence-to-Sequence Chatbot Response with Performance Analysis. | Chih-Wei Lee, Yau-Shian Wang, Tsung-Yuan Hsu, Kuan-Yu Chen, Hung-yi Lee, Lin-Shan Lee |
| 2018 | ICASSP | Domain Independent Key Term Extraction from Spoken Content Based on Context and Term Location Information in the Utterances. | Hsien-Chin Lin, Chi-Yu Yang, Hung-yi Lee, Lin-Shan Lee |
| 2018 | ICASSP | Transcribing Lyrics from Commercial Song Audio: the First Step Towards Singing Content Processing. | Che-Ping Tsai, Yi-Lin Tuan, Lin-Shan Lee |
| 2018 | ICASSP | Segmental Audio Word2Vec: Representing Utterances as Sequences of Vectors with Applications in Spoken Term Detection. | Yu-Hsuan Wang, Hung-yi Lee, Lin-Shan Lee |
| 2018 | Interspeech | Multi-target Voice Conversion without Parallel Data by Adversarially Learning Disentangled Audio Representations. | Ju-Chieh Chou, Cheng-chieh Yeh, Hung-yi Lee, Lin-Shan Lee |
| 2018 | Interspeech | Completely Unsupervised Phoneme Recognition by Adversarially Learning Mapping Relationships from Audio Embeddings. | Da-Rong Liu, Kuan-Yu Chen, Hung-yi Lee, Lin-Shan Lee |
| 2017 | ASRU | Personalized word representations carrying personalized semantics learned from social network posts. | Zih-Wei Lin, Tzu-Wei Sung, Hung-yi Lee, Lin-Shan Lee |
| 2017 | ICASSP | Personalized acoustic modeling by weakly supervised multi-task deep learning using acoustic tokens discovered from unlabeled data. | Cheng-Kuang Wei, Cheng-Tao Chung, Hung-yi Lee, Lin-Shan Lee |
| 2017 | Interspeech | Order-Preserving Abstractive Summarization for Spoken Content Based on Connectionist Temporal Classification. | Bo-Ru Lu, Frank Shyu, Yun-Nung Chen, Hung-yi Lee, Lin-Shan Lee |
| 2016 | Interspeech | Audio Word2Vec: Unsupervised Learning of Audio Segment Representations Using Sequence-to-Sequence Autoencoder. | Yu-An Chung, Chao-Chung Wu, Chia-Hao Shen, Hung-yi Lee, Lin-Shan Lee |
| 2016 | Interspeech | Towards Machine Comprehension of Spoken Content: Initial TOEFL Listening Comprehension Test by Machine. | Bo-Hsiang Tseng, Sheng-syun Shen, Hung-yi Lee, Lin-Shan Lee |
| 2016 | Interspeech | Interactive Spoken Content Retrieval by Deep Reinforcement Learning. | Yen-Chen Wu, Tzu-Hsiang Lin, Yang-De Chen, Hung-yi Lee, Lin-Shan Lee |
| 2015 | ASRU | An iterative deep learning framework for unsupervised discovery of speech features and linguistic units with applications on spoken term detection. | Cheng-Tao Chung, Cheng-Yu Tsai, Hsiang-Hung Lu, Chia-Hsiang Liu, Hung-yi Lee, Lin-Shan Lee |
| 2015 | ASRU | Towards structured deep neural network for automatic speech recognition. | Yi-Hsiu Liao, Hung-yi Lee, Lin-Shan Lee |
| 2015 | ASRU | Personalizing universal recurrent neural network language model with user characteristic features by social network crowdsourcing. | Bo-Hsiang Tseng, Hung-yi Lee, Lin-Shan Lee |
| 2015 | ICASSP | Enhancing automatically discovered multi-level acoustic patterns considering context consistency with applications in spoken term detection. | Cheng-Tao Chung, Wei-Ning Hsu, Cheng-Yi Lee, Lin-Shan Lee |
| 2015 | ICASSP | Enhancing sparse voice annotation for semantic retrieval of personal photos by continuous space word representations. | Yuan-ming Liou, Hung-tsung Lu, Yi-Sheng Fu, Winston H. Hsu, Lin-Shan Lee |
| 2015 | Interspeech | Semantic retrieval of personal photos using a deep autoencoder fusing visual features with speech annotations represented as word/paragraph vectors. | Hung-tsung Lu, Yuan-ming Liou, Hung-yi Lee, Lin-Shan Lee |
| 2015 | Interspeech | Structuring lectures in massive open online courses (MOOCs) for efficient learning by linking similar sections and predicting prerequisites. | Sheng-syun Shen, Hung-yi Lee, Shang-wen Li, Victor Zue, Lin-Shan Lee |
| 2015 | Interspeech | Personalized speech recognizer with keyword-based personalized lexicon and language model using word vector representations. | Ching-feng Yeh, Yuan-ming Liou, Hung-yi Lee, Lin-Shan Lee |
| 2014 | ICASSP | Unsupervised spoken term detection with spoken queries by multi-level acoustic patterns with varying model granularity. | Cheng-Tao Chung, Chun-an Chan, Lin-Shan Lee |
| 2014 | ICASSP | Transcribing code-switched bilingual lectures using deep neural networks with unit merging in acoustic modeling. | Ching-feng Yeh, Lin-Shan Lee |
| 2014 | Interspeech | Semantic retrieval of personal photos using matrix factorization and two-layer random walk fusing sparse speech annotations with visual features. | Yuan-ming Liou, Yi-Sheng Fu, Hung-yi Lee, Lin-Shan Lee |
| 2014 | Interspeech | Alignment of spoken utterances with slide content for easier learning with recorded lectures using structured support vector machine (SVM). | Han Lu, Sheng-syun Shen, Sz-Rung Shiang, Hung-yi Lee, Lin-Shan Lee |
| 2014 | Interspeech | Spoken question answering using tree-structured conditional random fields and two-layer random walk. | Sz-Rung Shiang, Hung-yi Lee, Lin-Shan Lee |
| 2013 | ASRU | Towards unsupervised semantic retrieval of spoken content with query expansion based on automatically discovered acoustic patterns. | Yun-Chiao Li, Hung-yi Lee, Cheng-Tao Chung, Chun-an Chan, Lin-Shan Lee |
| 2013 | ICASSP | Toward unsupervised model-based spoken term detection with spoken queries without annotated data. | Chun-an Chan, Cheng-Tao Chung, Yu-Hsin Kuo, Lin-Shan Lee |
| 2013 | ICASSP | Unsupervised discovery of linguistic structure including two-level acoustic patterns using three cascaded stages of iterative optimization. | Cheng-Tao Chung, Chun-an Chan, Lin-Shan Lee |
| 2013 | ICASSP | Unsupervised domain adaptation for spoken document summarization with structured support vector machine. | Hung-yi Lee, Yu-Yu Chou, Yow-Bang Wang, Lin-Shan Lee |
| 2013 | ICASSP | Enhancing query expansion for semantic retrieval of spoken content with automatically discovered acoustic patterns. | Hung-yi Lee, Yun-Chiao Li, Cheng-Tao Chung, Lin-Shan Lee |
| 2013 | ICASSP | A dialogue game framework with personalized training using reinforcement learning for computer-assisted language learning. | Pei-hao Su, Yow-Bang Wang, Tien-han Yu, Lin-Shan Lee |
| 2013 | ICASSP | Toward unsupervised discovery of pronunciation error patterns using universal phoneme posteriorgram for computer-assisted language learning. | Yow-Bang Wang, Lin-Shan Lee |
| 2013 | ICASSP | Interactive spoken content retrieval by extended query model and continuous state space Markov Decision Process. | Tsung-Hsien Wen, Hung-yi Lee, Pei-hao Su, Lin-Shan Lee |
| 2013 | Interspeech | Supervised spoken document summarization based on structured support vector machine with utterance clusters as hidden variables. | Sz-Rung Shiang, Hung-yi Lee, Lin-Shan Lee |
| 2013 | Interspeech | A recursive dialogue game framework with optimal Policy offering personalized computer-assisted language learning. | Pei-hao Su, Yow-Bang Wang, Tsung-Hsien Wen, Tien-han Yu, Lin-Shan Lee |
| 2013 | Interspeech | Recurrent neural network based language model personalization by social network crowdsourcing. | Tsung-Hsien Wen, Aaron Heidel, Hung-yi Lee, Yu Tsao, Lin-Shan Lee |
| 2013 | Interspeech | Speaking rate normalization with lattice-based context-dependent phoneme duration modeling for personalized speech recognizers on mobile devices. | Ching-feng Yeh, Hung-yi Lee, Lin-Shan Lee |
| 2012 | ICASSP | Two-dimensional frame-and-feature weighted Viterbi decoding for robust speech recognition. | Yang Chang, Lin-Shan Lee |
| 2012 | ICASSP | Unsupervised two-stage keyword extraction from spoken documents by topic coherence and support vector machine. | Yun-Nung Chen, Yu Huang, Hung-yi Lee, Lin-Shan Lee |
| 2012 | ICASSP | Utterance-level latent topic transition modeling for spoken documents and its application in automatic summarization. | Hung-yi Lee, Yun-Nung Chen, Lin-Shan Lee |
| 2012 | ICASSP | Semantic query expansion and context-based discriminative term modeling for spoken document retrieval. | Tsung-wei Tu, Hung-yi Lee, Yu-Yu Chou, Lin-Shan Lee |
| 2012 | ICASSP | Improved approaches of modeling and detecting Error Patterns with empirical analysis for Computer-Aided Pronunciation Training. | Yow-Bang Wang, Lin-Shan Lee |
| 2012 | ICASSP | Recognition of highly imbalanced code-mixed bilingual speech with frame-level language detection based on blurred posteriorgram. | Ching-feng Yeh, Aaron Heidel, Hung-yi Lee, Lin-Shan Lee |
| 2012 | Interspeech | Discriminative Fuzzy Clustering Maximum a Posterior Linear Regression for Speaker Adaptation. | Ting-Yao Hu, Yu Tsao, Lin-Shan Lee |
| 2012 | Interspeech | Open-Vocabulary Retrieval of Spoken Content with Shorter/Longer Queries Considering Word/Subword-based Acoustic Feature Similarity. | Hung-yi Lee, Po-wei Chou, Lin-Shan Lee |
| 2012 | Interspeech | Supervised Spoken Document Summarization jointly Considering Utterance Importance and Redundancy by Structured Support Vector Machine. | Hung-yi Lee, Yu-Yu Chou, Yow-Bang Wang, Lin-Shan Lee |
| 2012 | Interspeech | Error Pattern Detection Integrating Generative and Discriminative Learning for Computer-Aided Pronunciation Training. | Yow-Bang Wang, Lin-Shan Lee |
| 2012 | Interspeech | Interactive Spoken Content Retrieval with Different Types of Actions Optimized By a Markov Decision Process. | Tsung-Hsien Wen, Hung-yi Lee, Lin-Shan Lee |
| 2011 | ASRU | Improved spoken term detection using support vector machines with acoustic and context features from pseudo-relevance feedback. | Tsung-wei Tu, Hung-yi Lee, Lin-Shan Lee |
| 2011 | ICASSP | Integrating frame-based and segment-based dynamic time warping for unsupervised spoken term detection with spoken queries. | Chun-an Chan, Lin-Shan Lee |
| 2011 | ICASSP | Improved spoken term detection with graph-based re-ranking in feature space. | Yun-Nung Chen, Chia-Ping Chen, Hung-yi Lee, Chun-an Chan, Lin-Shan Lee |
| 2011 | ICASSP | Improved spoken term detection using support vector machines based on lattice context consistency. | Hung-yi Lee, Tsung-wei Tu, Chia-Ping Chen, Chao-Yu Huang, Lin-Shan Lee |
| 2011 | ICASSP | Multi-stream spectro-temporal and cepstral features based on data-driven hierarchical phoneme clusters. | Shang-wen Li, Liang-Che Sun, Lin-Shan Lee |
| 2011 | ICASSP | Bilingual acoustic modeling with state mapping and three-stage adaptation for transcribing unbalanced code-mixed lectures. | Ching-feng Yeh, Liang-Che Sun, Chao-Yu Huang, Lin-Shan Lee |
| 2011 | Interspeech | Unsupervised Hidden Markov Modeling of Spoken Queries for Spoken Term Detection without Speech Recognition. | Chun-an Chan, Lin-Shan Lee |
| 2011 | Interspeech | Spoken Lecture Summarization by Random Walk over a Graph Constructed with Automatically Extracted Key Terms. | Yun-Nung Chen, Yu Huang, Ching-feng Yeh, Lin-Shan Lee |
| 2011 | Interspeech | Improved Tonal Language Speech Recognition by Integrating Spectro-Temporal Evidence and Pitch Information with Properly Chosen Tonal Acoustic Units. | Shang-wen Li, Yow-Bang Wang, Liang-Che Sun, Lin-Shan Lee |
| 2011 | Interspeech | Bilingual Acoustic Model Adaptation by Unit Merging on Different Levels and Cross-Level Integration. | Ching-feng Yeh, Chao-Yu Huang, Lin-Shan Lee |
| 2010 | ICASSP | An initial attempt to improve spoken term detection by learning optimal weights for different indexing features. | Yu-Hui Chen, Chia-Chen Chou, Hung-yi Lee, Lin-Shan Lee |
| 2010 | ICASSP | Integrating recognition and retrieval with user feedback: A new framework for spoken term detection. | Hung-yi Lee, Lin-Shan Lee |
| 2010 | ICASSP | An initial attempt for phoneme recognition using Structured Support Vector Machine (SVM). | Hao Tang, Chao-Hong Meng, Lin-Shan Lee |
| 2010 | Interspeech | Unsupervised spoken-term detection with spoken queries using segment-based dynamic time warping. | Chun-an Chan, Lin-Shan Lee |
| 2010 | Interspeech | Improved spoken term detection by feature space pseudo-relevance feedback. | Chia-Ping Chen, Hung-yi Lee, Ching-feng Yeh, Lin-Shan Lee |
| 2010 | Interspeech | Improved spoken term detection by discriminative training of acoustic models based on user relevance feedback. | Hung-yi Lee, Chia-Ping Chen, Ching-feng Yeh, Lin-Shan Lee |
| 2010 | Interspeech | Improved phoneme recognition by integrating evidence from spectro-temporal and cepstral features. | Shang-wen Li, Liang-Che Sun, Lin-Shan Lee |
| 2010 | Interspeech | Mandarin tone recognition using affine-invariant prosodic features and tone posteriorgram. | Yow-Bang Wang, Lin-Shan Lee |
| 2009 | ACL | Discriminative Lexicon Adaptation for Improved Character Accuracy - A New Direction in Chinese Language Modeling. | Yi-Cheng Pan, Lin-Shan Lee, Sadaoki Furui |
| 2009 | ASRU | Voice-based information retrieval - how far are we from the text-based information retrieval ? | Lin-Shan Lee, Yi-Cheng Pan |
| 2009 | ASRU | Spoken term detection from bilingual spontaneous speech using code-switched lattice-based structures for words and subword units. | Hung-yi Lee, Yueh-Lien Tang, Hao Tang, Lin-Shan Lee |
| 2009 | ICASSP | Improved clustered hierarchical tandem system with bottom-up processing. | Shuo-Yiin Chang, Lin-Shan Lee |
| 2009 | ICASSP | Latent semantic retrieval of personal photos with sparse user annotation by fused image/speech/text features. | Yi-Sheng Fu, Chia-Yu Wan, Lin-Shan Lee |
| 2009 | ICASSP | Learning on demand - course lecture distillation by information extraction and semantic structuring for spoken documents. | Sheng-yi Kong, Miao-ru Wu, Che-Kuang Lin, Yi-Sheng Fu, Lin-Shan Lee |
| 2009 | ICASSP | Improved lattice-based spoken document retrieval by directly learning from the evaluation measures. | Chao-Hong Meng, Hung-yi Lee, Lin-Shan Lee |
| 2009 | Interspeech | Mandarin spontaneous narrative planning - prosodic evidence from national taiwan university lecture corpus. | Chiu-yu Tseng, Zhao-yu Su, Lin-Shan Lee |
| 2008 | ICASSP | Context dependent quantization for distributed and/or robust speech recognition. | Chia-Yu Wan, Yi Chen, Lin-Shan Lee |
| 2008 | Interspeech | Data-driven clustered hierarchical tandem system for LVCSR. | Shuo-Yiin Chang, Lin-Shan Lee |
| 2008 | Interspeech | Improved large vocabulary Mandarin speech recognition by selectively using tone information with a two-stage prosodic model. | Li-Wei Cheng, Lin-Shan Lee |
| 2008 | Interspeech | Confusion-based entropy-weighted decoding for robust speech recognition. | Yi Chen, Chia-Yu Wan, Lin-Shan Lee |
| 2008 | Interspeech | Evaluation of modulation spectrum equalization techniques for large vocabulary robust speech recognition. | Liang-Che Sun, Chang-Wen Hsu, Lin-Shan Lee |
| 2007 | ASRU | Robust speech recognition by properly utilizing reliable frames and segments in corrupted signals. | Yi Chen, Chia-Yu Wan, Lin-Shan Lee |
| 2007 | ASRU | Robust topic inference for latent semantic language model adaptation. | Aaron Heidel, Lin-Shan Lee |
| 2007 | ASRU | Analytical comparison between position specific posterior lattices and confusion networks based on words and subword units for spoken document indexing. | Yi-Cheng Pan, Hung-lin Chang, Lin-Shan Lee |
| 2007 | ASRU | Type-II dialogue systems for information access from unstructured knowledge sources. | Yi-Cheng Pan, Lin-Shan Lee |
| 2007 | ASRU | Modulation spectrum equalization for robust speech recognition. | Liang-Che Sun, Chang-Wen Hsu, Lin-Shan Lee |
| 2007 | ICASSP | Pronunciation Modeling for Spontaneous Speech Recognition using Latent Pronunciation Analysis (LPA) and Prior Knowledge. | Che-Kuang Lin, Lin-Shan Lee |
| 2007 | ICASSP | Three-Stage Error Concealment for Distributed Speech Recognition (DSR) with Histogram-Based Quantization (HQ) Under Noisy Environment. | Chia-Yu Wan, Yi Chen, Lin-Shan Lee |
| 2007 | Interspeech | Language model adaptation using latent dirichlet allocation and an efficient topic inference algorithm. | Aaron Heidel, Hung-An Chang, Lin-Shan Lee |
| 2007 | Interspeech | Extended powered cepstral normalization (p-CN) with range equalization for robust features in speech recognition. | Chang-Wen Hsu, Lin-Shan Lee |
| 2007 | Interspeech | Subword-based position specific posterior lattices (s-PSPL) for indexing speech information. | Yi-Cheng Pan, Hung-lin Chang, Berlin Chen, Lin-Shan Lee |
| 2007 | Interspeech | Lexicon adaptation with reduced character error (LARCE) - a new direction in Chinese language modeling. | Yi-Cheng Pan, Lin-Shan Lee |
| 2006 | ICASSP | Entropy-Based Feature Parameter Weighting for Robust Speech Recognition. | Yi Chen, Chia-Yu Wan, Lin-Shan Lee |
| 2006 | ICASSP | Improved Spoken Document Retrieval With Dynamic Key Term Lexicon and Probabilistic Latent Semantic Analysis (PLSA). | Ya-chao Hsieh, Yu-tsun Huang, Chien-Chih Wang, Lin-Shan Lee |
| 2006 | ICASSP | Improved Spoken Document Summarization Using Probabilistic Latent Semantic Analysis (PLSA). | Sheng-yi Kong, Lin-Shan Lee |
| 2006 | ICASSP | Joint Uncertainty Decoding (JUD) with Histogram-Based Quantization (HQ) for Robust and/or Distributed Speech Recognition. | Chia-Yu Wan, Lin-Shan Lee |
| 2006 | Interspeech | A new framework for system combination based on integrated hypothesis space. | I-Fan Chen, Lin-Shan Lee |
| 2006 | Interspeech | Extension and further analysis of higher order cepstral moment normalization (HOCMN) for robust features in speech recognition. | Chang-Wen Hsu, Lin-Shan Lee |
| 2006 | Interspeech | Powered cepstral normalization (p-CN) for robust features in speech recognition. | Chang-Wen Hsu, Lin-Shan Lee |
| 2006 | Interspeech | Prosodic modeling in large vocabulary Mandarin speech recognition. | Jui-Ting Huang, Lin-Shan Lee |
| 2006 | Interspeech | Feature analysis for emotion recognition from Mandarin speech considering the special characteristics of Chinese language. | Yi-Hao Kao, Lin-Shan Lee |
| 2006 | Interspeech | Multi-layered summarization of spoken document archives by information extraction and semantic structuring. | Lin-Shan Lee, Sheng-yi Kong, Yi-Cheng Pan, Yi-Sheng Fu, Yu-tsun Huang |
| 2006 | Interspeech | Latent prosodic modeling (LPM) for speech with applications in recognizing spontaneous Mandarin speech with disfluencies. | Che-Kuang Lin, Lin-Shan Lee |
| 2006 | Interspeech | Efficient interactive retrieval of spoken documents with key terms ranked by reinforcement learning. | Yi-Cheng Pan, Jia-Yu Chen, Yen-shin Lee, Yi-Sheng Fu, Lin-Shan Lee |
| 2005 | Interspeech | Energy-based frame selection for reliable feature normalization and transformation in robust speech recognition. | Yi Chen, Lin-Shan Lee |
| 2005 | Interspeech | Hierarchical topic organization and visual presentation of spoken documents using probabilistic latent semantic analysis (PLSA) for efficient retrieval/browsing applications. | Te-Hsuan Li, Ming-Han Lee, Berlin Chen, Lin-Shan Lee |
| 2005 | Interspeech | Improved spontaneous Mandarin speech recognition by disfluency interruption point (IP) detection using prosodic features. | Che-Kuang Lin, Lin-Shan Lee |
| 2005 | Interspeech | Histogram-based quantization (HQ) for robust and scalable distributed speech recognition. | Chia-Yu Wan, Lin-Shan Lee |
| 2004 | ICASSP | Efficient and robust distributed speech recognition (DSR) over wireless fading channels: 2D-DCT compression, iterative bit allocation, short BCH code and interleaving. | Wei-Hao Hsu, Lin-Shan Lee |
| 2004 | ICASSP | Higher order cepstral moment normalization (HOCMN) for robust speech recognition. | Chang-Wen Hsu, Lin-Shan Lee |
| 2004 | Interspeech | Improved speech enhancement by applying time-shift property of DFT on hankel matrices for signal subspace decomposition. | Gwo-hwa Ju, Lin-Shan Lee |
| 2004 | Interspeech | A new feature extraction front-end for robust speech recognition using progressive histogram equalization and multi-eigenvector temporal filtering. | Shang-nien Tsai, Lin-Shan Lee |
| 2003 | ICASSP | Data-driven temporal filters based on multi-eigenvectors for robust features in speech recognition. | Ni-Chun Wang, Jeih-Weih Hung, Lin-Shan Lee |
| 2003 | Interspeech | Improved Chinese broadcast news transcription by language modeling with temporally consistent training corpora and iterative phrase extraction. | Pi-Chuan Chang, Shuo-Peng Liao, Lin-Shan Lee |
| 2003 | Interspeech | Automatic title generation for Chinese spoken documents using an adaptive k nearest-neighbor approach. | Shun-Chuan Chen, Lin-Shan Lee |
| 2003 | Interspeech | Perceptually-constrained generalized singular value decomposition-based approach for enhancing speech corrupted by colored noise. | Gwo-hwa Ju, Lin-Shan Lee |
| 2003 | Interspeech | Speech enhancement and improved recognition accuracy by integrating wavelet transform and spectral subtraction algorithm. | Gwo-hwa Ju, Lin-Shan Lee |
| 2003 | Interspeech | Automatic title generation for Chinese spoken documents considering the special structure of the language. | Lin-Shan Lee, Shun-Chuan Chen |
| 2003 | Interspeech | Cross domain Chinese speech understanding and answering based on named-entity extraction. | Yun-Tien Lee, Shun-Chuan Chen, Lin-Shan Lee |
| 2003 | Interspeech | Why is the special structure of the language important for Chinese spoken language processing? - examples on spoken document retrieval, segmentation and summarization. | Lin-Shan Lee, Yuan Ho, Jia-fu Chen, Shun-Chuan Chen |
| 2002 | ICASSP | Data-driven temporal filters for robust features in speech recognition obtained via Minimum Classification Error (MCE). | Jeih-Weih Hung, Lin-Shan Lee |
| 2002 | Interspeech | Data-driven temporal filters obtained via different optimization criteria evaluated on Aurora2 database. | Jeih-Weih Hung, Lin-Shan Lee |
| 2002 | Interspeech | Speech enhancement based on generalized singular value decomposition approach. | Gwo-hwa Ju, Lin-Shan Lee |
| 2002 | Interspeech | Distributed Chinese keyword spotting and verification for spoken dialogues under wireless environment. | Yun-Tien Lee, Cheng-Huang Wu, Yumin Lee, Lin-Shan Lee |
| 2002 | Interspeech | Improved Chinese spoken document retrieval with hybrid modeling and data-driven indexing features. | Chun-Jen Wang, Berlin Chen, Lin-Shan Lee |
| 2001 | ICASSP | Rapid speaker adaptation using a priori knowledge by eigenspace analysis of MLLR parameters. | Nick J.-C. Wang, Sammy S.-M. Lee, Frank Seide, Lin-Shan Lee |
| 2001 | Interspeech | Credibility proof for speech content and speaker verification by fragile watermarking with consecutive frame-based processing. | Yiou-Wen Cheng, Lin-Shan Lee |
| 2001 | Interspeech | Improved spoken document retrieval by exploring extra acoustic and linguistic cues. | Berlin Chen, Hsin-Min Wang, Lin-Shan Lee |
| 2001 | Interspeech | An HMM/n-gram-based linguistic processing approach for Mandarin spoken document retrieval. | Berlin Chen, Hsin-Min Wang, Lin-Shan Lee |
| 2001 | Interspeech | Comparative analysis for data-driven temporal filters obtained via principal component analysis (PCA) and linear discriminant analysis (LDA) in speech recognition. | Jeih-Weih Hung, Hsin-Min Wang, Lin-Shan Lee |
| 2001 | Interspeech | Pronunciation variation analysis with respect to various linguistic levels and contextual conditions for Mandarin Chinese. | Ming-Yi Tsai, Fu-Chiang Chou, Lin-Shan Lee |
| 2001 | Interspeech | Segmental eigenvoice for rapid speaker adaptation. | Yu Tsao, Shang-Ming Lee, Fu-Chiang Chou, Lin-Shan Lee |
| 2001 | Interspeech | Eigen-MLLR coefficients as new feature parameters for speaker identification. | Nick J.-C. Wang, Wei-Ho Tsai, Lin-Shan Lee |
| 2000 | ICASSP | Retrieval of broadcast news speech in Mandarin Chinese collected in Taiwan using syllable-level statistical characteristics. | Berlin Chen, Hsin-Min Wang, Lin-Shan Lee |
| 2000 | ICASSP | Fundamental performance analysis for spoken dialogue systems based on a quantitative simulation approach. | Bor-Shen Lin, Lin-Shan Lee |
| 2000 | Interspeech | Fast speaker adaptation using eigenspace-based maximum likelihood linear regression. | Kuan-Ting Chen, Wen-Wei Liau, Hsin-Min Wang, Lin-Shan Lee |
| 2000 | Interspeech | Retrieval of mandarin broadcast news using spoken queries. | Berlin Chen, Hsin-Min Wang, Lin-Shan Lee |
| 2000 | Interspeech | Automatic metric-based speech segmentation for broadcast news via principal component analysis. | Jeih-Weih Hung, Hsin-Min Wang, Lin-Shan Lee |
| 2000 | Interspeech | MAT-2000 - design, collection, and validation of a Mandarin 2000-speaker telephone speech database. | Hsiao-Chuan Wang, Frank Seide, Chiu-yu Tseng, Lin-Shan Lee |
| 2000 | LREC | Live Lexicons and Dynamic Corpora Adapted to the Network Resources for Chinese Spoken Language Processing Applications in an Internet Era. | Lin-Shan Lee, Lee-Feng Chien |
| 1999 | ICASSP | Improved parallel model combination techniques with split Gaussian mixtures for speech recognition under noisy conditions. | Jeih-weih Hung, Jia-Lin Shen, Lin-Shan Lee |
| 1999 | ICASSP | A framework of performance evaluation and error analysis methodology for speech understanding systems. | Bor-Shen Lin, Lin-Shan Lee |
| 1999 | Interspeech | Selection of waveform units for corpus-based Mandarin speech synthesis based on decision trees and prosodic modification costs. | Fu-Chiang Chou, Chiu-yu Tseng, Lin-Shan Lee |
| 1999 | Interspeech | Phonetic state tied-mixture tone modeling for large vocabulary continuous Mandarin speech recognition. | Tai-Hsuan Ho, Chin-Jung Liu, Herman Sun, Ming-Yi Tsai, Lin-Shan Lee |
| 1999 | Interspeech | Consistent dialogue across concurrent topics based on an expert system model. | Bor-Shen Lin, Hsin-Min Wang, Lin-Shan Lee |
| 1998 | ICASSP | Improved search strategy for large vocabulary continuous Mandarin speech recognition. | Tai-Hsuan Ho, Kae-Cherng Yang, Kuo-Hsun Huang, Lin-Shan Lee |
| 1998 | ICASSP | Improved robustness for speech recognition under noisy conditions using correlated parallel model combination. | Jeih-Weih Hung, Jia-Lin Shen, Lin-Shan Lee |
| 1998 | ICASSP | Statistics-based segment pattern lexicon-a new direction for Chinese language modeling. | Kae-Cherng Yang, Tai-Hsuan Ho, Lee-Feng Chien, Lin-Shan Lee |
| 1998 | Interspeech | A*-admissible key-phrase spotting with sub-syllable level utterance verification. | Berlin Chen, Hsin-Min Wang, Lee-Feng Chien, Lin-Shan Lee |
| 1998 | Interspeech | Automatic segmental and prosodic labeling of Mandarin speech database. | Fu-Chiang Chou, Chiu-yu Tseng, Lin-Shan Lee |
| 1998 | Interspeech | Improved parallel model combination based on better domain transformation for speech recognition under noisy environments. | Jeih-Weih Hung, Jia-Lin Shen, Lin-Shan Lee |
| 1998 | Interspeech | Hierarchical tag-graph search for spontaneous speech understanding in spoken dialog systems. | Bor-Shen Lin, Berlin Chen, Hsin-Min Wang, Lin-Shan Lee |
| 1998 | Interspeech | Improved robust speech recognition considering signal correlation approximated by taylor series. | Jia-Lin Shen, Jeih-Weih Hung, Lin-Shan Lee |
| 1998 | Interspeech | Robust entropy-based endpoint detection for speech recognition in noisy environments. | Jia-Lin Shen, Jeih-Weih Hung, Lin-Shan Lee |
| 1998 | Interspeech | A syllable-based Chinese spoken dialogue system for telephone directory services primarily trained with a corpus. | Yen-Ju Yang, Lin-Shan Lee |
| 1997 | ICASSP | A multi-phase approach for fast spotting of large vocabulary Chinese keywords from Mandarin speech using prosodic information. | Bo-Ren Bai, Chiu-yu Tseng, Lin-Shan Lee |
| 1997 | ICASSP | Internet Chinese information retrieval using unconstrained Mandarin speech queries based on a client-server architecture and a PAT-tree-based language model. | Lee-Feng Chien, Sung-Chien Lin, Jenn-Chau Hong, Ming-Chiuan Chen, Hsin-Min Wang, Jia-Lin Shen, Keh-Jiann Chen, Lin-Shan Lee |
| 1997 | ICASSP | A Chinese text-to-speech system based on part-of-speech analysis, prosodic modeling and non-uniform units. | Fu-Chiang Chou, Chiu-yu Tseng, Keh-Jiann Chen, Lin-Shan Lee |
| 1997 | ICASSP | Syllable-based relevance feedback techniques for Mandarin voice record retrieval using speech queries. | Lin-Shan Lee, Bo-Ren Bai, Lee-Feng Chien |
| 1997 | Interspeech | Intelligent retrieval of very large Chinese dictionaries with speech queries. | Sung-Chien Lin, Lee-Feng Chien, Ming-Chiuan Chen, Lin-Shan Lee, Keh-Jiann Chen |
| 1997 | Interspeech | Chinese language model adaptation based on document classification and multiple domain-specific language models. | Sung-Chien Lin, Chi-Lung Tsai, Lee-Feng Chien, Keh-Jiann Chen, Lin-Shan Lee |
| 1996 | ICASSP | An efficient voice retrieval system for very-large-vocabulary Chinese textual databases with a clustered language model. | Sung-Chien Lin, Lee-Feng Chien, Keh-Jiann Chen, Lin-Shan Lee |
| 1996 | ICASSP | Fast and accurate recognition of very-large-vocabulary continuous Mandarin speech for Chinese language with improved segmental probability modeling. | Jia-Lin Shen, Lin-Shan Lee |
| 1996 | Interspeech | Very-large-vocabulary Mandarin voice message file retrieval using speech queries. | Bo-Ren Bai, Lee-Feng Chien, Lin-Shan Lee |
| 1996 | Interspeech | Automatic generation of prosodic structure for high quality Mandarin speech synthesis. | Fu-Chiang Chou, Chiu-yu Tseng, Lin-Shan Lee |
| 1996 | Interspeech | Use of prosodic information to integrate acoustic and linguistic knowledge in continuous Mandarin speech recognition with very large vocabulary. | Hung-Yun Hsieh, Ren-Yuan Lyu, Lin-Shan Lee |
| 1996 | Interspeech | Robust speech recognition features based on temporal trajectory filtering of frequency band spectrum. | Jia-Lin Shen, Wen-Liang Hwang, Lin-Shan Lee |
| 1996 | Interspeech | Speaker intention modeling for large vocabulary Mandarin spoken dialogues. | Yen-Ju Yang, Lee-Feng Chien, Lin-Shan Lee |
| 1995 | ICASSP | Golden Mandarin (III)-a user-adaptive prosodic-segment-based Mandarin dictation machine for Chinese language with very large vocabulary. | Ren-Yuan Lyu, Lee-Feng Chien, Shiao-Hong Hwang, Hung-Yun Hsieh, Rung-Chiuan Yang, Bo-Ren Bai, Jia-Chi Weng, Yen-Ju Yang, Shi-Wei Lin, Keh-Jiann Chen, Chiu-yu Tseng, Lin-Shan Lee |
| 1995 | ICASSP | Complete recognition of continuous Mandarin speech for Chinese language with very large vocabulary but limited training data. | Hsin-Min Wang, Jia-Lin Shen, Yen-Ju Yang, Chiu-yu Tseng, Lin-Shan Lee |
| 1995 | Interspeech | Large vocabulary, word-based Mandarin dictation system. | Jung-Kuei Chen, Lin-Shan Lee, Frank K. Soong |
| 1995 | Interspeech | Fast and accurate continuous speech recognition for Chinese language with very large vocabulary. | Tai-Hsuan Ho, Hsin-Min Wang, Lee-Feng Chien, Keh-Jiann Chen, Lin-Shan Lee |
| 1995 | Interspeech | A syllable-based very-large-vocabulary voice retrieval system for Chinese databases with textual attributes. | Sung-Chien Lin, Lee-Feng Chien, Keh-Jiann Chen, Lin-Shan Lee |
| 1995 | Interspeech | Unconstrained speech retrieval for Chinese document databases with very large vocabulary and unlimited domains. | Sung-Chien Lin, Lee-Feng Chien, Keh-Jiann Chen, Lin-Shan Lee |
| 1995 | Interspeech | A chernoff distance based segmental probability model (CD-SPM) approach for Mandarin syllable recognition. | Jia-Lin Shen, Lin-Shan Lee |
| 1994 | ICASSP | Large vocabulary word recognition based on tree-trellis search. | Jung-Kuei Chen, Frank K. Soong, Lin-Shan Lee |
| 1994 | ICASSP | An initial study on a segmental probability model approach to large-vocabulary continuous Mandarin speech recognition. | Jia-Lin Shen, Hsin-Min Wang, Bo-Ren Bai, Lin-Shan Lee |
| 1994 | Interspeech | Incremental speaker adaptation using phonetically balanced training sentences for Mandarin syllable recognition based on segmental probability models. | Jia-Lin Shen, Hsin-Min Wang, Ren-Yuan Lyu, Lin-Shan Lee |
| 1994 | Interspeech | An intelligent and efficient word-class-based Chinese language model for Mandarin speech recognition with very large vocabulary. | Yen-Ju Yang, Sung-Chien Lin, Lee-Feng Chien, Keh-Jiann Chen, Lin-Shan Lee |
| 1993 | ICASSP | Golden Mandarin (II)-an improved single-chip real-time Mandarin dictation machine for Chinese language with very large vocabulary. | Lin-Shan Lee, Chiu-yu Tseng, Keh-Jiann Chen, I-Jung Hung, Ming-Yu Lee, Lee-Feng Chien, Yumin Lee, Ren-Yuan Lyu, Hsin-Min Wang, Yung-Chuan Wu, Tung-Sheng Lin, Hung-Yan Gu, Chi-ping Nee, Chun-Yi Liao, Yeng-Ju Yang, Yuan-Cheng Chang, Rung-Chiung Yang |
| 1993 | ICASSP | A new framework for recognition of Mandarin syllables with tones using sub-syllabic units. | Chih-Heng Lin, Lin-Shan Lee, Pei-Yih Ting |
| 1991 | ACL | A Preference-first Language Processor Integrating the Unification Grammar and Markov Language Model for Speech Recognition Applications. | Lee-Feng Chien, Keh-Jiann Chen, Lin-Shan Lee |
| 1990 | COLING | An Augmented Chart Data Structure with Efficient Word Lattice Parsing Scheme In Speech Recognition Applications. | Lee-Feng Chien, Keh-Jiann Chen, Lin-Shan Lee |
| 1990 | ICASSP | An augmented chart parsing algorithm integrating unification grammar and Markov language model for continuous speech recognition. | Lee-Feng Chien, Keh-Jiann Chen, Lin-Shan Lee |
| 1990 | ICASSP | A real-time Mandarin dictation machine for Chinese language with unlimited texts and very large vocabulary. | Lin-Shan Lee, Chiu-yu Tseng, Hung-Yan Gu, Fu-hua Liu, Robert Chen-Hao Chang, Shew-Heng Hsieh, Chian-hung Chen |
| 1988 | ICASSP | New speech recognition approaches based upon finite state vector quantization with structural constraints. | Pei-Yih Ting, Chiu-yu Tseng, Lin-Shan Lee |
| 1987 | IJCAI | The Preliminary Results of a Mandarin Dictation Machine Based Upon Chinese Natural Language Analysis. | Lin-Shan Lee, Chiu-yu Tseng, Keh-Jiann Chen, James Huang |
| 1986 | AAAI | A Chinese Natural Language Processing System Based Upon the Theory of Empty Categories. | Long Ji Lin, Lin-Shan Lee, James Huang, Keh-Jiann Chen |
| 1986 | ICASSP | A Chinese text-to-speech system based upon a syllable concatenation model. | Ming Ouhyoung, Chin-jiang Shie, Chiu-yu Tseng, Lin-Shan Lee |
| 1981 | CRYPTO | Results on Sampling-based Scrambling for Secure Speech Communication. | Lin-Shan Lee, Ger-Chih Chou |