| 2009 | From speech to letters - using a novel neural network architecture for grapheme based ASR. | Florian Eyben, Martin Wllmer, Bjrn W. Schuller, Alex Graves |
| 2009 | Local and global models for spontaneous speech segment detection and characterization. | Richard Dufour, Yannick Estve, Paul Delglise, Frdric Bchet |
| 2009 | Automatic detection of vowel pronunciation errors using multiple information sources. | Joost van Doremalen, Catia Cucchiarini, Helmer Strik |
| 2009 | Transition features for CRF-based speech recognition and boundary detection. | Spiros Dimopoulos, Eric Fosler-Lussier, Chin-Hui Lee, Alexandros Potamianos |
| 2009 | Iterative decoding: A novel re-scoring framework for confusion networks. | Anoop Deoras, Frederick Jelinek |
| 2009 | The ESAT 2008 system for N-Best Dutch speech recognition benchmark. | Kris Demuynck, Antti Puurula, Dirk Van Compernolle, Patrick Wambacq |
| 2009 | Improving online incremental speaker adaptation with eigen feature space MLLR. | Xiaodong Cui, Jian Xue, Bowen Zhou |
| 2009 | Reinforcing language model for speech translation with auxiliary data. | Jia Cui, Yonggang Deng, Bowen Zhou |
| 2009 | Scaling shrinkage-based language models. | Stanley F. Chen, Lidia Mangu, Bhuvana Ramabhadran, Ruhi Sarikaya, Abhinav Sethy |
| 2009 | Large-margin feature adaptation for automatic speech recognition. | Chih-Chieh Cheng, Fei Sha, Lawrence K. Saul |
| 2009 | Improved vocabulary independent search with approximate match based on Conditional Random Fields. | Upendra V. Chaudhari, Michael Picheny |
| 2009 | Articulatory feature detection with Support Vector Machines for integration into ASR and phone recognition. | Upendra V. Chaudhari, Michael Picheny |
| 2009 | Online discriminative learning: theory and applications. | Nicol Cesa-Bianchi |
| 2009 | Any questions? Automatic question detection in meetings. | Kofi Boakye, Benot Favre, Dilek Hakkani-Tr |
| 2009 | Diagonal priors for full covariance speech recognition. | Peter Bell, Simon King |
| 2009 | Topic-based speaker recognition for German parliamentary speeches. | Doris Baum |
| 2009 | Comparing automatic rich transcription for Portuguese, Spanish and English Broadcast News. | Fernando Batista, Isabel Trancoso, Nuno J. Mamede |
| 2009 | Lattice-based lexical cues for word fragment detection in conversational speech. | Kartik Audhkhasi, Panayiotis G. Georgiou, Shrikanth S. Narayanan |
| 2009 | Extended Minimum Classification Error Training in Voice Activity Detection. | Takayuki Arakawa, Haitham Al-Hassanieh, Masanori Tsujikawa, Ryosuke Isotani |
| 2009 | Manipulation of consonants in natural speech. | Jont B. Allen, Feipeng Li |
| 2009 | An improved perceptual speech enhancement technique employing a psychoacoustically motivated weighting factor. | Md. Jahangir Alam, Sid-Ahmed Selouani, Douglas D. O'Shaughnessy |
| 2009 | Pronunciation modeling for dialectal arabic speech recognition. | Hassan Al-Haj, Roger Hsiao, Ian R. Lane, Alan W. Black, Alex Waibel |
| 2007 | Improving lecture speech summarization using rhetorical information. | Justin Jian Zhang, Ricky Ho Yin Chan, Pascale Fung |
| 2007 | Interpolation of lost speech segments using LP-HNM model with codebook-mapping post-processing. | Esfandiar Zavarehei, Saeed Vaseghi |
| 2007 | A compact semidefinite programming (SDP) formulation for large margin estimation of HMMS in speech recognition. | Yan Yin, Hui Jiang |