| 2013 | Hybrid speech recognition with Deep Bidirectional LSTM. | Alex Graves, Navdeep Jaitly, Abdel-rahman Mohamed |
| 2013 | DNN acoustic modeling with modular multi-lingual feature extraction networks. | Jonas Gehring, Quoc Bao Nguyen, Florian Metze, Alex Waibel |
| 2013 | Using web text to improve keyword spotting in speech. | Ankur Gandhe, Long Qin, Florian Metze, Alexander I. Rudnicky, Ian R. Lane, Matthias Eck |
| 2013 | ASR for electro-laryngeal speech. | Anna Katharina Fuchs, Juan Andres Morales-Cordovilla, Martin Hagmller |
| 2013 | Expert-based reward shaping and exploration scheme for boosting policy learning of dialogue management. | Emmanuel Ferreira, Fabrice Lefvre |
| 2013 | Lightly supervised automatic subtitling of weather forecasts. | Joris Driesen, Steve Renals |
| 2013 | Combining stochastic average gradient and Hessian-free optimization for sequence training of deep neural networks. | Pierre L. Dognin, Vaibhava Goel |
| 2013 | Porting concepts from DNNs back to GMMs. | Kris Demuynck, Fabian Triefenbach |
| 2013 | Barge-in effects in Bayesian dialogue act recognition and simulation. | Heriberto Cuayhuitl, Nina Dethlefs, Helen Wright Hastie, Oliver Lemon |
| 2013 | Using proxies for OOV keywords in the keyword search task. | Guoguo Chen, Oguz Yilmaz, Jan Trmal, Daniel Povey, Sanjeev Khudanpur |
| 2013 | Unsupervised induction and filling of semantic slots for spoken dialogue systems using frame-semantic parsing. | Yun-Nung Chen, William Yang Wang, Alexander I. Rudnicky |
| 2013 | Effective pseudo-relevance feedback for language modeling in speech recognition. | Berlin Chen, Yi-Wen Chen, Kuan-Yu Chen, Ea-Ee Jan |
| 2013 | Deep maxout neural networks for speech recognition. | Meng Cai, Yongzhe Shi, Jia Liu |
| 2013 | Hierarchical neural networks and enhanced class posteriors for social signal classification. | Raymond Brueckner, Bjrn W. Schuller |
| 2013 | On-line adaptation of semantic models for spoken language understanding. | Ali Orkan Bayer, Giuseppe Riccardi |
| 2013 | Vector Taylor series based HMM adaptation for generalized cepstrum in noisy environment. | Soonho Baek, Hong-Goo Kang |
| 2013 | A propagation approach to modelling the joint distributions of clean and corrupted speech in the Mel-Cepstral domain. | Ramn Fernandez Astudillo |
| 2011 | Linear versus mel frequency cepstral coefficients for speaker recognition. | Xinhui Zhou, Daniel Garcia-Romero, Ramani Duraiswami, Carol Y. Espy-Wilson, Shihab A. Shamma |
| 2011 | Speaker adaptation based on speaker-dependent eigenphone estimation. | Wen-Lin Zhang, Wei-Qiang Zhang, Bi-Cheng Li |
| 2011 | Unsupervised learning in cross-corpus acoustic emotion recognition. | Zixing Zhang, Felix Weninger, Martin Wllmer, Bjrn W. Schuller |
| 2011 | Analyzing conversations using rich phrase patterns. | Bin Zhang, Alex Marin, Brian Hutchinson, Mari Ostendorf |
| 2011 | Detection-based accented speech recognition using articulatory features. | Chao Zhang, Yi Liu, Chin-Hui Lee |
| 2011 | Extending noise robust structured support vector machines to larger vocabulary tasks. | Shi-Xiong Zhang, Mark J. F. Gales |
| 2011 | Evaluating prosodic features for automated scoring of non-native read speech. | Klaus Zechner, Xiaoming Xi, Lei Chen |
| 2011 | Automatic detection of "g-dropping" in American English using forced alignment. | Jiahong Yuan, Mark Y. Liberman |