| 2015 | Detecting actionable items in meetings by convolutional deep structured semantic models. | Yun-Nung Chen, Dilek Hakkani-Tr, Xiaodong He |
| 2015 | Sparse non-negative matrix language modeling for geo-annotated query session data. | Ciprian Chelba, Noam Shazeer |
| 2015 | Discriminative training of context-dependent language model scaling factors and interpolation weights. | Shuangyu Chang, Abhik Lahiri, Issac Alphonso, Barlas Oguz, Michael Levit, Benot Dumoulin |
| 2015 | A universal model for flexible item selection in conversational dialogs. | Asli Celikyilmaz, Zhaleh Feizollahi, Dilek Hakkani-Tr, Ruhi Sarikaya |
| 2015 | High-performance Swahili keyword search with very limited language pack: The THUEE system for the OpenKWS15 evaluation. | Meng Cai, Zhiqiang Lv, Cheng Lu, Jian Kang, Like Hui, Zhuo Zhang, Jia Liu |
| 2015 | Spoken language translation graphs re-decoding using automatic quality assessment. | Laurent Besacier, Benjamin Lecouteux, Ngoc-Quang Luong, Ngoc-Tien Le |
| 2015 | The MGB challenge: Evaluating multi-genre broadcast media recognition. | Peter Bell, Mark J. F. Gales, Thomas Hain, Jonathan Kilgour, Pierre Lanchantin, Xunying Liu, Andrew McParland, Steve Renals, Oscar Saz, Mirjam Wester, Philip C. Woodland |
| 2015 | The third 'CHiME' speech separation and recognition challenge: Dataset, task and baselines. | Jon Barker, Ricard Marxer, Emmanuel Vincent, Shinji Watanabe |
| 2015 | Open-domain personalized dialog system using user-interested topics in system responses. | Jeesoo Bang, Sangdo Han, Kyusong Lee, Gary Geunbae Lee |
| 2015 | Combining spectral feature mapping and multi-channel model-based source separation for noise-robust automatic speech recognition. | Deblin Bagchi, Michael I. Mandel, Zhongqiu Wang, Yanzhang He, Andrew R. Plummer, Eric Fosler-Lussier |
| 2015 | Multi-reference WER for evaluating ASR for languages with no orthographic rules. | Ahmed M. Ali, Walid Magdy, Peter Bell, Steve Renals |
| 2015 | A system for automatic alignment of broadcast media captions using weighted finite-state transducers. | Peter Bell, Steve Renals |
| 2013 | Compact acoustic modeling based on acoustic manifold using a mixture of factor analyzers. | Wen-Lin Zhang, Bi-Cheng Li, Wei-Qiang Zhang |
| 2013 | Convolutional neural network based triangular CRF for joint intent detection and slot filling. | Puyang Xu, Ruhi Sarikaya |
| 2013 | The TAO of ATWV: Probing the mysteries of keyword search performance. | Steven Wegmann, Arlo Faria, Adam Janin, Korbinian Riedhammer, Nelson Morgan |
| 2013 | Context-dependent modelling of deep neural network using logistic regression. | Guangsen Wang, Khe Chai Sim |
| 2013 | A hierarchical system for word discovery exploiting DTW-based initialization. | Oliver Walter, Timo Korthals, Reinhold Haeb-Umbach, Bhiksha Raj |
| 2013 | The second 'CHiME' speech separation and recognition challenge: An overview of challenge systems and outcomes. | Emmanuel Vincent, Jon Barker, Shinji Watanabe, Jonathan Le Roux, Francesco Nesta, Marco Matassoni |
| 2013 | Semi-supervised training of Deep Neural Networks. | Karel Vesel, Mirko Hannemann, Luks Burget |
| 2013 | Learning a subword vocabulary based on unigram likelihood. | Matti Varjokallio, Mikko Kurimo, Sami Virpioja |
| 2013 | An SVD-based scheme for MFCC compression in distributed speech recognition system. | Azzedine Touazi, Mohamed Debyeche |
| 2013 | A generalized discriminative training framework for system combination. | Yuuki Tachioka, Shinji Watanabe, Jonathan Le Roux, John R. Hershey |
| 2013 | Hybrid acoustic models for distant and multichannel large vocabulary speech recognition. | Pawel Swietojanski, Arnab Ghoshal, Steve Renals |
| 2013 | Semantic entity detection from multiple ASR hypotheses within the WFST framework. | Jan Svec, Pavel Ircing, Lubos Smdl |
| 2013 | Automatic model complexity control for generalized variable parameter HMMs. | Rongfeng Su, Xunying Liu, Lan Wang |