| 2015 | Topic-space based setup of a neural network for theme identification of highly imperfect transcriptions. | Mohamed Morchid, Richard Dufour, Georges Linars |
| 2015 | RNNDROP: A novel dropout for RNNS in ASR. | Taesup Moon, Heeyoul Choi, Hoshik Lee, Inchul Song |
| 2015 | Deep bi-directional recurrent networks over spectral windows. | Abdel-rahman Mohamed, Frank Seide, Dong Yu, Jasha Droppo, Andreas Stolcke, Geoffrey Zweig, Gerald Penn |
| 2015 | Adaptive selection from multiple response candidates in example-based dialogue. | Masahiro Mizukami, Hideaki Kizuki, Toshio Nomura, Graham Neubig, Koichiro Yoshino, Sakriani Sakti, Tomoki Toda, Satoshi Nakamura |
| 2015 | Improving robustness against reverberation for automatic speech recognition. | Vikramjit Mitra, Julien van Hout, Wen Wang, Martin Graciarena, Mitchell McLaren, Horacio Franco, Dimitra Vergyri |
| 2015 | Time-frequency convolutional networks for robust speech recognition. | Vikramjit Mitra, Horacio Franco |
| 2015 | The 2015 sheffield system for longitudinal diarisation of broadcast media. | Rosanna Milner, Oscar Saz, Salil Deena, Mortaza Doulaty, Raymond W. M. Ng, Thomas Hain |
| 2015 | EESEN: End-to-end speech recognition using deep RNN models and WFST-based decoding. | Yajie Miao, Mohammad Gowayyed, Florian Metze |
| 2015 | Analysis of factors affecting system performance in the ASpIRE challenge. | Jennifer Melot, Nicolas Malyska, Jessica Ray, Wade Shen |
| 2015 | Utterance classification in speech-to-speech translation for zero-resource languages in the hospital administration domain. | Lara J. Martin, Andrew Wilkinson, Sai Sumanth Miryala, Vivian Robison, Alan W. Black |
| 2015 | Exploiting synchrony spectra and deep neural networks for noise-robust automatic speech recognition. | Ning Ma, Ricard Marxer, Jon Barker, Guy J. Brown |
| 2015 | Uncertainty estimation of DNN classifiers. | Sri Harish Reddy Mallidi, Tetsuji Ogawa, Hynek Hermansky |
| 2015 | Improved system fusion for keyword search. | Zhiqiang Lv, Meng Cai, Cheng Lu, Jian Kang, Like Hui, Wei-Qiang Zhang, Jia Liu |
| 2015 | Naturalness and rapport in a pitch adaptive learning companion. | Nichola Lubold, Heather Pon-Barry, Erin Walker |
| 2015 | A study of social-affective communication: Automatic prediction of emotion triggers and responses in television talk shows. | Nurul Lubis, Sakriani Sakti, Graham Neubig, Koichiro Yoshino, Tomoki Toda, Satoshi Nakamura |
| 2015 | Phonetic unit selection for cross-lingual query-by-example spoken term detection. | Paula Lopez-Otero, Laura Doco Fernndez, Carmen Garca-Mateo |
| 2015 | Acoustic modeling with neural graph embeddings. | Yuzong Liu, Katrin Kirchhoff |
| 2015 | Natural language understanding for partial queries. | Xiaohu Liu, Asli Celikyilmaz, Ruhi Sarikaya |
| 2015 | LSTM time and frequency recurrence for automatic speech recognition. | Jinyu Li, Abdelrahman Mohamed, Geoffrey Zweig, Yifan Gong |
| 2015 | Towards structured deep neural network for automatic speech recognition. | Yi-Hsiu Liao, Hung-yi Lee, Lin-Shan Lee |
| 2015 | Speaker intonation adaptation for transforming text-to-speech synthesis speaker identity. | Mahsa Sadat Elyasi Langarani, Jan P. H. van Santen |
| 2015 | The development of the cambridge university alignment systems for the multi-genre broadcast challenge. | Pierre Lanchantin, Mark J. F. Gales, Penny Karanasou, Xunying Liu, Yanmin Qian, Linlin Wang, Philip C. Woodland, Chao Zhang |
| 2015 | Implementation of generic positive-negative tracker in extensible dialog system. | Sangjun Koo, Seonghan Ryu, Gary Geunbae Lee |
| 2015 | Speaker diarisation and longitudinal linking in multi-genre broadcast data. | Penny Karanasou, Mark J. F. Gales, Pierre Lanchantin, Xunying Liu, Yanmin Qian, Linlin Wang, Philip C. Woodland, Chao Zhang |
| 2015 | Training data pseudo-shuffling and direct decoding framework for recurrent neural network based acoustic modeling. | Naoyuki Kanda, Mitsuyoshi Tachimori, Xugang Lu, Hisashi Kawai |