| 2015 | Boosted acoustic model learning and hypotheses rescoring on the CHiME-3 task. | Shahab Jalalvand, Daniele Falavigna, Marco Matassoni, Piergiorgio Svaizer, Maurizio Omologo |
| 2015 | Different word representations and their combination for proper name retrieval from diachronic documents. | Irina Illina, Dominique Fohr |
| 2015 | Robust speech recognition in unknown reverberant and noisy conditions. | Roger Hsiao, Jeff Z. Ma, William Hartmann, Martin Karafit, Frantisek Grzl, Luks Burget, Igor Szke, Jan Cernock, Shinji Watanabe, Zhuo Chen, Sri Harish Reddy Mallidi, Hynek Hermansky, Stavros Tsakalidis, Richard M. Schwartz |
| 2015 | The MERL/SRI system for the 3RD CHiME challenge using beamforming, robust feature extraction, and advanced speech recognition. | Takaaki Hori, Zhuo Chen, Hakan Erdogan, John R. Hershey, Jonathan Le Roux, Vikramjit Mitra, Shinji Watanabe |
| 2015 | Towards utterance-based neural network adaptation in acoustic modeling. | Ivan Himawan, Petr Motlcek, Marc Ferras Font, Srikanth R. Madikeri |
| 2015 | BLSTM supported GEV beamformer front-end for the 3RD CHiME challenge. | Jahn Heymann, Lukas Drude, Aleksej Chinaev, Reinhold Haeb-Umbach |
| 2015 | Deep multimodal semantic embeddings for speech and images. | David F. Harwath, James R. Glass |
| 2015 | The Automatic Speech recogition In Reverberant Environments (ASpIRE) challenge. | Mary Harper |
| 2015 | CRIM and LIUM approaches for multi-genre broadcast media transcription. | Vishwa Gupta, Paul Delglise, Gilles Boulianne, Yannick Estve, Sylvain Meignier, Anthony Rousseau |
| 2015 | Spectral learning with non negative probabilities for finite state automaton. | Hadrien Glaude, Cyrille Enderli, Olivier Pietquin |
| 2015 | Deep bottleneck features for i-vector based text-independent speaker verification. | Sina Hamidi Ghalehjegh, Richard C. Rose |
| 2015 | Policy committee for adaptation in multi-domain spoken dialogue systems. | Milica Gasic, Nikola Mrksic, Pei-hao Su, David Vandyke, Tsung-Hsien Wen, Steve J. Young |
| 2015 | Unified ASR system using LGM-based source separation, noise-robust feature extraction, and word hypothesis selection. | Yusuke Fujita, Ryoichi Takashima, Takeshi Homma, Rintaro Ikeshita, Yohei Kawaguchi, Takashi Sumiyoshi, Takashi Endo, Masahito Togami |
| 2015 | Improving data selection for low-resource STT and KWS. | Thiago Fraga-Silva, Antoine Laurent, Jean-Luc Gauvain, Lori Lamel, Viet Bac Le, Abdelkhalek Messaoudi |
| 2015 | Applying deep learning to answer selection: A study and an open task. | Minwei Feng, Bing Xiang, Michael R. Glass, Lidan Wang, Bowen Zhou |
| 2015 | An information fusion approach to recognizing microphone array speech in the CHiME-3 challenge based on a deep learning framework. | Jun Du, Qing Wang, Yanhui Tu, Xiao Bao, Li-Rong Dai, Chin-Hui Lee |
| 2015 | Latent Dirichlet Allocation based organisation of broadcast media archives for deep neural network adaptation. | Mortaza Doulaty, Oscar Saz, Raymond W. M. Ng, Thomas Hain |
| 2015 | The NAIST ASR system for the 2015 Multi-Genre Broadcast challenge: On combination of deep learning systems using a rank-score function. | Quoc Truong Do, Michael Heck, Sakriani Sakti, Graham Neubig, Tomoki Toda, Satoshi Nakamura |
| 2015 | Automatic prosody prediction for Chinese speech synthesis using BLSTM-RNN and embedding features. | Chuang Ding, Lei Xie, Jie Yan, Weini Zhang, Yang Liu |
| 2015 | Single and multi-channel approaches for distant speech recognition under noisy reverberant conditions: I2R'S system description for the ASpIRE challenge. | Jonathan William Dennis, Tran Huy Dat |
| 2015 | Structured discriminative models using deep neural-network features. | Rogier C. van Dalen, Jingzhou Yang, Haipeng Wang, Anton Ragni, Chao Zhang, Mark J. F. Gales |
| 2015 | Multilingual representations for low resource speech recognition and keyword search. | Jia Cui, Brian Kingsbury, Bhuvana Ramabhadran, Abhinav Sethy, Kartik Audhkhasi, Xiaodong Cui, Ellen Kislal, Lidia Mangu, Markus Nubaum-Thom, Michael Picheny, Zoltn Tske, Pavel Golik, Ralf Schlter, Hermann Ney, Mark J. F. Gales, Kate M. Knill, Anton Ragni, Haipeng Wang, Philip C. Woodland |
| 2015 | An iterative deep learning framework for unsupervised discovery of speech features and linguistic units with applications on spoken term detection. | Cheng-Tao Chung, Cheng-Yu Tsai, Hsiang-Hung Lu, Chia-Hsiang Liu, Hung-yi Lee, Lin-Shan Lee |
| 2015 | Incorporating paragraph embeddings and density peaks clustering for spoken document summarization. | Kuan-Yu Chen, Kai-Wun Shih, Shih-Hung Liu, Berlin Chen, Hsin-Min Wang |
| 2015 | Investigation of back-off based interpolation between recurrent neural network and n-gram language models. | Xie Chen, Xunying Liu, Mark J. F. Gales, Philip C. Woodland |