| 2013 | The IBM keyword search system for the DARPA RATS program. | Lidia Mangu, Hagen Soltau, Hong-Kwang Kuo, George Saon |
| 2013 | Acoustic data-driven pronunciation lexicon for large vocabulary speech recognition. | Liang Lu, Arnab Ghoshal, Steve Renals |
| 2013 | Multi-stream temporally varying weight regression for cross-lingual speech recognition. | Shilin Liu, Khe Chai Sim |
| 2013 | Query understanding enhanced by hierarchical parsing structures. | Jingjing Liu, Panupong Pasupat, Yining Wang, Scott Cyphers, James R. Glass |
| 2013 | Improving robustness of deep neural networks via spectral masking for automatic speech recognition. | Bo Li, Khe Chai Sim |
| 2013 | Towards unsupervised semantic retrieval of spoken content with query expansion based on automatically discovered acoustic patterns. | Yun-Chiao Li, Hung-yi Lee, Cheng-Tao Chung, Chun-an Chan, Lin-Shan Lee |
| 2013 | Large scale deep neural network acoustic modeling with semi-supervised training data for YouTube video transcription. | Hank Liao, Erik McDermott, Andrew W. Senior |
| 2013 | Fixed-dimensional acoustic embeddings of variable-length segments in low-resource settings. | Keith D. Levin, Katharine Henry, Aren Jansen, Karen Livescu |
| 2013 | Emotion recognition from spontaneous speech using Hidden Markov models with deep belief networks. | Duc Le, Emily Mower Provost |
| 2013 | Modified splice and its extension to non-stereo data for noise robust speech recognition. | D. S. Pavan Kumar, N. Vishnu Prasad, Vikas Joshi, Srinivasan Umesh |
| 2013 | Acoustic characteristics related to the perceptual pitch in whispered vowels. | Hideaki Konno, Hideo Kanemitsu, Nobuyuki Takahashi, Mineichi Kudo |
| 2013 | Investigation of multilingual deep neural networks for spoken term detection. | Kate M. Knill, Mark J. F. Gales, Shakti P. Rath, Philip C. Woodland, Chao Zhang, Shi-Xiong Zhang |
| 2013 | Automatic sentiment extraction from YouTube videos. | Lakshmish Kaushik, Abhijeet Sangwan, John H. L. Hansen |
| 2013 | Discriminative piecewise linear transformation based on deep learning for noise robust automatic speech recognition. | Yosuke Kashiwagi, Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose |
| 2013 | Score normalization and system combination for improved keyword spotting. | Damianos Karakos, Richard M. Schwartz, Stavros Tsakalidis, Le Zhang, Shivesh Ranjan, Tim Ng, Roger Hsiao, Guruprasad Saikumar, Ivan Bulyko, Long Nguyen, John Makhoul, Frantisek Grzl, Mirko Hannemann, Martin Karafit, Igor Szke, Karel Vesel, Lori Lamel, Viet Bac Le |
| 2013 | Elastic spectral distortion for low resource speech recognition with deep neural networks. | Naoyuki Kanda, Ryu Takeda, Yasunari Obuchi |
| 2013 | Phonetic and anthropometric conditioning of MSA-KST cognitive impairment characterization system. | Alexei V. Ivanov, Shahab Jalalvand, Roberto Gretter, Daniele Falavigna |
| 2013 | Impact of deep MLP architecture on different acoustic modeling techniques for under-resourced speech recognition. | David Imseng, Petr Motlcek, Philip N. Garner, Herv Bourlard |
| 2013 | Learning state labels for sparse classification of speech with matrix deconvolution. | Antti Hurmalainen, Tuomas Virtanen |
| 2013 | Accelerating recurrent neural network training via two stage classes and parallelization. | Zhiheng Huang, Geoffrey Zweig, Michael Levit, Benot Dumoulin, Barlas Oguz, Shawn Chang |
| 2013 | Discriminative semi-supervised training for keyword search in low resource languages. | Roger Hsiao, Tim Ng, Frantisek Grzl, Damianos Karakos, Stavros Tsakalidis, Long Nguyen, Richard M. Schwartz |
| 2013 | Dialogue management for leading the conversation in persuasive dialogue systems. | Takuya Hiraoka, Yuki Yamauchi, Graham Neubig, Sakriani Sakti, Tomoki Toda, Satoshi Nakamura |
| 2013 | Unsupervised word segmentation from noisy input. | Jahn Heymann, Oliver Walter, Reinhold Haeb-Umbach, Bhiksha Raj |
| 2013 | Acoustic unit discovery and pronunciation generation from a grapheme-based lexicon. | William Hartmann, Anindya Roy, Lori Lamel, Jean-Luc Gauvain |
| 2013 | Semi-supervised bootstrapping approach for neural network feature extractor training. | Frantisek Grzl, Martin Karafit |