| 2007 | Speech recognition with localized time-frequency pattern detectors. | Ken Schutte, James R. Glass |
| 2007 | Comparing one and two-stage acoustic modeling in the recognition of emotion in speech. | Bjrn W. Schuller, Bogdan Vlasenko, Ricardo Minguez, Gerhard Rigoll, Andreas Wendemuth |
| 2007 | Error simulation for training statistical dialogue systems. | Jost Schatzmann, Blaise Thomson, Steve J. Young |
| 2007 | Lattice-based Viterbi decoding techniques for speech translation. | George Saon, Michael Picheny |
| 2007 | Phonological feature based variable frame rate scheme for improved speech recognition. | Abhijeet Sangwan, John H. L. Hansen |
| 2007 | Broad phonetic class recognition in a Hidden Markov model framework using extended Baum-Welch transformations. | Tara N. Sainath, Dimitri Kanevsky, Bhuvana Ramabhadran |
| 2007 | Advances in Arabic broadcast news transcription at RWTH. | David Rybach, Stefan Hahn, Christian Gollan, Ralf Schlter, Hermann Ney |
| 2007 | The LIMSI QAst systems: Comparison between human and automatic rules generation for question-answering on speech transcriptions. | Sophie Rosset, Olivier Galibert, Gilles Adda, ric Bilinski |
| 2007 | Towards robust automatic evaluation of pathologic telephone speech. | Korbinian Riedhammer, Georg Stemmer, Tino Haderlein, Maria Schuster, Frank Rosanowski, Elmar Nth, Andreas K. Maier |
| 2007 | Recognition and understanding of meetings the AMI and AMIDA projects. | Steve Renals, Thomas Hain, Herv Bourlard |
| 2007 | A multi-layer architecture for semi-synchronous event-driven dialogue management. | Antoine Raux, Maxine Esknazi |
| 2007 | The IBM 2007 speech transcription system for European parliamentary speeches. | Bhuvana Ramabhadran, Olivier Siohan, Abhinav Sethy |
| 2007 | Non-native speech databases. | Martin Raab, Rainer Gruhn, Elmar Nth |
| 2007 | Random discriminant structure analysis for automatic recognition of connected vowels. | Yu Qiao, Satoshi Asakawa, Nobuaki Minematsu |
| 2007 | Type-II dialogue systems for information access from unstructured knowledge sources. | Yi-Cheng Pan, Lin-Shan Lee |
| 2007 | Analytical comparison between position specific posterior lattices and confusion networks based on words and subword units for spoken document indexing. | Yi-Cheng Pan, Hung-lin Chang, Lin-Shan Lee |
| 2007 | Implicit user-adaptive system engagement in speech, pen and multimodal interfaces. | Sharon L. Oviatt |
| 2007 | Efficient use of overlap information in speaker diarization. | Scott Otterson, Mari Ostendorf |
| 2007 | Refine bigram PLSA model by assigning latent topics unevenly. | Jiazhong Nie, Runxin Li, Dingsheng Luo, Xihong Wu |
| 2007 | Automatic detection of contrastive elements in spontaneous speech. | Ani Nenkova, Dan Jurafsky |
| 2007 | Extensible speech recognition system using proxy-agent. | Teppei Nakano, Shinya Fujie, Tetsunori Kobayashi |
| 2007 | Joint decoding of multiple speech patterns for robust speech recognition. | Nishanth Ulhas Nair, T. V. Sreenivas |
| 2007 | Spoken language understanding with kernels for syntactic/semantic structures. | Alessandro Moschitti, Giuseppe Riccardi, Christian Raymond |
| 2007 | Spoken language understanding: a survey. | Renato de Mori |
| 2007 | Exploiting complementary aspects of phonological features in automatic speech recognition. | Parya Momayyez, James Waterhouse, Richard Rose |