Skip to content

Hagen Soltau

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

58

Venues

7

Active years

1998–2025

Best venue rank

A*

Where they publish

Papers

58 indexed papers, newest first.

YearVenueTitleAuthors
2025CVPRLearning Visual Composition through Improved Semantic Guidance.Austin Stone, Hagen Soltau, Robert Geirhos, Xi Yi, Ye Xia, Bingyi Cao, Kaifeng Chen, Abhijit Ogale, Jonathon Shlens
2024ICASSPRetrieval Augmented End-to-End Spoken Dialog Models.Mingqiu Wang, Izhak Shafran, Hagen Soltau, Wei Han, Yuan Cao, Dian Yu, Laurent El Shafey
2023ASRUDetecting Speech Abnormalities With a Perceiver-Based Sequence Classifier that Leverages a Universal Speech Model.Hagen Soltau, Izhak Shafran, Alex Ottenwess, Joseph R. Duffy, Rene L. Utianski, Leland R. Barnard, John L. Stricker, Daniela A. Wiepert, David T. Jones, Hugo Botha
2023ASRUSLM: Bridge the Thin Gap Between Speech and Text Foundation Models.Mingqiu Wang, Wei Han, Izhak Shafran, Zelin Wu, Chung-Cheng Chiu, Yuan Cao, Nanxin Chen, Yu Zhang, Hagen Soltau, Paul K. Rubenstein, Lukas Zilka, Dian Yu, Golan Pundak, Nikhil Siddhartha, Johan Schalkwyk, Yonghui Wu
2023EMNLPAnyTOD: A Programmable Task-Oriented Dialog System.Jeffrey Zhao, Yuan Cao, Raghav Gupta, Harrison Lee, Abhinav Rastogi, Mingqiu Wang, Hagen Soltau, Izhak Shafran, Yonghui Wu
2023InterspeechSpeech Aware Dialog System Technology Challenge (DSTC11).Hagen Soltau, Izhak Shafran, Mingqiu Wang, Abhinav Rastogi, Jeffrey Zhao, Ye Jia, Wei Han, Yuan Cao, Aramys Miranda
2022EMNLPKnowledge-grounded Dialog State Tracking.Dian Yu, Mingqiu Wang, Yuan Cao, Laurent El Shafey, Izhak Shafran, Hagen Soltau
2022InterspeechRNN Transducers for Named Entity Recognition with constraints on alignment for understanding medical conversations.Hagen Soltau, Izhak Shafran, Mingqiu Wang, Laurent El Shafey
2022NAACLUnsupervised Slot Schema Induction for Task-oriented Dialog.Dian Yu, Mingqiu Wang, Yuan Cao, Izhak Shafran, Laurent El Shafey, Hagen Soltau
2021ASRUWord-Level Confidence Estimation for RNN Transducers.Mingqiu Wang, Hagen Soltau, Laurent El Shafey, Izhak Shafran
2021InterspeechUnderstanding Medical Conversations: Rich Transcription, Confidence Scores & Information Extraction.Hagen Soltau, Mingqiu Wang, Izhak Shafran, Laurent El Shafey
2020LRECThe Medical Scribe: Corpus Development and Model Performance Analyses.Izhak Shafran, Nan Du, Linh Tran, Amanda Perry, Lauren Keyes, Mark Knichel, Ashley Domin, Lei Huang, Yuhui Chen, Gang Li, Mingqiu Wang, Laurent El Shafey, Hagen Soltau, Justin S. Paul
2019ASRUMonotonic Recurrent Neural Network Transducer and Decoding Strategies.Anshuman Tripathi, Han Lu, Hasim Sak, Hagen Soltau
2019InterspeechJoint Speech Recognition and Speaker Diarization via Sequence Transduction.Laurent El Shafey, Hagen Soltau, Izhak Shafran
2017ASRUReducing the computational complexity for whole word models.Hagen Soltau, Hank Liao, Hasim Sak
2017InterspeechNeural Speech Recognizer: Acoustic-to-Word LSTM Model for Large Vocabulary Speech Recognition.Hagen Soltau, Hank Liao, Hasim Sak
2014ICASSPOut-of-vocabulary word detection in a speech-to-speech translation system.Hong-Kwang Kuo, Ellen Eide Kislal, Lidia Mangu, Hagen Soltau, Toms Beran
2014ICASSPEfficient spoken term detection using confusion networks.Lidia Mangu, Brian Kingsbury, Hagen Soltau, Hong-Kwang Kuo, Michael Picheny
2014ICASSPProgress in dynamic network decoding.David Nolden, Hagen Soltau, Hermann Ney
2014ICASSPA comparison of two optimization techniques for sequence discriminative training of deep neural networks.George Saon, Hagen Soltau
2014ICASSPJoint training of convolutional and non-convolutional neural networks.Hagen Soltau, George Saon, Tara N. Sainath
2014ICASSPAnalyzing convolutional neural networks for speech activity detection in mismatched acoustic conditions.Samuel Thomas, Sriram Ganapathy, George Saon, Hagen Soltau
2014InterspeechRemoving redundancy from lattices.David Nolden, Hagen Soltau, Daniel Povey, Pegah Ghahremani, Lidia Mangu, Hermann Ney
2014InterspeechUnfolded recurrent neural networks for speech recognition.George Saon, Hagen Soltau, Ahmad Emami, Michael Picheny
2013ASRUThe IBM keyword search system for the DARPA RATS program.Lidia Mangu, Hagen Soltau, Hong-Kwang Kuo, George Saon
2013ASRUImprovements to Deep Convolutional Neural Networks for LVCSR.Tara N. Sainath, Brian Kingsbury, Abdel-rahman Mohamed, George E. Dahl, George Saon, Hagen Soltau, Toms Beran, Aleksandr Y. Aravkin, Bhuvana Ramabhadran
2013ASRUSpeaker adaptation of neural network acoustic models using i-vectors.George Saon, Hagen Soltau, David Nahamoo, Michael Picheny
2013ICASSPExploiting diversity for spoken term detection.Lidia Mangu, Hagen Soltau, Hong-Kwang Kuo, Brian Kingsbury, George Saon
2013ICASSPMorpheme-based feature-rich language models using Deep Neural Networks for LVCSR of Egyptian Arabic.Amr El-Desoky Mousa, Hong-Kwang Jeff Kuo, Lidia Mangu, Hagen Soltau
2013InterspeechThe IBM speech activity detection system for the DARPA RATS program.George Saon, Samuel Thomas, Hagen Soltau, Sriram Ganapathy, Brian Kingsbury
2013InterspeechNeural network acoustic models for the DARPA RATS program.Hagen Soltau, Hong-Kwang Kuo, Lidia Mangu, George Saon, Toms Beran
2012InterspeechScalable Minimum Bayes Risk Training of Deep Neural Network Acoustic Models Using Distributed Hessian-free Optimization.Brian Kingsbury, Tara N. Sainath, Hagen Soltau
2011ASRUThe IBM 2011 GALE Arabic speech transcription system.Lidia Mangu, Hong-Kwang Kuo, Stephen M. Chu, Brian Kingsbury, George Saon, Hagen Soltau, Fadi Biadsy
2011ASRUFrom Modern Standard Arabic to Levantine ASR: Leveraging GALE for dialects.Hagen Soltau, Lidia Mangu, Fadi Biadsy
2011ICASSPThe IBM 2009 GALE Arabic speech transcription system.Brian Kingsbury, Hagen Soltau, George Saon, Stephen M. Chu, Hong-Kwang Kuo, Lidia Mangu, Suman V. Ravuri, Nelson Morgan, Adam Janin
2010ICASSPA comparative study on system combination schemes for LVCSR.Chengyuan Ma, Hong-Kwang Jeff Kuo, Hagen Soltau, Xiaodong Cui, Upendra V. Chaudhari, Lidia Mangu, Chin-Hui Lee
2010ICASSPThe IBM 2008 GALE Arabic speech transcription system.George Saon, Hagen Soltau, Upendra V. Chaudhari, Stephen M. Chu, Brian Kingsbury, Hong-Kwang Kuo, Lidia Mangu, Daniel Povey
2010InterspeechDecoding with shrinkage-based language models.Ahmad Emami, Stanley F. Chen, Abraham Ittycheriah, Hagen Soltau, Bing Zhao
2010InterspeechBoosting systems for LVCSR.George Saon, Hagen Soltau
2009ASRUDynamic network decoding revisited.Hagen Soltau, George Saon
2009ICASSPLarge margin semi-tied covariance transforms for discriminative training.George Saon, Daniel Povey, Hagen Soltau
2008InterspeechFast speaker adaptive training for speech recognition.Daniel Povey, Hong-Kwang Jeff Kuo, Hagen Soltau
2007ICASSPThe IBM 2006 Gale Arabic ASR System.Hagen Soltau, George Saon, Brian Kingsbury, Hong-Kwang Jeff Kuo, Lidia Mangu, Daniel Povey, Geoffrey Zweig
2005ICASSPfMPE: Discriminatively Trained Features for Speech Recognition.Daniel Povey, Brian Kingsbury, Lidia Mangu, George Saon, Hagen Soltau, Geoffrey Zweig
2005ICASSPThe IBM 2004 Conversational Telephony System for Rich Transcription.Hagen Soltau, Brian Kingsbury, Lidia Mangu, Daniel Povey, George Saon, Geoffrey Zweig
2004ICASSPThe 2003 ISL rich transcription system for conversational telephony speech.Hagen Soltau, Hua Yu, Florian Metze, Christian Fgen, Qin Jin, Szu-Chen Stan Jou
2002ICASSPEfficient language model lookahead through polymorphic linguistic context assignment.Hagen Soltau, Florian Metze, Christian Fgen, Alex Waibel
2002InterspeechCompensating for hyperarticulation by modeling articulatory properties.Hagen Soltau, Florian Metze, Alex Waibel
2001ICASSPSpeaker compensation with sine-log all-pass transforms.John W. McDonough, Florian Metze, Hagen Soltau, Alex Waibel
2001ICASSPThe ISL evaluation system for Verbmobil-II.Hagen Soltau, Thomas Schaaf, Florian Metze, Alex Waibel
2001ICASSPAdvances in automatic meeting record creation and access.Alex Waibel, Michael Bett, Florian Metze, Klaus Ries, Thomas Schaaf, Tanja Schultz, Hagen Soltau, Hua Yu, Klaus Zechner
2001InterspeechSpeech recognition over netmeeting connections.Florian Metze, John W. McDonough, Hagen Soltau
2001NAACLAdvances in meeting recognition.Alex Waibel, Hua Yu, Tanja Schultz, Yue Pan, Michael Bett, Martin Westphal, Hagen Soltau, Thomas Schaaf, Florian Metze
2000ICASSPConfidence measure based language identification.Florian Metze, Thomas Kemp, Thomas Schaaf, Tanja Schultz, Hagen Soltau
2000ICASSPSpecialized acoustic models for hyperarticulated speech.Hagen Soltau, Alex Waibel
2000InterspeechPhone dependent modeling of hyperarticulated effects#.Hagen Soltau, Alex Waibel
1998ICASSPRecognition of music types.Hagen Soltau, Tanja Schultz, Martin Westphal, Alex Waibel
1998InterspeechOn the influence of hyperarticulated speech on recognition performance.Hagen Soltau, Alex Waibel