| 2024 | Interspeech | FLEURS-R: A Restored Multilingual Speech Corpus for Generation Tasks. | Min Ma, Yuma Koizumi, Shigeki Karita, Heiga Zen, Jason Riesa, Haruko Ishikawa, Michiel Bacchiani |
| 2023 | Interspeech | LibriTTS-R: A Restored Multi-Speaker Text-to-Speech Corpus. | Yuma Koizumi, Heiga Zen, Shigeki Karita, Yifan Ding, Kohei Yatabe, Nobuyuki Morioka, Michiel Bacchiani, Yu Zhang, Wei Han, Ankur Bapna |
| 2022 | ICASSP | Knowledge Transfer from Large-Scale Pretrained Language Models to End-To-End Speech Recognizers. | Yotaro Kubo, Shigeki Karita, Michiel Bacchiani |
| 2022 | Interspeech | SNRi Target Training for Joint Speech Enhancement and Recognition. | Yuma Koizumi, Shigeki Karita, Arun Narayanan, Sankaran Panchapagesan, Michiel Bacchiani |
| 2022 | Interspeech | SpecGrad: Diffusion Probabilistic Model based Neural Vocoder with Adaptive Noise Spectral Shaping. | Yuma Koizumi, Heiga Zen, Kohei Yatabe, Nanxin Chen, Michiel Bacchiani |
| 2020 | ICASSP | Joint Phoneme-Grapheme Model for End-To-End Speech Recognition. | Yotaro Kubo, Michiel Bacchiani |
| 2018 | ICASSP | State-of-the-Art Speech Recognition with Sequence-to-Sequence Models. | Chung-Cheng Chiu, Tara N. Sainath, Yonghui Wu, Rohit Prabhavalkar, Patrick Nguyen, Zhifeng Chen, Anjuli Kannan, Ron J. Weiss, Kanishka Rao, Ekaterina Gonina, Navdeep Jaitly, Bo Li, Jan Chorowski, Michiel Bacchiani |
| 2018 | ICASSP | Performance of Mask Based Statistical Beamforming in a Smart Home Scenario. | Jahn Heymann, Michiel Bacchiani, Tara N. Sainath |
| 2018 | ICASSP | Sound Source Separation Using Phase Difference and Reliable Mask Selection Selection. | Chanwoo Kim, Anjali Menon, Michiel Bacchiani, Richard M. Stern |
| 2018 | ICASSP | Spectral Distortion Model for Training Phase-Sensitive Deep-Neural Networks for Far-Field Speech Recognition. | Chanwoo Kim, Tara N. Sainath, Arun Narayanan, Ananya Misra, Rajeev C. Nongpiur, Michiel Bacchiani |
| 2018 | ICASSP | Multi-Dialect Speech Recognition with a Single Sequence-to-Sequence Model. | Bo Li, Tara N. Sainath, Khe Chai Sim, Michiel Bacchiani, Eugene Weinstein, Patrick Nguyen, Zhifeng Chen, Yanghui Wu, Kanishka Rao |
| 2018 | ICASSP | Sampled Connectionist Temporal Classification. | Ehsan Variani, Tom Bagby, Kamel Lahouel, Erik McDermott, Michiel Bacchiani |
| 2018 | Interspeech | Efficient Implementation of the Room Simulator for Training Deep Neural Network Acoustic Models. | Chanwoo Kim, Ehsan Variani, Arun Narayanan, Michiel Bacchiani |
| 2018 | Interspeech | Domain Adaptation Using Factorized Hidden Layer for Robust Automatic Speech Recognition. | Khe Chai Sim, Arun Narayanan, Ananya Misra, Anshuman Tripathi, Golan Pundak, Tara N. Sainath, Parisa Haghani, Bo Li, Michiel Bacchiani |
| 2017 | ASRU | Improving the efficiency of forward-backward algorithm using batched computation in TensorFlow. | Khe Chai Sim, Arun Narayanan, Tom Bagby, Tara N. Sainath, Michiel Bacchiani |
| 2017 | Interspeech | Generation of Large-Scale Simulated Utterances in Virtual Rooms to Train Deep-Neural Networks for Far-Field Speech Recognition in Google Home. | Chanwoo Kim, Ananya Misra, Kean K. Chin, Thad Hughes, Arun Narayanan, Tara N. Sainath, Michiel Bacchiani |
| 2017 | Interspeech | Acoustic Modeling for Google Home. | Bo Li, Tara N. Sainath, Arun Narayanan, Joe Caroselli, Michiel Bacchiani, Ananya Misra, Izhak Shafran, Hasim Sak, Golan Pundak, Kean K. Chin, Khe Chai Sim, Ron J. Weiss, Kevin W. Wilson, Ehsan Variani, Chanwoo Kim, Olivier Siohan, Mitchel Weintraub, Erik McDermott, Richard Rose, Matt Shannon |
| 2017 | Interspeech | End-to-End Training of Acoustic Models for Large Vocabulary Continuous Speech Recognition with TensorFlow. | Ehsan Variani, Tom Bagby, Erik McDermott, Michiel Bacchiani |
| 2016 | ICASSP | Factored spatial and spectral multichannel raw waveform CLDNNs. | Tara N. Sainath, Ron J. Weiss, Kevin W. Wilson, Arun Narayanan, Michiel Bacchiani |
| 2016 | Interspeech | Neural Network Adaptive Beamforming for Robust Multichannel Speech Recognition. | Bo Li, Tara N. Sainath, Ron J. Weiss, Kevin W. Wilson, Michiel Bacchiani |
| 2016 | Interspeech | Reducing the Computational Complexity of Multimicrophone Acoustic Models with Integrated Feature Extraction. | Tara N. Sainath, Arun Narayanan, Ron J. Weiss, Ehsan Variani, Kevin W. Wilson, Michiel Bacchiani, Izhak Shafran |
| 2016 | Interspeech | Complex Linear Projection (CLP): A Discriminative Approach to Joint Feature Extraction and Acoustic Modeling. | Ehsan Variani, Tara N. Sainath, Izhak Shafran, Michiel Bacchiani |
| 2015 | ASRU | Speaker location and microphone spacing invariant acoustic modeling from raw multichannel waveforms. | Tara N. Sainath, Ron J. Weiss, Kevin W. Wilson, Arun Narayanan, Michiel Bacchiani, Andrew W. Senior |
| 2015 | Interspeech | Large vocabulary automatic speech recognition for children. | Hank Liao, Golan Pundak, Olivier Siohan, Melissa K. Carroll, Noah Coccaro, Qi-Ming Jiang, Tara N. Sainath, Andrew W. Senior, Franoise Beaufays, Michiel Bacchiani |
| 2014 | ICASSP | Context dependent state tying for speech recognition using deep neural network acoustic models. | Michiel Bacchiani, David Rybach |
| 2014 | ICASSP | Asynchronous stochastic optimization for sequence training of deep neural networks. | Georg Heigold, Erik McDermott, Vincent Vanhoucke, Andrew W. Senior, Michiel Bacchiani |
| 2014 | ICASSP | GMM-free DNN acoustic model training. | Andrew W. Senior, Georg Heigold, Michiel Bacchiani, Hank Liao |
| 2014 | Interspeech | Asynchronous, online, GMM-free training of a context dependent acoustic model for speech recognition. | Michiel Bacchiani, Andrew W. Senior, Georg Heigold |
| 2014 | Interspeech | Robust speech recognition using temporal masking and thresholding algorithm. | Chanwoo Kim, Kean K. Chin, Michiel Bacchiani, Richard M. Stern |
| 2014 | Interspeech | Asynchronous stochastic optimization for sequence training of deep neural networks: towards big data. | Erik McDermott, Georg Heigold, Pedro J. Moreno, Andrew W. Senior, Michiel Bacchiani |
| 2013 | ICASSP | Rapid adaptation for mobile speech applications. | Michiel Bacchiani |
| 2013 | Interspeech | ivector-based acoustic data selection. | Olivier Siohan, Michiel Bacchiani |
| 2011 | Interspeech | Discriminative Features for Language Identification. | Christopher Alberti, Michiel Bacchiani |
| 2010 | Interspeech | Decision tree state clustering with word and syllable features. | Hank Liao, Christopher Alberti, Michiel Bacchiani, Olivier Siohan |
| 2009 | ICASSP | An audio indexing system for election video material. | Christopher Alberti, Michiel Bacchiani, Ari Bezman, Ciprian Chelba, Anastassia Drofa, Hank Liao, Pedro J. Moreno, Ted Power, Arnaud Sahuguet, Maria Shugrina, Olivier Siohan |
| 2009 | ICASSP | Restoring punctuation and capitalization in transcribed speech. | Agustn Gravano, Martin Jansche, Michiel Bacchiani |
| 2008 | ICASSP | Deploying GOOG-411: Early lessons in data, measurement, and testing. | Michiel Bacchiani, Franoise Beaufays, Johan Schalkwyk, Mike Schuster, Brian Strope |
| 2008 | ICASSP | Confidence scores for acoustic model adaptation. | Christian Gollan, Michiel Bacchiani |
| 2005 | Interspeech | Fast vocabulary-independent audio search using path-based graph indexing. | Olivier Siohan, Michiel Bacchiani |
| 2004 | ICASSP | Meta-data conditional language modeling. | Michiel Bacchiani, Brian Roark |
| 2004 | ICASSP | Improved name recognition with meta-data dependent name networks. | Sameer Maskey, Michiel Bacchiani, Brian Roark, Richard Sproat |
| 2004 | NAACL | Language Model Adaptation with MAP Estimation and the Perceptron Algorithm. | Michiel Bacchiani, Brian Roark, Murat Saraclar |
| 2003 | ICASSP | Unsupervised language model adaptation. | Michiel Bacchiani, Brian Roark |
| 2003 | NAACL | Supervised and unsupervised PCFG adaptation to novel domains. | Brian Roark, Michiel Bacchiani |
| 2002 | CHI | SCANMail: a voicemail interface that makes speech browsable, readable and searchable. | Steve Whittaker, Julia Hirschberg, Brian Amento, Litza A. Stark, Michiel Bacchiani, Philip L. Isenhour, Larry Stead, Gary Zamchick, Aaron E. Rosenberg |
| 2002 | Interspeech | Combining maximum likelihood and maximum a posteriori estimation for detailed acoustic modeling of context dependency. | Michiel Bacchiani |
| 2001 | ICASSP | Automatic transcription of voicemail at AT&T. | Michiel Bacchiani |
| 2001 | Interspeech | SCANMail: browsing and searching speech data by content. | Julia Hirschberg, Michiel Bacchiani, Donald Hindle, Philip L. Isenhour, Aaron E. Rosenberg, Litza A. Stark, Larry Stead, Steve Whittaker, Gary Zamchick |
| 2001 | Interspeech | Caller identification for the SCANMail voicemail browser. | Aaron E. Rosenberg, Julia Hirschberg, Michiel Bacchiani, Sarangarajan Parthasarathy, Philip L. Isenhour, Larry Stead |
| 2001 | NAACL | SCANMail: Audio Navigation in the Voicemail Domain. | Michiel Bacchiani, Julia Hirschberg, Aaron E. Rosenberg, Steve Whittaker, Donald Hindle, Philip L. Isenhour, Matt Jones, Litza A. Stark, Gary Zamchick |
| 2000 | Interspeech | Using maximum likelihood linear regression for segment clustering and speaker identification. | Michiel Bacchiani |
| 1998 | Interspeech | Using automatically-derived acoustic sub-word units in large vocabulary speech recognition. | Michiel Bacchiani, Mari Ostendorf |
| 1996 | ICASSP | Design of a speech recognition system based on acoustically derived segmental units. | Michiel Bacchiani, Mari Ostendorf, Yoshinori Sagisaka, Kuldip K. Paliwal |
| 1996 | Interspeech | Speech recognition based on acoustically derived segment units. | Toshiaki Fukada, Michiel Bacchiani, Kuldip K. Paliwal, Yoshinori Sagisaka |
| 1995 | Interspeech | Minimum classification error training algorithm for feature extractor and pattern classifier in speech recognition. | Kuldip K. Paliwal, Michiel Bacchiani, Yoshinori Sagisaka |
| 1994 | ICASSP | Optimization of time-frequency masking filters using the minimum classification error criterion. | Michiel Bacchiani, Kiyoaki Aikawa |