| 2024 | ICDE | LightLT: A Lightweight Representation Quantization Framework for Long-Tail Data. | Haoyu Wang, Ruirui Li, Zhengyang Wang, Xianfeng Tang, Danqing Zhang, Monica Xiao Cheng, Bing Yin, Jasha Droppo, Suhang Wang, Jing Gao |
| 2023 | ICASSP | Federated Self-Learning with Weak Supervision for Speech Recognition. | Milind Rao, Gopinath Chennupati, Gautam Tiwari, Anit Kumar Sahu, Anirudh Raju, Ariya Rastrow, Jasha Droppo |
| 2023 | Interspeech | Diffusion-based accent modelling in speech synthesis. | Kamil Deja, Georgi Tinchev, Marta Czarnowska, Marius Cotescu, Jasha Droppo |
| 2022 | ICASSP | Improving Fairness in Speaker Verification via Group-Adapted Fusion Network. | Hua Shen, Yuguang Yang, Guoli Sun, Ryan Langman, Eunjung Han, Jasha Droppo, Andreas Stolcke |
| 2022 | ICASSP | Improved Representation Learning For Acoustic Event Classification Using Tree-Structured Ontology. | Arman Zharmagambetov, Qingming Tang, Chieh-Chi Kao, Qin Zhang, Ming Sun, Viktor Rozgic, Jasha Droppo, Chao Wang |
| 2022 | Interspeech | Adversarial Reweighting for Speaker Verification Fairness. | Minho Jin, Chelsea Ju, Zeya Chen, Yi-Chieh Liu, Jasha Droppo, Andreas Stolcke |
| 2022 | Interspeech | Reducing Geographic Disparities in Automatic Speech Recognition via Elastic Weight Consolidation. | Viet Anh Trinh, Pegah Ghahremani, Brian John King, Jasha Droppo, Andreas Stolcke, Roland Maas |
| 2022 | KDD | ILASR: Privacy-Preserving Incremental Learning for Automatic Speech Recognition at Production Scale. | Gopinath Chennupati, Milind Rao, Gurpreet Chadha, Aaron Eakin, Anirudh Raju, Gautam Tiwari, Anit Kumar Sahu, Ariya Rastrow, Jasha Droppo, Andy Oberlin, Buddha Nandanoor, Prahalad Venkataramanan, Zheng Wu, Pankaj Sitpure |
| 2021 | ICASSP | Top-Down Attention in End-to-End Spoken Language Understanding. | Yixin Chen, Weiyi Lu, Alejandro Mottini, Li Erran Li, Jasha Droppo, Zheng Du, Belinda Zeng |
| 2021 | ICASSP | Joint ASR and Language Identification Using RNN-T: An Efficient Approach to Dynamic Language Switching. | Surabhi Punjabi, Harish Arsikere, Zeynab Raeesy, Chander Chandak, Nikhil Bhave, Ankish Bansal, Markus Mller, Sergio Murillo, Ariya Rastrow, Andreas Stolcke, Jasha Droppo, Sri Garimella, Roland Maas, Mat Hans, Athanasios Mouchtaris, Siegfried Kunzmann |
| 2021 | ICASSP | DO as I Mean, Not as I Say: Sequence Loss Training for Spoken Language Understanding. | Milind Rao, Pranav Dheram, Gautam Tiwari, Anirudh Raju, Jasha Droppo, Ariya Rastrow, Andreas Stolcke |
| 2021 | ICASSP | Exploring the application of synthetic audio in training keyword spotters. | Andrew Werchniak, Roberto Barra-Chicote, Yuriy Mishchenko, Jasha Droppo, Jeff Condal, Peng Liu, Anish Shah |
| 2021 | Interspeech | Scaling Laws for Acoustic Models. | Jasha Droppo, Oguz Elibol |
| 2021 | Interspeech | SynthASR: Unlocking Synthetic Data for Speech Recognition. | Amin Fazel, Wei Yang, Yulan Liu, Roberto Barra-Chicote, Yixiong Meng, Roland Maas, Jasha Droppo |
| 2021 | Interspeech | Detection of Lexical Stress Errors in Non-Native (L2) English with Data Augmentation and Attention. | Daniel Korzekwa, Roberto Barra-Chicote, Szymon Zaporowski, Grzegorz Beringer, Jaime Lorenzo-Trueba, Alicja Serafinowicz, Jasha Droppo, Thomas Drugman, Bozena Kostek |
| 2021 | Interspeech | Scaling Effect of Self-Supervised Speech Models. | Jie Pu, Yuguang Yang, Ruirui Li, Oguz Elibol, Jasha Droppo |
| 2021 | Interspeech | Listen with Intent: Improving Speech Recognition with Audio-to-Intent Front-End. | Swayambhu Nath Ray, Minhua Wu, Anirudh Raju, Pegah Ghahremani, Raghavendra Bilgi, Milind Rao, Harish Arsikere, Ariya Rastrow, Andreas Stolcke, Jasha Droppo |
| 2021 | Interspeech | wav2vec-C: A Self-Supervised Model for Speech Representation Learning. | Samik Sadhu, Di He, Che-Wei Huang, Sri Harish Mallidi, Minhua Wu, Ariya Rastrow, Andreas Stolcke, Jasha Droppo, Roland Maas |
| 2021 | Interspeech | Evaluating the Vulnerability of End-to-End Automatic Speech Recognition Models to Membership Inference Attacks. | Muhammad A. Shah, Joseph Szurley, Markus Mller, Athanasios Mouchtaris, Jasha Droppo |
| 2021 | Interspeech | CoDERT: Distilling Encoder Representations with Co-Learning for Transducer-Based Speech Recognition. | Rupak Vignesh Swaminathan, Brian John King, Grant P. Strimel, Jasha Droppo, Athanasios Mouchtaris |
| 2021 | Interspeech | Improving Multi-Speaker TTS Prosody Variance with a Residual Encoder and Normalizing Flows. | Ivn Valls-Prez, Julian Roth, Grzegorz Beringer, Roberto Barra-Chicote, Jasha Droppo |
| 2020 | Interspeech | Efficient Minimum Word Error Rate Training of RNN-Transducer for End-to-End Speech Recognition. | Jinxi Guo, Gautam Tiwari, Jasha Droppo, Maarten Van Segbroeck, Che-Wei Huang, Andreas Stolcke, Roland Maas |
| 2019 | ICASSP | Single-channel Speech Extraction Using Speaker Inventory and Attention Network. | Xiong Xiao, Zhuo Chen, Takuya Yoshioka, Hakan Erdogan, Changliang Liu, Dimitrios Dimitriadis, Jasha Droppo, Yifan Gong |
| 2018 | ICASSP | Sequence Modeling in Unsupervised Single-Channel Overlapped Speech Recognition. | Zhehuai Chen, Jasha Droppo |
| 2018 | ICASSP | The Microsoft 2017 Conversational Speech Recognition System. | Wayne Xiong, Lingfeng Wu, Fil Alleva, Jasha Droppo, Xuedong Huang, Andreas Stolcke |
| 2017 | ASRU | Acoustic-to-word model without OOV. | Jinyu Li, Guoli Ye, Rui Zhao, Jasha Droppo, Yifan Gong |
| 2017 | ICASSP | The microsoft 2016 conversational speech recognition system. | Wayne Xiong, Jasha Droppo, Xuedong Huang, Frank Seide, Mike Seltzer, Andreas Stolcke, Dong Yu, Geoffrey Zweig |
| 2017 | ICASSP | Advances in all-neural speech recognition. | Geoffrey Zweig, Chengzhu Yu, Jasha Droppo, Andreas Stolcke |
| 2017 | Interspeech | Comparing Human and Machine Errors in Conversational Speech Transcription. | Andreas Stolcke, Jasha Droppo |
| 2016 | ICASSP | Self-stabilized deep neural network. | Pegah Ghahremani, Jasha Droppo |
| 2016 | ICASSP | Linearly augmented deep neural network. | Pegah Ghahremani, Jasha Droppo, Michael L. Seltzer |
| 2016 | ICASSP | Exploiting LSTM structure in deep neural networks for speech recognition. | Tianxing He, Jasha Droppo |
| 2016 | ICASSP | Parallelizing WFST speech decoders. | Charith Mendis, Jasha Droppo, Saeed Maleki, Madanlal Musuvathi, Todd Mytkowicz, Geoffrey Zweig |
| 2016 | Interspeech | Deep Convolutional Neural Networks with Layer-Wise Context Expansion and Attention. | Dong Yu, Wayne Xiong, Jasha Droppo, Andreas Stolcke, Guoli Ye, Jinyu Li, Geoffrey Zweig |
| 2015 | ASRU | Deep bi-directional recurrent networks over spectral windows. | Abdel-rahman Mohamed, Frank Seide, Dong Yu, Jasha Droppo, Andreas Stolcke, Geoffrey Zweig, Gerald Penn |
| 2015 | ICASSP | Improving speech recognition in reverberation using a room-aware deep neural network and multi-task learning. | Ritwik Giri, Michael L. Seltzer, Jasha Droppo, Dong Yu |
| 2015 | ICASSP | Speech recognition with prediction-adaptation-correction recurrent neural networks. | Yu Zhang, Dong Yu, Michael L. Seltzer, Jasha Droppo |
| 2014 | ICASSP | Phone sequence modeling with recurrent neural networks. | Nicolas Boulanger-Lewandowski, Jasha Droppo, Mike Seltzer, Dong Yu |
| 2014 | ICASSP | On parallelizability of stochastic gradient descent for speech DNNS. | Frank Seide, Hao Fu, Jasha Droppo, Gang Li, Dong Yu |
| 2014 | ICASSP | Single-channel mixed speech recognition using deep neural networks. | Chao Weng, Dong Yu, Michael L. Seltzer, Jasha Droppo |
| 2014 | Interspeech | 1-bit stochastic gradient descent and its application to data-parallel distributed training of speech DNNs. | Frank Seide, Hao Fu, Jasha Droppo, Gang Li, Dong Yu |
| 2014 | Interspeech | An introduction to computational networks and the computational network toolkit (invited talk). | Dong Yu, Adam Eversole, Michael L. Seltzer, Kaisheng Yao, Brian Guenter, Oleksii Kuchaiev, Frank Seide, Huaming Wang, Jasha Droppo, Zhiheng Huang, Geoffrey Zweig, Christopher J. Rossbach, Jon Currey |
| 2013 | ICASSP | Multi-task learning in deep neural networks for improved phoneme recognition. | Michael L. Seltzer, Jasha Droppo |
| 2012 | ICASSP | A chunk-based phonetic score for mobile voice search. | Rohit Prabhavalkar, Jasha Droppo |
| 2011 | ICASSP | Joint encoding of the waveform and speech recognition features using a transform codec. | Xing Fan, Michael L. Seltzer, Jasha Droppo, Henrique S. Malvar, Alex Acero |
| 2011 | ICASSP | Learning non-parametric models of pronunciation. | Brian Hutchinson, Jasha Droppo |
| 2011 | Interspeech | Automatically Optimizing Utterance Classification Performance without Human in the Loop. | Yun-Cheng Ju, Jasha Droppo |
| 2010 | ICASSP | Context dependent phonetic string edit distance for automatic speech recognition. | Jasha Droppo, Alex Acero |
| 2010 | ICASSP | Information retrieval methods for automatic speech recognition. | Xiaoqiang Xiao, Jasha Droppo, Alex Acero |
| 2010 | Interspeech | Continuous speech recognition with a TF-IDF acoustic model. | Geoffrey Zweig, Patrick Nguyen, Jasha Droppo, Alex Acero |
| 2009 | ICASSP | Experimenting with a global decision tree for state clustering in automatic speech recognition systems. | Jasha Droppo, Alex Acero |
| 2008 | ICASSP | Speech enhancement using a pitch predictive model. | Luis Buera, Jasha Droppo, Alex Acero |
| 2008 | ICASSP | Robust design of wideband loudspeaker arrays. | Ivan Tashev, Jasha Droppo, Michael L. Seltzer, Alex Acero |
| 2008 | ICASSP | A minimum-mean-square-error noise reduction algorithm on Mel-frequency cepstra for robust speech recognition. | Dong Yu, Li Deng, Jasha Droppo, Jian Wu, Yifan Gong, Alex Acero |
| 2008 | Interspeech | Towards a non-parametric acoustic model: an acoustic decision tree for observation probability calculation. | Jasha Droppo, Michael L. Seltzer, Alex Acero, Yu-Hsiang Bosco Chiu |
| 2007 | ICASSP | Maximum Entropy Confidence Estimation for Speech Recognition. | Christopher White, Jasha Droppo, Alex Acero, Julian Odell |
| 2007 | Interspeech | A fine pitch model for speech. | Jasha Droppo, Alex Acero |
| 2006 | ICASSP | Joint Discriminative Front End and Back End Training for Improved Speech Recognition Accuracy. | Jasha Droppo, Alex Acero |
| 2005 | ICASSP | Leakage Model and Teeth Clack Removal for Air- and Bone-Conductive Integrated Microphones. | Zicheng Liu, Amar Subramanya, Zhengyou Zhang, Jasha Droppo, Alex Acero |
| 2005 | Interspeech | Maximum mutual information SPLICE transform for seen and unseen conditions. | Jasha Droppo, Alex Acero |
| 2005 | Interspeech | Robust bandwidth extension of noise-corrupted narrowband speech. | Michael L. Seltzer, Alex Acero, Jasha Droppo |
| 2005 | Interspeech | A graphical model for multi-sensory speech processing in air-and-bone conductive microphones. | Amarnag Subramanya, Zhengyou Zhang, Zicheng Liu, Jasha Droppo, Alex Acero |
| 2004 | ICASSP | Noise robust speech recognition with a switching linear dynamic model. | Jasha Droppo, Alex Acero |
| 2004 | ICASSP | Multi-sensory microphones for robust speech detection, enhancement and recognition. | Zhengyou Zhang, Zicheng Liu, Mike Sinclair, Alex Acero, Li Deng, Jasha Droppo, Xuedong Huang, Yanli Zheng |
| 2004 | MMSP | Direct filtering for air- and bone-conductive microphones. | Zicheng Liu, Zhengyou Zhang, Alejandro Acero, Jasha Droppo, Xuedong Huang |
| 2003 | ICASSP | Incremental Bayes learning with prior evolution for tracking nonstationary noise statistics from noisy speech data. | Li Deng, Jasha Droppo, Alex Acero |
| 2003 | Interspeech | A comparison of three non-linear observation models for noisy speech features. | Jasha Droppo, Li Deng, Alex Acero |
| 2003 | Interspeech | A harmonic-model-based front end for robust speech recognition. | Michael L. Seltzer, Jasha Droppo, Alex Acero |
| 2002 | ICASSP | A Bayesian approach to speech feature enhancement using the dynamic cepstral prior. | Li Deng, Jasha Droppo, Alex Acero |
| 2002 | ICASSP | Uncertainty decoding with SPLICE for noise robust speech recognition. | Jasha Droppo, Alex Acero, Li Deng |
| 2002 | Interspeech | Sequential MAP noise estimation and a phase-sensitive model of the acoustic environment. | Li Deng, Jasha Droppo, Alex Acero |
| 2002 | Interspeech | Exploiting variances in robust feature extraction based on a parametric model of speech distortion. | Li Deng, Jasha Droppo, Alex Acero |
| 2002 | Interspeech | Noise from corrupted speech log mel-spectral energies. | Jasha Droppo, Alex Acero, Li Deng |
| 2002 | Interspeech | Evaluation of SPLICE on the Aurora 2 and 3 tasks. | Jasha Droppo, Li Deng, Alex Acero |
| 2001 | ICASSP | High-performance robust speech recognition using stereo training data. | Li Deng, Alex Acero, Li Jiang, Jasha Droppo, Xuedong Huang |
| 2001 | ICASSP | Efficient on-line acoustic environment estimation for FCDCN in a continuous speech recognition system. | Jasha Droppo, Alex Acero, Li Deng |
| 2001 | ICASSP | MiPad: a multimodal interaction prototype. | Xuedong Huang, Alex Acero, Ciprian Chelba, Li Deng, Jasha Droppo, Doug Duchene, Joshua Goodman, Hsiao-Wuen Hon, Derek Jacoby, Li Jiang, Ricky Loynd, Milind Mahajan, Peter Mau, Scott Meredith, Salman Mughal, Salvado Neto, Mike Plumpe, Kuansan Steury, Gina Venolia, Kuansan Wang, Ye-Yi Wang |
| 2001 | Interspeech | Evaluation of the SPLICE algorithm on the Aurora2 database. | Jasha Droppo, Li Deng, Alex Acero |