| 2025 | ASRU | Iterative Feedback in the Online Active Learning Paradigm. | Mark Lindsey, Francis Kubala, Richard M. Stern |
| 2025 | ICASSP | A Unified Metric for Simultaneous Evaluation of Error Rate and Annotation Cost. | Mark Lindsey, Francis Kubala, Richard M. Stern |
| 2023 | ASRU | Reducing the Cost of Spoof Detection Labeling using Mixed-Strategy Active Learning and Pretrained Models. | Mark Lindsey, Nathaniel R. Robinson, Francis Kubala, Richard M. Stern |
| 2023 | ICASSP | Unsupervised Voice Type Discrimination Score Adaptation Using X-Vector Clusters. | Mark Lindsey, Tyler Vuong, Richard M. Stern |
| 2023 | Interspeech | Respiratory distress estimation in human-robot interaction scenario. | Eduardo Alvarado, Nicols Grgeda, Alejandro Luzanto, Rodrigo Mah, Jorge Wuth, Laura Mendoza, Richard M. Stern, Nstor Becerra Yoma |
| 2022 | Interspeech | Improved Modulation-Domain Loss for Neural-Network-based Speech Enhancement. | Tyler Vuong, Richard M. Stern |
| 2021 | ICASSP | A Modulation-Domain Loss for Neural-Network-Based Real-Time Speech Enhancement. | Tyler Vuong, Yangyang Xia, Richard M. Stern |
| 2021 | Interspeech | The Application of Learnable STRF Kernels to the 2021 Fearless Steps Phase-03 SAD Challenge. | Tyler Vuong, Yangyang Xia, Richard M. Stern |
| 2021 | Interspeech | Temporal Context in Speech Emotion Recognition. | Yangyang Xia, Li-Wei Chen, Alexander Rudnicky, Richard M. Stern |
| 2020 | Interspeech | Learnable Spectro-Temporal Receptive Fields for Robust Voice Type Discrimination. | Tyler Vuong, Yangyang Xia, Richard M. Stern |
| 2019 | ICASSP | Robust Recognition of Reverberant and Noisy Speech Using Coherence-based Processing. | Anjali Menon, Chanwoo Kim, Richard M. Stern |
| 2018 | ICASSP | Sound Source Separation Using Phase Difference and Reliable Mask Selection Selection. | Chanwoo Kim, Anjali Menon, Michiel Bacchiani, Richard M. Stern |
| 2018 | Interspeech | A Priori SNR Estimation Based on a Recurrent Neural Network for Robust Speech Enhancement. | Yangyang Xia, Richard M. Stern |
| 2018 | ISNN | A Comparative Study of Spatial Speech Separation Techniques to Improve Speech Recognition. | Xinhui Zhou, Chiman Kwan, Bulent Ayhan, Chanwoo Kim, Kshitiz Kumar, Richard M. Stern |
| 2017 | ASRU | Binaural processing for robust recognition of degraded speech. | Anjali Menon, Chanwoo Kim, Umpei Kurokawa, Richard M. Stern |
| 2017 | Interspeech | Robust Speech Recognition Based on Binaural Auditory Processing. | Anjali Menon, Chanwoo Kim, Richard M. Stern |
| 2017 | Interspeech | Robustness Over Time-Varying Channels in DNN-HMM ASR Based Human-Robot Interaction. | Jos Novoa, Jorge Wuth, Juan Pablo Escudero, Josu Fredes, Rodrigo Mah, Richard M. Stern, Nstor Becerra Yoma |
| 2016 | Interspeech | Fusion Strategies for Robust Speech Recognition and Keyword Spotting for Channel- and Noise-Degraded Speech. | Vikramjit Mitra, Julien van Hout, Wen Wang, Chris Bartels, Horacio Franco, Dimitra Vergyri, Abeer Alwan, Adam Janin, John H. L. Hansen, Richard M. Stern, Abhijeet Sangwan, Nelson Morgan |
| 2016 | Interspeech | The Use of Locally Normalized Cepstral Coefficients (LNCC) to Improve Speaker Recognition Accuracy in Highly Reverberant Rooms. | Vctor Poblete, Juan Pablo Escudero, Josu Fredes, Jos Novoa, Richard M. Stern, Simon King, Nstor Becerra Yoma |
| 2015 | ICASSP | Efficient audio declipping using regularized least squares. | Mark J. Harvilla, Richard M. Stern |
| 2015 | ICASSP | Towards machines that know when they do not know: Summary of work done at 2014 Frederick Jelinek Memorial Workshop. | Hynek Hermansky, Luks Burget, Jordan Cohen, Emmanuel Dupoux, Naomi Feldman, John Godfrey, Sanjeev Khudanpur, Matthew Maciejewski, Sri Harish Reddy Mallidi, Anjali Menon, Tetsuji Ogawa, Vijayaditya Peddinti, Richard C. Rose, Richard M. Stern, Matthew Wiesner, Karel Vesel |
| 2015 | Interspeech | Robustness to additive noise of locally-normalized cepstral coefficients in speaker verification. | Josu Fredes, Jos Novoa, Vctor Poblete, Simon King, Richard M. Stern, Nstor Becerra Yoma |
| 2015 | Interspeech | Robust parameter estimation for audio declipping in noise. | Mark J. Harvilla, Richard M. Stern |
| 2014 | ICASSP | An analysis of binaural spectro-temporal masking as nonlinear beamforming. | Amir R. Moghimi, Richard M. Stern |
| 2014 | Interspeech | Least squares signal declipping for robust speech recognition. | Mark J. Harvilla, Richard M. Stern |
| 2014 | Interspeech | Robust speech recognition using temporal masking and thresholding algorithm. | Chanwoo Kim, Kean K. Chin, Michiel Bacchiani, Richard M. Stern |
| 2014 | Interspeech | Post-masking: a hybrid approach to array processing for speech recognition. | Amir R. Moghimi, Bhiksha Raj, Richard M. Stern |
| 2014 | Interspeech | Robust speech recognition in reverberant environments using subband-based steady-state monaural and binaural suppression. | Hyung-Min Park, Matthew Maciejewski, Chanwoo Kim, Richard M. Stern |
| 2013 | Interspeech | Optimization of sigmoidal rate-level function based on acoustic features. | Vctor Poblete, Nstor Becerra Yoma, Richard M. Stern |
| 2012 | ICASSP | Histogram-based subband powerwarping and spectral averaging for robust speech recognition under matched and multistyle training. | Mark Harvilla, Richard M. Stern |
| 2012 | ICASSP | Two-microphone source separation algorithm based on statistical modeling of angle distributions. | Chanwoo Kim, Charbel El Khawand, Richard M. Stern |
| 2012 | ICASSP | Power-Normalized Cepstral Coefficients (PNCC) for robust speech recognition. | Chanwoo Kim, Richard M. Stern |
| 2011 | ICASSP | Binaural sound source separation motivated by auditory processing. | Chanwoo Kim, Kshitiz Kumar, Richard M. Stern |
| 2011 | ICASSP | Delta-spectral cepstral coefficients for robust speech recognition. | Kshitiz Kumar, Chanwoo Kim, Richard M. Stern |
| 2011 | ICASSP | An iterative least-squares technique for dereverberation. | Kshitiz Kumar, Bhiksha Raj, Rita Singh, Richard M. Stern |
| 2011 | ICASSP | Gammatone sub-band magnitude-domain dereverberation for ASR. | Kshitiz Kumar, Rita Singh, Bhiksha Raj, Richard M. Stern |
| 2010 | ICASSP | A hybrid physical and statistical dynamic articulatory framework incorporating analysis-by-synthesis for improved phone classification. | Ziad Al Bawab, Bhiksha Raj, Richard M. Stern |
| 2010 | ICASSP | Learning-based auditory encoding for robust speech recognition. | Yu-Hsiang Bosco Chiu, Bhiksha Raj, Richard M. Stern |
| 2010 | ICASSP | Feature extraction for robust speech recognition based on maximizing the sharpness of the power distribution and on power flooring. | Chanwoo Kim, Richard M. Stern |
| 2010 | ICASSP | Maximum-likelihood-based cepstral inverse filtering for blind speech dereverberation. | Kshitiz Kumar, Richard M. Stern |
| 2010 | Interspeech | Nonlinear enhancement of onset for robust speech recognition. | Chanwoo Kim, Richard M. Stern |
| 2010 | Interspeech | Automatic selection of thresholds for signal separation algorithms based on interaural delay. | Chanwoo Kim, Richard M. Stern, Kiwan Eom, Jaewon Lee |
| 2009 | ASRU | Robust speech recognition using a Small Power Boosting algorithm. | Chanwoo Kim, Kshitiz Kumar, Richard M. Stern |
| 2009 | ASRU | Power function-based power distribution normalization algorithm for robust speech recognition. | Chanwoo Kim, Richard M. Stern |
| 2009 | ICASSP | Minimum variance modulation filter for robust speech recognition. | Yu-Hsiang Bosco Chiu, Richard M. Stern |
| 2009 | Interspeech | Deriving vocal tract shapes from electromagnetic articulograph data via geometric adaptation and matching. | Ziad Al Bawab, Lorenzo Turicchia, Richard M. Stern, Bhiksha Raj |
| 2009 | Interspeech | Unsupervised training scheme with non-stereo data for empirical feature vector compensation. | Luis Buera, Antonio Miguel, Alfonso Ortega, Eduardo Lleida, Richard M. Stern |
| 2009 | Interspeech | Towards fusion of feature extraction and acoustic model training: a top down process for robust speech recognition. | Yu-Hsiang Bosco Chiu, Bhiksha Raj, Richard M. Stern |
| 2009 | Interspeech | Speaker segmentation and clustering for simultaneously presented speech. | Lingyun Gu, Richard M. Stern |
| 2009 | Interspeech | Signal separation for robust speech recognition based on phase difference information obtained in the frequency domain. | Chanwoo Kim, Kshitiz Kumar, Bhiksha Raj, Richard M. Stern |
| 2009 | Interspeech | Feature extraction for robust speech recognition using a power-law nonlinearity and power-bias subtraction. | Chanwoo Kim, Richard M. Stern |
| 2008 | ICASSP | Analysis-by-synthesis features for speech recognition. | Ziad Al Bawab, Bhiksha Raj, Richard M. Stern |
| 2008 | ICASSP | Single-channel speech separation based on modulation frequency. | Lingyun Gu, Richard M. Stern |
| 2008 | ICASSP | Environment-invariant compensation for reverberation using linear post-filtering for minimum distortion. | Kshitiz Kumar, Richard M. Stern |
| 2008 | Interspeech | Analysis of physiologically-motivated signal processing for robust speech recognition. | Yu-Hsiang Bosco Chiu, Richard M. Stern |
| 2008 | Interspeech | Robust signal-to-noise ratio estimation based on waveform amplitude distribution analysis. | Chanwoo Kim, Richard M. Stern |
| 2007 | ICASSP | Profile View Lip Reading. | Kshitiz Kumar, Tsuhan Chen, Richard M. Stern |
| 2007 | ICASSP | Missing Feature Speech Recognition using Dereverberation and Echo Suppression in Reverberant Environments. | Hyung-Min Park, Richard M. Stern |
| 2007 | Interspeech | "polyaural" array processing for automatic speech recognition in degraded environments. | Richard M. Stern, Evandro B. Gouva, Govindarajan Thattai |
| 2006 | ICASSP | Band-Independent Mask Estimation for Missing-Feature Reconstruction in the Presence of Unknown Background Noise. | Wooil Kim, Richard M. Stern |
| 2006 | ICASSP | Spatial Separation of Speech Signals Using Continuously-Variable Masks Estimated From Comparisons of Zero Crossings. | Hyung-Min Park, Richard M. Stern |
| 2006 | Interspeech | An integrated approach to improve speech recognition rate for non-native speakers. | Yunbin Deng, Xiaokun Li, Chiman Kwan, Roger Xu, Bhiksha Raj, Richard M. Stern, David Williamson |
| 2006 | Interspeech | Physiologically-motivated synchrony-based processing for robust automatic speech recognition. | Chanwoo Kim, Yu-Hsiang Bosco Chiu, Richard M. Stern |
| 2006 | Interspeech | Voting for two speaker segmentation. | Narayanaswamy Balakrishnan, Rashmi Gangadharaiah, Richard M. Stern |
| 2005 | Interspeech | Environment-independent mask estimation for missing-feature reconstruction. | Wooil Kim, Richard M. Stern, Hanseok Ko |
| 2004 | ICASSP | Feature generation based on maximum normalized acoustic likelihood for improved speech recognition. | Xiang Li, Richard M. Stern |
| 2004 | ICASSP | On tracking noise with linear dynamical system models. | Bhiksha Raj, Rita Singh, Richard M. Stern |
| 2004 | ICASSP | Parameter sharing in subband likelihood-maximizing beamforming for speech recognition using microphone arrays. | Michael L. Seltzer, Richard M. Stern |
| 2004 | Interspeech | Parallel feature generation based on maximizing normalized acoustic likelihood. | Xiang Li, Richard M. Stern |
| 2003 | ICASSP | Training of stream weights for the decoding of speech using parallel feature streams. | Xiang Li, Richard M. Stern |
| 2003 | ICASSP | Subband parameter optimization of microphone arrays for speech recognition in reverberant environments. | Michael L. Seltzer, Richard M. Stern |
| 2003 | Interspeech | Feature generation based on maximum classification probability for improved speech recognition. | Xiang Li, Richard M. Stern |
| 2003 | Interspeech | Duration normalization and hypothesis combination for improved spontaneous speech recognition. | Jon P. Nedel, Richard M. Stern |
| 2003 | Interspeech | Normalization of time-derivative parameters using histogram equalization. | Yasunari Obuchi, Richard M. Stern |
| 2002 | ICASSP | Speech recognizer-based microphone array processing for robust hands-free speech recognition. | Michael L. Seltzer, Bhiksha Raj, Richard M. Stern |
| 2002 | Interspeech | Combining search spaces of heterogeneous recognizers for improved speech recogniton. | Xiang Li, Rita Singh, Richard M. Stern |
| 2001 | ICASSP | Duration normalization for improved recognition of spontaneous and read speech via missing feature methods. | Jon P. Nedel, Richard M. Stern |
| 2001 | ICASSP | Speech in Noisy Environments: robust automatic segmentation, feature extraction, and hypothesis combination. | Rita Singh, Michael L. Seltzer, Bhiksha Raj, Richard M. Stern |
| 2000 | ICASSP | Inter-class MLLR for speaker adaptation. | Sam-Joo Doh, Richard M. Stern |
| 2000 | ICASSP | Automatic generation of phone sets and lexical transcriptions. | Rita Singh, Bhiksha Raj, Richard M. Stern |
| 2000 | Interspeech | Using class weighting in inter-class MLLR. | Sam-Joo Doh, Richard M. Stern |
| 2000 | Interspeech | Instantaneous-distortion based weighted acoustic modeling for robust recognition of coded speech. | Juan M. Huerta, Richard M. Stern |
| 2000 | Interspeech | Phone transition acoustic modeling: application to speaker independent and spontaneous speech systems. | Jon P. Nedel, Rita Singh, Richard M. Stern |
| 2000 | Interspeech | Automatic subword unit refinement for spontaneous speech recognition via phone splitting. | Jon P. Nedel, Rita Singh, Richard M. Stern |
| 2000 | Interspeech | Reconstruction of damaged spectrographic features for robust speech recognition. | Bhiksha Raj, Michael L. Seltzer, Richard M. Stern |
| 2000 | Interspeech | Classifier-based mask estimation for missing feature methods of robust speech recognition. | Michael L. Seltzer, Bhiksha Raj, Richard M. Stern |
| 2000 | Interspeech | Structured redefinition of sound units by merging and splitting for improved speech recognition. | Rita Singh, Bhiksha Raj, Richard M. Stern |
| 1999 | ICASSP | Automatic clustering and generation of contextual questions for tied states in hidden Markov models. | Rita Singh, Bhiksha Raj, Richard M. Stern |
| 1999 | Interspeech | Domain adduced state tying for cross-domain acoustic modelling. | Rita Singh, Bhiksha Raj, Richard M. Stern |
| 1998 | Interspeech | Speech recognition from GSM codec parameters. | Juan M. Huerta, Richard M. Stern |
| 1998 | Interspeech | Inference of missing spectrographic features for robust speech recognition. | Bhiksha Raj, Rita Singh, Richard M. Stern |
| 1997 | ICASSP | The effects of background music on speech recognition accuracy. | Bhiksha Raj, Vipul N. Parikh, Richard M. Stern |
| 1997 | Interspeech | Speaker normalization through formant-based warping of the frequency scale. | Evandro B. Gouva, Richard M. Stern |
| 1997 | Interspeech | Compensation for environmental and speaker variability by normalization of pole locations. | Juan M. Huerta, Richard M. Stern |
| 1996 | ICASSP | A vector Taylor series approach for environment-independent speech recognition. | Pedro J. Moreno, Bhiksha Raj, Richard M. Stern |
| 1996 | Interspeech | Cepstral compensation by polynomial approximation for environment-independent speech recognition. | Bhiksha Raj, Evandro Bacci Gouva, Pedro J. Moreno, Richard M. Stern |
| 1995 | ICASSP | Multivariate-Gaussian-based cepstral normalization for robust speech recognition. | Pedro J. Moreno, Bhiksha Raj, Evandro B. Gouva, Richard M. Stern |
| 1995 | ICASSP | On the effects of speech rate in large vocabulary speech recognition systems. | Matthew A. Siegler, Richard M. Stern |
| 1995 | Interspeech | A unified approach for robust speech recognition. | Pedro J. Moreno, Bhiksha Raj, Richard M. Stern |
| 1994 | ICASSP | Environment normalization for robust speech recognition using direct cepstral comparison. | Fu-Hua Liu, Richard M. Stern, Alejandro Acero, Pedro J. Moreno |
| 1994 | ICASSP | Sources of degradation of speech recognition in the telephone network. | Pedro J. Moreno, Richard M. Stern |
| 1994 | Interspeech | Robust speech recognition in the automobile. | Nobutoshi Hanai, Richard M. Stern |
| 1994 | Interspeech | Environmental robustness in automatic speech recognition using physiologic ally-motivated signal processing. | Yoshiaki Ohshima, Richard M. Stern |
| 1994 | Interspeech | Signal processing for robust speech recognition. | Richard M. Stern, Fu-Hua Liu, Pedro J. Moreno, Alejandro Acero |
| 1994 | NAACL | Signal Processing for Robust Speech Recognition. | Fu-Hua Liu, Pedro J. Moreno, Richard M. Stern, Alejandro Acero |
| 1993 | ICASSP | Multi-microphone correlation-based processing for robust speech recognition. | Thomas M. Sullivan, Richard M. Stern |
| 1993 | NAACL | Efficient Cepstral Normalization For Robust Speech Recognition. | Fu-Hua Liu, Richard M. Stern, Xuedong Huang, Alejandro Acero |
| 1992 | ICASSP | Efficient joint compensation of speech for the effects of additive noise and linear filtering. | Fu-Hua Liu, Alejandro Acero, Richard M. Stern |
| 1992 | Interspeech | Multiple approaches to robust speech recognition. | Richard M. Stern, Fu-Hua Liu, Yoshiaki Ohshima, Thomas M. Sullivan, Alejandro Acero |
| 1992 | NAACL | Multiple Approaches to Robust Speech Recognition. | Richard M. Stern, Fu-Hua Liu, Yoshiaki Ohshima, Thomas M. Sullivan, Alejandro Acero |
| 1992 | NAACL | Speech Understanding in Open Tasks. | Wayne H. Ward, Sunil lssar, Xuedong Huang, Hsiao-Wuen Hon, Mei-Yuh Hwang, Sheryl Young, Michael Matessa, Fu-Hua Liu, Richard M. Stern |
| 1991 | ICASSP | Robust speech recognition by normalization of the acoustic space. | Alejandro Acero, Richard M. Stern |
| 1991 | ICASSP | Speaker adaptation in continuous speech recognition via estimation of correlated mean vectors. | William A. Rozzi, Richard M. Stern |
| 1990 | ICASSP | Environmental robustness in automatic speech recognition. | Alejandro Acero, Richard M. Stern |
| 1990 | Interspeech | Acoustical pre-processing for robust spoken language systems. | Alejandro Acero, Richard M. Stern |
| 1990 | NAACL | Towards Environment-Independent Spoken Language Systems. | Alejandro Acero, Richard M. Stern |
| 1990 | NAACL | Overview of the Third DARPA Speech and Natural Language Workshop. | Richard M. Stern |
| 1989 | NAACL | ACOUSTICAL PRE-PROCESSING FOR ROBUST SPEECH RECOGNITION. | Richard M. Stern, Alejandro Acero |
| 1988 | ICASSP | Parsing spoken phrases despite missing words. | Wayne H. Ward, Alexander G. Hauptmann, Richard M. Stern, Thomas Chanak |
| 1987 | ICASSP | Sentence parsing with weak grammatical constraints. | Richard M. Stern, Wayne H. Ward, Alexander G. Hauptmann, Juan Leon |
| 1984 | ICASSP | Unsupervised adaptation to new speakers in feature-based letter recognition. | Mosh J. Lasry, Richard M. Stern |
| 1983 | ICASSP | Feature-based speaker-independent recognition of isolated english letters. | Ronald A. Cole, Richard M. Stern, Michael S. Phillips, Scott M. Brill, Andrew P. Pilant, Philippe Specker |
| 1983 | ICASSP | Dynamic speaker adaptation for isolated letter recognition using MAP estimation. | Richard M. Stern, Mosh J. Lasry |