Skip to content

Richard M. Stern

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

123

Venues

5

Active years

1983–2025

Best venue rank

A

Where they publish

Papers

123 indexed papers, newest first.

YearVenueTitleAuthors
2025ASRUIterative Feedback in the Online Active Learning Paradigm.Mark Lindsey, Francis Kubala, Richard M. Stern
2025ICASSPA Unified Metric for Simultaneous Evaluation of Error Rate and Annotation Cost.Mark Lindsey, Francis Kubala, Richard M. Stern
2023ASRUReducing the Cost of Spoof Detection Labeling using Mixed-Strategy Active Learning and Pretrained Models.Mark Lindsey, Nathaniel R. Robinson, Francis Kubala, Richard M. Stern
2023ICASSPUnsupervised Voice Type Discrimination Score Adaptation Using X-Vector Clusters.Mark Lindsey, Tyler Vuong, Richard M. Stern
2023InterspeechRespiratory distress estimation in human-robot interaction scenario.Eduardo Alvarado, Nicols Grgeda, Alejandro Luzanto, Rodrigo Mah, Jorge Wuth, Laura Mendoza, Richard M. Stern, Nstor Becerra Yoma
2022InterspeechImproved Modulation-Domain Loss for Neural-Network-based Speech Enhancement.Tyler Vuong, Richard M. Stern
2021ICASSPA Modulation-Domain Loss for Neural-Network-Based Real-Time Speech Enhancement.Tyler Vuong, Yangyang Xia, Richard M. Stern
2021InterspeechThe Application of Learnable STRF Kernels to the 2021 Fearless Steps Phase-03 SAD Challenge.Tyler Vuong, Yangyang Xia, Richard M. Stern
2021InterspeechTemporal Context in Speech Emotion Recognition.Yangyang Xia, Li-Wei Chen, Alexander Rudnicky, Richard M. Stern
2020InterspeechLearnable Spectro-Temporal Receptive Fields for Robust Voice Type Discrimination.Tyler Vuong, Yangyang Xia, Richard M. Stern
2019ICASSPRobust Recognition of Reverberant and Noisy Speech Using Coherence-based Processing.Anjali Menon, Chanwoo Kim, Richard M. Stern
2018ICASSPSound Source Separation Using Phase Difference and Reliable Mask Selection Selection.Chanwoo Kim, Anjali Menon, Michiel Bacchiani, Richard M. Stern
2018InterspeechA Priori SNR Estimation Based on a Recurrent Neural Network for Robust Speech Enhancement.Yangyang Xia, Richard M. Stern
2018ISNNA Comparative Study of Spatial Speech Separation Techniques to Improve Speech Recognition.Xinhui Zhou, Chiman Kwan, Bulent Ayhan, Chanwoo Kim, Kshitiz Kumar, Richard M. Stern
2017ASRUBinaural processing for robust recognition of degraded speech.Anjali Menon, Chanwoo Kim, Umpei Kurokawa, Richard M. Stern
2017InterspeechRobust Speech Recognition Based on Binaural Auditory Processing.Anjali Menon, Chanwoo Kim, Richard M. Stern
2017InterspeechRobustness Over Time-Varying Channels in DNN-HMM ASR Based Human-Robot Interaction.Jos Novoa, Jorge Wuth, Juan Pablo Escudero, Josu Fredes, Rodrigo Mah, Richard M. Stern, Nstor Becerra Yoma
2016InterspeechFusion Strategies for Robust Speech Recognition and Keyword Spotting for Channel- and Noise-Degraded Speech.Vikramjit Mitra, Julien van Hout, Wen Wang, Chris Bartels, Horacio Franco, Dimitra Vergyri, Abeer Alwan, Adam Janin, John H. L. Hansen, Richard M. Stern, Abhijeet Sangwan, Nelson Morgan
2016InterspeechThe Use of Locally Normalized Cepstral Coefficients (LNCC) to Improve Speaker Recognition Accuracy in Highly Reverberant Rooms.Vctor Poblete, Juan Pablo Escudero, Josu Fredes, Jos Novoa, Richard M. Stern, Simon King, Nstor Becerra Yoma
2015ICASSPEfficient audio declipping using regularized least squares.Mark J. Harvilla, Richard M. Stern
2015ICASSPTowards machines that know when they do not know: Summary of work done at 2014 Frederick Jelinek Memorial Workshop.Hynek Hermansky, Luks Burget, Jordan Cohen, Emmanuel Dupoux, Naomi Feldman, John Godfrey, Sanjeev Khudanpur, Matthew Maciejewski, Sri Harish Reddy Mallidi, Anjali Menon, Tetsuji Ogawa, Vijayaditya Peddinti, Richard C. Rose, Richard M. Stern, Matthew Wiesner, Karel Vesel
2015InterspeechRobustness to additive noise of locally-normalized cepstral coefficients in speaker verification.Josu Fredes, Jos Novoa, Vctor Poblete, Simon King, Richard M. Stern, Nstor Becerra Yoma
2015InterspeechRobust parameter estimation for audio declipping in noise.Mark J. Harvilla, Richard M. Stern
2014ICASSPAn analysis of binaural spectro-temporal masking as nonlinear beamforming.Amir R. Moghimi, Richard M. Stern
2014InterspeechLeast squares signal declipping for robust speech recognition.Mark J. Harvilla, Richard M. Stern
2014InterspeechRobust speech recognition using temporal masking and thresholding algorithm.Chanwoo Kim, Kean K. Chin, Michiel Bacchiani, Richard M. Stern
2014InterspeechPost-masking: a hybrid approach to array processing for speech recognition.Amir R. Moghimi, Bhiksha Raj, Richard M. Stern
2014InterspeechRobust speech recognition in reverberant environments using subband-based steady-state monaural and binaural suppression.Hyung-Min Park, Matthew Maciejewski, Chanwoo Kim, Richard M. Stern
2013InterspeechOptimization of sigmoidal rate-level function based on acoustic features.Vctor Poblete, Nstor Becerra Yoma, Richard M. Stern
2012ICASSPHistogram-based subband powerwarping and spectral averaging for robust speech recognition under matched and multistyle training.Mark Harvilla, Richard M. Stern
2012ICASSPTwo-microphone source separation algorithm based on statistical modeling of angle distributions.Chanwoo Kim, Charbel El Khawand, Richard M. Stern
2012ICASSPPower-Normalized Cepstral Coefficients (PNCC) for robust speech recognition.Chanwoo Kim, Richard M. Stern
2011ICASSPBinaural sound source separation motivated by auditory processing.Chanwoo Kim, Kshitiz Kumar, Richard M. Stern
2011ICASSPDelta-spectral cepstral coefficients for robust speech recognition.Kshitiz Kumar, Chanwoo Kim, Richard M. Stern
2011ICASSPAn iterative least-squares technique for dereverberation.Kshitiz Kumar, Bhiksha Raj, Rita Singh, Richard M. Stern
2011ICASSPGammatone sub-band magnitude-domain dereverberation for ASR.Kshitiz Kumar, Rita Singh, Bhiksha Raj, Richard M. Stern
2010ICASSPA hybrid physical and statistical dynamic articulatory framework incorporating analysis-by-synthesis for improved phone classification.Ziad Al Bawab, Bhiksha Raj, Richard M. Stern
2010ICASSPLearning-based auditory encoding for robust speech recognition.Yu-Hsiang Bosco Chiu, Bhiksha Raj, Richard M. Stern
2010ICASSPFeature extraction for robust speech recognition based on maximizing the sharpness of the power distribution and on power flooring.Chanwoo Kim, Richard M. Stern
2010ICASSPMaximum-likelihood-based cepstral inverse filtering for blind speech dereverberation.Kshitiz Kumar, Richard M. Stern
2010InterspeechNonlinear enhancement of onset for robust speech recognition.Chanwoo Kim, Richard M. Stern
2010InterspeechAutomatic selection of thresholds for signal separation algorithms based on interaural delay.Chanwoo Kim, Richard M. Stern, Kiwan Eom, Jaewon Lee
2009ASRURobust speech recognition using a Small Power Boosting algorithm.Chanwoo Kim, Kshitiz Kumar, Richard M. Stern
2009ASRUPower function-based power distribution normalization algorithm for robust speech recognition.Chanwoo Kim, Richard M. Stern
2009ICASSPMinimum variance modulation filter for robust speech recognition.Yu-Hsiang Bosco Chiu, Richard M. Stern
2009InterspeechDeriving vocal tract shapes from electromagnetic articulograph data via geometric adaptation and matching.Ziad Al Bawab, Lorenzo Turicchia, Richard M. Stern, Bhiksha Raj
2009InterspeechUnsupervised training scheme with non-stereo data for empirical feature vector compensation.Luis Buera, Antonio Miguel, Alfonso Ortega, Eduardo Lleida, Richard M. Stern
2009InterspeechTowards fusion of feature extraction and acoustic model training: a top down process for robust speech recognition.Yu-Hsiang Bosco Chiu, Bhiksha Raj, Richard M. Stern
2009InterspeechSpeaker segmentation and clustering for simultaneously presented speech.Lingyun Gu, Richard M. Stern
2009InterspeechSignal separation for robust speech recognition based on phase difference information obtained in the frequency domain.Chanwoo Kim, Kshitiz Kumar, Bhiksha Raj, Richard M. Stern
2009InterspeechFeature extraction for robust speech recognition using a power-law nonlinearity and power-bias subtraction.Chanwoo Kim, Richard M. Stern
2008ICASSPAnalysis-by-synthesis features for speech recognition.Ziad Al Bawab, Bhiksha Raj, Richard M. Stern
2008ICASSPSingle-channel speech separation based on modulation frequency.Lingyun Gu, Richard M. Stern
2008ICASSPEnvironment-invariant compensation for reverberation using linear post-filtering for minimum distortion.Kshitiz Kumar, Richard M. Stern
2008InterspeechAnalysis of physiologically-motivated signal processing for robust speech recognition.Yu-Hsiang Bosco Chiu, Richard M. Stern
2008InterspeechRobust signal-to-noise ratio estimation based on waveform amplitude distribution analysis.Chanwoo Kim, Richard M. Stern
2007ICASSPProfile View Lip Reading.Kshitiz Kumar, Tsuhan Chen, Richard M. Stern
2007ICASSPMissing Feature Speech Recognition using Dereverberation and Echo Suppression in Reverberant Environments.Hyung-Min Park, Richard M. Stern
2007Interspeech"polyaural" array processing for automatic speech recognition in degraded environments.Richard M. Stern, Evandro B. Gouva, Govindarajan Thattai
2006ICASSPBand-Independent Mask Estimation for Missing-Feature Reconstruction in the Presence of Unknown Background Noise.Wooil Kim, Richard M. Stern
2006ICASSPSpatial Separation of Speech Signals Using Continuously-Variable Masks Estimated From Comparisons of Zero Crossings.Hyung-Min Park, Richard M. Stern
2006InterspeechAn integrated approach to improve speech recognition rate for non-native speakers.Yunbin Deng, Xiaokun Li, Chiman Kwan, Roger Xu, Bhiksha Raj, Richard M. Stern, David Williamson
2006InterspeechPhysiologically-motivated synchrony-based processing for robust automatic speech recognition.Chanwoo Kim, Yu-Hsiang Bosco Chiu, Richard M. Stern
2006InterspeechVoting for two speaker segmentation.Narayanaswamy Balakrishnan, Rashmi Gangadharaiah, Richard M. Stern
2005InterspeechEnvironment-independent mask estimation for missing-feature reconstruction.Wooil Kim, Richard M. Stern, Hanseok Ko
2004ICASSPFeature generation based on maximum normalized acoustic likelihood for improved speech recognition.Xiang Li, Richard M. Stern
2004ICASSPOn tracking noise with linear dynamical system models.Bhiksha Raj, Rita Singh, Richard M. Stern
2004ICASSPParameter sharing in subband likelihood-maximizing beamforming for speech recognition using microphone arrays.Michael L. Seltzer, Richard M. Stern
2004InterspeechParallel feature generation based on maximizing normalized acoustic likelihood.Xiang Li, Richard M. Stern
2003ICASSPTraining of stream weights for the decoding of speech using parallel feature streams.Xiang Li, Richard M. Stern
2003ICASSPSubband parameter optimization of microphone arrays for speech recognition in reverberant environments.Michael L. Seltzer, Richard M. Stern
2003InterspeechFeature generation based on maximum classification probability for improved speech recognition.Xiang Li, Richard M. Stern
2003InterspeechDuration normalization and hypothesis combination for improved spontaneous speech recognition.Jon P. Nedel, Richard M. Stern
2003InterspeechNormalization of time-derivative parameters using histogram equalization.Yasunari Obuchi, Richard M. Stern
2002ICASSPSpeech recognizer-based microphone array processing for robust hands-free speech recognition.Michael L. Seltzer, Bhiksha Raj, Richard M. Stern
2002InterspeechCombining search spaces of heterogeneous recognizers for improved speech recogniton.Xiang Li, Rita Singh, Richard M. Stern
2001ICASSPDuration normalization for improved recognition of spontaneous and read speech via missing feature methods.Jon P. Nedel, Richard M. Stern
2001ICASSPSpeech in Noisy Environments: robust automatic segmentation, feature extraction, and hypothesis combination.Rita Singh, Michael L. Seltzer, Bhiksha Raj, Richard M. Stern
2000ICASSPInter-class MLLR for speaker adaptation.Sam-Joo Doh, Richard M. Stern
2000ICASSPAutomatic generation of phone sets and lexical transcriptions.Rita Singh, Bhiksha Raj, Richard M. Stern
2000InterspeechUsing class weighting in inter-class MLLR.Sam-Joo Doh, Richard M. Stern
2000InterspeechInstantaneous-distortion based weighted acoustic modeling for robust recognition of coded speech.Juan M. Huerta, Richard M. Stern
2000InterspeechPhone transition acoustic modeling: application to speaker independent and spontaneous speech systems.Jon P. Nedel, Rita Singh, Richard M. Stern
2000InterspeechAutomatic subword unit refinement for spontaneous speech recognition via phone splitting.Jon P. Nedel, Rita Singh, Richard M. Stern
2000InterspeechReconstruction of damaged spectrographic features for robust speech recognition.Bhiksha Raj, Michael L. Seltzer, Richard M. Stern
2000InterspeechClassifier-based mask estimation for missing feature methods of robust speech recognition.Michael L. Seltzer, Bhiksha Raj, Richard M. Stern
2000InterspeechStructured redefinition of sound units by merging and splitting for improved speech recognition.Rita Singh, Bhiksha Raj, Richard M. Stern
1999ICASSPAutomatic clustering and generation of contextual questions for tied states in hidden Markov models.Rita Singh, Bhiksha Raj, Richard M. Stern
1999InterspeechDomain adduced state tying for cross-domain acoustic modelling.Rita Singh, Bhiksha Raj, Richard M. Stern
1998InterspeechSpeech recognition from GSM codec parameters.Juan M. Huerta, Richard M. Stern
1998InterspeechInference of missing spectrographic features for robust speech recognition.Bhiksha Raj, Rita Singh, Richard M. Stern
1997ICASSPThe effects of background music on speech recognition accuracy.Bhiksha Raj, Vipul N. Parikh, Richard M. Stern
1997InterspeechSpeaker normalization through formant-based warping of the frequency scale.Evandro B. Gouva, Richard M. Stern
1997InterspeechCompensation for environmental and speaker variability by normalization of pole locations.Juan M. Huerta, Richard M. Stern
1996ICASSPA vector Taylor series approach for environment-independent speech recognition.Pedro J. Moreno, Bhiksha Raj, Richard M. Stern
1996InterspeechCepstral compensation by polynomial approximation for environment-independent speech recognition.Bhiksha Raj, Evandro Bacci Gouva, Pedro J. Moreno, Richard M. Stern
1995ICASSPMultivariate-Gaussian-based cepstral normalization for robust speech recognition.Pedro J. Moreno, Bhiksha Raj, Evandro B. Gouva, Richard M. Stern
1995ICASSPOn the effects of speech rate in large vocabulary speech recognition systems.Matthew A. Siegler, Richard M. Stern
1995InterspeechA unified approach for robust speech recognition.Pedro J. Moreno, Bhiksha Raj, Richard M. Stern
1994ICASSPEnvironment normalization for robust speech recognition using direct cepstral comparison.Fu-Hua Liu, Richard M. Stern, Alejandro Acero, Pedro J. Moreno
1994ICASSPSources of degradation of speech recognition in the telephone network.Pedro J. Moreno, Richard M. Stern
1994InterspeechRobust speech recognition in the automobile.Nobutoshi Hanai, Richard M. Stern
1994InterspeechEnvironmental robustness in automatic speech recognition using physiologic ally-motivated signal processing.Yoshiaki Ohshima, Richard M. Stern
1994InterspeechSignal processing for robust speech recognition.Richard M. Stern, Fu-Hua Liu, Pedro J. Moreno, Alejandro Acero
1994NAACLSignal Processing for Robust Speech Recognition.Fu-Hua Liu, Pedro J. Moreno, Richard M. Stern, Alejandro Acero
1993ICASSPMulti-microphone correlation-based processing for robust speech recognition.Thomas M. Sullivan, Richard M. Stern
1993NAACLEfficient Cepstral Normalization For Robust Speech Recognition.Fu-Hua Liu, Richard M. Stern, Xuedong Huang, Alejandro Acero
1992ICASSPEfficient joint compensation of speech for the effects of additive noise and linear filtering.Fu-Hua Liu, Alejandro Acero, Richard M. Stern
1992InterspeechMultiple approaches to robust speech recognition.Richard M. Stern, Fu-Hua Liu, Yoshiaki Ohshima, Thomas M. Sullivan, Alejandro Acero
1992NAACLMultiple Approaches to Robust Speech Recognition.Richard M. Stern, Fu-Hua Liu, Yoshiaki Ohshima, Thomas M. Sullivan, Alejandro Acero
1992NAACLSpeech Understanding in Open Tasks.Wayne H. Ward, Sunil lssar, Xuedong Huang, Hsiao-Wuen Hon, Mei-Yuh Hwang, Sheryl Young, Michael Matessa, Fu-Hua Liu, Richard M. Stern
1991ICASSPRobust speech recognition by normalization of the acoustic space.Alejandro Acero, Richard M. Stern
1991ICASSPSpeaker adaptation in continuous speech recognition via estimation of correlated mean vectors.William A. Rozzi, Richard M. Stern
1990ICASSPEnvironmental robustness in automatic speech recognition.Alejandro Acero, Richard M. Stern
1990InterspeechAcoustical pre-processing for robust spoken language systems.Alejandro Acero, Richard M. Stern
1990NAACLTowards Environment-Independent Spoken Language Systems.Alejandro Acero, Richard M. Stern
1990NAACLOverview of the Third DARPA Speech and Natural Language Workshop.Richard M. Stern
1989NAACLACOUSTICAL PRE-PROCESSING FOR ROBUST SPEECH RECOGNITION.Richard M. Stern, Alejandro Acero
1988ICASSPParsing spoken phrases despite missing words.Wayne H. Ward, Alexander G. Hauptmann, Richard M. Stern, Thomas Chanak
1987ICASSPSentence parsing with weak grammatical constraints.Richard M. Stern, Wayne H. Ward, Alexander G. Hauptmann, Juan Leon
1984ICASSPUnsupervised adaptation to new speakers in feature-based letter recognition.Mosh J. Lasry, Richard M. Stern
1983ICASSPFeature-based speaker-independent recognition of isolated english letters.Ronald A. Cole, Richard M. Stern, Michael S. Phillips, Scott M. Brill, Andrew P. Pilant, Philippe Specker
1983ICASSPDynamic speaker adaptation for isolated letter recognition using MAP estimation.Richard M. Stern, Mosh J. Lasry