Skip to content

Carol Y. Espy-Wilson

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

97

Venues

7

Active years

1986–2025

Best venue rank

A*

Where they publish

Papers

97 indexed papers, newest first.

YearVenueTitleAuthors
2025ASRUAcoustic to Articulatory Speech Inversion for Children with Velopharyngeal Insufficiency.Saba Tabatabaee, Suzanne Boyce, Liran Oren, Mark Tiede, Carol Y. Espy-Wilson
2025HRISpeaking with Robots in Noisy Environments.Shuubham Ojha, Felix Gervits, Carol Y. Espy-Wilson
2025ICASSPCPT-Boosted Wav2vec2.0: Towards Noise Robust Speech Recognition for Classroom Environments.Ahmed Adel Attia, Dorottya Demszky, Tollop gnrm, Jing Liu, Carol Y. Espy-Wilson
2025ICASSPSpeech-Based Estimation of Schizophrenia Severity Using Feature Fusion.Gowtham Premananth, Carol Y. Espy-Wilson
2025ICASSPSelf-supervised Multimodal Speech Representations for the Assessment of Schizophrenia Symptoms.Gowtham Premananth, Carol Y. Espy-Wilson
2025InterspeechFrom Weak Labels to Strong Results: Utilizing 5, 000 Hours of Noisy Classroom Transcripts with Minimal Accurate Data.Ahmed Adel Attia, Dorottya Demszky, Jing Liu, Carol Y. Espy-Wilson
2025InterspeechSubtyping Speech Errors in Childhood Speech Sound Disorders with Acoustic-to-Articulatory Speech Inversion.Nina R. Benway, Saba Tabatabaee, Benjamin Munson, Jonathan Preston, Carol Y. Espy-Wilson
2025InterspeechSpeech Kinematic Analysis from Acoustics: Scientific, Clinical and Practical Applications.Carol Y. Espy-Wilson
2025InterspeechAnalyzing the Impact of Accent on English Speech: Acoustic and Articulatory Perspectives.Gowtham Premananth, Vinith Kugathasan, Carol Y. Espy-Wilson
2025InterspeechMultimodal Biomarkers for Schizophrenia: Towards Individual Symptom Severity Estimation.Gowtham Premananth, Philip Resnik, Sonia Bansal, Deanna L. Kelly, Carol Y. Espy-Wilson
2025InterspeechEnhancing Acoustic-to-Articulatory Speech Inversion by Incorporating Nasality.Saba Tabatabaee, Suzanne Boyce, Liran Oren, Mark Tiede, Carol Y. Espy-Wilson
2025InterspeechFT-Boosted SV: Towards Noise Robust Speaker Verification for English Speaking Classroom Environments.Saba Tabatabaee, Jing Liu, Carol Y. Espy-Wilson
2024AIESKid-Whisper: Towards Bridging the Performance Gap in Automatic Speech Recognition for Children VS. Adults.Ahmed Adel Attia, Jing Liu, Wei Ai, Dorottya Demszky, Carol Y. Espy-Wilson
2024InterspeechExamining Vocal Tract Coordination in Childhood Apraxia of Speech with Acoustic-to-Articulatory Speech Inversion Feature Sets.Nina R. Benway, Jonathan L. Preston, Carol Y. Espy-Wilson
2024InterspeechA Multimodal Framework for the Assessment of the Schizophrenia Spectrum.Gowtham Premananth, Yashish M. Siriwardena, Philip Resnik, Sonia Bansal, Deanna L. Kelly, Carol Y. Espy-Wilson
2024InterspeechAccent Conversion with Articulatory Representations.Yashish M. Siriwardena, Nathan Swedlow, Audrey Howard, Evan Gitterman, Dan Darcy, Carol Y. Espy-Wilson, Andrea Fanelli
2023ICASSPMasked Autoencoders are Articulatory Learners.Ahmed Adel Attia, Carol Y. Espy-Wilson
2023ICASSPThe Secret Source : Incorporating Source Features to Improve Acoustic-To-Articulatory Speech Inversion.Yashish M. Siriwardena, Carol Y. Espy-Wilson
2023InterspeechEnhancing Speech Articulation Analysis Using A Geometric Transformation of the X-ray Microbeam Dataset.Ahmed Adel Attia, Mark Tiede, Carol Y. Espy-Wilson
2023InterspeechAcoustic-to-Articulatory Speech Inversion Features for Mispronunciation Detection of /ɹ/ in Child Speech Sound Disorders.Nina R. Benway, Yashish M. Siriwardena, Jonathan L. Preston, Elaine Hitchcock, Tara McAllister Byun, Carol Y. Espy-Wilson
2023InterspeechSpeaker-independent Speech Inversion for Estimation of Nasalance.Yashish M. Siriwardena, Carol Y. Espy-Wilson, Suzanne Boyce, Mark Tiede, Liran Oren
2023InterspeechLearning to Compute the Articulatory Representations of Speech with the MIRRORNET.Yashish M. Siriwardena, Carol Y. Espy-Wilson, Shihab A. Shamma
2022ICASSPHarmonicity Plays a Critical Role in DNN Based Versus in Biologically-Inspired Monaural Speech Segregation Systems.Rahil Parikh, Ilya Kavalerov, Carol Y. Espy-Wilson, Shihab A. Shamma
2022ICASSPMultimodal Depression Classification using Articulatory Coordination Features and Hierarchical Attention Based text Embeddings.Nadee Seneviratne, Carol Y. Espy-Wilson
2022InterspeechAn Empirical Analysis on the Vulnerabilities of End-to-End Speech Segregation Models.Rahil Parikh, Gaspar Rochette, Carol Y. Espy-Wilson, Shihab A. Shamma
2022InterspeechAcoustic To Articulatory Speech Inversion Using Multi-Resolution Spectro-Temporal Representations Of Speech Signals.Rahil Parikh, Nadee Seneviratne, Ganesh Sivaraman, Shihab A. Shamma, Carol Y. Espy-Wilson
2022InterspeechMultimodal Depression Severity Score Prediction Using Articulatory Coordination Features and Hierarchical Attention Based Text Embeddings.Nadee Seneviratne, Carol Y. Espy-Wilson
2022InterspeechAcoustic-to-articulatory Speech Inversion with Multi-task Learning.Yashish M. Siriwardena, Ganesh Sivaraman, Carol Y. Espy-Wilson
2021ICMIMultimodal Approach for Assessing Neuromotor Coordination in Schizophrenia Using Convolutional Neural Networks.Yashish M. Siriwardena, Carol Y. Espy-Wilson, Chris Kitchen, Deanna L. Kelly
2021InterspeechSpeech Based Depression Severity Level Classification Using a Multi-Stage Dilated CNN-LSTM Model.Nadee Seneviratne, Carol Y. Espy-Wilson
2021InterspeechGeneralized Dilated CNN Models for Depression Detection Using Inverted Vocal Tract Variables.Nadee Seneviratne, Carol Y. Espy-Wilson
2020InterspeechExtended Study on the Use of Vocal Tract Variables to Quantify Neuromotor Coordination in Depression.Nadee Seneviratne, James R. Williamson, Adam C. Lammert, Thomas F. Quatieri, Carol Y. Espy-Wilson
2019InterspeechAssessing Neuromotor Coordination in Depression Using Inverted Vocal Tract Variables.Carol Y. Espy-Wilson, Adam C. Lammert, Nadee Seneviratne, Thomas F. Quatieri
2019InterspeechMulti-Modal Learning for Speech Emotion Recognition: An Analysis and Comparison of ASR Outputs with Ground Truth Transcription.Saurabh Sahu, Vikramjit Mitra, Nadee Seneviratne, Carol Y. Espy-Wilson
2019InterspeechMulti-Corpus Acoustic-to-Articulatory Speech Inversion.Nadee Seneviratne, Ganesh Sivaraman, Carol Y. Espy-Wilson
2018ICASSPSemi-Supervised and Transfer Learning Approaches for Low Resource Sentiment Classification.Rahul Gupta, Saurabh Sahu, Carol Y. Espy-Wilson, Shrikanth S. Narayanan
2018ICASSPSmoothing Model Predictions Using Adversarial Training Procedures for Speech Based Emotion Recognition.Saurabh Sahu, Rahul Gupta, Ganesh Sivaraman, Carol Y. Espy-Wilson
2018InterspeechOn Enhancing Speech Emotion Recognition Using Generative Adversarial Networks.Saurabh Sahu, Rahul Gupta, Carol Y. Espy-Wilson
2018InterspeechNoise Robust Acoustic to Articulatory Speech Inversion.Nadee Seneviratne, Ganesh Sivaraman, Vikramjit Mitra, Carol Y. Espy-Wilson
2017ICASSPJoint modeling of articulatory and acoustic spaces for continuous speech recognition tasks.Vikramjit Mitra, Ganesh Sivaraman, Chris Bartels, Hosung Nam, Wen Wang, Carol Y. Espy-Wilson, Dimitra Vergyri, Horacio Franco
2017InterspeechAn Affect Prediction Approach Through Depression Severity Parameter Incorporation in Neural Networks.Rahul Gupta, Saurabh Sahu, Carol Y. Espy-Wilson, Shrikanth S. Narayanan
2017InterspeechAdversarial Auto-Encoders for Speech Based Emotion Recognition.Saurabh Sahu, Rahul Gupta, Ganesh Sivaraman, Wael AbdAlmageed, Carol Y. Espy-Wilson
2017InterspeechAnalysis of Acoustic-to-Articulatory Speech Inversion Across Different Accents and Languages.Ganesh Sivaraman, Carol Y. Espy-Wilson, Martijn Wieling
2016InterspeechSpeech Features for Depression Detection.Saurabh Sahu, Carol Y. Espy-Wilson
2016InterspeechVocal Tract Length Normalization for Speaker Independent Acoustic-to-Articulatory Speech Inversion.Ganesh Sivaraman, Vikramjit Mitra, Hosung Nam, Mark K. Tiede, Carol Y. Espy-Wilson
2015InterspeechAnalysis of coarticulated speech using estimated articulatory trajectories.Ganesh Sivaraman, Vikramjit Mitra, Mark K. Tiede, Elliot Saltzman, Louis Goldstein, Carol Y. Espy-Wilson
2014ICASSPArticulatory features from deep neural networks and their role in speech recognition.Vikramjit Mitra, Ganesh Sivaraman, Hosung Nam, Carol Y. Espy-Wilson, Elliot Saltzman
2013ICASSPA cine MRI-based study of sibilant fricatives production in post-glossectomy speakers.Xinhui Zhou, Jonghye Woo, Maureen L. Stone, Carol Y. Espy-Wilson
2012ICASSPMulticondition training of Gaussian PLDA models in i-vector space for noise and reverberation robust speaker recognition.Daniel Garcia-Romero, Xinhui Zhou, Carol Y. Espy-Wilson
2012InterspeechAutomatic intelligibility assessment of pathologic speech in head and neck cancer based on auditory-inspired spectro-temporal modulations.Xinhui Zhou, Daniel Garcia-Romero, Nima Mesgarani, Maureen L. Stone, Carol Y. Espy-Wilson, Shihab A. Shamma
2011ASRURobust speech recognition using articulatory gestures in a Dynamic Bayesian Network framework.Vikramjit Mitra, Hosung Nam, Carol Y. Espy-Wilson
2011ASRULinear versus mel frequency cepstral coefficients for speaker recognition.Xinhui Zhou, Daniel Garcia-Romero, Ramani Duraiswami, Carol Y. Espy-Wilson, Shihab A. Shamma
2011ICASSPGesture-based Dynamic Bayesian Network for noise robust speech recognition.Vikramjit Mitra, Hosung Nam, Carol Y. Espy-Wilson, Elliot Saltzman, Louis Goldstein
2011ICASSPSpeech inversion: Benefits of tract variables over pellet trajectories.Vikramjit Mitra, Hosung Nam, Carol Y. Espy-Wilson, Elliot Saltzman, Louis Goldstein
2011InterspeechAnalysis of i-vector Length Normalization in Speaker Recognition Systems.Daniel Garcia-Romero, Carol Y. Espy-Wilson
2011InterspeechAutomatic Speech Codec Identification with Applications to Tampering Detection of Speech Recordings.Jingting Zhou, Daniel Garcia-Romero, Carol Y. Espy-Wilson
2011InterspeechA Comparative Acoustic Study on Speech of Glossectomy Patients and Normal Subjects.Xinhui Zhou, Maureen L. Stone, Carol Y. Espy-Wilson
2010ICASSPAutomatic acquisition device identification from speech recordings.Daniel Garcia-Romero, Carol Y. Espy-Wilson
2010ICASSPAn MRI-based articulatory and acoustic study of lateral sound in American English.Xinhui Zhou, Carol Y. Espy-Wilson, Mark Tiede, Suzanne Boyce
2010InterspeechRobust word recognition using articulatory trajectories and gestures.Vikramjit Mitra, Hosung Nam, Carol Y. Espy-Wilson, Elliot Saltzman, Louis Goldstein
2010InterspeechA procedure for estimating gestural scores from natural speech.Hosung Nam, Vikramjit Mitra, Mark Tiede, Elliot Saltzman, Louis Goldstein, Carol Y. Espy-Wilson, Mark Hasegawa-Johnson
2009ICASSPFrom acoustics to Vocal Tract time functions.Vikramjit Mitra, I. Ycel zbek, Hosung Nam, Xinhui Zhou, Carol Y. Espy-Wilson
2009ICASSPAn algorithm for speech segregation of co-channel speech.Srikanth Vishnubhotla, Carol Y. Espy-Wilson
2009InterspeechA noise-type and level-dependent MPO-based speech enhancement architecture with variable frame analysis for noise-robust speech recognition.Vikramjit Mitra, Bengt J. Borgstrom, Carol Y. Espy-Wilson, Abeer Alwan
2009InterspeechNoise robustness of tract variables and their application to speech recognition.Vikramjit Mitra, Hosung Nam, Carol Y. Espy-Wilson, Elliot Saltzman, Louis Goldstein
2008ICASSPLanguage detection in audio content analysis.Vikramjit Mitra, Daniel Garcia-Romero, Carol Y. Espy-Wilson
2008InterspeechIntersession variability in speaker recognition: a behind the scene analysis.Daniel Garcia-Romero, Carol Y. Espy-Wilson
2008InterspeechLanguage and genre detection in audio content analysis.Vikramjit Mitra, Daniel Garcia-Romero, Carol Y. Espy-Wilson
2008InterspeechAn algorithm for multi-pitch tracking in co-channel speech.Srikanth Vishnubhotla, Carol Y. Espy-Wilson
2007InterspeechLandmark-based approach to speech recognition: an alternative to HMMs.Carol Y. Espy-Wilson, Tarun Pruthi, Amit Juneja, Om Deshmukh
2007InterspeechA semi-automatic approach for speaker mining of tapped telephone conversations.Sandeep Manocha, Carol Y. Espy-Wilson
2007InterspeechAcoustic parameters for the automatic detection of vowel nasalization.Tarun Pruthi, Carol Y. Espy-Wilson
2007InterspeechAn articulatory and acoustic study of "retroflex" and "bunched" american English rhotic sound based on MRI.Xinhui Zhou, Carol Y. Espy-Wilson, Mark Tiede, Suzanne Boyce
2006InterspeechModified phase opponency based solution to the speech separation challenge.Om Deshmukh, Carol Y. Espy-Wilson
2006InterspeechSpeech enhancement using modified phase opponency model.Om Deshmukh, Carol Y. Espy-Wilson
2006InterspeechA new set of features for text-independent speaker identification.Carol Y. Espy-Wilson, Sandeep Manocha, Srikanth Vishnubhotla
2006InterspeechAn MRI based study of the acoustic effects of sinus cavities and its application to speaker recognition.Tarun Pruthi, Carol Y. Espy-Wilson
2006InterspeechAutomatic detection of irregular phonation in continuous speech.Srikanth Vishnubhotla, Carol Y. Espy-Wilson
2005ICASSPModeling of the Front Cavity and Sublingual Space in American English Rhotic Sounds.Zhaoyan Zhang, Carol Y. Espy-Wilson, Suzanne Boyce, Mark Tiede
2005InterspeechSpeech enhancement using auditory phase opponency model.Om Deshmukh, Carol Y. Espy-Wilson
2004ICASSPA novel method for computation of periodicity, aperiodicity and pitch of speech signals.Om Deshmukh, Jawahar Singh, Carol Y. Espy-Wilson
2003ICASSPA measure of aperiodicity and periodicity in speech.Om Deshmukh, Carol Y. Espy-Wilson
2003IJCNNSpeech segmentation using probabilistic phonetic feature hierarchy and support vector machines.Amit Juneja, Carol Y. Espy-Wilson
2003InterspeechAcoustic modeling of american English lateral approximants.Zhaoyan Zhang, Carol Y. Espy-Wilson, Mark Tiede
2002ICASSPAcoustic-phonetic speech parameters for speaker-independent speech recognition.Om Deshmukh, Carol Y. Espy-Wilson, Amit Juneja
2002ICASSPAn event-based acoustic-phonetic approach for speech segmentation and E-set recognition.Amit Juneja, Om Deshmukh, Carol Y. Espy-Wilson
2000InterspeechDetection of speech landmarks using temporal cues.Ariel Salomon, Carol Y. Espy-Wilson
2000InterspeechA new strategy of formant tracking based on dynamic programming.Kun Xia, Carol Y. Espy-Wilson
1999InterspeechImprovement of electrolaryngeal speech by introducing normal excitation information.Kun Ma, Pelin Demirel, Carol Y. Espy-Wilson, Joel MacAuslan
1999InterspeechAutomatic detection of manner events based on temporal parameters.Ariel Salomon, Carol Y. Espy-Wilson
1997InterspeechThe design of acoustic parameters for speaker-independent speech recognition.Nabil N. Bitar, Carol Y. Espy-Wilson
1997InterspeechAcoustic modelling of American English /r/.Carol Y. Espy-Wilson, Shrikanth S. Narayanan, Suzanne Boyce, Abeer Alwan
1996ICASSPKnowledge-based parameters for HMM speech recognition.Nabil N. Bitar, Carol Y. Espy-Wilson
1996InterspeechCoarticulatory stability in american English /r/.Suzanne Boyce, Carol Y. Espy-Wilson
1996InterspeechEnhancement of alaryngeal speech by adaptive filtering.Carol Y. Espy-Wilson, Venkatesh R. Chari, Caroline B. Huang
1995InterspeechSpeech parameterization based on phonetic features: application to speech recognition.Nabil N. Bitar, Carol Y. Espy-Wilson
1986ICASSPA phonetically based semivowel recognition system.Carol Y. Espy-Wilson