| 2025 | ASRU | Acoustic to Articulatory Speech Inversion for Children with Velopharyngeal Insufficiency. | Saba Tabatabaee, Suzanne Boyce, Liran Oren, Mark Tiede, Carol Y. Espy-Wilson |
| 2025 | HRI | Speaking with Robots in Noisy Environments. | Shuubham Ojha, Felix Gervits, Carol Y. Espy-Wilson |
| 2025 | ICASSP | CPT-Boosted Wav2vec2.0: Towards Noise Robust Speech Recognition for Classroom Environments. | Ahmed Adel Attia, Dorottya Demszky, Tollop gnrm, Jing Liu, Carol Y. Espy-Wilson |
| 2025 | ICASSP | Speech-Based Estimation of Schizophrenia Severity Using Feature Fusion. | Gowtham Premananth, Carol Y. Espy-Wilson |
| 2025 | ICASSP | Self-supervised Multimodal Speech Representations for the Assessment of Schizophrenia Symptoms. | Gowtham Premananth, Carol Y. Espy-Wilson |
| 2025 | Interspeech | From Weak Labels to Strong Results: Utilizing 5, 000 Hours of Noisy Classroom Transcripts with Minimal Accurate Data. | Ahmed Adel Attia, Dorottya Demszky, Jing Liu, Carol Y. Espy-Wilson |
| 2025 | Interspeech | Subtyping Speech Errors in Childhood Speech Sound Disorders with Acoustic-to-Articulatory Speech Inversion. | Nina R. Benway, Saba Tabatabaee, Benjamin Munson, Jonathan Preston, Carol Y. Espy-Wilson |
| 2025 | Interspeech | Speech Kinematic Analysis from Acoustics: Scientific, Clinical and Practical Applications. | Carol Y. Espy-Wilson |
| 2025 | Interspeech | Analyzing the Impact of Accent on English Speech: Acoustic and Articulatory Perspectives. | Gowtham Premananth, Vinith Kugathasan, Carol Y. Espy-Wilson |
| 2025 | Interspeech | Multimodal Biomarkers for Schizophrenia: Towards Individual Symptom Severity Estimation. | Gowtham Premananth, Philip Resnik, Sonia Bansal, Deanna L. Kelly, Carol Y. Espy-Wilson |
| 2025 | Interspeech | Enhancing Acoustic-to-Articulatory Speech Inversion by Incorporating Nasality. | Saba Tabatabaee, Suzanne Boyce, Liran Oren, Mark Tiede, Carol Y. Espy-Wilson |
| 2025 | Interspeech | FT-Boosted SV: Towards Noise Robust Speaker Verification for English Speaking Classroom Environments. | Saba Tabatabaee, Jing Liu, Carol Y. Espy-Wilson |
| 2024 | AIES | Kid-Whisper: Towards Bridging the Performance Gap in Automatic Speech Recognition for Children VS. Adults. | Ahmed Adel Attia, Jing Liu, Wei Ai, Dorottya Demszky, Carol Y. Espy-Wilson |
| 2024 | Interspeech | Examining Vocal Tract Coordination in Childhood Apraxia of Speech with Acoustic-to-Articulatory Speech Inversion Feature Sets. | Nina R. Benway, Jonathan L. Preston, Carol Y. Espy-Wilson |
| 2024 | Interspeech | A Multimodal Framework for the Assessment of the Schizophrenia Spectrum. | Gowtham Premananth, Yashish M. Siriwardena, Philip Resnik, Sonia Bansal, Deanna L. Kelly, Carol Y. Espy-Wilson |
| 2024 | Interspeech | Accent Conversion with Articulatory Representations. | Yashish M. Siriwardena, Nathan Swedlow, Audrey Howard, Evan Gitterman, Dan Darcy, Carol Y. Espy-Wilson, Andrea Fanelli |
| 2023 | ICASSP | Masked Autoencoders are Articulatory Learners. | Ahmed Adel Attia, Carol Y. Espy-Wilson |
| 2023 | ICASSP | The Secret Source : Incorporating Source Features to Improve Acoustic-To-Articulatory Speech Inversion. | Yashish M. Siriwardena, Carol Y. Espy-Wilson |
| 2023 | Interspeech | Enhancing Speech Articulation Analysis Using A Geometric Transformation of the X-ray Microbeam Dataset. | Ahmed Adel Attia, Mark Tiede, Carol Y. Espy-Wilson |
| 2023 | Interspeech | Acoustic-to-Articulatory Speech Inversion Features for Mispronunciation Detection of /ɹ/ in Child Speech Sound Disorders. | Nina R. Benway, Yashish M. Siriwardena, Jonathan L. Preston, Elaine Hitchcock, Tara McAllister Byun, Carol Y. Espy-Wilson |
| 2023 | Interspeech | Speaker-independent Speech Inversion for Estimation of Nasalance. | Yashish M. Siriwardena, Carol Y. Espy-Wilson, Suzanne Boyce, Mark Tiede, Liran Oren |
| 2023 | Interspeech | Learning to Compute the Articulatory Representations of Speech with the MIRRORNET. | Yashish M. Siriwardena, Carol Y. Espy-Wilson, Shihab A. Shamma |
| 2022 | ICASSP | Harmonicity Plays a Critical Role in DNN Based Versus in Biologically-Inspired Monaural Speech Segregation Systems. | Rahil Parikh, Ilya Kavalerov, Carol Y. Espy-Wilson, Shihab A. Shamma |
| 2022 | ICASSP | Multimodal Depression Classification using Articulatory Coordination Features and Hierarchical Attention Based text Embeddings. | Nadee Seneviratne, Carol Y. Espy-Wilson |
| 2022 | Interspeech | An Empirical Analysis on the Vulnerabilities of End-to-End Speech Segregation Models. | Rahil Parikh, Gaspar Rochette, Carol Y. Espy-Wilson, Shihab A. Shamma |
| 2022 | Interspeech | Acoustic To Articulatory Speech Inversion Using Multi-Resolution Spectro-Temporal Representations Of Speech Signals. | Rahil Parikh, Nadee Seneviratne, Ganesh Sivaraman, Shihab A. Shamma, Carol Y. Espy-Wilson |
| 2022 | Interspeech | Multimodal Depression Severity Score Prediction Using Articulatory Coordination Features and Hierarchical Attention Based Text Embeddings. | Nadee Seneviratne, Carol Y. Espy-Wilson |
| 2022 | Interspeech | Acoustic-to-articulatory Speech Inversion with Multi-task Learning. | Yashish M. Siriwardena, Ganesh Sivaraman, Carol Y. Espy-Wilson |
| 2021 | ICMI | Multimodal Approach for Assessing Neuromotor Coordination in Schizophrenia Using Convolutional Neural Networks. | Yashish M. Siriwardena, Carol Y. Espy-Wilson, Chris Kitchen, Deanna L. Kelly |
| 2021 | Interspeech | Speech Based Depression Severity Level Classification Using a Multi-Stage Dilated CNN-LSTM Model. | Nadee Seneviratne, Carol Y. Espy-Wilson |
| 2021 | Interspeech | Generalized Dilated CNN Models for Depression Detection Using Inverted Vocal Tract Variables. | Nadee Seneviratne, Carol Y. Espy-Wilson |
| 2020 | Interspeech | Extended Study on the Use of Vocal Tract Variables to Quantify Neuromotor Coordination in Depression. | Nadee Seneviratne, James R. Williamson, Adam C. Lammert, Thomas F. Quatieri, Carol Y. Espy-Wilson |
| 2019 | Interspeech | Assessing Neuromotor Coordination in Depression Using Inverted Vocal Tract Variables. | Carol Y. Espy-Wilson, Adam C. Lammert, Nadee Seneviratne, Thomas F. Quatieri |
| 2019 | Interspeech | Multi-Modal Learning for Speech Emotion Recognition: An Analysis and Comparison of ASR Outputs with Ground Truth Transcription. | Saurabh Sahu, Vikramjit Mitra, Nadee Seneviratne, Carol Y. Espy-Wilson |
| 2019 | Interspeech | Multi-Corpus Acoustic-to-Articulatory Speech Inversion. | Nadee Seneviratne, Ganesh Sivaraman, Carol Y. Espy-Wilson |
| 2018 | ICASSP | Semi-Supervised and Transfer Learning Approaches for Low Resource Sentiment Classification. | Rahul Gupta, Saurabh Sahu, Carol Y. Espy-Wilson, Shrikanth S. Narayanan |
| 2018 | ICASSP | Smoothing Model Predictions Using Adversarial Training Procedures for Speech Based Emotion Recognition. | Saurabh Sahu, Rahul Gupta, Ganesh Sivaraman, Carol Y. Espy-Wilson |
| 2018 | Interspeech | On Enhancing Speech Emotion Recognition Using Generative Adversarial Networks. | Saurabh Sahu, Rahul Gupta, Carol Y. Espy-Wilson |
| 2018 | Interspeech | Noise Robust Acoustic to Articulatory Speech Inversion. | Nadee Seneviratne, Ganesh Sivaraman, Vikramjit Mitra, Carol Y. Espy-Wilson |
| 2017 | ICASSP | Joint modeling of articulatory and acoustic spaces for continuous speech recognition tasks. | Vikramjit Mitra, Ganesh Sivaraman, Chris Bartels, Hosung Nam, Wen Wang, Carol Y. Espy-Wilson, Dimitra Vergyri, Horacio Franco |
| 2017 | Interspeech | An Affect Prediction Approach Through Depression Severity Parameter Incorporation in Neural Networks. | Rahul Gupta, Saurabh Sahu, Carol Y. Espy-Wilson, Shrikanth S. Narayanan |
| 2017 | Interspeech | Adversarial Auto-Encoders for Speech Based Emotion Recognition. | Saurabh Sahu, Rahul Gupta, Ganesh Sivaraman, Wael AbdAlmageed, Carol Y. Espy-Wilson |
| 2017 | Interspeech | Analysis of Acoustic-to-Articulatory Speech Inversion Across Different Accents and Languages. | Ganesh Sivaraman, Carol Y. Espy-Wilson, Martijn Wieling |
| 2016 | Interspeech | Speech Features for Depression Detection. | Saurabh Sahu, Carol Y. Espy-Wilson |
| 2016 | Interspeech | Vocal Tract Length Normalization for Speaker Independent Acoustic-to-Articulatory Speech Inversion. | Ganesh Sivaraman, Vikramjit Mitra, Hosung Nam, Mark K. Tiede, Carol Y. Espy-Wilson |
| 2015 | Interspeech | Analysis of coarticulated speech using estimated articulatory trajectories. | Ganesh Sivaraman, Vikramjit Mitra, Mark K. Tiede, Elliot Saltzman, Louis Goldstein, Carol Y. Espy-Wilson |
| 2014 | ICASSP | Articulatory features from deep neural networks and their role in speech recognition. | Vikramjit Mitra, Ganesh Sivaraman, Hosung Nam, Carol Y. Espy-Wilson, Elliot Saltzman |
| 2013 | ICASSP | A cine MRI-based study of sibilant fricatives production in post-glossectomy speakers. | Xinhui Zhou, Jonghye Woo, Maureen L. Stone, Carol Y. Espy-Wilson |
| 2012 | ICASSP | Multicondition training of Gaussian PLDA models in i-vector space for noise and reverberation robust speaker recognition. | Daniel Garcia-Romero, Xinhui Zhou, Carol Y. Espy-Wilson |
| 2012 | Interspeech | Automatic intelligibility assessment of pathologic speech in head and neck cancer based on auditory-inspired spectro-temporal modulations. | Xinhui Zhou, Daniel Garcia-Romero, Nima Mesgarani, Maureen L. Stone, Carol Y. Espy-Wilson, Shihab A. Shamma |
| 2011 | ASRU | Robust speech recognition using articulatory gestures in a Dynamic Bayesian Network framework. | Vikramjit Mitra, Hosung Nam, Carol Y. Espy-Wilson |
| 2011 | ASRU | Linear versus mel frequency cepstral coefficients for speaker recognition. | Xinhui Zhou, Daniel Garcia-Romero, Ramani Duraiswami, Carol Y. Espy-Wilson, Shihab A. Shamma |
| 2011 | ICASSP | Gesture-based Dynamic Bayesian Network for noise robust speech recognition. | Vikramjit Mitra, Hosung Nam, Carol Y. Espy-Wilson, Elliot Saltzman, Louis Goldstein |
| 2011 | ICASSP | Speech inversion: Benefits of tract variables over pellet trajectories. | Vikramjit Mitra, Hosung Nam, Carol Y. Espy-Wilson, Elliot Saltzman, Louis Goldstein |
| 2011 | Interspeech | Analysis of i-vector Length Normalization in Speaker Recognition Systems. | Daniel Garcia-Romero, Carol Y. Espy-Wilson |
| 2011 | Interspeech | Automatic Speech Codec Identification with Applications to Tampering Detection of Speech Recordings. | Jingting Zhou, Daniel Garcia-Romero, Carol Y. Espy-Wilson |
| 2011 | Interspeech | A Comparative Acoustic Study on Speech of Glossectomy Patients and Normal Subjects. | Xinhui Zhou, Maureen L. Stone, Carol Y. Espy-Wilson |
| 2010 | ICASSP | Automatic acquisition device identification from speech recordings. | Daniel Garcia-Romero, Carol Y. Espy-Wilson |
| 2010 | ICASSP | An MRI-based articulatory and acoustic study of lateral sound in American English. | Xinhui Zhou, Carol Y. Espy-Wilson, Mark Tiede, Suzanne Boyce |
| 2010 | Interspeech | Robust word recognition using articulatory trajectories and gestures. | Vikramjit Mitra, Hosung Nam, Carol Y. Espy-Wilson, Elliot Saltzman, Louis Goldstein |
| 2010 | Interspeech | A procedure for estimating gestural scores from natural speech. | Hosung Nam, Vikramjit Mitra, Mark Tiede, Elliot Saltzman, Louis Goldstein, Carol Y. Espy-Wilson, Mark Hasegawa-Johnson |
| 2009 | ICASSP | From acoustics to Vocal Tract time functions. | Vikramjit Mitra, I. Ycel zbek, Hosung Nam, Xinhui Zhou, Carol Y. Espy-Wilson |
| 2009 | ICASSP | An algorithm for speech segregation of co-channel speech. | Srikanth Vishnubhotla, Carol Y. Espy-Wilson |
| 2009 | Interspeech | A noise-type and level-dependent MPO-based speech enhancement architecture with variable frame analysis for noise-robust speech recognition. | Vikramjit Mitra, Bengt J. Borgstrom, Carol Y. Espy-Wilson, Abeer Alwan |
| 2009 | Interspeech | Noise robustness of tract variables and their application to speech recognition. | Vikramjit Mitra, Hosung Nam, Carol Y. Espy-Wilson, Elliot Saltzman, Louis Goldstein |
| 2008 | ICASSP | Language detection in audio content analysis. | Vikramjit Mitra, Daniel Garcia-Romero, Carol Y. Espy-Wilson |
| 2008 | Interspeech | Intersession variability in speaker recognition: a behind the scene analysis. | Daniel Garcia-Romero, Carol Y. Espy-Wilson |
| 2008 | Interspeech | Language and genre detection in audio content analysis. | Vikramjit Mitra, Daniel Garcia-Romero, Carol Y. Espy-Wilson |
| 2008 | Interspeech | An algorithm for multi-pitch tracking in co-channel speech. | Srikanth Vishnubhotla, Carol Y. Espy-Wilson |
| 2007 | Interspeech | Landmark-based approach to speech recognition: an alternative to HMMs. | Carol Y. Espy-Wilson, Tarun Pruthi, Amit Juneja, Om Deshmukh |
| 2007 | Interspeech | A semi-automatic approach for speaker mining of tapped telephone conversations. | Sandeep Manocha, Carol Y. Espy-Wilson |
| 2007 | Interspeech | Acoustic parameters for the automatic detection of vowel nasalization. | Tarun Pruthi, Carol Y. Espy-Wilson |
| 2007 | Interspeech | An articulatory and acoustic study of "retroflex" and "bunched" american English rhotic sound based on MRI. | Xinhui Zhou, Carol Y. Espy-Wilson, Mark Tiede, Suzanne Boyce |
| 2006 | Interspeech | Modified phase opponency based solution to the speech separation challenge. | Om Deshmukh, Carol Y. Espy-Wilson |
| 2006 | Interspeech | Speech enhancement using modified phase opponency model. | Om Deshmukh, Carol Y. Espy-Wilson |
| 2006 | Interspeech | A new set of features for text-independent speaker identification. | Carol Y. Espy-Wilson, Sandeep Manocha, Srikanth Vishnubhotla |
| 2006 | Interspeech | An MRI based study of the acoustic effects of sinus cavities and its application to speaker recognition. | Tarun Pruthi, Carol Y. Espy-Wilson |
| 2006 | Interspeech | Automatic detection of irregular phonation in continuous speech. | Srikanth Vishnubhotla, Carol Y. Espy-Wilson |
| 2005 | ICASSP | Modeling of the Front Cavity and Sublingual Space in American English Rhotic Sounds. | Zhaoyan Zhang, Carol Y. Espy-Wilson, Suzanne Boyce, Mark Tiede |
| 2005 | Interspeech | Speech enhancement using auditory phase opponency model. | Om Deshmukh, Carol Y. Espy-Wilson |
| 2004 | ICASSP | A novel method for computation of periodicity, aperiodicity and pitch of speech signals. | Om Deshmukh, Jawahar Singh, Carol Y. Espy-Wilson |
| 2003 | ICASSP | A measure of aperiodicity and periodicity in speech. | Om Deshmukh, Carol Y. Espy-Wilson |
| 2003 | IJCNN | Speech segmentation using probabilistic phonetic feature hierarchy and support vector machines. | Amit Juneja, Carol Y. Espy-Wilson |
| 2003 | Interspeech | Acoustic modeling of american English lateral approximants. | Zhaoyan Zhang, Carol Y. Espy-Wilson, Mark Tiede |
| 2002 | ICASSP | Acoustic-phonetic speech parameters for speaker-independent speech recognition. | Om Deshmukh, Carol Y. Espy-Wilson, Amit Juneja |
| 2002 | ICASSP | An event-based acoustic-phonetic approach for speech segmentation and E-set recognition. | Amit Juneja, Om Deshmukh, Carol Y. Espy-Wilson |
| 2000 | Interspeech | Detection of speech landmarks using temporal cues. | Ariel Salomon, Carol Y. Espy-Wilson |
| 2000 | Interspeech | A new strategy of formant tracking based on dynamic programming. | Kun Xia, Carol Y. Espy-Wilson |
| 1999 | Interspeech | Improvement of electrolaryngeal speech by introducing normal excitation information. | Kun Ma, Pelin Demirel, Carol Y. Espy-Wilson, Joel MacAuslan |
| 1999 | Interspeech | Automatic detection of manner events based on temporal parameters. | Ariel Salomon, Carol Y. Espy-Wilson |
| 1997 | Interspeech | The design of acoustic parameters for speaker-independent speech recognition. | Nabil N. Bitar, Carol Y. Espy-Wilson |
| 1997 | Interspeech | Acoustic modelling of American English /r/. | Carol Y. Espy-Wilson, Shrikanth S. Narayanan, Suzanne Boyce, Abeer Alwan |
| 1996 | ICASSP | Knowledge-based parameters for HMM speech recognition. | Nabil N. Bitar, Carol Y. Espy-Wilson |
| 1996 | Interspeech | Coarticulatory stability in american English /r/. | Suzanne Boyce, Carol Y. Espy-Wilson |
| 1996 | Interspeech | Enhancement of alaryngeal speech by adaptive filtering. | Carol Y. Espy-Wilson, Venkatesh R. Chari, Caroline B. Huang |
| 1995 | Interspeech | Speech parameterization based on phonetic features: application to speech recognition. | Nabil N. Bitar, Carol Y. Espy-Wilson |
| 1986 | ICASSP | A phonetically based semivowel recognition system. | Carol Y. Espy-Wilson |