| 2026 | ACL | Cross-Modal Taxonomic Generalization in (Vision-) Language Models. | Tianyang Xu, Marcelo Sandoval-Castaeda, Karen Livescu, Greg Shakhnarovich, Kanishka Misra |
| 2025 | ACL | SignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale. | Shester Gueuwou, Xiaodan Du, Greg Shakhnarovich, Karen Livescu |
| 2025 | ACL | SHuBERT: Self-Supervised Sign Language Representation Learning via Multi-Stream Cluster Prediction. | Shester Gueuwou, Xiaodan Du, Greg Shakhnarovich, Karen Livescu, Alexander H. Liu |
| 2025 | ASRU | Flow-SLM: Joint Learning of Linguistic and Acoustic Information for Spoken Language Modeling. | Ju-Chieh Chou, Jiawei Zhou, Karen Livescu |
| 2025 | ASRU | Transcribe, Translate, or Transliterate: An Investigation of Intermediate Representations in Spoken Language Models. | Tollop gnrm, Christopher D. Manning, Dan Jurafsky, Karen Livescu |
| 2025 | ICASSP | Constructing Datasets From Public Police Body Camera Footage. | Jamie Rosas-Smith, Martijn Bartelds, Ruizhe Huang, Leibny Paola Garca-Perera, Karen Livescu, Dan Jurafsky, Anjalie Field |
| 2025 | ICLR | Chunk-Distilled Language Modeling. | Yanhong Li, Karen Livescu, Jiawei Zhou |
| 2025 | Interspeech | The ML-SUPERB 2.0 Challenge: Towards Inclusive ASR Benchmarking for All Language Varieties. | William Chen, Chutong Meng, Jiatong Shi, Martijn Bartelds, Shih-Heng Wang, Hsiu-Hsuan Wang, Rafael Mosquera, Sara Hincapie, Dan Jurafsky, Antonis Anastasopoulos, Hung-yi Lee, Karen Livescu, Shinji Watanabe |
| 2024 | ACL | On the Evaluation of Speech Foundation Models for Spoken Language Understanding. | Siddhant Arora, Ankita Pasad, Chung-Ming Chien, Jionghao Han, Roshan S. Sharma, Jee-weon Jung, Hira Dhamyal, William Chen, Suwon Shon, Hung-yi Lee, Karen Livescu, Shinji Watanabe |
| 2024 | ACL | Structured Tree Alignment for Evaluation of (Speech) Constituency Parsing. | Freda Shi, Kevin Gimpel, Karen Livescu |
| 2024 | EMNLP | Towards Robust Speech Representation Learning for Thousands of Languages. | William Chen, Wangyou Zhang, Yifan Peng, Xinjian Li, Jinchuan Tian, Jiatong Shi, Xuankai Chang, Soumi Maiti, Karen Livescu, Shinji Watanabe |
| 2024 | ICASSP | AV2WAV: Diffusion-Based Re-Synthesis from Continuous Self-Supervised Features for Audio-Visual Speech Enhancement. | Ju-Chieh Chou, Chung-Ming Chien, Karen Livescu |
| 2024 | ICASSP | Generative Context-Aware Fine-Tuning of Self-Supervised Speech Models. | Suwon Shon, Kwangyoun Kim, Prashant Sridhar, Yi-Te Hsu, Shinji Watanabe, Karen Livescu |
| 2024 | Interspeech | Self-Supervised Speech Representations are More Phonetic than Semantic. | Kwanghee Choi, Ankita Pasad, Tomohiko Nakamura, Satoru Fukayama, Karen Livescu, Shinji Watanabe |
| 2024 | Interspeech | Convolution-Augmented Parameter-Efficient Fine-Tuning for Speech Recognition. | Kwangyoun Kim, Suwon Shon, Yi-Te Hsu, Prashant Sridhar, Karen Livescu, Shinji Watanabe |
| 2024 | Interspeech | ML-SUPERB 2.0: Benchmarking Multilingual Speech Models Across Modeling Constraints, Languages, and Datasets. | Jiatong Shi, Shih-Heng Wang, William Chen, Martijn Bartelds, Vanya Bannihatti Kumar, Jinchuan Tian, Xuankai Chang, Dan Jurafsky, Karen Livescu, Hung-yi Lee, Shinji Watanabe |
| 2024 | Interspeech | DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding. | Suwon Shon, Kwangyoun Kim, Yi-Te Hsu, Prashant Sridhar, Shinji Watanabe, Karen Livescu |
| 2024 | Interspeech | On the Effects of Heterogeneous Data Sources on Speech-to-Text Foundation Models. | Jinchuan Tian, Yifan Peng, William Chen, Kwanghee Choi, Karen Livescu, Shinji Watanabe |
| 2024 | NAACL | UniverSLU: Universal Spoken Language Understanding for Diverse Tasks with Natural Language Instructions. | Siddhant Arora, Hayato Futami, Jee-weon Jung, Yifan Peng, Roshan S. Sharma, Yosuke Kashiwagi, Emiru Tsunoo, Karen Livescu, Shinji Watanabe |
| 2023 | ACL | SLUE Phase-2: A Benchmark Suite of Diverse Spoken Language Understanding Tasks. | Suwon Shon, Siddhant Arora, Chyi-Jiunn Lin, Ankita Pasad, Felix Wu, Roshan S. Sharma, Wei-Lun Wu, Hung-yi Lee, Karen Livescu, Shinji Watanabe |
| 2023 | ASRU | Few-Shot Spoken Language Understanding Via Joint Speech-Text Models. | Chung-Ming Chien, Mingjiamei Zhang, Ju-Chieh Chou, Karen Livescu |
| 2023 | ASRU | Audio-Visual Neural Syntax Acquisition. | Cheng-I Jeff Lai, Freda Shi, Puyuan Peng, Yoon Kim, Kevin Gimpel, Shiyu Chang, Yung-Sung Chuang, Saurabhchand Bhati, David D. Cox, David Harwath, Yang Zhang, Karen Livescu, James R. Glass |
| 2023 | EMNLP | Toward Joint Language Modeling for Speech Units and Text. | Ju-Chieh Chou, Chung-Ming Chien, Wei-Ning Hsu, Karen Livescu, Arun Babu, Alexis Conneau, Alexei Baevski, Michael Auli |
| 2023 | ICASSP | Comparative Layer-Wise Analysis of Self-Supervised Speech Models. | Ankita Pasad, Bowen Shi, Karen Livescu |
| 2023 | ICASSP | Context-Aware Fine-Tuning of Self-Supervised Speech Models. | Suwon Shon, Felix Wu, Kwangyoun Kim, Prashant Sridhar, Karen Livescu, Shinji Watanabe |
| 2022 | AAAI | Chess as a Testbed for Language Model State Tracking. | Shubham Toshniwal, Sam Wiseman, Karen Livescu, Kevin Gimpel |
| 2022 | ACL | Searching for fingerspelled content in American Sign Language. | Bowen Shi, Diane Brentari, Greg Shakhnarovich, Karen Livescu |
| 2022 | ACL | Substructure Distribution Projection for Zero-Shot Cross-Lingual Dependency Parsing. | Freda Shi, Kevin Gimpel, Karen Livescu |
| 2022 | EMNLP | Open-Domain Sign Language Translation Learned from Online Video. | Bowen Shi, Diane Brentari, Gregory Shakhnarovich, Karen Livescu |
| 2022 | EMNLP | Baked-in State Probing. | Shubham Toshniwal, Sam Wiseman, Karen Livescu, Kevin Gimpel |
| 2022 | ICASSP | SLUE: New Benchmark Tasks For Spoken Language Understanding Evaluation on Natural Speech. | Suwon Shon, Ankita Pasad, Felix Wu, Pablo Brusco, Yoav Artzi, Karen Livescu, Kyu Jeong Han |
| 2022 | NAACL | On the Use of External Data for Spoken Named Entity Recognition. | Ankita Pasad, Felix Wu, Suwon Shon, Karen Livescu, Kyu Jeong Han |
| 2021 | ACL | Substructure Substitution: Structured Data Augmentation for NLP. | Haoyue Shi, Karen Livescu, Kevin Gimpel |
| 2021 | ASRU | Layer-Wise Analysis of a Self-Supervised Speech Representation Model. | Ankita Pasad, Ju-Chieh Chou, Karen Livescu |
| 2021 | CVPR | Fingerspelling Detection in American Sign Language. | Bowen Shi, Diane Brentari, Greg Shakhnarovich, Karen Livescu |
| 2021 | Interspeech | Learning Speech Models from Multi-Modal Data. | Karen Livescu |
| 2020 | ACL | Discrete Latent Variable Representations for Low-Resource Text Classification. | Shuning Jin, Sam Wiseman, Karl Stratos, Karen Livescu |
| 2020 | ACL | PeTra: A Sparsely Supervised Memory Model for People Tracking. | Shubham Toshniwal, Allyson Ettinger, Kevin Gimpel, Karen Livescu |
| 2020 | EMNLP | On the Role of Supervision in Unsupervised Constituency Parsing. | Haoyue Shi, Karen Livescu, Kevin Gimpel |
| 2020 | EMNLP | Learning to Ignore: Long Document Coreference with Bounded Memory Neural Networks. | Shubham Toshniwal, Sam Wiseman, Allyson Ettinger, Karen Livescu, Kevin Gimpel |
| 2020 | ICASSP | Unsupervised Pre-Training of Bidirectional Speech Encoders via Masked Reconstruction. | Weiran Wang, Qingming Tang, Karen Livescu |
| 2020 | Interspeech | Multilingual Jointly Trained Acoustic and Written Word Embeddings. | Yushi Hu, Shane Settle, Karen Livescu |
| 2019 | ACL | Visually Grounded Neural Syntax Acquisition. | Haoyue Shi, Jiayuan Mao, Kevin Gimpel, Karen Livescu |
| 2019 | ICASSP | Semantic Query-by-example Speech Search Using Visual Grounding. | Herman Kamper, Aristotelis Anastassiou, Karen Livescu |
| 2019 | ICASSP | Acoustically Grounded Word Embeddings for Improved Acoustics-to-word Speech Recognition. | Shane Settle, Kartik Audhkhasi, Karen Livescu, Michael Picheny |
| 2019 | ICCV | Fingerspelling Recognition in the Wild With Iterative Visual Attention. | Bowen Shi, Aurora Martinez Del Rio, Jonathan Keane, Diane Brentari, Greg Shakhnarovich, Karen Livescu |
| 2019 | Interspeech | Pre-Trained Text Embeddings for Enhanced Text-to-Speech Synthesis. | Tomoki Hayashi, Shinji Watanabe, Tomoki Toda, Kazuya Takeda, Shubham Toshniwal, Karen Livescu |
| 2019 | Interspeech | On the Contributions of Visual and Textual Supervision in Low-Resource Semantic Speech Retrieval. | Ankita Pasad, Bowen Shi, Herman Kamper, Karen Livescu |
| 2019 | NAACL | Pre-training on high-resource speech recognition improves low-resource speech-to-text translation. | Sameer Bansal, Herman Kamper, Karen Livescu, Adam Lopez, Sharon Goldwater |
| 2018 | CVPR | Semantic Speech Retrieval With a Visually Grounded Model of Untranscribed Speech. | Herman Kamper, Gregory Shakhnarovich, Karen Livescu |
| 2018 | EMNLP | Variational Sequential Labelers for Semi-Supervised Learning. | Mingda Chen, Qingming Tang, Karen Livescu, Kevin Gimpel |
| 2018 | ICASSP | A Study of All-Convolutional Encoders for Connectionist Temporal Classification. | Kalpesh Krishna, Liang Lu, Kevin Gimpel, Karen Livescu |
| 2018 | ICASSP | Acoustic Feature Learning Using Cross-Domain Articulatory Measurements. | Qingming Tang, Weiran Wang, Karen Livescu |
| 2018 | Interspeech | Low-Resource Speech-to-Text Translation. | Sameer Bansal, Herman Kamper, Karen Livescu, Adam Lopez, Sharon Goldwater |
| 2018 | NAACL | Parsing Speech: a Neural Approach to Integrating Lexical and Acoustic-Prosodic Information. | Trang Tran, Shubham Toshniwal, Mohit Bansal, Kevin Gimpel, Karen Livescu, Mari Ostendorf |
| 2017 | ASRU | An embedded segmental K-means model for unsupervised segmentation and clustering of speech. | Herman Kamper, Karen Livescu, Sharon Goldwater |
| 2017 | ASRU | Multitask training with unlabeled data for end-to-end sign language fingerspelling recognition. | Bowen Shi, Karen Livescu |
| 2017 | ICLR | Multi-view Recurrent Neural Acoustic Word Embeddings. | Wanjia He, Weiran Wang, Karen Livescu |
| 2017 | Interspeech | Visually Grounded Learning of Keyword Prediction from Untranscribed Speech. | Herman Kamper, Shane Settle, Gregory Shakhnarovich, Karen Livescu |
| 2017 | Interspeech | Query-by-Example Search with Discriminative Neural Acoustic Word Embeddings. | Shane Settle, Keith D. Levin, Herman Kamper, Karen Livescu |
| 2017 | Interspeech | Acoustic Feature Learning via Deep Variational Canonical Correlation Analysis. | Qingming Tang, Weiran Wang, Karen Livescu |
| 2017 | Interspeech | Multitask Learning with Low-Level Auxiliary Tasks for Encoder-Decoder Based Speech Recognition. | Shubham Toshniwal, Hao Tang, Liang Lu, Karen Livescu |
| 2016 | EMNLP | Charagram: Embedding Words and Sentences via Character n-grams. | John Wieting, Mohit Bansal, Kevin Gimpel, Karen Livescu |
| 2016 | ICASSP | Signer-independent fingerspelling recognition with deep neural network adaptation. | Taehwan Kim, Weiran Wang, Hao Tang, Karen Livescu |
| 2016 | ICASSP | Deep convolutional acoustic word embeddings using word-pair side information. | Herman Kamper, Weiran Wang, Karen Livescu |
| 2016 | ICML | Nonparametric Canonical Correlation Analysis. | Tomer Michaeli, Weiran Wang, Karen Livescu |
| 2016 | Interspeech | Efficient Segmental Cascades for Speech Recognition. | Hao Tang, Weiran Wang, Kevin Gimpel, Karen Livescu |
| 2016 | Interspeech | Triphone State-Tying via Deep Canonical Correlation Analysis. | Weiran Wang, Hao Tang, Karen Livescu |
| 2015 | ASRU | Discriminative segmental cascades for feature-rich phone recognition. | Hao Tang, Weiran Wang, Kevin Gimpel, Karen Livescu |
| 2015 | ICASSP | Unsupervised learning of acoustic features via deep canonical correlation analysis. | Weiran Wang, Raman Arora, Karen Livescu, Jeff A. Bilmes |
| 2015 | ICML | On Deep Multi-View Representation Learning. | Weiran Wang, Raman Arora, Karen Livescu, Jeff A. Bilmes |
| 2015 | NAACL | Deep Multilingual Correlation for Improved Word Embeddings. | Ang Lu, Weiran Wang, Mohit Bansal, Kevin Gimpel, Karen Livescu |
| 2014 | ACL | Tailoring Continuous Word Representations for Dependency Parsing. | Mohit Bansal, Kevin Gimpel, Karen Livescu |
| 2014 | ICASSP | Multi-view learning with supervision for transformed bottleneck features. | Raman Arora, Karen Livescu |
| 2014 | Interspeech | A comparison of training approaches for discriminative segmental models. | Hao Tang, Kevin Gimpel, Karen Livescu |
| 2013 | ASRU | Fixed-dimensional acoustic embeddings of variable-length segments in low-resource settings. | Keith D. Levin, Katharine Henry, Aren Jansen, Karen Livescu |
| 2013 | ICASSP | Multi-view CCA-based acoustic features for phonetic recognition across speakers and domains. | Raman Arora, Karen Livescu |
| 2013 | ICASSP | Discriminative articulatory models for spoken term detection in low-resource conversational settings. | Rohit Prabhavalkar, Karen Livescu, Eric Fosler-Lussier, Joseph Keshet |
| 2013 | ICCV | Fingerspelling Recognition with Semi-Markov Conditional Random Fields. | Taehwan Kim, Gregory Shakhnarovich, Karen Livescu |
| 2013 | ICML | Deep Canonical Correlation Analysis. | Galen Andrew, Raman Arora, Jeff A. Bilmes, Karen Livescu |
| 2013 | Interspeech | Discriminative training of WFST factors with application to pronunciation modeling. | Preethi Jyothi, Eric Fosler-Lussier, Karen Livescu |
| 2012 | ACL | Discriminative Pronunciation Modeling: A Large-Margin, Feature-Rich Approach. | Hao Tang, Joseph Keshet, Karen Livescu |
| 2012 | Interspeech | Discriminatively learning factorized finite state pronunciation models from dynamic Bayesian networks. | Preethi Jyothi, Eric Fosler-Lussier, Karen Livescu |
| 2011 | ASRU | A factored conditional random field model for articulatory feature forced transcription. | Rohit Prabhavalkar, Eric Fosler-Lussier, Karen Livescu |
| 2011 | ICASSP | Lexical access experiments with context-dependent articulatory feature-based models. | Preethi Jyothi, Karen Livescu, Eric Fosler-Lussier |
| 2011 | Interspeech | Nearest Neighbors with Learned Distances for Phonetic Frame Classification. | John Labiak, Karen Livescu |
| 2011 | Interspeech | Articulatory Feature Classification Using Nearest Neighbors. | Arild Brandrud Nss, Karen Livescu, Rohit Prabhavalkar |
| 2010 | Interspeech | Modeling pronunciation variation with context-dependent articulatory feature decision trees. | Samuel R. Bowman, Karen Livescu |
| 2010 | Interspeech | Audio-visual anticipatory coarticulation modeling by human and machine. | Louis H. Terry, Karen Livescu, Janet B. Pierrehumbert, Aggelos K. Katsaggelos |
| 2009 | ASRU | Multi-view learning of acoustic features for speaker recognition. | Karen Livescu, Mark Stoehr |
| 2009 | ICASSP | On the phonetic information in ultrasonic microphone signals. | Karen Livescu, Bo Zhu, James R. Glass |
| 2009 | ICML | Multi-view clustering via canonical correlation analysis. | Kamalika Chaudhuri, Sham M. Kakade, Karen Livescu, Karthik Sridharan |
| 2007 | ASRU | Monolingual and crosslingual comparison of tandem features derived from articulatory and phone MLPS. | zgr etin, Mathew Magimai-Doss, Karen Livescu, Arthur Kantor, Simon King, Chris D. Bartels, Joe Frankel |
| 2007 | ICASSP | An Articulatory Feature-Based Tandem Approach and Factored Observation Modeling. | zgr etin, Arthur Kantor, Simon King, Chris D. Bartels, Mathew Magimai-Doss, Joe Frankel, Karen Livescu |
| 2007 | ICASSP | Manual Transcription of Conversational Speech at the Articulatory Feature Level. | Karen Livescu, Ari Bezman, Nash M. Borges, Lisa Yung, zgr etin, Joe Frankel, Simon King, Mathew Magimai-Doss, Xuemin Chi, Lisa Lavoie |
| 2007 | ICASSP | Articulatory Feature-Based Methods for Acoustic and Audio-Visual Speech Recognition: Summary from the 2006 JHU Summer workshop. | Karen Livescu, zgr etin, Mark Hasegawa-Johnson, Simon King, Chris D. Bartels, Nash M. Borges, Arthur Kantor, Partha Lal, Lisa Yung, Ari Bezman, Stephen Dawson-Haggerty, Bronwyn Woods, Joe Frankel, Mathew Magimai-Doss, Kate Saenko |
| 2007 | Interspeech | Articulatory feature classifiers trained on 2000 hours of telephone speech. | Joe Frankel, Mathew Magimai-Doss, Simon King, Karen Livescu, zgr etin |
| 2005 | ICASSP | Landmark-Based Speech Recognition: Report of the 2004 Johns Hopkins Summer Workshop. | Mark Hasegawa-Johnson, James Baker, Sarah Borys, Ken Chen, Emily Coogan, Steven Greenberg, Amit Juneja, Katrin Kirchhoff, Karen Livescu, Srividya Mohan, Jennifer Muller, M. Kemal Snmez, Tianyu Wang |
| 2005 | ICASSP | Production domain modeling of pronunciation for visual speech recognition. | Kate Saenko, Karen Livescu, James R. Glass, Trevor Darrell |
| 2005 | ICCV | Visual Speech Recognition with Loosely Synchronized Feature Streams. | Kate Saenko, Karen Livescu, Michael Siracusa, Kevin W. Wilson, James R. Glass, Trevor Darrell |
| 2004 | Interspeech | Feature-based pronunciation modeling with trainable asynchrony probabilities. | Karen Livescu, James R. Glass |
| 2004 | NAACL | Feature-based Pronunciation Modeling for Speech Recognition. | Karen Livescu, James R. Glass |
| 2003 | Interspeech | Hidden feature models for speech recognition using dynamic Bayesian networks. | Karen Livescu, James R. Glass, Jeff A. Bilmes |
| 2002 | ICASSP | Structurally discriminative graphical models for automatic speech recognition - results from the 2001 Johns Hopkins Summer Workshop. | Geoffrey Zweig, Jeff A. Bilmes, Thomas Richardson, Karim Filali, Karen Livescu, Peng Xu, Kirk Jackson, Yigal Brandman, Eric D. Sandness, Eva Holtz, Jerry Torres, Bill Byrne |
| 2001 | Interspeech | Segment-based recognition on the phonebook task: initial results and observations on duration modeling. | Karen Livescu, James R. Glass |
| 2000 | ICASSP | Lexical modeling of non-native speech for automatic speech recognition. | Karen Livescu, James R. Glass |