Skip to content

Karen Livescu

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

106

Venues

11

Active years

2000–2026

Best venue rank

A*

Where they publish

Papers

106 indexed papers, newest first.

YearVenueTitleAuthors
2026ACLCross-Modal Taxonomic Generalization in (Vision-) Language Models.Tianyang Xu, Marcelo Sandoval-Castaeda, Karen Livescu, Greg Shakhnarovich, Kanishka Misra
2025ACLSignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale.Shester Gueuwou, Xiaodan Du, Greg Shakhnarovich, Karen Livescu
2025ACLSHuBERT: Self-Supervised Sign Language Representation Learning via Multi-Stream Cluster Prediction.Shester Gueuwou, Xiaodan Du, Greg Shakhnarovich, Karen Livescu, Alexander H. Liu
2025ASRUFlow-SLM: Joint Learning of Linguistic and Acoustic Information for Spoken Language Modeling.Ju-Chieh Chou, Jiawei Zhou, Karen Livescu
2025ASRUTranscribe, Translate, or Transliterate: An Investigation of Intermediate Representations in Spoken Language Models.Tollop gnrm, Christopher D. Manning, Dan Jurafsky, Karen Livescu
2025ICASSPConstructing Datasets From Public Police Body Camera Footage.Jamie Rosas-Smith, Martijn Bartelds, Ruizhe Huang, Leibny Paola Garca-Perera, Karen Livescu, Dan Jurafsky, Anjalie Field
2025ICLRChunk-Distilled Language Modeling.Yanhong Li, Karen Livescu, Jiawei Zhou
2025InterspeechThe ML-SUPERB 2.0 Challenge: Towards Inclusive ASR Benchmarking for All Language Varieties.William Chen, Chutong Meng, Jiatong Shi, Martijn Bartelds, Shih-Heng Wang, Hsiu-Hsuan Wang, Rafael Mosquera, Sara Hincapie, Dan Jurafsky, Antonis Anastasopoulos, Hung-yi Lee, Karen Livescu, Shinji Watanabe
2024ACLOn the Evaluation of Speech Foundation Models for Spoken Language Understanding.Siddhant Arora, Ankita Pasad, Chung-Ming Chien, Jionghao Han, Roshan S. Sharma, Jee-weon Jung, Hira Dhamyal, William Chen, Suwon Shon, Hung-yi Lee, Karen Livescu, Shinji Watanabe
2024ACLStructured Tree Alignment for Evaluation of (Speech) Constituency Parsing.Freda Shi, Kevin Gimpel, Karen Livescu
2024EMNLPTowards Robust Speech Representation Learning for Thousands of Languages.William Chen, Wangyou Zhang, Yifan Peng, Xinjian Li, Jinchuan Tian, Jiatong Shi, Xuankai Chang, Soumi Maiti, Karen Livescu, Shinji Watanabe
2024ICASSPAV2WAV: Diffusion-Based Re-Synthesis from Continuous Self-Supervised Features for Audio-Visual Speech Enhancement.Ju-Chieh Chou, Chung-Ming Chien, Karen Livescu
2024ICASSPGenerative Context-Aware Fine-Tuning of Self-Supervised Speech Models.Suwon Shon, Kwangyoun Kim, Prashant Sridhar, Yi-Te Hsu, Shinji Watanabe, Karen Livescu
2024InterspeechSelf-Supervised Speech Representations are More Phonetic than Semantic.Kwanghee Choi, Ankita Pasad, Tomohiko Nakamura, Satoru Fukayama, Karen Livescu, Shinji Watanabe
2024InterspeechConvolution-Augmented Parameter-Efficient Fine-Tuning for Speech Recognition.Kwangyoun Kim, Suwon Shon, Yi-Te Hsu, Prashant Sridhar, Karen Livescu, Shinji Watanabe
2024InterspeechML-SUPERB 2.0: Benchmarking Multilingual Speech Models Across Modeling Constraints, Languages, and Datasets.Jiatong Shi, Shih-Heng Wang, William Chen, Martijn Bartelds, Vanya Bannihatti Kumar, Jinchuan Tian, Xuankai Chang, Dan Jurafsky, Karen Livescu, Hung-yi Lee, Shinji Watanabe
2024InterspeechDiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding.Suwon Shon, Kwangyoun Kim, Yi-Te Hsu, Prashant Sridhar, Shinji Watanabe, Karen Livescu
2024InterspeechOn the Effects of Heterogeneous Data Sources on Speech-to-Text Foundation Models.Jinchuan Tian, Yifan Peng, William Chen, Kwanghee Choi, Karen Livescu, Shinji Watanabe
2024NAACLUniverSLU: Universal Spoken Language Understanding for Diverse Tasks with Natural Language Instructions.Siddhant Arora, Hayato Futami, Jee-weon Jung, Yifan Peng, Roshan S. Sharma, Yosuke Kashiwagi, Emiru Tsunoo, Karen Livescu, Shinji Watanabe
2023ACLSLUE Phase-2: A Benchmark Suite of Diverse Spoken Language Understanding Tasks.Suwon Shon, Siddhant Arora, Chyi-Jiunn Lin, Ankita Pasad, Felix Wu, Roshan S. Sharma, Wei-Lun Wu, Hung-yi Lee, Karen Livescu, Shinji Watanabe
2023ASRUFew-Shot Spoken Language Understanding Via Joint Speech-Text Models.Chung-Ming Chien, Mingjiamei Zhang, Ju-Chieh Chou, Karen Livescu
2023ASRUAudio-Visual Neural Syntax Acquisition.Cheng-I Jeff Lai, Freda Shi, Puyuan Peng, Yoon Kim, Kevin Gimpel, Shiyu Chang, Yung-Sung Chuang, Saurabhchand Bhati, David D. Cox, David Harwath, Yang Zhang, Karen Livescu, James R. Glass
2023EMNLPToward Joint Language Modeling for Speech Units and Text.Ju-Chieh Chou, Chung-Ming Chien, Wei-Ning Hsu, Karen Livescu, Arun Babu, Alexis Conneau, Alexei Baevski, Michael Auli
2023ICASSPComparative Layer-Wise Analysis of Self-Supervised Speech Models.Ankita Pasad, Bowen Shi, Karen Livescu
2023ICASSPContext-Aware Fine-Tuning of Self-Supervised Speech Models.Suwon Shon, Felix Wu, Kwangyoun Kim, Prashant Sridhar, Karen Livescu, Shinji Watanabe
2022AAAIChess as a Testbed for Language Model State Tracking.Shubham Toshniwal, Sam Wiseman, Karen Livescu, Kevin Gimpel
2022ACLSearching for fingerspelled content in American Sign Language.Bowen Shi, Diane Brentari, Greg Shakhnarovich, Karen Livescu
2022ACLSubstructure Distribution Projection for Zero-Shot Cross-Lingual Dependency Parsing.Freda Shi, Kevin Gimpel, Karen Livescu
2022EMNLPOpen-Domain Sign Language Translation Learned from Online Video.Bowen Shi, Diane Brentari, Gregory Shakhnarovich, Karen Livescu
2022EMNLPBaked-in State Probing.Shubham Toshniwal, Sam Wiseman, Karen Livescu, Kevin Gimpel
2022ICASSPSLUE: New Benchmark Tasks For Spoken Language Understanding Evaluation on Natural Speech.Suwon Shon, Ankita Pasad, Felix Wu, Pablo Brusco, Yoav Artzi, Karen Livescu, Kyu Jeong Han
2022NAACLOn the Use of External Data for Spoken Named Entity Recognition.Ankita Pasad, Felix Wu, Suwon Shon, Karen Livescu, Kyu Jeong Han
2021ACLSubstructure Substitution: Structured Data Augmentation for NLP.Haoyue Shi, Karen Livescu, Kevin Gimpel
2021ASRULayer-Wise Analysis of a Self-Supervised Speech Representation Model.Ankita Pasad, Ju-Chieh Chou, Karen Livescu
2021CVPRFingerspelling Detection in American Sign Language.Bowen Shi, Diane Brentari, Greg Shakhnarovich, Karen Livescu
2021InterspeechLearning Speech Models from Multi-Modal Data.Karen Livescu
2020ACLDiscrete Latent Variable Representations for Low-Resource Text Classification.Shuning Jin, Sam Wiseman, Karl Stratos, Karen Livescu
2020ACLPeTra: A Sparsely Supervised Memory Model for People Tracking.Shubham Toshniwal, Allyson Ettinger, Kevin Gimpel, Karen Livescu
2020EMNLPOn the Role of Supervision in Unsupervised Constituency Parsing.Haoyue Shi, Karen Livescu, Kevin Gimpel
2020EMNLPLearning to Ignore: Long Document Coreference with Bounded Memory Neural Networks.Shubham Toshniwal, Sam Wiseman, Allyson Ettinger, Karen Livescu, Kevin Gimpel
2020ICASSPUnsupervised Pre-Training of Bidirectional Speech Encoders via Masked Reconstruction.Weiran Wang, Qingming Tang, Karen Livescu
2020InterspeechMultilingual Jointly Trained Acoustic and Written Word Embeddings.Yushi Hu, Shane Settle, Karen Livescu
2019ACLVisually Grounded Neural Syntax Acquisition.Haoyue Shi, Jiayuan Mao, Kevin Gimpel, Karen Livescu
2019ICASSPSemantic Query-by-example Speech Search Using Visual Grounding.Herman Kamper, Aristotelis Anastassiou, Karen Livescu
2019ICASSPAcoustically Grounded Word Embeddings for Improved Acoustics-to-word Speech Recognition.Shane Settle, Kartik Audhkhasi, Karen Livescu, Michael Picheny
2019ICCVFingerspelling Recognition in the Wild With Iterative Visual Attention.Bowen Shi, Aurora Martinez Del Rio, Jonathan Keane, Diane Brentari, Greg Shakhnarovich, Karen Livescu
2019InterspeechPre-Trained Text Embeddings for Enhanced Text-to-Speech Synthesis.Tomoki Hayashi, Shinji Watanabe, Tomoki Toda, Kazuya Takeda, Shubham Toshniwal, Karen Livescu
2019InterspeechOn the Contributions of Visual and Textual Supervision in Low-Resource Semantic Speech Retrieval.Ankita Pasad, Bowen Shi, Herman Kamper, Karen Livescu
2019NAACLPre-training on high-resource speech recognition improves low-resource speech-to-text translation.Sameer Bansal, Herman Kamper, Karen Livescu, Adam Lopez, Sharon Goldwater
2018CVPRSemantic Speech Retrieval With a Visually Grounded Model of Untranscribed Speech.Herman Kamper, Gregory Shakhnarovich, Karen Livescu
2018EMNLPVariational Sequential Labelers for Semi-Supervised Learning.Mingda Chen, Qingming Tang, Karen Livescu, Kevin Gimpel
2018ICASSPA Study of All-Convolutional Encoders for Connectionist Temporal Classification.Kalpesh Krishna, Liang Lu, Kevin Gimpel, Karen Livescu
2018ICASSPAcoustic Feature Learning Using Cross-Domain Articulatory Measurements.Qingming Tang, Weiran Wang, Karen Livescu
2018InterspeechLow-Resource Speech-to-Text Translation.Sameer Bansal, Herman Kamper, Karen Livescu, Adam Lopez, Sharon Goldwater
2018NAACLParsing Speech: a Neural Approach to Integrating Lexical and Acoustic-Prosodic Information.Trang Tran, Shubham Toshniwal, Mohit Bansal, Kevin Gimpel, Karen Livescu, Mari Ostendorf
2017ASRUAn embedded segmental K-means model for unsupervised segmentation and clustering of speech.Herman Kamper, Karen Livescu, Sharon Goldwater
2017ASRUMultitask training with unlabeled data for end-to-end sign language fingerspelling recognition.Bowen Shi, Karen Livescu
2017ICLRMulti-view Recurrent Neural Acoustic Word Embeddings.Wanjia He, Weiran Wang, Karen Livescu
2017InterspeechVisually Grounded Learning of Keyword Prediction from Untranscribed Speech.Herman Kamper, Shane Settle, Gregory Shakhnarovich, Karen Livescu
2017InterspeechQuery-by-Example Search with Discriminative Neural Acoustic Word Embeddings.Shane Settle, Keith D. Levin, Herman Kamper, Karen Livescu
2017InterspeechAcoustic Feature Learning via Deep Variational Canonical Correlation Analysis.Qingming Tang, Weiran Wang, Karen Livescu
2017InterspeechMultitask Learning with Low-Level Auxiliary Tasks for Encoder-Decoder Based Speech Recognition.Shubham Toshniwal, Hao Tang, Liang Lu, Karen Livescu
2016EMNLPCharagram: Embedding Words and Sentences via Character n-grams.John Wieting, Mohit Bansal, Kevin Gimpel, Karen Livescu
2016ICASSPSigner-independent fingerspelling recognition with deep neural network adaptation.Taehwan Kim, Weiran Wang, Hao Tang, Karen Livescu
2016ICASSPDeep convolutional acoustic word embeddings using word-pair side information.Herman Kamper, Weiran Wang, Karen Livescu
2016ICMLNonparametric Canonical Correlation Analysis.Tomer Michaeli, Weiran Wang, Karen Livescu
2016InterspeechEfficient Segmental Cascades for Speech Recognition.Hao Tang, Weiran Wang, Kevin Gimpel, Karen Livescu
2016InterspeechTriphone State-Tying via Deep Canonical Correlation Analysis.Weiran Wang, Hao Tang, Karen Livescu
2015ASRUDiscriminative segmental cascades for feature-rich phone recognition.Hao Tang, Weiran Wang, Kevin Gimpel, Karen Livescu
2015ICASSPUnsupervised learning of acoustic features via deep canonical correlation analysis.Weiran Wang, Raman Arora, Karen Livescu, Jeff A. Bilmes
2015ICMLOn Deep Multi-View Representation Learning.Weiran Wang, Raman Arora, Karen Livescu, Jeff A. Bilmes
2015NAACLDeep Multilingual Correlation for Improved Word Embeddings.Ang Lu, Weiran Wang, Mohit Bansal, Kevin Gimpel, Karen Livescu
2014ACLTailoring Continuous Word Representations for Dependency Parsing.Mohit Bansal, Kevin Gimpel, Karen Livescu
2014ICASSPMulti-view learning with supervision for transformed bottleneck features.Raman Arora, Karen Livescu
2014InterspeechA comparison of training approaches for discriminative segmental models.Hao Tang, Kevin Gimpel, Karen Livescu
2013ASRUFixed-dimensional acoustic embeddings of variable-length segments in low-resource settings.Keith D. Levin, Katharine Henry, Aren Jansen, Karen Livescu
2013ICASSPMulti-view CCA-based acoustic features for phonetic recognition across speakers and domains.Raman Arora, Karen Livescu
2013ICASSPDiscriminative articulatory models for spoken term detection in low-resource conversational settings.Rohit Prabhavalkar, Karen Livescu, Eric Fosler-Lussier, Joseph Keshet
2013ICCVFingerspelling Recognition with Semi-Markov Conditional Random Fields.Taehwan Kim, Gregory Shakhnarovich, Karen Livescu
2013ICMLDeep Canonical Correlation Analysis.Galen Andrew, Raman Arora, Jeff A. Bilmes, Karen Livescu
2013InterspeechDiscriminative training of WFST factors with application to pronunciation modeling.Preethi Jyothi, Eric Fosler-Lussier, Karen Livescu
2012ACLDiscriminative Pronunciation Modeling: A Large-Margin, Feature-Rich Approach.Hao Tang, Joseph Keshet, Karen Livescu
2012InterspeechDiscriminatively learning factorized finite state pronunciation models from dynamic Bayesian networks.Preethi Jyothi, Eric Fosler-Lussier, Karen Livescu
2011ASRUA factored conditional random field model for articulatory feature forced transcription.Rohit Prabhavalkar, Eric Fosler-Lussier, Karen Livescu
2011ICASSPLexical access experiments with context-dependent articulatory feature-based models.Preethi Jyothi, Karen Livescu, Eric Fosler-Lussier
2011InterspeechNearest Neighbors with Learned Distances for Phonetic Frame Classification.John Labiak, Karen Livescu
2011InterspeechArticulatory Feature Classification Using Nearest Neighbors.Arild Brandrud Nss, Karen Livescu, Rohit Prabhavalkar
2010InterspeechModeling pronunciation variation with context-dependent articulatory feature decision trees.Samuel R. Bowman, Karen Livescu
2010InterspeechAudio-visual anticipatory coarticulation modeling by human and machine.Louis H. Terry, Karen Livescu, Janet B. Pierrehumbert, Aggelos K. Katsaggelos
2009ASRUMulti-view learning of acoustic features for speaker recognition.Karen Livescu, Mark Stoehr
2009ICASSPOn the phonetic information in ultrasonic microphone signals.Karen Livescu, Bo Zhu, James R. Glass
2009ICMLMulti-view clustering via canonical correlation analysis.Kamalika Chaudhuri, Sham M. Kakade, Karen Livescu, Karthik Sridharan
2007ASRUMonolingual and crosslingual comparison of tandem features derived from articulatory and phone MLPS.zgr etin, Mathew Magimai-Doss, Karen Livescu, Arthur Kantor, Simon King, Chris D. Bartels, Joe Frankel
2007ICASSPAn Articulatory Feature-Based Tandem Approach and Factored Observation Modeling.zgr etin, Arthur Kantor, Simon King, Chris D. Bartels, Mathew Magimai-Doss, Joe Frankel, Karen Livescu
2007ICASSPManual Transcription of Conversational Speech at the Articulatory Feature Level.Karen Livescu, Ari Bezman, Nash M. Borges, Lisa Yung, zgr etin, Joe Frankel, Simon King, Mathew Magimai-Doss, Xuemin Chi, Lisa Lavoie
2007ICASSPArticulatory Feature-Based Methods for Acoustic and Audio-Visual Speech Recognition: Summary from the 2006 JHU Summer workshop.Karen Livescu, zgr etin, Mark Hasegawa-Johnson, Simon King, Chris D. Bartels, Nash M. Borges, Arthur Kantor, Partha Lal, Lisa Yung, Ari Bezman, Stephen Dawson-Haggerty, Bronwyn Woods, Joe Frankel, Mathew Magimai-Doss, Kate Saenko
2007InterspeechArticulatory feature classifiers trained on 2000 hours of telephone speech.Joe Frankel, Mathew Magimai-Doss, Simon King, Karen Livescu, zgr etin
2005ICASSPLandmark-Based Speech Recognition: Report of the 2004 Johns Hopkins Summer Workshop.Mark Hasegawa-Johnson, James Baker, Sarah Borys, Ken Chen, Emily Coogan, Steven Greenberg, Amit Juneja, Katrin Kirchhoff, Karen Livescu, Srividya Mohan, Jennifer Muller, M. Kemal Snmez, Tianyu Wang
2005ICASSPProduction domain modeling of pronunciation for visual speech recognition.Kate Saenko, Karen Livescu, James R. Glass, Trevor Darrell
2005ICCVVisual Speech Recognition with Loosely Synchronized Feature Streams.Kate Saenko, Karen Livescu, Michael Siracusa, Kevin W. Wilson, James R. Glass, Trevor Darrell
2004InterspeechFeature-based pronunciation modeling with trainable asynchrony probabilities.Karen Livescu, James R. Glass
2004NAACLFeature-based Pronunciation Modeling for Speech Recognition.Karen Livescu, James R. Glass
2003InterspeechHidden feature models for speech recognition using dynamic Bayesian networks.Karen Livescu, James R. Glass, Jeff A. Bilmes
2002ICASSPStructurally discriminative graphical models for automatic speech recognition - results from the 2001 Johns Hopkins Summer Workshop.Geoffrey Zweig, Jeff A. Bilmes, Thomas Richardson, Karim Filali, Karen Livescu, Peng Xu, Kirk Jackson, Yigal Brandman, Eric D. Sandness, Eva Holtz, Jerry Torres, Bill Byrne
2001InterspeechSegment-based recognition on the phonebook task: initial results and observations on duration modeling.Karen Livescu, James R. Glass
2000ICASSPLexical modeling of non-native speech for automatic speech recognition.Karen Livescu, James R. Glass