Skip to content

Ann Lee

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

28

Venues

9

Active years

2011–2025

Best venue rank

A*

Where they publish

Papers

28 indexed papers, newest first.

YearVenueTitleAuthors
2025ASRUMeta Audiobox Aesthetics: Unified Automatic Assessment for Speech, Music and Sound.Andros Tjandra, Yi-Chiao Wu, Baishan Guo, John Hoffman, Brian Ellis, Apoorv Vyas, Bowen Shi, Sanyuan Chen, Matt Le, Nick Zacharov, Carleigh Wood, Ann Lee, Wei-Ning Hsu
2024ACLTextless Acoustic Model with Self-Supervised Distillation for Noise-Robust Expressive Speech-to-Speech Translation.Min-Jae Hwang, Ilia Kulikov, Benjamin N. Peloquin, Hongyu Gong, Peng-Jen Chen, Ann Lee
2024ICASSPM2BART: Multilingual and Multimodal Encoder-Decoder Pre-Training for Any-to-Any Machine Translation.Peng-Jen Chen, Bowen Shi, Kelvin Niu, Ann Lee, Wei-Ning Hsu
2023ACLSpeech-to-Speech Translation for a Real-world Unwritten Language.Peng-Jen Chen, Kevin Tran, Yilin Yang, Jingfei Du, Justine Kao, Yu-An Chung, Paden Tomasello, Paul-Ambroise Duquenne, Holger Schwenk, Hongyu Gong, Hirofumi Inaguma, Sravya Popuri, Changhan Wang, Juan Pino, Wei-Ning Hsu, Ann Lee
2023ACLSpeechMatrix: A Large-Scale Mined Corpus of Multilingual Speech-to-Speech Translations.Paul-Ambroise Duquenne, Hongyu Gong, Ning Dong, Jingfei Du, Ann Lee, Vedanuj Goswami, Changhan Wang, Juan Pino, Benot Sagot, Holger Schwenk
2023ACLUnitY: Two-pass Direct Speech-to-speech Translation with Discrete Units.Hirofumi Inaguma, Sravya Popuri, Ilia Kulikov, Peng-Jen Chen, Changhan Wang, Yu-An Chung, Yun Tang, Ann Lee, Shinji Watanabe, Juan Pino
2023CHIDesigning for Peer-Led Critical Pedagogies in Computer-Mediated Support Groups for Home Care Workers.Anthony Poon, Lourdes Guerrero, Julia Loughman, Matthew Luebke, Ann Lee, Madeline R. Sterling, Nicola Dell
2023ICASSPA Holistic Cascade System, Benchmark, and Human Evaluation Protocol for Expressive Speech-to-Speech Translation.Wen-Chin Huang, Benjamin N. Peloquin, Justine Kao, Changhan Wang, Hongyu Gong, Elizabeth Salesky, Yossi Adi, Ann Lee, Peng-Jen Chen
2023ICASSPBridging Speech and Textual Pre-Trained Models With Unsupervised ASR.Jiatong Shi, Chan-Jan Hsu, Ho-Lam Chung, Dongji Gao, Paola Garca, Shinji Watanabe, Ann Lee, Hung-Yi Lee
2023ICASSPEnhancing Speech-To-Speech Translation with Multiple TTS Targets.Jiatong Shi, Yun Tang, Ann Lee, Hirofumi Inaguma, Changhan Wang, Juan Pino, Shinji Watanabe
2022ACLText-Free Prosody-Aware Generative Spoken Language Modeling.Eugene Kharitonov, Ann Lee, Adam Polyak, Yossi Adi, Jade Copet, Kushal Lakhotia, Tu Anh Nguyen, Morgane Rivire, Abdelrahman Mohamed, Emmanuel Dupoux, Wei-Ning Hsu
2022ACLDirect Speech-to-Speech Translation With Discrete Units.Ann Lee, Peng-Jen Chen, Changhan Wang, Jiatao Gu, Sravya Popuri, Xutai Ma, Adam Polyak, Yossi Adi, Qing He, Yun Tang, Juan Pino, Wei-Ning Hsu
2022ICMLFlashlight: Enabling Innovation in Tools for Machine Learning.Jacob D. Kahn, Vineel Pratap, Tatiana Likhomanenko, Qiantong Xu, Awni Y. Hannun, Jeff Cai, Paden Tomasello, Ann Lee, Edouard Grave, Gilad Avidov, Benoit Steiner, Vitaliy Liptchinsky, Gabriel Synnaeve, Ronan Collobert
2022InterspeechEnhanced Direct Speech-to-Speech Translation Using Self-supervised Pre-training and Data Augmentation.Sravya Popuri, Peng-Jen Chen, Changhan Wang, Juan Pino, Yossi Adi, Jiatao Gu, Wei-Ning Hsu, Ann Lee
2022NAACLTextless Speech-to-Speech Translation on Real Data.Ann Lee, Hongyu Gong, Paul-Ambroise Duquenne, Holger Schwenk, Peng-Jen Chen, Changhan Wang, Sravya Popuri, Yossi Adi, Juan Miguel Pino, Jiatao Gu, Wei-Ning Hsu
2021ACLDiscriminative Reranking for Neural Machine Translation.Ann Lee, Michael Auli, Marc'Aurelio Ranzato
2021ACLVoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation.Changhan Wang, Morgane Rivire, Ann Lee, Anne Wu, Chaitanya Talnikar, Daniel Haziza, Mary Williamson, Juan Miguel Pino, Emmanuel Dupoux
2021EMNLPfairseq S\^2: A Scalable and Integrable Speech Synthesis Toolkit.Changhan Wang, Wei-Ning Hsu, Yossi Adi, Adam Polyak, Ann Lee, Peng-Jen Chen, Jiatao Gu, Juan Pino
2021InterspeechRobust wav2vec 2.0: Analyzing Domain Shift in Self-Supervised Pre-Training.Wei-Ning Hsu, Anuroop Sriram, Alexei Baevski, Tatiana Likhomanenko, Qiantong Xu, Vineel Pratap, Jacob Kahn, Ann Lee, Ronan Collobert, Gabriel Synnaeve, Michael Auli
2020ICASSPSelf-Training for End-to-End Speech Recognition.Jacob Kahn, Ann Lee, Awni Y. Hannun
2019InterspeechSequence-to-Sequence Speech Recognition with Time-Depth Separable Convolutions.Awni Y. Hannun, Ann Lee, Qiantong Xu, Ronan Collobert
2016ICASSPPersonalized mispronunciation detection and diagnosis based on unsupervised error pattern discovery.Ann Lee, Nancy F. Chen, James R. Glass
2016InterspeechExploiting Depth and Highway Connections in Convolutional Recurrent Deep Neural Networks for Speech Recognition.Wei-Ning Hsu, Yu Zhang, Ann Lee, James R. Glass
2015InterspeechMispronunciation detection without nonnative training data.Ann Lee, James R. Glass
2014InterspeechContext-dependent pronunciation error pattern discovery with limited annotations.Ann Lee, James R. Glass
2013ICASSPMispronunciation detection via dynamic time warping on deep belief network-based posteriorgrams.Ann Lee, Yaodong Zhang, James R. Glass
2012InterspeechSentence Detection Using Multiple Annotations.Ann Lee, James R. Glass
2011MMSPAutomatic highlights extraction for drama video using music emotion and human face features.Keng-Sheng Lin, Ann Lee, Yi-Hsuan Yang, Cheng-Te Lee, Homer H. Chen