Ann Lee
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
28
Venues
9
Active years
2011–2025
Best venue rank
A*
Where they publish
Papers
28 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ASRU | Meta Audiobox Aesthetics: Unified Automatic Assessment for Speech, Music and Sound. | Andros Tjandra, Yi-Chiao Wu, Baishan Guo, John Hoffman, Brian Ellis, Apoorv Vyas, Bowen Shi, Sanyuan Chen, Matt Le, Nick Zacharov, Carleigh Wood, Ann Lee, Wei-Ning Hsu |
| 2024 | ACL | Textless Acoustic Model with Self-Supervised Distillation for Noise-Robust Expressive Speech-to-Speech Translation. | Min-Jae Hwang, Ilia Kulikov, Benjamin N. Peloquin, Hongyu Gong, Peng-Jen Chen, Ann Lee |
| 2024 | ICASSP | M2BART: Multilingual and Multimodal Encoder-Decoder Pre-Training for Any-to-Any Machine Translation. | Peng-Jen Chen, Bowen Shi, Kelvin Niu, Ann Lee, Wei-Ning Hsu |
| 2023 | ACL | Speech-to-Speech Translation for a Real-world Unwritten Language. | Peng-Jen Chen, Kevin Tran, Yilin Yang, Jingfei Du, Justine Kao, Yu-An Chung, Paden Tomasello, Paul-Ambroise Duquenne, Holger Schwenk, Hongyu Gong, Hirofumi Inaguma, Sravya Popuri, Changhan Wang, Juan Pino, Wei-Ning Hsu, Ann Lee |
| 2023 | ACL | SpeechMatrix: A Large-Scale Mined Corpus of Multilingual Speech-to-Speech Translations. | Paul-Ambroise Duquenne, Hongyu Gong, Ning Dong, Jingfei Du, Ann Lee, Vedanuj Goswami, Changhan Wang, Juan Pino, Benot Sagot, Holger Schwenk |
| 2023 | ACL | UnitY: Two-pass Direct Speech-to-speech Translation with Discrete Units. | Hirofumi Inaguma, Sravya Popuri, Ilia Kulikov, Peng-Jen Chen, Changhan Wang, Yu-An Chung, Yun Tang, Ann Lee, Shinji Watanabe, Juan Pino |
| 2023 | CHI | Designing for Peer-Led Critical Pedagogies in Computer-Mediated Support Groups for Home Care Workers. | Anthony Poon, Lourdes Guerrero, Julia Loughman, Matthew Luebke, Ann Lee, Madeline R. Sterling, Nicola Dell |
| 2023 | ICASSP | A Holistic Cascade System, Benchmark, and Human Evaluation Protocol for Expressive Speech-to-Speech Translation. | Wen-Chin Huang, Benjamin N. Peloquin, Justine Kao, Changhan Wang, Hongyu Gong, Elizabeth Salesky, Yossi Adi, Ann Lee, Peng-Jen Chen |
| 2023 | ICASSP | Bridging Speech and Textual Pre-Trained Models With Unsupervised ASR. | Jiatong Shi, Chan-Jan Hsu, Ho-Lam Chung, Dongji Gao, Paola Garca, Shinji Watanabe, Ann Lee, Hung-Yi Lee |
| 2023 | ICASSP | Enhancing Speech-To-Speech Translation with Multiple TTS Targets. | Jiatong Shi, Yun Tang, Ann Lee, Hirofumi Inaguma, Changhan Wang, Juan Pino, Shinji Watanabe |
| 2022 | ACL | Text-Free Prosody-Aware Generative Spoken Language Modeling. | Eugene Kharitonov, Ann Lee, Adam Polyak, Yossi Adi, Jade Copet, Kushal Lakhotia, Tu Anh Nguyen, Morgane Rivire, Abdelrahman Mohamed, Emmanuel Dupoux, Wei-Ning Hsu |
| 2022 | ACL | Direct Speech-to-Speech Translation With Discrete Units. | Ann Lee, Peng-Jen Chen, Changhan Wang, Jiatao Gu, Sravya Popuri, Xutai Ma, Adam Polyak, Yossi Adi, Qing He, Yun Tang, Juan Pino, Wei-Ning Hsu |
| 2022 | ICML | Flashlight: Enabling Innovation in Tools for Machine Learning. | Jacob D. Kahn, Vineel Pratap, Tatiana Likhomanenko, Qiantong Xu, Awni Y. Hannun, Jeff Cai, Paden Tomasello, Ann Lee, Edouard Grave, Gilad Avidov, Benoit Steiner, Vitaliy Liptchinsky, Gabriel Synnaeve, Ronan Collobert |
| 2022 | Interspeech | Enhanced Direct Speech-to-Speech Translation Using Self-supervised Pre-training and Data Augmentation. | Sravya Popuri, Peng-Jen Chen, Changhan Wang, Juan Pino, Yossi Adi, Jiatao Gu, Wei-Ning Hsu, Ann Lee |
| 2022 | NAACL | Textless Speech-to-Speech Translation on Real Data. | Ann Lee, Hongyu Gong, Paul-Ambroise Duquenne, Holger Schwenk, Peng-Jen Chen, Changhan Wang, Sravya Popuri, Yossi Adi, Juan Miguel Pino, Jiatao Gu, Wei-Ning Hsu |
| 2021 | ACL | Discriminative Reranking for Neural Machine Translation. | Ann Lee, Michael Auli, Marc'Aurelio Ranzato |
| 2021 | ACL | VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation. | Changhan Wang, Morgane Rivire, Ann Lee, Anne Wu, Chaitanya Talnikar, Daniel Haziza, Mary Williamson, Juan Miguel Pino, Emmanuel Dupoux |
| 2021 | EMNLP | fairseq S\^2: A Scalable and Integrable Speech Synthesis Toolkit. | Changhan Wang, Wei-Ning Hsu, Yossi Adi, Adam Polyak, Ann Lee, Peng-Jen Chen, Jiatao Gu, Juan Pino |
| 2021 | Interspeech | Robust wav2vec 2.0: Analyzing Domain Shift in Self-Supervised Pre-Training. | Wei-Ning Hsu, Anuroop Sriram, Alexei Baevski, Tatiana Likhomanenko, Qiantong Xu, Vineel Pratap, Jacob Kahn, Ann Lee, Ronan Collobert, Gabriel Synnaeve, Michael Auli |
| 2020 | ICASSP | Self-Training for End-to-End Speech Recognition. | Jacob Kahn, Ann Lee, Awni Y. Hannun |
| 2019 | Interspeech | Sequence-to-Sequence Speech Recognition with Time-Depth Separable Convolutions. | Awni Y. Hannun, Ann Lee, Qiantong Xu, Ronan Collobert |
| 2016 | ICASSP | Personalized mispronunciation detection and diagnosis based on unsupervised error pattern discovery. | Ann Lee, Nancy F. Chen, James R. Glass |
| 2016 | Interspeech | Exploiting Depth and Highway Connections in Convolutional Recurrent Deep Neural Networks for Speech Recognition. | Wei-Ning Hsu, Yu Zhang, Ann Lee, James R. Glass |
| 2015 | Interspeech | Mispronunciation detection without nonnative training data. | Ann Lee, James R. Glass |
| 2014 | Interspeech | Context-dependent pronunciation error pattern discovery with limited annotations. | Ann Lee, James R. Glass |
| 2013 | ICASSP | Mispronunciation detection via dynamic time warping on deep belief network-based posteriorgrams. | Ann Lee, Yaodong Zhang, James R. Glass |
| 2012 | Interspeech | Sentence Detection Using Multiple Annotations. | Ann Lee, James R. Glass |
| 2011 | MMSP | Automatic highlights extraction for drama video using music emotion and human face features. | Keng-Sheng Lin, Ann Lee, Yi-Hsuan Yang, Cheng-Te Lee, Homer H. Chen |