| 2025 | COLING | Transformer-based Speech Model Learns Well as Infants and Encodes Abstractions through Exemplars in the Poverty of the Stimulus Environment. | Yi Yang, Yiming Wang, Jiahong Yuan |
| 2025 | ICASSP | The USTC System for EEG-Music Emotion Recognition Challenge. | Jiaxin Chen, Yiming Wang, Yin-Long Liu, Rui Feng, Jiahong Yuan, Zhen-Hua Ling |
| 2025 | ICASSP | Cross-Lingual Speech Emotion Recognition: Humans vs. Self-Supervised Models. | Zhichen Han, Tianqi Geng, Hui Feng, Jiahong Yuan, Korin Richmond, Yuanchao Li |
| 2025 | ICMI | Beyond Mimicry: Auditing Human Bias to Build Fairer AI for Alzheimer's Assessment. | Liu He, Rui Feng, Xinran Han, Yin-Long Liu, Jiahong Yuan |
| 2024 | EMNLP | Automated Tone Transcription and Clustering with Tone2Vec. | Yi Yang, Yiming Wang, Zhiqiang Tang, Jiahong Yuan |
| 2023 | ICASSP | The Ustc System for Adress-m Challenge. | Kangdi Mei, Xinyun Ding, Yinlong Liu, Zhiqiang Guo, Feiyang Xu, Xin Li, Tuya Naren, Jiahong Yuan, Zhenhua Ling |
| 2023 | Interspeech | Improved Contextualized Speech Representations for Tonal Analysis. | Jiahong Yuan, Xingyu Cai, Kenneth Church |
| 2022 | ICASSP | Text2video: Text-Driven Talking-Head Video Synthesis with Personalized Phoneme - Pose Dictionary. | Sibo Zhang, Jiahong Yuan, Miao Liao, Liangjun Zhang |
| 2022 | ICLR | W-CTC: a Connectionist Temporal Classification Loss with Wild Cards. | Xingyu Cai, Jiahong Yuan, Yuchen Bian, Guangxu Xun, Jiaji Huang, Kenneth Church |
| 2021 | ASRU | Decoupling Recognition and Transcription in Mandarin ASR. | Jiahong Yuan, Xingyu Cai, Dongji Gao, Renjie Zheng, Liang Huang, Kenneth Church |
| 2021 | ICASSP | Speaking Rate and Tonal Realization in Mandarin Chinese: What Can We Learn From Large Speech Corpora? | Jiahong Yuan, Kenneth Church |
| 2021 | ICASSP | Pause-Encoded Language Models for Recognition of Alzheimer's Disease and Emotion. | Jiahong Yuan, Xingyu Cai, Kenneth Church |
| 2021 | Interspeech | Speech Emotion Recognition with Multi-Task Learning. | Xingyu Cai, Jiahong Yuan, Renjie Zheng, Liang Huang, Kenneth Church |
| 2021 | NAACL | On Attention Redundancy: A Comprehensive Study. | Yuchen Bian, Jiaji Huang, Xingyu Cai, Jiahong Yuan, Kenneth Church |
| 2020 | EMNLP | Fluent and Low-latency Simultaneous Speech-to-Speech Translation with Self-adaptive Training. | Renjie Zheng, Mingbo Ma, Baigong Zheng, Kaibo Liu, Jiahong Yuan, Kenneth Church, Liang Huang |
| 2020 | ICASSP | Detection and Analysis of T/D Deletion in Librispeech. | Jiahong Yuan, Hui Lin, Yang Liu |
| 2020 | Interspeech | LAIX Corpus of Chinese Learner English: Towards a Benchmark for L2 English ASR. | Yanhong Wang, Huan Luan, Jiahong Yuan, Bin Wang, Hui Lin |
| 2020 | Interspeech | Disfluencies and Fine-Tuning Pre-Trained Language Models for Detection of Alzheimer's Disease. | Jiahong Yuan, Yuchen Bian, Xingyu Cai, Jiaji Huang, Zheng Ye, Kenneth Church |
| 2019 | ICASSP | Classification of Chinese Dialect Regions from L2 English Speech. | Jiahong Yuan, Zhengqiang Rao, Hui Lin, Yang Liu |
| 2019 | Interspeech | On the Role of Style in Parsing Speech with Neural Models. | Trang Tran, Jiahong Yuan, Yang Liu, Mari Ostendorf |
| 2018 | Interspeech | GlobalTIMIT: Acoustic-Phonetic Datasets for the World's Languages. | Nattanun Chanchaochai, Christopher Cieri, Japhet Debrah, Hongwei Ding, Yue Jiang, Sishi Liao, Mark Y. Liberman, Jonathan Wright, Jiahong Yuan, Juhong Zhan, Yuqing Zhan |
| 2018 | Interspeech | Pitch Characteristics of L2 English Speech by Chinese Speakers: A Large-scale Study. | Jiahong Yuan, Qiusi Dong, Fei Wu, Huan Luan, Xiaofei Yang, Hui Lin, Yang Liu |
| 2016 | Interspeech | The Rhythmic Constraint on Prosodic Boundaries in Mandarin Chinese Based on Corpora of Silent Reading and Speech Perception. | Wei Lai, Jiahong Yuan, Ya Li, Xiaoying Xu, Mark Y. Liberman |
| 2016 | Interspeech | Phoneme, Phone Boundary, and Tone in Automatic Scoring of Mandarin Proficiency. | Jiahong Yuan, Mark Y. Liberman |
| 2015 | Interspeech | Investigating consonant reduction in Mandarin Chinese with improved forced alignment. | Jiahong Yuan, Mark Y. Liberman |
| 2014 | ICASSP | Mandarin tone classification without pitch tracking. | Neville Ryant, Jiahong Yuan, Mark Y. Liberman |
| 2014 | ICASSP | Highly accurate phonetic segmentation using boundary correction models and system fusion. | Andreas Stolcke, Neville Ryant, Vikramjit Mitra, Jiahong Yuan, Wen Wang, Mark Y. Liberman |
| 2014 | ICASSP | Automatic phonetic segmentation in Mandarin Chinese: Boundary models, glottal features and tone. | Jiahong Yuan, Neville Ryant, Mark Y. Liberman |
| 2013 | ICASSP | Using multiple versions of speech input in phone recognition. | Mark Y. Liberman, Jiahong Yuan, Andreas Stolcke, Wen Wang, Vikramjit Mitra |
| 2013 | ICASSP | Articulatory trajectories for large-vocabulary speech recognition. | Vikramjit Mitra, Wen Wang, Andreas Stolcke, Hosung Nam, Colleen Richey, Jiahong Yuan, Mark Y. Liberman |
| 2013 | ICASSP | Scale-space expansion of acoustic features improves speech event detection. | Neville Ryant, Jiahong Yuan, Mark Y. Liberman |
| 2013 | Interspeech | Speech activity detection on youtube using deep neural networks. | Neville Ryant, Mark Y. Liberman, Jiahong Yuan |
| 2013 | Interspeech | The spectral dynamics of vowels in Mandarin Chinese. | Jiahong Yuan |
| 2013 | Interspeech | Automatic phonetic segmentation using boundary models. | Jiahong Yuan, Neville Ryant, Mark Y. Liberman, Andreas Stolcke, Vikramjit Mitra, Wen Wang |
| 2013 | NAACL | A Cross-language Study on Automatic Speech Disfluency Detection. | Wen Wang, Andreas Stolcke, Jiahong Yuan, Mark Y. Liberman |
| 2011 | ASRU | Automatic detection of "g-dropping" in American English using forced alignment. | Jiahong Yuan, Mark Y. Liberman |
| 2010 | ICASSP | Robust speaking rate estimation using broad phonetic class recognition. | Jiahong Yuan, Mark Y. Liberman |
| 2010 | Interspeech | Linguistic rhythm in foreign accent. | Jiahong Yuan |
| 2010 | Interspeech | F | Jiahong Yuan, Mark Y. Liberman |
| 2009 | Interspeech | Comparison of vowel structures of Japanese and English in articulatory and auditory spaces. | Jianwu Dang, Mark Tiede, Jiahong Yuan |
| 2009 | Interspeech | Investigating /l/ variation in English through forced alignment. | Jiahong Yuan, Mark Y. Liberman |
| 2008 | Interspeech | Covariations of English segmental durations across speakers. | Jiahong Yuan |
| 2008 | Interspeech | Different roles of pitch and duration in distinguishing word stress in English. | Jiahong Yuan, Stephen Isard, Mark Y. Liberman |
| 2007 | Interspeech | A corpus study of the 3 | Yiya Chen, Jiahong Yuan |
| 2007 | Interspeech | Perception of disfluency: language differences and listener bias. | Catherine Lai, Kyle Gorman, Jiahong Yuan, Mark Y. Liberman |
| 2006 | Interspeech | Towards an integrated understanding of speaking rate in conversation. | Jiahong Yuan, Mark Y. Liberman, Christopher Cieri |
| 2005 | Interspeech | Pitch accent prediction: effects of genre and speaker. | Jiahong Yuan, Jason M. Brenier, Daniel Jurafsky |
| 2002 | Interspeech | The acoustic realization of anger, fear, joy and sadness in Chinese. | Jiahong Yuan, Liqin Shen, Fangxin Chen |