| 2025 | Interspeech | Label Semantic-Driven Contrastive Learning for Speech Emotion Recognition. | Jiaxi Hu, Leyuan Qu, Haoxun Li, Taihao Li |
| 2025 | Interspeech | EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis. | Haoxun Li, Leyuan Qu, Jiaxi Hu, Taihao Li |
| 2024 | ICASSP | Improving Speech Emotion Recognition with Unsupervised Speaking Style Transfer. | Leyuan Qu, Wei Wang, Cornelius Weber, Pengcheng Yue, Taihao Li, Stefan Wermter |
| 2024 | ICASSP | Multi-Modal Emotion Recognition Using Multiple Acoustic Features and Dual Cross-Modal Transformer. | Yanfeng Wu, Pengcheng Yue, Leyuan Qu, Taihao Li, Yu-Ping Ruan |
| 2022 | LREC | A Multimodal German Dataset for Automatic Lip Reading Systems and Transfer Learning. | Gerald Schwiebert, Cornelius Weber, Leyuan Qu, Henrique Siqueira, Stefan Wermter |
| 2021 | ASRU | Hearing Faces: Target Speaker Text-to-Speech Synthesis from a Face. | Bjrn Plster, Cornelius Weber, Leyuan Qu, Stefan Wermter |
| 2020 | ICANN | Variational Autoencoder with Global- and Medium Timescale Auxiliaries for Emotion Recognition from Speech. | Hussam Almotlak, Cornelius Weber, Leyuan Qu, Stefan Wermter |
| 2020 | Interspeech | Multimodal Target Speech Separation with Voice and Face References. | Leyuan Qu, Cornelius Weber, Stefan Wermter |
| 2019 | Interspeech | LipSound: Neural Mel-Spectrogram Reconstruction for Lip Reading. | Leyuan Qu, Cornelius Weber, Stefan Wermter |
| 2018 | ICANN | Combining Articulatory Features with End-to-End Learning in Speech Recognition. | Leyuan Qu, Cornelius Weber, Egor Lakomkin, Johannes Twiefel, Stefan Wermter |
| 2016 | ICASSP | Landmark of Mandarin nasal codas and its application in pronunciation error detection. | Yanlu Xie, Mark Hasegawa-Johnson, Leyuan Qu, Jinsong Zhang |