| 2025 | IJCNN | ISL-MED: A General Iterative Self-Learning Framework for Speech Complex Emotion Detection. | Xiaolong Wu, Xinxin Luo, Chang Feng, Hankiz Yilahun, Mingxing Xu, Askar Hamdulla, Thomas Fang Zheng |
| 2024 | ICASSP | Enhancing Quantised End-to-End ASR Models Via Personalisation. | Qiuming Zhao, Guangzhi Sun, Chao Zhang, Mingxing Xu, Thomas Fang Zheng |
| 2024 | ICONIP | Advancing Respiratory Sound Classification: Integration of Audio Spectrogram Transformer with ConnectMix and NEFTune Augmentation. | Runze Huang, Mingxing Xu, Thomas Fang Zheng |
| 2024 | ICONIP | Emotional Atmosphere Soft Label for Emotion Recognition in Conversations. | Xiaolong Wu, Chang Feng, Hankiz Yilahun, Mingxing Xu, Askar Hamdulla, Thomas Fang Zheng |
| 2024 | Interspeech | A Joint Noise Disentanglement and Adversarial Training Framework for Robust Speaker Verification. | Xujiang Xing, Mingxing Xu, Thomas Fang Zheng |
| 2024 | Interspeech | SAML: Speaker Adaptive Mixture of LoRA Experts for End-to-End ASR. | Qiuming Zhao, Guangzhi Sun, Chao Zhang, Mingxing Xu, Thomas Fang Zheng |
| 2024 | Interspeech | Whisper-PMFA: Partial Multi-Scale Feature Aggregation for Speaker Verification using Whisper Models. | Yiyang Zhao, Shuai Wang, Guangzhi Sun, Zehua Chen, Chao Zhang, Mingxing Xu, Thomas Fang Zheng |
| 2023 | VCIP | Robust Point Cloud Classification With Permutohedral Lattice-based Representation. | Li Ding, Mingxing Xu, Wenrui Dai, Cewu Lu, Weisheng Hu, Lin Zhang, Junfeng Du, Hongkai Xiong |
| 2021 | EMNLP | Learning from Multiple Noisy Augmented Data Sets for Better Cross-Lingual Spoken Language Understanding. | Yingmei Guo, Linjun Shou, Jian Pei, Ming Gong, Mingxing Xu, Zhiyong Wu, Daxin Jiang |
| 2021 | Interspeech | Cross-Database Replay Detection in Terminal-Dependent Speaker Verification. | Xingliang Cheng, Mingxing Xu, Thomas Fang Zheng |
| 2020 | IJCNLP | FERNet: Fine-grained Extraction and Reasoning Network for Emotion Recognition in Dialogues. | Yingmei Guo, Zhiyong Wu, Mingxing Xu |
| 2019 | ACII | Multi-Scale Convolutional Recurrent Neural Network with Ensemble Method for Weakly Labeled Sound Event Detection. | Yingmei Guo, Mingxing Xu, Zhiyong Wu, Jianming Wu, Bin Su |
| 2018 | Interspeech | Emotion Recognition from Variable-Length Speech Segments Using Deep Learning on Spectrograms. | Xi Ma, Zhiyong Wu, Jia Jia, Mingxing Xu, Helen Meng, Lianhong Cai |
| 2018 | Interspeech | Imbalance Learning-based Framework for Fear Recognition in the MediaEval Emotional Impact of Movies Task. | Xiaotong Zhang, Xingliang Cheng, Mingxing Xu, Thomas Fang Zheng |
| 2017 | ICASSP | Learning cross-lingual knowledge with multilingual BLSTM for emphasis detection with limited training data. | Yishuang Ning, Zhiyong Wu, Runnan Li, Jia Jia, Mingxing Xu, Helen M. Meng, Lianhong Cai |
| 2017 | ICASSP | Speaker segmentation using deep speaker vectors for fast speaker change scenarios. | Renyu Wang, Mingliang Gu, Lantian Li, Mingxing Xu, Thomas Fang Zheng |
| 2017 | Interspeech | Speech Emotion Recognition with Emotion-Pair Based Framework Considering Emotion Distribution Information in Dimensional Emotion Space. | Xi Ma, Zhiyong Wu, Jia Jia, Mingxing Xu, Helen Meng, Lianhong Cai |
| 2016 | ICASSP | A deep bidirectional long short-term memory based multi-scale approach for music dynamic emotion prediction. | Xinxing Li, Haishu Xianyu, Jiashen Tian, Wenxiao Chen, Fanhang Meng, Mingxing Xu, Lianhong Cai |
| 2016 | ICASSP | Question detection from acoustic features using recurrent neural network with gated recurrent unit. | Yaodong Tang, Yuchen Huang, Zhiyong Wu, Helen Meng, Mingxing Xu, Lianhong Cai |
| 2016 | ICASSP | SVR based double-scale regression for dynamic emotion prediction in music. | Haishu Xianyu, Xinxing Li, Wenxiao Chen, Fanhang Meng, Jiashen Tian, Mingxing Xu, Lianhong Cai |
| 2016 | Interspeech | Combining CNN and BLSTM to Extract Textual and Acoustic Features for Recognizing Stances in Mandarin Ideological Debate Competition. | Linchuan Li, Zhiyong Wu, Mingxing Xu, Helen M. Meng, Lianhong Cai |
| 2016 | Interspeech | Analysis on Gated Recurrent Unit Based Question Detection Approach. | Yaodong Tang, Zhiyong Wu, Helen M. Meng, Mingxing Xu, Lianhong Cai |
| 2014 | IJCNN | Improved keyword spotting system by optimizing posterior confidence measure vector using feed-forward neural network. | Yuchen Liu, Mingxing Xu, Lianhong Cai |
| 2012 | Interspeech | Compensation of Intrinsic Variability with Factor Analysis Modeling for Robust Speaker Verification. | Sheng Chen, Mingxing Xu |
| 2009 | Interspeech | ANN based decision fusion for speech emotion recognition. | Lu Xu, Mingxing Xu, Dali Yang |
| 2008 | ACL | Lyric-based Song Sentiment Classification with Sentiment Vector Space Model. | Yunqing Xia, Linlin Wang, Kam-Fai Wong, Mingxing Xu |
| 2006 | ICASSP | Cohort-Based Speaker Model Synthesis for Channel Robust Speaker Recognition. | Wei Wu, Thomas Fang Zheng, Mingxing Xu |
| 2003 | Interspeech | Using word confidence measure for OOV words detection in a spontaneous spoken dialog system. | Hui Sun, Guoliang Zhang, Fang Zheng, Mingxing Xu |
| 2002 | Interspeech | Improved katz smoothing for language modeling in speech recogniton. | Genqing Wu, Fang Zheng, Wenhu Wu, Mingxing Xu, Ling Jin |
| 2001 | ICASSP | Topic Forest: a plan-based dialog management structure. | Xiaojun Wu, Fang Zheng, Mingxing Xu |
| 2001 | Interspeech | Robust parsing in spoken dialogue systems. | Pengju Yan, Fang Zheng, Mingxing Xu |
| 2000 | Interspeech | Language understanding component for Chinese dialogue system. | Yinfei Huang, Fang Zheng, Mingxing Xu, Pengju Yan, Wenhu Wu |
| 2000 | Interspeech | An equivalent-class based MMI learning method for MGCPM. | Chunhua Luo, Fang Zheng, Mingxing Xu |
| 2000 | Interspeech | Semi-continuous segmental probability modeling for continuous speech recognition. | Jiyong Zhang, Fang Zheng, Mingxing Xu, Ditang Fang |
| 1999 | Interspeech | An effective scoring method for speaking skill evaluation system. | Zhanjiang Song, Fang Zheng, Mingxing Xu, Wenhu Wu |
| 1999 | Interspeech | A fast and effective state decoding algorithm. | Mingxing Xu, Fang Zheng, Wenhu Wu |
| 1999 | Interspeech | Easytalk: a large-vocabulary speaker-independent Chinese dictation machine. | Fang Zheng, Zhanjiang Song, Mingxing Xu, Jian Wu, Yinfei Huang, Wenhu Wu, Cheng Bi |