| 2025 | ICASSP | MHSDB: A Comprehensive Benchmark for Multimodal Humor and Sarcasm Detection Leveraging Foundation Models. | Zhongren Dong, Donghao Wang, Ciqiang Chen, Dong-Yan Huang, Zixing Zhang |
| 2019 | ACII | Speech Emotion Recognition using Spectral Normalized CycleGAN. | Wan Ding, Dong-Yan Huang, Danqing Luo, Yuexian Zou |
| 2017 | ICMI | Audio-visual emotion recognition using deep transfer learning and multiple temporal models. | Xi Ouyang, Shigenori Kawaai, Ester Gue Hua Goh, Shengmei Shen, Wan Ding, Huaiping Ming, Dong-Yan Huang |
| 2017 | Interspeech | Multimodal Prediction of Affective Dimensions via Fusing Multiple Regression Techniques. | Dong-Yan Huang, Wan Ding, Mingyu Xu, Huaiping Ming, Minghui Dong, Xinguo Yu, Haizhou Li |
| 2017 | Interspeech | Denoising Recurrent Neural Network for Deep Bidirectional LSTM Based Voice Conversion. | Jie Wu, Dong-Yan Huang, Lei Xie, Haizhou Li |
| 2016 | ICASSP | Combining multiple kernel models for automatic intelligibility detection of pathological speech. | Dong-Yan Huang, Minghui Dong, Haizhou Li |
| 2016 | ICASSP | Exemplar-based sparse representation of timbre and prosody for voice conversion. | Huaiping Ming, Dong-Yan Huang, Lei Xie, Shaofei Zhang, Minghui Dong, Haizhou Li |
| 2016 | ICMI | Audio and face video emotion recognition in the wild using deep neural networks and small datasets. | Wan Ding, Mingyu Xu, Dong-Yan Huang, Weisi Lin, Minghui Dong, Xinguo Yu, Haizhou Li |
| 2016 | Interspeech | Deep Bidirectional LSTM Modeling of Timbre and Prosody for Emotional Voice Conversion. | Huaiping Ming, Dong-Yan Huang, Lei Xie, Jie Wu, Minghui Dong, Haizhou Li |
| 2015 | ACII | Fundamental frequency modeling using wavelets for emotional voice conversion. | Huaiping Ming, Dong-Yan Huang, Minghui Dong, Haizhou Li, Lei Xie, Shaofei Zhang |
| 2015 | Interspeech | A real-time variable-q non-stationary Gabor transform for pitch shifting. | Dong-Yan Huang, Minghui Dong, Haizhou Li |
| 2015 | Interspeech | An alternating optimization approach for phase retrieval. | Huaiping Ming, Dong-Yan Huang, Lei Xie, Haizhou Li, Minghui Dong |
| 2015 | Interspeech | Regularized non-negative matrix factorization using alternating direction method of multipliers and its application to source separation. | Shaofei Zhang, Dong-Yan Huang, Lei Xie, Engsiong Chng, Haizhou Li, Minghui Dong |
| 2014 | ICASSP | Intelligibility detection of pathological speech using asymmetric sparse kernel partial least squares classifier. | Dong-Yan Huang, Minghui Dong, Haizhou Li |
| 2014 | Interspeech | I | Minghui Dong, Siu Wa Lee, Haizhou Li, Paul Y. Chan, Xuejian Peng, Jochen Walter Ehnes, Dong-Yan Huang |
| 2012 | Interspeech | Detecting Intelligibility by Linear Dimensionality Reduction and Normalized Voice Quality Hierarchical Features. | Dong-Yan Huang, Yongwei Zhu, Dajun Wu, Rongshan Yu |
| 2012 | ISCAS | A comparison of SVM and asymmetric SIMPLS in emotion recognition from naturalistic dialogues. | Dong-Yan Huang, Wei Sun |
| 2011 | Interspeech | Speaker State Classification Based on Fusion of Asymmetric SIMPLS and Support Vector Machines. | Dong-Yan Huang, Shuzhi Sam Ge, Zhengchen Zhang |
| 2009 | ISCAS | The Misadjustment of the Cascaded LMS Prediction Filter. | Dong-Yan Huang, Susanto Rahardja |
| 2009 | MMSP | Biologically inspired algorithm for enhancement of speech intelligibility over telephone channel. | Dong-Yan Huang, Susanto Rahardja, Ee Ping Ong |
| 2008 | VTC | Convergence Performance of the Cascaded RLS-LMS Prediction. | Dong-Yan Huang, Susanto Rahardja |
| 2005 | ICASSP | Characterization of a cascade LMS predictor. | Dong-Yan Huang, Xinrong Su, Arumugam Nallanathan |
| 2005 | ICASSP | Software simulation tools on forward error correction schemes for the wireless transmission of MPEG4 AAC audio bitstreams. | Wai Choong Wong, Zheng Yang Chin, Dong-Yan Huang |
| 2005 | ISCAS | A performance bound for a cascade LMS predictor. | Dong-Yan Huang, Xinrong Su |
| 2004 | ICASSP | Performance analysis of an RLS-LMS algorithm for lossless audio compression. | Dong-Yan Huang |
| 2004 | ISCAS | Speech pitch detection in noisy environment using multi-rate adaptive lossless FIR filters. | Dong-Yan Huang, Weisi Lin, Susanto Rahardja |
| 1999 | ISCAS | Implementation of the MPEG-4 advanced audio coding encoder on ADSP-21060 SHARC. | Dong-Yan Huang, Xuesong Gong, Daqing Zhou, Miki Toshio, Sanae Hotani |
| 1997 | ICASSP | Comparison of two eigenstructure algorithms for lossless multirate filter optimization. | Dong-Yan Huang, Phillip A. Regalia, Maurice G. Bellanger |
| 1995 | ICASSP | Attainable error bounds in multirate adaptive lossless FIR filters. | Phillip A. Regalia, Dong-Yan Huang |