| 2025 | ASRU | KAN-AST: Kolmogorov-Arnold Network based Audio Spectrogram Transformer for Audio Classification. | Phuong Tuan Dat, Tran Huy Dat |
| 2025 | ASRU | Utilizing Kolmogorov-Arnold Network in Self-Supervised Learning for Speaker Diarization. | Minh Vu, Phuong Tuan Dat, Kah Kuan Teh, Van Tuan Nguyen, Tran Huy Dat |
| 2025 | AVSS | XLSR-Kanformer: A KAN-Intergrated model for Synthetic Speech Detection. | Tuan Dat Phuong, Tran Huy Dat |
| 2025 | ICASSP | Automatic Speech Recognition and Spoken Language Understanding of Maritime Radio Communications: A case study with Singapore data. | Phuong Dat, Jayakrishnan Melur Madhathil, Tran Huy Dat |
| 2019 | ICASSP | Embedding Physical Augmentation and Wavelet Scattering Transform to Generative Adversarial Networks for Audio Classification with Limited Training Resources. | Kah Kuan Teh, Tran Huy Dat |
| 2019 | Interspeech | The I2R's ASR System for the VOiCES from a Distance Challenge 2019. | Tze Yuang Chong, Kye Min Tan, Kah Kuan Teh, Chang Huai You, Hanwu Sun, Tran Huy Dat |
| 2019 | Interspeech | The I2R's ASR System for the VOiCES from a Distance Challenge 2019. | Tze Yuang Chong, Kye Min Tan, Kah Kuan Teh, Chang Huai You, Hanwu Sun, Tran Huy Dat |
| 2019 | Interspeech | The I2R's Submission to VOiCES Distance Speaker Recognition Challenge 2019. | Hanwu Sun, Kah Kuan Teh, Ivan Kukanov, Tran Huy Dat |
| 2017 | Interspeech | Data Augmentation, Missing Feature Mask and Kernel Classification for Through-the-Wall Acoustic Surveillance. | Tran Huy Dat, Wen Zheng Terence Ng, Yi Ren Leng |
| 2017 | Interspeech | An Integrated Solution for Snoring Sound Classification Using Bhattacharyya Distance Based GMM Supervectors with SVM, Feature Selection with Random Forest and Spectrogram with CNN. | Tin Lay Nwe, Tran Huy Dat, Wen Zheng Terence Ng, Bin Ma |
| 2016 | ICASSP | A comparative study of multi-channel processing methods for noisy automatic speech recognition in urban environments. | Tran Huy Dat, Jonathan William Dennis, Yi Ren Leng, Wen Zheng Terence Ng |
| 2015 | ASRU | Single and multi-channel approaches for distant speech recognition under noisy reverberant conditions: I2R'S system description for the ASpIRE challenge. | Jonathan William Dennis, Tran Huy Dat |
| 2015 | ICASSP | Combining robust spike coding with spiking neural networks for sound event classification. | Jonathan William Dennis, Tran Huy Dat, Haizhou Li |
| 2015 | Interspeech | Spiking neural networks and the generalised hough transform for speech pattern detection. | Jonathan William Dennis, Tran Huy Dat, Haizhou Li |
| 2014 | ICASSP | Generalized Gaussian Distribution Kullback-Leibler kernel for robust sound event recognition. | Tran Huy Dat, Wen Zheng Terence Ng, Jonathan William Dennis, Yi Ren Leng |
| 2014 | ICASSP | A discriminatively trained Hough Transform for frame-level phoneme recognition. | Jonathan William Dennis, Tran Huy Dat, Haizhou Li, Engsiong Chng |
| 2014 | Interspeech | Analysis of spectrogram image methods for sound event classification. | Jonathan William Dennis, Tran Huy Dat, Chng Eng Siong |
| 2013 | ICASSP | Temporal coding of local spectrogram features for robust sound recognition. | Jonathan William Dennis, Qiang Yu, Huajin Tang, Tran Huy Dat, Haizhou Li |
| 2013 | ICOST | Evaluation of the Pet Robot CuDDler Using Godspeed Questionnaire. | Yeow Kee Tan, Alvin Hong Yee Wong, Chern Yuen Anthony Wong, Tran Anh Dung, Adrian Hwang Jian Tay, Dilip Kumar Limbu, Tran Huy Dat, Weng Zheng Ng, Rui Yan, Benedict Tay Tiong Chee |
| 2012 | Interspeech | Overlapping Sound Event Recognition using Local Spectrogram Features with the Generalised Hough Transform. | Jonathan William Dennis, Tran Huy Dat, Engsiong Chng |
| 2012 | Interspeech | Using Blob Detection in Missing Feature Linear-Frequency Cepstral Coefficients for Robust Sound Event Recognition. | Yi Ren Leng, Tran Huy Dat |
| 2011 | ICASSP | Probabilistic distance SVM with Hellinger-Exponential Kernel for sound event classification. | Tran Huy Dat, Haizhou Li |
| 2011 | ICASSP | Jump Function Kolmogorov for overlapping audio event classification. | Tran Huy Dat, Haizhou Li |
| 2011 | Interspeech | Image Representation of the Subband Power Distribution for Robust Sound Classification. | Jonathan William Dennis, Tran Huy Dat, Haizhou Li |
| 2011 | Interspeech | Semi-Supervised Tree Support Vector Machine for Online Cough Recognition. | Huynh Thai Hoa, An Vu Tran, Tran Huy Dat |
| 2011 | Interspeech | Alternative Frequency Scale Cepstral Coefficient for Robust Sound Event Recognition. | Yi Ren Leng, Tran Huy Dat, Norihide Kitaoka, Haizhou Li |
| 2010 | ICASSP | Feature integration for heart sound biometrics. | Tran Huy Dat, Yi Ren Leng, Haizhou Li |
| 2010 | Interspeech | Selective gammatone filterbank feature for robust sound event recognition. | Yi Ren Leng, Tran Huy Dat, Norihide Kitaoka, Haizhou Li |
| 2009 | ICASSP | Sound event classification based on Feature Integration, Recursive Feature Elimination and Structured Classification. | Tran Huy Dat, Haizhou Li |
| 2008 | ICASSP | Jump function komogorov and its application for audio stream segmentation and classification. | Tran Huy Dat, Haizhou Li |
| 2008 | Interspeech | Speaker identification in noise mismatch conditions based on jump function Kolmogorov analysis in wavelet domain. | Tran Huy Dat, Haizhou Li |
| 2007 | ICASSP | Feature Selection Based on Fisher Ratio and Mutual Information Analyses for Robust Brain Computer Interface. | Tran Huy Dat, Cuntai Guan |
| 2006 | ICASSP | Multichannel Speech Enhancement Based on Speech Spectral Magnitude Estimation Using Generalized Gamma Prior Distribution. | Tran Huy Dat, Kazuya Takeda, Fumitada Itakura |
| 2005 | ICASSP | Generalized gamma modeling of speech and its online estimation for speech enhancement. | Tran Huy Dat, Kazuya Takeda, Fumitada Itakura |
| 2005 | ICASSP | SNR and Local Noise Power Estimations Based on Gaussian Mixture Modeling on the Log-Power Domain. | Kazuya Takeda, Tran Huy Dat, Hiroshi Fujimura, Fumitada Itakura |
| 2005 | ICDE | A speech enhancement system based on data clustering and cumulative histogram equalization. | Tran Huy Dat, Kazuya Takeda, Fumitada Itakura |
| 2004 | Interspeech | Speech enhancement based on magnitude estimation using the gamma prior. | Weifeng Li, Kazuya Takeda, Fumitada Itakura, Tran Huy Dat |