| 2025 | INLG | Effectiveness of Chain-of-Thought in Distilling Reasoning Capability from Large Language Models. | Cong-Thanh Do, Rama Sanand Doddipatla, Kate M. Knill |
| 2023 | ASRU | Towards a Unified End-to-End Language Understanding System for Speech and Text Inputs. | Mohan Li, Catalin Zorila, Cong-Thanh Do, Rama Doddipatla |
| 2023 | ICASSP | Cumulative Attention Based Streaming Transformer ASR with Internal Language Model Joint Training and Rescoring. | Mohan Li, Cong-Thanh Do, Rama Doddipatla |
| 2023 | Interspeech | Domain Adaptive Self-supervised Training of Automatic Speech Recognition. | Cong-Thanh Do, Rama Doddipatla, Mohan Li, Thomas Hain |
| 2022 | Interspeech | Multiple-hypothesis RNN-T Loss for Unsupervised Fine-tuning and Self-training of Neural Transducer. | Cong-Thanh Do, Mohan Li, Rama Doddipatla |
| 2021 | ICASSP | Multiple-Hypothesis CTC-Based Semi-Supervised Adaptation of End-to-End Speech Recognition. | Cong-Thanh Do, Rama Doddipatla, Thomas Hain |
| 2021 | ICASSP | Train Your Classifier First: Cascade Neural Networks Training from Upper Layers to Lower Layers. | Shucong Zhang, Cong-Thanh Do, Rama Doddipatla, Erfan Loweimi, Peter Bell, Steve Renals |
| 2020 | ICASSP | Learning Noise Invariant Features Through Transfer Learning For Robust End-to-End Speech Recognition. | Shucong Zhang, Cong-Thanh Do, Rama Doddipatla, Steve Renals |
| 2019 | ICASSP | Subband Temporal Envelope Features and Data Augmentation for End-to-end Recognition of Distant Conversational Speech. | Cong-Thanh Do |
| 2018 | Interspeech | Weighting Time-Frequency Representation of Speech Using Auditory Saliency for Automatic Speech Recognition. | Cong-Thanh Do, Yannis Stylianou |
| 2017 | Interspeech | Improved Automatic Speech Recognition Using Subband Temporal Envelope Features and Time-Delay Neural Network Denoising Autoencoder. | Cong-Thanh Do, Yannis Stylianou |
| 2014 | Interspeech | Objective evaluation of HMM-based speech synthesis system using kullback-leibler divergence. | Cong-Thanh Do, Marc Evrard, A. Leman, Christophe d'Alessandro, Albert Rilliard, J.-L. Crebouw |
| 2013 | Interspeech | Augmenting short-term cepstral features with long-term discriminative features for speaker verification of telephone data. | Cong-Thanh Do, Claude Barras, Viet Bac Le, Achintya Kumar Sarkar |
| 2012 | Interspeech | Cochlear implant-like processing of speech signal for speaker verification. | Cong-Thanh Do, Claude Barras |
| 2010 | Interspeech | Recognizing cochlear implant-like spectrally reduced speech with HMM-based ASR: experiments with MFCCs and PLP coefficients. | Cong-Thanh Do, Dominique Pastor, Gal Le Lan, Andr Goalic |