| 2025 | ASRU | Evaluation of LLMs in Speech is Often Flawed: Test Set Contamination in Large Language Models for Speech Recognition. | Yuan Tseng, Titouan Parcollet, Rogier C. van Dalen, Shucong Zhang, Sourav Bhattacharya |
| 2025 | ASRU | Benchmarking Rotary Position Embeddings for Automatic Speech Recognition. | Shucong Zhang, Titouan Parcollet, Rogier C. van Dalen, Sourav Bhattacharya |
| 2025 | ICASSP | Linear Time Complexity Conformers with SummaryMixing for Streaming Speech Recognition. | Titouan Parcollet, Rogier van Dalen, Shucong Zhang, Sourav Bhattacharya |
| 2025 | Interspeech | Robust Unsupervised Adaptation of a Speech Recogniser Using Entropy Minimisation and Speaker Codes. | Rogier C. van Dalen, Shucong Zhang, Titouan Parcollet, Sourav Bhattacharya |
| 2025 | Interspeech | Loquacious Set: 25, 000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use. | Titouan Parcollet, Yuan Tseng, Shucong Zhang, Rogier C. van Dalen |
| 2024 | Interspeech | SummaryMixing: A Linear-Complexity Alternative to Self-Attention for Speech Recognition and Understanding. | Titouan Parcollet, Rogier van Dalen, Shucong Zhang, Sourav Bhattacharya |
| 2024 | Interspeech | Linear-Complexity Self-Supervised Learning for Speech Processing. | Shucong Zhang, Titouan Parcollet, Rogier van Dalen, Sourav Bhattacharya |
| 2023 | Interspeech | On the (In)Efficiency of Acoustic Feature Extractors for Self-Supervised Speech Representation Learning. | Titouan Parcollet, Shucong Zhang, Rogier van Dalen, Alberto Gil C. P. Ramos, Sourav Bhattacharya |
| 2023 | Interspeech | Real-Time Personalised Speech Enhancement Transformers with Dynamic Cross-attended Speaker Representations. | Shucong Zhang, Malcolm Chadwick, Alberto Gil C. P. Ramos, Titouan Parcollet, Rogier van Dalen, Sourav Bhattacharya |
| 2022 | ICASSP | Transformer-Based Streaming ASR with Cumulative Attention. | Mohan Li, Shucong Zhang, Catalin Zorila, Rama Doddipatla |
| 2021 | ICASSP | Train Your Classifier First: Cascade Neural Networks Training from Upper Layers to Lower Layers. | Shucong Zhang, Cong-Thanh Do, Rama Doddipatla, Erfan Loweimi, Peter Bell, Steve Renals |
| 2021 | Interspeech | Stochastic Attention Head Removal: A Simple and Effective Method for Improving Transformer Based ASR Models. | Shucong Zhang, Erfan Loweimi, Peter Bell, Steve Renals |
| 2020 | ICASSP | Learning Noise Invariant Features Through Transfer Learning For Robust End-to-End Speech Recognition. | Shucong Zhang, Cong-Thanh Do, Rama Doddipatla, Steve Renals |
| 2019 | ICASSP | Windowed Attention Mechanisms for Speech Recognition. | Shucong Zhang, Erfan Loweimi, Peter Bell, Steve Renals |
| 2019 | Interspeech | Trainable Dynamic Subsampling for End-to-End Speech Recognition. | Shucong Zhang, Erfan Loweimi, Yumo Xu, Peter Bell, Steve Renals |