| 2019 | ASRU | A Cross-Corpus Study on Speech Emotion Recognition. | Rosanna Milner, Md Asif Jalal, Raymond W. M. Ng, Thomas Hain |
| 2019 | ICASSP | Teacher-student Training for Acoustic Event Detection Using Audioset. | Ruibo Shi, Raymond W. M. Ng, Pawel Swietojanski |
| 2018 | CoNLL | Multi-Modal Sequence Fusion via Recursive Attention for Emotion Recognition. | Rory Beard, Ritwik Das, Raymond W. M. Ng, P. G. Keerthana Gopalakrishnan, Luka Eerens, Pawel Swietojanski, Ondrej Miksik |
| 2017 | ASRU | Exploring the use of acoustic embeddings in neural machine translation. | Salil Deena, Raymond W. M. Ng, Pranava Swaroop Madhyastha, Lucia Specia, Thomas Hain |
| 2017 | ICASSP | Shefce: A Cantonese-English bilingual speech corpus for pronunciation assessment. | Raymond W. M. Ng, Alvin C. M. Kwan, Tan Lee, Thomas Hain |
| 2017 | Interspeech | Semi-Supervised Adaptation of RNNLMs by Fine-Tuning with Domain-Specific Auxiliary Features. | Salil Deena, Raymond W. M. Ng, Pranava Swaroop Madhyastha, Lucia Specia, Thomas Hain |
| 2016 | ICASSP | Groupwise learning for ASR k-best list reranking in spoken language translation. | Raymond W. M. Ng, Kashif Shah, Lucia Specia, Thomas Hain |
| 2016 | Interspeech | Automatic Genre and Show Identification of Broadcast Media. | Mortaza Doulaty, Oscar Saz, Raymond W. M. Ng, Thomas Hain |
| 2016 | Interspeech | webASR 2 - Improved Cloud Based Speech Technology. | Thomas Hain, Jeremy Christian, Oscar Saz, Salil Deena, Madina Hasan, Raymond W. M. Ng, Rosanna Milner, Mortaza Doulaty, Yulan Liu |
| 2016 | Interspeech | Combining Weak Tokenisers for Phonotactic Language Recognition in a Resource-Constrained Setting. | Raymond W. M. Ng, Bhusan Chettri, Thomas Hain |
| 2015 | ASRU | Latent Dirichlet Allocation based organisation of broadcast media archives for deep neural network adaptation. | Mortaza Doulaty, Oscar Saz, Raymond W. M. Ng, Thomas Hain |
| 2015 | ASRU | The 2015 sheffield system for longitudinal diarisation of broadcast media. | Rosanna Milner, Oscar Saz, Salil Deena, Mortaza Doulaty, Raymond W. M. Ng, Thomas Hain |
| 2015 | ASRU | The 2015 sheffield system for transcription of Multi-Genre Broadcast media. | Oscar Saz, Mortaza Doulaty, Salil Deena, Rosanna Milner, Raymond W. M. Ng, Madina Hasan, Yulan Liu, Thomas Hain |
| 2015 | EMNLP | Investigating Continuous Space Language Models for Machine Translation Quality Estimation. | Kashif Shah, Raymond W. M. Ng, Fethi Bougares, Lucia Specia |
| 2015 | ICASSP | Quality estimation for asr k-best list rescoring in spoken language translation. | Raymond W. M. Ng, Kashif Shah, Wilker Aziz, Lucia Specia, Thomas Hain |
| 2015 | Interspeech | A study on the stability and effectiveness of features in quality estimation for spoken language translation. | Raymond W. M. Ng, Kashif Shah, Lucia Specia, Thomas Hain |
| 2013 | ICASSP | Adaptation of lecture speech recognition system with machine translation output. | Raymond W. M. Ng, Thomas Hain, Trevor Cohn |
| 2012 | ICASSP | Syllable: A self-contained unit to model pronunciation variation. | Raymond W. M. Ng, Keikichi Hirose |
| 2012 | Interspeech | An alignment matching method to explore pseudosyllable properties across different corpora. | Raymond W. M. Ng, Thomas Hain, Keikichi Hirose |
| 2011 | ICASSP | Score fusion and calibration in multiple language detectors with large performance variation. | Raymond W. M. Ng, Cheung-Chi Leung, Tan Lee, Bin Ma, Haizhou Li |
| 2010 | ICASSP | Prosodic attribute model for spoken language identification. | Raymond W. M. Ng, Cheung-Chi Leung, Tan Lee, Bin Ma, Haizhou Li |
| 2010 | Interspeech | Towards long-range prosodic attribute modeling for language recognition. | Raymond W. M. Ng, Cheung-Chi Leung, Ville Hautamki, Tan Lee, Bin Ma, Haizhou Li |
| 2006 | Interspeech | Towards automatic parameter extraction of command-response model for Cantonese. | Raymond W. M. Ng, Tan Lee, Wentao Gu |