Skip to content

Mark J. F. Gales

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

274

Venues

12

Active years

1992–2025

Best venue rank

A*

Where they publish

Papers

274 indexed papers, newest first.

YearVenueTitleAuthors
2025ACLSkillAggregation: Reference-free LLM-Dependent Aggregation.Guangzhi Sun, Anmol Kagrecha, Potsawee Manakul, Philip C. Woodland, Mark J. F. Gales
2025COLINGFinetuning LLMs for Comparative Assessment Tasks.Vatsal Raina, Adian Liusie, Mark J. F. Gales
2025EMNLPUniversal Acoustic Adversarial Attacks for Flexible Control of Speech-LLMs.Rao Ma, Mengjie Qian, Vyas Raina, Mark J. F. Gales, Kate M. Knill
2025EMNLPUnlearning vs. Obfuscation: Are We Truly Removing Knowledge?Guangzhi Sun, Potsawee Manakul, Xiao Zhan, Mark J. F. Gales
2025InterspeechAssessment of L2 Oral Proficiency using Speech Large Language Models.Rao Ma, Mengjie Qian, Siyuan Tang, Stefano Bann, Kate M. Knill, Mark J. F. Gales
2025InterspeechTraining Articulatory Inversion Models for Interspeaker Consistency.Charles McGhee, Mark J. F. Gales, Kate M. Knill
2025InterspeechScaling and Prompting for Improved End-to-End Spoken Grammatical Error Correction.Mengjie Qian, Rao Ma, Stefano Bann, Kate M. Knill, Mark J. F. Gales
2025NAACLCross-Lingual Transfer Learning for Speech Translation.Rao Ma, Mengjie Qian, Yassir Fathullah, Siyuan Tang, Mark J. F. Gales, Kate M. Knill
2025UAIGeneralised Probabilistic Modelling and Improved Uncertainty Estimation in Comparative LLM-as-a-judge.Yassir Fathullah, Mark J. F. Gales
2024ACLTeacher-Student Training for Debiasing: General Permutation Debiasing for Large Language Models.Adian Liusie, Yassir Fathullah, Mark J. F. Gales
2024ACLAn Information-Theoretic Approach to Analyze NLP Classification Tasks.Luran Wang, Mark J. F. Gales, Vatsal Raina
2024COLINGIs It Possible to Modify Text to a Target Readability Level? An Initial Investigation Using Zero-Shot Large Language Models.Asma Farajidizaji, Vatsal Raina, Mark J. F. Gales
2024EACLWho Needs Decoders? Efficient Estimation of Sequence-Level Attributes with Proxies.Yassir Fathullah, Puria Radmard, Adian Liusie, Mark J. F. Gales
2024EACLLLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models.Adian Liusie, Potsawee Manakul, Mark J. F. Gales
2024EMNLPLLM Task Interference: An Initial Study on the Impact of Task-Switch in Conversational History.Akash Gupta, Ivaxi Sheth, Vyas Raina, Mark J. F. Gales, Mario Fritz
2024EMNLPEfficient LLM Comparative Assessment: A Product of Experts Framework for Pairwise Comparisons.Adian Liusie, Vatsal Raina, Yassir Fathullah, Mark J. F. Gales
2024EMNLPIs LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment.Vyas Raina, Adian Liusie, Mark J. F. Gales
2024EMNLPMuting Whisper: A Universal Acoustic Adversarial Attack on Speech Foundation Models.Vyas Raina, Rao Ma, Charles McGhee, Kate M. Knill, Mark J. F. Gales
2024ICASSPTowards End-to-End Spoken Grammatical Error Correction.Stefano Bann, Rao Ma, Mengjie Qian, Kate M. Knill, Mark J. F. Gales
2024InterspeechOn the Usefulness of Speaker Embeddings for Speaker Retrieval in the Wild: A Comparative Study of x-vector and ECAPA-TDNN Models.Erfan Loweimi, Mengjie Qian, Kate M. Knill, Mark J. F. Gales
2024InterspeechHighly Intelligible Speaker-Independent Articulatory Synthesis.Charles McGhee, Kate M. Knill, Mark J. F. Gales
2024InterspeechLearn and Don't Forget: Adding a New Language to ASR Foundation Models.Mengjie Qian, Siyuan Tang, Rao Ma, Kate M. Knill, Mark J. F. Gales
2024NAACLEfficient Sample-Specific Encoder Perturbations.Yassir Fathullah, Mark J. F. Gales
2024NAACLInvestigating the Emergent Audio Classification Ability of ASR Foundation Models.Rao Ma, Adian Liusie, Mark J. F. Gales, Kate M. Knill
2024NAACLWaterJudge: Quality-Detection Trade-off when Watermarking Large Language Models.Piotr Molenda, Adian Liusie, Mark J. F. Gales
2023EMNLPSelfCheckGPT: Zero-Resource Black-Box Hallucination Detection for Generative Large Language Models.Potsawee Manakul, Adian Liusie, Mark J. F. Gales
2023ICASSPEnsemble Prosody Prediction For Expressive Speech Synthesis.Tian Huey Teh, Vivian Hu, Devang S. Ram Mohan, Zack Hodari, Christopher G. R. Wallis, Toms Gmez Ibarrondo, Alexandra Torresquintero, James Leoni, Mark J. F. Gales, Simon King
2023IJCNLPMitigating Word Bias in Zero-shot Prompt-based Classifiers.Adian Liusie, Potsawee Manakul, Mark J. F. Gales
2023IJCNLPMQAG: Multiple-choice Question Answering and Generation for Assessing Information Consistency in Summarization.Potsawee Manakul, Adian Liusie, Mark J. F. Gales
2023IJCNLPMinimum Bayes' Risk Decoding for System Combination of Grammatical Error Correction Systems.Vyas Raina, Mark J. F. Gales
2023InterspeechMulti-Head State Space Model for Speech Recognition.Yassir Fathullah, Chunyang Wu, Yuan Shangguan, Junteng Jia, Wenhan Xiong, Jay Mahadeokar, Chunxi Liu, Yangyang Shi, Ozlem Kalinli, Mike Seltzer, Mark J. F. Gales
2023InterspeechN-best T5: Robust ASR Error Correction using Multiple Input Hypotheses and Constrained Decoding Space.Rao Ma, Mark J. F. Gales, Kate M. Knill, Mengjie Qian
2023InterspeechAdapting an Unadaptable ASR System.Rao Ma, Mengjie Qian, Mark J. F. Gales, Kate M. Knill
2023InterspeechSpeak & Improve: L2 English Speaking Practice Tool.Diane Nicholls, Kate M. Knill, Mark J. F. Gales, Anton Ragni, Paul Ricketts
2023UAILogit-based ensemble distribution distillation for robust autoregressive sequence uncertainties.Yassir Fathullah, Guoxuan Xia, Mark J. F. Gales
2022ACLAnswer Uncertainty and Unanswerability in Multiple-Choice Machine Reading Comprehension.Vatsal Raina, Mark J. F. Gales
2022IJCNLPAnalyzing Biases to Spurious Correlations in Text Classification Tasks.Adian Liusie, Vatsal Raina, Vyas Raina, Mark J. F. Gales
2022IJCNLPGrammatical Error Correction Systems for Automated Assessment: Are They Susceptible to Universal Adversarial Attacks?Vyas Raina, Yiting Lu, Mark J. F. Gales
2022InterspeechView-Specific Assessment of L2 Spoken English.Stefano Bann, Bhanu Balusu, Mark J. F. Gales, Kate M. Knill, Konstantinos Kyriakopoulos
2022NAACLResidue-Based Natural Language Adversarial Attack Detection.Vyas Raina, Mark J. F. Gales
2022UAISelf-distribution distillation: efficient uncertainty estimation.Yassir Fathullah, Mark J. F. Gales
2021ACLLong-Span Summarization via Local Attention and Content Selection.Potsawee Manakul, Mark J. F. Gales
2021EMNLPSparsity and Sentence Structure in Encoder-Decoder Attention of Summarization Systems.Potsawee Manakul, Mark J. F. Gales
2021ICASSPEnsemble Distillation Approaches for Grammatical Error Correction.Yassir Fathullah, Mark J. F. Gales, Andrey Malinin
2021ICASSPEfficient Use of End-to-End Data in Spoken Language Processing.Yiting Lu, Yu Wang, Mark J. F. Gales
2021ICASSPAnalysing Bias in Spoken Language Assessment Using Concept Activation Vectors.Xizi Wei, Mark J. F. Gales, Kate M. Knill
2021ICLRUncertainty Estimation in Autoregressive Structured Prediction.Andrey Malinin, Mark J. F. Gales
2021InterspeechDeliberation-Based Multi-Pass Speech Synthesis.Qingyun Dou, Xixin Wu, Moquan Wan, Yiting Lu, Mark J. F. Gales
2020ICASSPConfidence Estimation for Black Box Automatic Speech Recognition Systems Using Lattice Recurrent Neural Networks.Alexandros Kastanos, Anton Ragni, Mark J. F. Gales
2020ICLREnsemble Distribution Distillation.Andrey Malinin, Bruno Mlodozeniec, Mark J. F. Gales
2020InterspeechAttention Forcing for Speech Synthesis.Qingyun Dou, Joshua Efiong, Mark J. F. Gales
2020InterspeechNon-Native Children's Automatic Speech Recognition: The INTERSPEECH 2020 Shared Task ALTA Systems.Kate M. Knill, Linlin Wang, Yu Wang, Xixin Wu, Mark J. F. Gales
2020InterspeechAutomatic Detection of Accent and Lexical Pronunciation Errors in Spontaneous Non-Native English Speech.Konstantinos Kyriakopoulos, Kate M. Knill, Mark J. F. Gales
2020InterspeechSpoken Language 'Grammatical Error Correction'.Yiting Lu, Mark J. F. Gales, Yu Wang
2020InterspeechAbstractive Spoken Document Summarization Using Hierarchical Model with Multi-Stage Attention Diversity Optimization.Potsawee Manakul, Mark J. F. Gales, Linlin Wang
2020InterspeechUniversal Adversarial Attacks on Spoken Language Assessment Systems.Vyas Raina, Mark J. F. Gales, Kate M. Knill
2020InterspeechEnsemble Approaches for Uncertainty in Spoken Language Assessment.Xixin Wu, Kate M. Knill, Mark J. F. Gales, Andrey Malinin
2019ASRULearning Between Different Teacher and Student Models in ASR.Jeremy Heng Meng Wong, Mark J. F. Gales, Yu Wang
2019ICASSPAutomatic Grammatical Error Detection of Non-native Spoken Learner English.Kate M. Knill, Mark J. F. Gales, P. P. Manakul, Andrew Caines
2019ICASSPBi-directional Lattice Recurrent Neural Networks for Confidence Estimation.Qiujia Li, Preben Ness, Anton Ragni, Mark J. F. Gales
2019InterspeechA Deep Learning Approach to Automatic Characterisation of Rhythm in Non-Native English Speech.Konstantinos Kyriakopoulos, Kate M. Knill, Mark J. F. Gales
2019InterspeechImpact of ASR Performance on Spoken Grammatical Error Detection.Yiting Lu, Mark J. F. Gales, Kate M. Knill, P. P. Manakul, Linlin Wang, Yu Wang
2018ICASSPPhonetic and Graphemic Systems for Multi-Genre Broadcast Transcription.Yu Wang, Xie Chen, Mark J. F. Gales, Anton Ragni, Jeremy Heng Meng Wong
2018InterspeechActive Memory Networks for Language Modeling.Oscar Chen, Anton Ragni, Mark J. F. Gales, Xie Chen
2018InterspeechImpact of ASR Performance on Free Speaking Language Assessment.Kate M. Knill, Mark J. F. Gales, Konstantinos Kyriakopoulos, Andrey Malinin, Anton Ragni, Yu Wang, Andrew Caines
2018InterspeechA Deep Learning Approach to Assessing Non-native Pronunciation of English Using Phone Distances.Konstantinos Kyriakopoulos, Kate M. Knill, Mark J. F. Gales
2018InterspeechAutomatic Speech Recognition System Development in the "Wild".Anton Ragni, Mark J. F. Gales
2018InterspeechWaveform-Based Speaker Representations for Speech Synthesis.Moquan Wan, Gilles Degottex, Mark J. F. Gales
2018InterspeechSpeaker Adaptation and Adaptive Training for Jointly Optimised Tandem Systems.Yu Wang, Chao Zhang, Mark J. F. Gales, Philip C. Woodland
2017ACLIncorporating Uncertainty into Deep Learning for Spoken Language Assessment.Andrey Malinin, Anton Ragni, Kate M. Knill, Mark J. F. Gales
2017ASRUFuture word contexts in neural network language models.Xie Chen, X. Liu, Anton Ragni, Y. Wang, Mark J. F. Gales
2017ASRUA hierarchical attention based model for off-topic spontaneous spoken response detection.Andrey Malinin, Kate M. Knill, Mark J. F. Gales
2017ASRUIntegrated speaker-adaptive speech synthesis.Moquan Wan, Gilles Degottex, Mark J. F. Gales
2017ASRUMulti-task ensembles with teacher-student training.Jeremy Heng Meng Wong, Mark J. F. Gales
2017ICASSPRecurrent neural network language models for keyword search.Xie Chen, Anton Ragni, J. Vasilakes, Xunying Liu, Kate M. Knill, Mark J. F. Gales
2017ICASSPMorph-to-word transduction for accurate and efficient automatic speech recognition and keyword search.Anton Ragni, Danielle Saunders, P. Zahemszky, J. Vasilakes, Mark J. F. Gales, Kate M. Knill
2017ICASSPStimulated training for automatic speech recognition and keyword search in limited resource conditions.Anton Ragni, Chunyang Wu, Mark J. F. Gales, J. Vasilakes, Kate M. Knill
2017InterspeechInvestigating Bidirectional Recurrent Neural Network Language Models for Speech Recognition.Xie Chen, Anton Ragni, Xunying Liu, Mark J. F. Gales
2017InterspeechUse of Graphemic Lexicons for Spoken Language Assessment.Kate M. Knill, Mark J. F. Gales, Konstantinos Kyriakopoulos, Anton Ragni, Yu Wang
2017InterspeechStudent-Teacher Training with Diverse Decision Tree Ensembles.Jeremy Heng Meng Wong, Mark J. F. Gales
2017InterspeechDeep Activation Mixture Model for Speech Recognition.Chunyang Wu, Mark J. F. Gales
2016ACLOff-topic Response Detection for Spontaneous Spoken English Assessment.Andrey Malinin, Rogier C. van Dalen, Kate M. Knill, Yu Wang, Mark J. F. Gales
2016ICASSPCUED-RNNLM - An open-source toolkit for efficient training and evaluation of recurrent neural network language models.Xie Chen, Xunying Liu, Y. Qian, Mark J. F. Gales, Philip C. Woodland
2016ICASSPImproved DNN-based segmentation for multi-genre broadcast audio.Linlin Wang, Chao Zhang, Philip C. Woodland, Mark J. F. Gales, Panagiota Karanasou, Pierre Lanchantin, Xunying Liu, Yanmin Qian
2016ICASSPCombining i-vector representation and structured neural networks for rapid adaptation.Chunyang Wu, Penny Karanasou, Mark J. F. Gales
2016ICASSPSystem combination with log-linear models.J. Yang, Chao Zhang, Anton Ragni, Mark J. F. Gales, Philip C. Woodland
2016InterspeechIncorporating a Generative Front-End Layer to Deep Neural Network for Noise Robust Automatic Speech Recognition.Souvik Kundu, Khe Chai Sim, Mark J. F. Gales
2016InterspeechSelection of Multi-Genre Broadcast Data for the Training of Automatic Speech Recognition Systems.Pierre Lanchantin, Mark J. F. Gales, Penny Karanasou, Xunying Liu, Yanman Qian, Linlin Wang, Philip C. Woodland, Chao Zhang
2016InterspeechMulti-Language Neural Network Language Models.Anton Ragni, Edgar Dakin, Xie Chen, Mark J. F. Gales, Kate M. Knill
2016InterspeechSequence Student-Teacher Training of Deep Neural Networks.Jeremy Heng Meng Wong, Mark J. F. Gales
2016InterspeechStimulated Deep Neural Network for Speech Recognition.Chunyang Wu, Penny Karanasou, Mark J. F. Gales, Khe Chai Sim
2016InterspeechLog-Linear System Combination Using Structured Support Vector Machines.Jingzhou Yang, Anton Ragni, Mark J. F. Gales, Kate M. Knill
2016SIGdialTowards Using Conversations with Spoken Dialogue Systems in the Automated Assessment of Non-Native Speakers of English.Diane J. Litman, Steve J. Young, Mark J. F. Gales, Kate M. Knill, Karen Ottewell, Rogier C. van Dalen, David Vandyke
2015ASRUThe MGB challenge: Evaluating multi-genre broadcast media recognition.Peter Bell, Mark J. F. Gales, Thomas Hain, Jonathan Kilgour, Pierre Lanchantin, Xunying Liu, Andrew McParland, Steve Renals, Oscar Saz, Mirjam Wester, Philip C. Woodland
2015ASRUInvestigation of back-off based interpolation between recurrent neural network and n-gram language models.Xie Chen, Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2015ASRUMultilingual representations for low resource speech recognition and keyword search.Jia Cui, Brian Kingsbury, Bhuvana Ramabhadran, Abhinav Sethy, Kartik Audhkhasi, Xiaodong Cui, Ellen Kislal, Lidia Mangu, Markus Nubaum-Thom, Michael Picheny, Zoltn Tske, Pavel Golik, Ralf Schlter, Hermann Ney, Mark J. F. Gales, Kate M. Knill, Anton Ragni, Haipeng Wang, Philip C. Woodland
2015ASRUStructured discriminative models using deep neural-network features.Rogier C. van Dalen, Jingzhou Yang, Haipeng Wang, Anton Ragni, Chao Zhang, Mark J. F. Gales
2015ASRUSpeaker diarisation and longitudinal linking in multi-genre broadcast data.Penny Karanasou, Mark J. F. Gales, Pierre Lanchantin, Xunying Liu, Yanmin Qian, Linlin Wang, Philip C. Woodland, Chao Zhang
2015ASRUThe development of the cambridge university alignment systems for the multi-genre broadcast challenge.Pierre Lanchantin, Mark J. F. Gales, Penny Karanasou, Xunying Liu, Yanmin Qian, Linlin Wang, Philip C. Woodland, Chao Zhang
2015ASRUImproving the interpretability of deep neural networks with stimulated learning.Shawn Tan, Khe Chai Sim, Mark J. F. Gales
2015ASRUCambridge university transcription systems for the multi-genre broadcast challenge.Philip C. Woodland, Xunying Liu, Yanmin Qian, Chao Zhang, Mark J. F. Gales, Penny Karanasou, Pierre Lanchantin, Linlin Wang
2015ICASSPImproving the training and evaluation efficiency of recurrent neural network language models.Xie Chen, Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2015ICASSPRecurrent neural network language model training with noise contrastive estimation for speech recognition.Xie Chen, Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2015ICASSPImproving multiple-crowd-sourced transcriptions using a speech recogniser.Rogier C. van Dalen, Kate M. Knill, Pirros Tsiakoulis, Mark J. F. Gales
2015ICASSPRobust excitation-based features for Automatic Speech Recognition.Thomas Drugman, Yannis Stylianou, Langzhou Chen, Xie Chen, Mark J. F. Gales
2015ICASSPUnicode-based graphemic systems for limited resource languages.Mark J. F. Gales, Kate M. Knill, Anton Ragni
2015ICASSPParaphrastic recurrent neural network language models.Xunying Liu, Xie Chen, Mark J. F. Gales, Philip C. Woodland
2015ICASSPA language space representation for speech recognition.Anton Ragni, Mark J. F. Gales, Kate M. Knill
2015ICASSPMulti-basis adaptive neural network for rapid adaptation in speech recognition.Chunyang Wu, Mark J. F. Gales
2015InterspeechRecurrent neural network language model adaptation for multi-genre broadcast speech recognition.Xie Chen, Tian Tan, Xunying Liu, Pierre Lanchantin, M. Wan, Mark J. F. Gales, Philip C. Woodland
2015InterspeechAnnotating large lattices with the exact word error.Rogier C. van Dalen, Mark J. F. Gales
2015InterspeechI-vector estimation using informative priors for adaptation of deep neural networks.Penny Karanasou, Mark J. F. Gales, Philip C. Woodland
2015InterspeechReconstructing voices within the multiple-average-voice-model framework.Pierre Lanchantin, Christophe Veaux, Mark J. F. Gales, Simon King, Junichi Yamagishi
2015InterspeechThe Cambridge University 2014 BOLT conversational telephone Mandarin Chinese LVCSR system for speech translation.Xunying Liu, Federico Flego, Linlin Wang, Chao Zhang, Mark J. F. Gales, Philip C. Woodland
2015InterspeechImproving speech recognition and keyword search for low resource languages using web data.Gideon Mendels, Erica Cooper, Victor Soto, Julia Hirschberg, Mark J. F. Gales, Kate M. Knill, Anton Ragni, Haipeng Wang
2015InterspeechJoint decoding of tandem and hybrid systems for improved keyword spotting on low resource languages.Haipeng Wang, Anton Ragni, Mark J. F. Gales, Kate M. Knill, Philip C. Woodland, Chao Zhang
2014ICASSPSpeaker dependent expression predictor from text: Expressiveness and transplantation.Langzhou Chen, Norbert Braunschweiler, Mark J. F. Gales
2014ICASSPMultiple-average-voice-based speech synthesis.Pierre Lanchantin, Mark J. F. Gales, Simon King, Junichi Yamagishi
2014ICASSPParaphrastic neural network language models.Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2014ICASSPEfficient lattice rescoring using recurrent neural network language models.Xunying Liu, Yongqiang Wang, Xie Chen, Mark J. F. Gales, Philip C. Woodland
2014ICASSPCluster adaptive training of average voice models.Vincent Wan, Javier Latorre, Kayoko Yanagisawa, Mark J. F. Gales, Yannis Stylianou
2014ICASSPInfinite structured support vector machines for speech recognition.Jingzhou Yang, Rogier C. van Dalen, Shi-Xiong Zhang, Mark J. F. Gales
2014ICASSPImpact of single-microphone dereverberation on DNN-based meeting transcription systems.Takuya Yoshioka, Xie Chen, Mark J. F. Gales
2014ICASSPInvestigation of unsupervised adaptation of DNN acoustic models with filter bank input.Takuya Yoshioka, Anton Ragni, Mark J. F. Gales
2014InterspeechAn initial investigation of long-term adaptation for meeting transcription.Xie Chen, Mark J. F. Gales, Kate M. Knill, Catherine Breslin, Langzhou Chen, K. K. Chin, Vincent Wan
2014InterspeechEfficient GPU-based training of recurrent neural network language models using spliced sentence bunch.Xie Chen, Yongqiang Wang, Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2014InterspeechAdaptation of deep neural network acoustic models using factorised i-vectors.Penny Karanasou, Yongqiang Wang, Mark J. F. Gales, Philip C. Woodland
2014InterspeechLanguage independent and unsupervised acoustic models for speech recognition and keyword spotting.Kate M. Knill, Mark J. F. Gales, Anton Ragni, Shakti P. Rath
2014InterspeechGenerating multiple-accent pronunciations for TTS using joint sequence model interpolation.BalaKrishna Kolluru, Vincent Wan, Javier Latorre, Kayoko Yanagisawa, Mark J. F. Gales
2014InterspeechSpeech intonation for TTS: study on evaluation methodology.Javier Latorre, Kayoko Yanagisawa, Vincent Wan, BalaKrishna Kolluru, Mark J. F. Gales
2014InterspeechData augmentation for low resource languages.Anton Ragni, Kate M. Knill, Shakti P. Rath, Mark J. F. Gales
2014InterspeechCombining tandem and hybrid systems for improved speech recognition and keyword spotting on low resource languages.Shakti P. Rath, Kate M. Knill, Anton Ragni, Mark J. F. Gales
2014InterspeechNoise-robust TTS speaker adaptation with statistics smoothing.Kayoko Yanagisawa, Langzhou Chen, Mark J. F. Gales
2013ASRUInvestigation of multilingual deep neural networks for spoken term detection.Kate M. Knill, Mark J. F. Gales, Shakti P. Rath, Philip C. Woodland, Chao Zhang, Shi-Xiong Zhang
2013ICASSPIntegrated automatic expression prediction and speech synthesis from text.Langzhou Chen, Mark J. F. Gales, Norbert Braunschweiler, Masami Akamine, Kate M. Knill
2013ICASSPEfficient decoding with generative score-spaces using the expectation semiring.Rogier C. van Dalen, Anton Ragni, Mark J. F. Gales
2013ICASSPA high-performance Cantonese keyword search system.Brian Kingsbury, Jia Cui, Xiaodong Cui, Mark J. F. Gales, Kate M. Knill, Jonathan Mamou, Lidia Mangu, David Nolden, Michael Picheny, Bhuvana Ramabhadran, Ralf Schlter, Abhinav Sethy, Philip C. Woodland
2013ICASSPTraining a supra-segmental parametric F0 model without interpolating F0.Javier Latorre, Mark J. F. Gales, Kate M. Knill, Masami Akamine
2013ICASSPParaphrastic language models and combination with neural network language models.Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2013ICASSPComplex cepstrum analysis based on the minimum mean squared error.Ranniery Maia, Masami Akamine, Mark J. F. Gales
2013ICASSPSystem combination and score normalization for spoken term detection.Jonathan Mamou, Jia Cui, Xiaodong Cui, Mark J. F. Gales, Brian Kingsbury, Kate M. Knill, Lidia Mangu, David Nolden, Michael Picheny, Bhuvana Ramabhadran, Ralf Schlter, Abhinav Sethy, Philip C. Woodland
2013ICASSPA confidence-based approach for improving keyword hypothesis scores.Matthew Stephen Seigel, Philip C. Woodland, Mark J. F. Gales
2013ICASSPTandem system adaptation using multiple linear feature transforms.Yongqiang Wang, Mark J. F. Gales
2013ICASSPKernelized log linear models for continuous speech recognition.Shi-Xiong Zhang, Mark J. F. Gales
2013InterspeechAutomatic Transcription of Multi-genre Media Archives.Pierre Lanchantin, Peter Bell, Mark J. F. Gales, Thomas Hain, Xunying Liu, Yanhua Long, Jennifer Quinnell, Steve Renals, Oscar Saz, Matthew Stephen Seigel, Pawel Swietojanski, Philip C. Woodland
2013InterspeechCross-domain paraphrasing for improving language modelling using out-of-domain data.Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2013InterspeechImproving lightly supervised training for broadcast transcription.Yanhua Long, Mark J. F. Gales, Pierre Lanchantin, Xunying Liu, Matthew Stephen Seigel, Philip C. Woodland
2013InterspeechMinimum mean squared error based warped complex cepstrum analysis for statistical parametric speech synthesis.Ranniery Maia, Mark J. F. Gales, Yannis Stylianou, Masami Akamine
2013InterspeechPhoto-realistic expressive text to talking head synthesis.Vincent Wan, Robert Anderson, Art Blokland, Norbert Braunschweiler, Langzhou Chen, BalaKrishna Kolluru, Javier Latorre, Ranniery Maia, Bjrn Stenger, Kayoko Yanagisawa, Yannis Stylianou, Masami Akamine, Mark J. F. Gales, Roberto Cipolla
2013InterspeechAn explicit independence constraint for factorised adaptation in speech recognition.Yongqiang Wang, Mark J. F. Gales
2013InterspeechInfinite support vector machines in speech recognition.Jingzhou Yang, Rogier C. van Dalen, Mark J. F. Gales
2012ICASSPUnsupervised clustering of emotion and voice styles for expressive TTS.Florian Eyben, Sabine Buchholz, Norbert Braunschweiler, Javier Latorre, Vincent Wan, Mark J. F. Gales, Kate M. Knill
2012ICASSPFactor analysis based VTS discriminative adaptive training.Federico Flego, Mark J. F. Gales
2012ICASSPComplex cepstrum as phase information in statistical parametric speech synthesis.Ranniery Maia, Masami Akamine, Mark J. F. Gales
2012ICASSPInference algorithms for generative score-spaces.Anton Ragni, Mark J. F. Gales
2012InterspeechExploring Rich Expressive Information from Audiobook Data Using Cluster Adaptive Training.Langzhou Chen, Mark J. F. Gales, Vincent Wan, Javier Latorre, Masami Akamine
2012InterspeechModel-Based Approaches for Degraded Channel Modelling in Robust ASR.Mark J. F. Gales, Federico Flego
2012InterspeechSpeech factorization for HMM-TTS based on cluster adaptive training.Javier Latorre, Vincent Wan, Mark J. F. Gales, Langzhou Chen, K. K. Chin, Kate M. Knill, Masami Akamine
2012InterspeechParaphrastic Language Models.Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2012InterspeechRapid Nonlinear Speaker Adaptation for Large-Vocabulary Continuous Speech Recognition.Zoi Roupakia, Anton Ragni, Mark J. F. Gales
2012InterspeechModel-based approaches to adaptive training in reverberant environments.Yongqiang Wang, Mark J. F. Gales
2012InterspeechCombining multiple high quality corpora for improving HMM-TTS.Vincent Wan, Javier Latorre, K. K. Chin, Langzhou Chen, Mark J. F. Gales, Heiga Zen, Kate M. Knill, Masami Akamine
2011ASRUA variational perspective on noise-robust speech recognition.Rogier C. van Dalen, Mark J. F. Gales
2011ASRUDerivative kernels for noise robust ASR.Anton Ragni, Mark J. F. Gales
2011ASRUImproving reverberant VTS for hands-free robust speech recognition.Yongqiang Wang, Mark J. F. Gales
2011ASRUExtending noise robust structured support vector machines to larger vocabulary tasks.Shi-Xiong Zhang, Mark J. F. Gales
2011ICASSPConstrained discriminative mapping transforms for unsupervised speaker adaptation.Langzhou Chen, Mark J. F. Gales, K. K. Chin
2011ICASSPRapid joint speaker and noise compensation for robust speech recognition.K. K. Chin, Haitian Xu, Mark J. F. Gales, Catherine Breslin, Kate M. Knill
2011ICASSPContinuous F0 in the source-excitation generation for HMM-based TTS: Do we need voiced/unvoiced classification?Javier Latorre, Mark J. F. Gales, Sabine Buchholz, Kate M. Knill, Masatsune Tamura, Yamato Ohtani, Masami Akamine
2011ICASSPSpeaker and noise factorisation on the AURORA4 task.Yongqiang Wang, Mark J. F. Gales
2011ICASSPDecision tree-based context clustering based on cross validation and hierarchical priors.Heiga Zen, Mark J. F. Gales
2011InterspeechIntegrated Online Speaker Clustering and Adaptation.Catherine Breslin, K. K. Chin, Mark J. F. Gales, Kate M. Knill
2011InterspeechImproving LVCSR System Combination Using Neural Network Language Model Cross Adaptation.Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2011InterspeechGraphone Model Interpolation and Arabic Pronunciation Generation.T. Li, Philip C. Woodland, Frank Diehl, Mark J. F. Gales
2011InterspeechMultipulse Sequences for Residual Signal Modeling.Ranniery Maia, Heiga Zen, Kate M. Knill, Mark J. F. Gales, Sabine Buchholz
2011InterspeechGaussian Process Experts for Voice Conversion.Nicholas Pilkington, Heiga Zen, Mark J. F. Gales
2011InterspeechStructured Support Vector Machines for Noise Robust Continuous Speech Recognition.Shi-Xiong Zhang, Mark J. F. Gales
2010ICASSPLanguage model combination and adaptation usingweighted finite state transducers.Xunying Liu, Mark J. F. Gales, Jim L. Hieronymus, Philip C. Woodland
2010ICASSPRecent improvements to the Cambridge Arabic Speech-to-Text systems.Marcus Tomalin, Frank Diehl, Mark J. F. Gales, Junho Park, Philip C. Woodland
2010ICASSPStatistical parametric speech synthesis based on product of experts.Heiga Zen, Mark J. F. Gales, Yoshihiko Nankaku, Keiichi Tokuda
2010InterspeechLightly supervised recognition for automatic alignment of large coherent speech recordings.Norbert Braunschweiler, Mark J. F. Gales, Sabine Buchholz
2010InterspeechPrior information for rapid speaker adaptation.Catherine Breslin, K. K. Chin, Mark J. F. Gales, Kate M. Knill, Haitian Xu
2010InterspeechAsymptotically exact noise-corrupted speech likelihoods.Rogier C. van Dalen, Mark J. F. Gales
2010InterspeechCanonical state models for automatic speech recognition.Mark J. F. Gales, Kai Yu
2010InterspeechTraining a parametric-based logF0 model with the minimum generation error criterion.Javier Latorre, Mark J. F. Gales, Heiga Zen
2010InterspeechLanguage model cross adaptation for LVCSR system combination.Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2010InterspeechImproved neural network based language modelling and adaptation.Junho Park, Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2009ASRUDiscriminative adaptive training with VTS and JUD.Federico Flego, Mark J. F. Gales
2009ASRUAcoustic modelling for speech recognition: Hidden Markov models and beyond?Mark J. F. Gales
2009ASRUSupport vector machines for noise robust ASR.Mark J. F. Gales, Anton Ragni, H. AlDamarki, C. Gautier
2009ASRUImproving joint uncertainty decoding performance by predictive methods for noise robust speech recognition.Haitian Xu, Mark J. F. Gales, K. K. Chin
2009ICASSPExtended VTS for noise-robust speech recognition.Rogier C. van Dalen, Mark J. F. Gales
2009ICASSPIncremental predictive and adaptive noise compensation.Federico Flego, Mark J. F. Gales
2009ICASSPCombining VTS model compensation and support vector machines.Mark J. F. Gales, Federico Flego
2009ICASSPTraining and adapting MLP features for Arabic speech recognition.Junho Park, Frank Diehl, Mark J. F. Gales, Marcus Tomalin, Philip C. Woodland
2009ICASSPBayesian discriminative adaptation for speech recognition.Chandra Kant Raut, Mark J. F. Gales
2009InterspeechTransforming features to compensate speech recogniser models for noise.Rogier C. van Dalen, Federico Flego, Mark J. F. Gales
2009InterspeechMorphological analysis and decomposition for Arabic speech-to-text systems.Frank Diehl, Mark J. F. Gales, Marcus Tomalin, Philip C. Woodland
2009InterspeechIncremental adaptation with VTS and joint adaptively trained systems.Federico Flego, Mark J. F. Gales
2009InterspeechExploiting Chinese character models to improve speech recognition performance.Jim L. Hieronymus, Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2009InterspeechAdaptive training with noisy constrained maximum likelihood linear regression for noise robust speech recognition.D. K. Kim, Mark J. F. Gales
2009InterspeechUse of contexts in language model interpolation and adaptation.Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2009InterspeechVariational dynamic kernels for speaker verification.Chris Longworth, Rogier C. van Dalen, Mark J. F. Gales
2009InterspeechEfficient generation and use of MLP features for Arabic speech recognition.Junho Park, Frank Diehl, Mark J. F. Gales, Marcus Tomalin, Philip C. Woodland
2008ICASSPPhonetic pronunciations for arabic speech-to-text systems.Frank Diehl, Mark J. F. Gales, Marcus Tomalin, Philip C. Woodland
2008ICASSPMultiple kernel learning for speaker verification.Chris Longworth, Mark J. F. Gales
2008ICASSPUnsupervised discriminative adaptation using discriminative mapping transforms.Kai Yu, Mark J. F. Gales, Philip C. Woodland
2008InterspeechCovariance modelling for noise-robust speech recognition.Rogier C. van Dalen, Mark J. F. Gales
2008InterspeechDiscriminative classifiers with generative kernels for noise robust ASR.Mark J. F. Gales, Chris Longworth
2008InterspeechContext dependent language model adaptation.Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2008InterspeechA generalised derivative kernel for speaker verification.Chris Longworth, Mark J. F. Gales
2008InterspeechAdaptive training using discriminative mapping transforms.Chandra Kant Raut, Kai Yu, Mark J. F. Gales
2007ASRUPredictive linear transforms for noise robust speech recognition.Mark J. F. Gales, Rogier C. van Dalen
2007ASRUDevelopment of a phonetic system for large vocabulary Arabic speech recognition.Mark J. F. Gales, Frank Diehl, Chandra Kant Raut, Marcus Tomalin, Philip C. Woodland, Kai Yu
2007ASRUDiscriminative language model adaptation for Mandarin broadcast speech transcription and translation.Xunying Liu, William J. Byrne, Mark J. F. Gales, Adri de Gispert, Marcus Tomalin, Philip C. Woodland, Kai Yu
2007ICASSPComplementary System Generation using Directed Decision Trees.Catherine Breslin, Mark J. F. Gales
2007ICASSPSpeech Recognition System Combination for Machine Translation.Mark J. F. Gales, Xunying Liu, Rohit Sinha, Philip C. Woodland, Kai Yu, Spyros Matsoukas, Tim Ng, Kham Nguyen, Long Nguyen, Jean-Luc Gauvain, Lori Lamel, Abdelkhalek Messaoudi
2007ICASSPAdaptive Training with Joint Uncertainty Decoding for Robust Recognition of Noisy Data.Hank Liao, Mark J. F. Gales
2007ICASSPConsensus Network Decoding for Statistical Machine Translation System Combination.Khe Chai Sim, William J. Byrne, Mark J. F. Gales, Hichem Sahbi, Philip C. Woodland
2007ICASSPImproving Speech Transcription for Mandarin-English Translation.Marcus Tomalin, Mark J. F. Gales, X. Andrew Liu, Khe Chai Sim, Rohit Sinha, Lan Wang, Philip C. Woodland, Kai Yu
2007ICASSPUnsupervised Training for Mandarin Broadcast News and Conversation Transcription.Lan Wang, Mark J. F. Gales, Philip C. Woodland
2007InterspeechBuilding multiple complementary systems using directed decision trees.Catherine Breslin, Mark J. F. Gales
2007InterspeechDerivative and parametric kernels for speaker verification.Chris Longworth, Mark J. F. Gales
2007InterspeechUnsupervised training with directed manual transcription for recognising Mandarin broadcast audio.Kai Yu, Mark J. F. Gales, Philip C. Woodland
2006ICASSPAugmented Statistical Models for Speech Recognition.Martin I. Layton, Mark J. F. Gales
2006ICASSPThe Cu-Htk Mandarin Broadcast News Transcription System.Rohit Sinha, Mark J. F. Gales, Do Yeong Kim, X. Andrew Liu, Khe Chai Sim, Philip C. Woodland
2006ICASSPIncremental Adaptation using Bayesian Inference.Kai Yu, Mark J. F. Gales
2006InterspeechGenerating complementary systems for speech recognition.Catherine Breslin, Mark J. F. Gales
2006InterspeechIssues with uncertainty decoding for noise robust speech recognition.Hank Liao, Mark J. F. Gales
2006InterspeechDiscriminative adaptation for speaker verification.Chris Longworth, Mark J. F. Gales
2005ICASSPTraining LVCSR Systems on Thousands of Hours of Data.Gunnar Evermann, Ho Yin Chan, Mark J. F. Gales, Bin Jia, David Mrva, Philip C. Woodland, Kai Yu
2005ICASSPDevelopment of the CUHTK 2004 Mandarin Conversational Telephone Speech Transcription System.Mark J. F. Gales, Bin Jia, X. Andrew Liu, Khe Chai Sim, Philip C. Woodland, Kai Yu
2005ICASSPDevelopment of the CU-HTK 2004 Broadcast News Transcription Systems.Do Yeong Kim, Ho Yin Chan, Gunnar Evermann, Mark J. F. Gales, David Mrva, Khe Chai Sim, Philip C. Woodland
2005ICASSPInvestigation of Acoustic Modeling Techniques for LVCSR Systems.Xunying Liu, Mark J. F. Gales, Khe Chai Sim, Kai Yu
2005ICASSPAdaptation of Precision Matrix Models on Large Vocabulary Continuous Speech Recognition.Khe Chai Sim, Mark J. F. Gales
2005InterspeechJoint uncertainty decoding for noise robust speech recognition.Hank Liao, Mark J. F. Gales
2005InterspeechTemporally varying model parameters for large vocabulary continuous speech recognition.Khe Chai Sim, Mark J. F. Gales
2005InterspeechThe Cambridge University March 2005 speaker diarisation system.Rohit Sinha, S. E. Tranter, Mark J. F. Gales, Philip C. Woodland
2004ICASSPDevelopment of the 2003 CU-HTK conversational telephone speech transcription system.Gunnar Evermann, Ho Yin Chan, Mark J. F. Gales, Thomas Hain, Xunying Liu, David Mrva, Lan Wang, Philip C. Woodland
2004ICASSPModel complexity control and compression using discriminative growth functions.Xunying Liu, Mark J. F. Gales
2004ICASSPRao-Blackwellised Gibbs sampling for switching linear dynamical systems.Antti-Veikko I. Rosti, Mark J. F. Gales
2004ICASSPBasis superposition precision matrix modelling for large vocabulary continuous speech recognition.Khe Chai Sim, Mark J. F. Gales
2004ICASSPAdaptive training using structured transforms.Kai Yu, Mark J. F. Gales
2004InterspeechUsing VTLN for broadcast news transcription.Do Yeong Kim, Srinivasan Umesh, Mark J. F. Gales, Thomas Hain, Philip C. Woodland
2003ICASSPProduct of Gaussians and multiple stream systems.S. S. Airey, Mark J. F. Gales
2003ICASSPPorting: SwitchBoard to the VoiceMail task.Mark J. F. Gales, Yuan Dong, Daniel Povey, Philip C. Woodland
2003ICASSPAutomatic complexity control for HLDA systems.Xunying Liu, Mark J. F. Gales, Philip C. Woodland
2003ICASSPDiscriminative map for acoustic model adaptation.Daniel Povey, Philip C. Woodland, Mark J. F. Gales
2003InterspeechProduct of Gaussians as a distributed representation for speech recognition.S. S. Airey, Mark J. F. Gales
2003InterspeechMMI-MAP and MPE-MAP for acoustic model adaptation.Daniel Povey, Mark J. F. Gales, Do Yeong Kim, Philip C. Woodland
2002ICASSPImproved cross-task recognition using MMIE training.Ricardo de Crdoba, Philip C. Woodland, Mark J. F. Gales
2002ICASSPThe HMM error model.Mark J. F. Gales
2002ICASSPFactor analysed hidden Markov models.Antti-Veikko I. Rosti, Mark J. F. Gales
2002ICASSPUsing SVMS and discriminative models for speech recognition.Nathan D. Smith, Mark J. F. Gales
2002InterspeechCombining a Gaussian mixture model front end with MFCC parameters.Matthew N. Stuttle, Mark J. F. Gales
2001ICASSPMultiple-cluster adaptive training schemes.Mark J. F. Gales
2001InterspeechA mixture of Gaussians front end for speech recognition.Matthew N. Stuttle, Mark J. F. Gales
2000ICASSPRapid likelihood calculation of subspace clustered Gaussian components.Anuradha Aiyer, Mark J. F. Gales, Michael A. Picheny
2000InterspeechTranscription of broadcast news with a time constraint: IBM's 10xRT HUB4 system.Ellen Eide, Benot Maison, Dimitri Kanevsky, Peder A. Olsen, Scott Saobing Chen, Lidia Mangu, Mark J. F. Gales, Miroslav Novak, Ramesh A. Gopinath
1999ICASSPRecent improvements to IBM's speech recognition system for automatic transcription of broadcast news.Scott Saobing Chen, Ellen Eide, Mark J. F. Gales, Ramesh A. Gopinath, Dimitri Kanevsky, Peder A. Olsen
1999InterspeechTail distribution modelling using the richter and power exponential distributions.Mark J. F. Gales, Peder A. Olsen
1998ICASSPSemi-tied covariance matrices.Mark J. F. Gales
1998InterspeechCluster adaptive training for speech recognition.Mark J. F. Gales
1997ICASSPBroadcast news transcription using HTK.Philip C. Woodland, Mark J. F. Gales, David Pye, Steve J. Young
1997InterspeechTransformation smoothing for speaker and environmental adaptation.Mark J. F. Gales
1997InterspeechA comparative study of methods for phonetic decision-tree state clustering.Harriet J. Nock, Mark J. F. Gales, Steve J. Young
1996InterspeechVariance compensation within the MLLR framework for robust speech recognition and speaker adaptation.Mark J. F. Gales, David Pye, Philip C. Woodland
1996InterspeechUse of Gaussian selection in large vocabulary continuous speech recognition using HMMs.Kate M. Knill, Mark J. F. Gales, Steve J. Young
1996InterspeechIterative unsupervised adaptation using maximum likelihood linear regression.Philip C. Woodland, David Pye, Mark J. F. Gales
1995InterspeechThe application of parallel model combination to a large vocabulary dictation task.Mark J. F. Gales, Steve J. Young
1994InterspeechParallel model combination on a noise corrupted resource management task.Mark J. F. Gales, Steve J. Young
1993InterspeechHMM recognition in noise using parallel model combination.Mark J. F. Gales, Steve J. Young
1993InterspeechSegmental hidden Markov models.Mark J. F. Gales, Steve J. Young
1992ICASSPAn improved approach to the hidden Markov model decomposition of speech and noise.Mark J. F. Gales, Steve J. Young