| 2016 | ICASSP | Deep neural network-guided unit selection synthesis. | Thomas Merritt, Robert A. J. Clark, Zhizheng Wu, Junichi Yamagishi, Simon King |
| 2016 | ICASSP | Wavelet-based decomposition of F0 as a secondary task for DNN-based speech synthesis with multi-task learning. | Manuel Sam Ribeiro, Oliver Watts, Junichi Yamagishi, Robert A. J. Clark |
| 2016 | Interspeech | The SIWIS Database: A Multilingual Speech Database with Acted Emphasis. | Jean-Philippe Goldman, Pierre-Edouard Honnet, Robert A. J. Clark, Philip N. Garner, Maria Ivanova, Alexandros Lazaridis, Hui Liang, Tiago Macedo, Beat Pfister, Manuel Sam Ribeiro, Eric Wehrli, Junichi Yamagishi |
| 2015 | ICASSP | A multi-level representation of f0 using the continuous wavelet transform and the Discrete Cosine Transform. | Manuel Sam Ribeiro, Robert A. J. Clark |
| 2015 | Interspeech | A perceptual investigation of wavelet-based decomposition of f0 for text-to-speech synthesis. | Manuel Sam Ribeiro, Junichi Yamagishi, Robert A. J. Clark |
| 2014 | Interspeech | Simple | Robert A. J. Clark |
| 2014 | Interspeech | Speech synthesis reactive to dynamic noise environmental conditions. | Susana Palmaz Lpez-Pelez, Robert A. J. Clark |
| 2014 | Interspeech | Unsupervised language filtering using the latent dirichlet allocation. | Wei Zhang, Robert A. J. Clark, Yongyuan Wang |
| 2013 | ICASSP | Lightly supervised GMM VAD to use audiobook for speech synthesiser. | Yoshitaka Mamiya, Junichi Yamagishi, Oliver Watts, Robert A. J. Clark, Simon King, Adriana Stan |
| 2013 | Interspeech | Simple | Robert A. J. Clark |
| 2013 | Interspeech | TUNDRA: a multilingual corpus of found data for TTS research created with light supervision. | Adriana Stan, Oliver Watts, Yoshitaka Mamiya, Mircea Giurgiu, Robert A. J. Clark, Junichi Yamagishi, Simon King |
| 2012 | Interspeech | Towards Hierarchical Prosodic Prominence Generation in TTS Synthesis. | Leonardo Badino, Robert A. J. Clark |
| 2012 | Interspeech | Asymmetries in the perception of synthesized speech. | Anna C. Janska, Erich Schrger, Thomas Jacobsen, Robert A. J. Clark |
| 2010 | Interspeech | Native and non-native speaker judgements on the quality of synthesized speech. | Anna C. Janska, Robert A. J. Clark |
| 2010 | Interspeech | On generating combilex pronunciations via morphological analysis. | Korin Richmond, Robert A. J. Clark, Susan Fitt |
| 2009 | Interspeech | Identification of contrast and its emphatic realization in HMM based speech synthesis. | Leonardo Badino, J. Sebastian Andersson, Junichi Yamagishi, Robert A. J. Clark |
| 2009 | Interspeech | Robust LTS rules with the Combilex speech technology lexicon. | Korin Richmond, Robert A. J. Clark, Susan Fitt |
| 2008 | Interspeech | Including pitch accent optionality in unit selection text-to-speech synthesis. | Leonardo Badino, Robert A. J. Clark, Volker Strom |
| 2007 | Interspeech | Modelling prominence and emphasis improves unit-selection synthesis. | Volker Strom, Ani Nenkova, Robert A. J. Clark, Yolanda Vazquez-Alvarez, Jason M. Brenier, Simon King, Dan Jurafsky |
| 2006 | Interspeech | Joint prosodic and segmental unit selection speech synthesis. | Robert A. J. Clark, Simon King |
| 2006 | Interspeech | Expressive prosody for unit-selection speech synthesis. | Volker Strom, Robert A. J. Clark, Simon King |
| 2005 | Interspeech | Multisyn voices from ARCTIC data for the blizzard challenge. | Robert A. J. Clark, Korin Richmond, Simon King |
| 2005 | Interspeech | Informed blending of databases for emotional speech synthesis. | Gregor Hofer, Korin Richmond, Robert A. J. Clark |
| 2005 | Interspeech | Multidimensional scaling of listener responses to synthetic speech. | Catherine Mayo, Robert A. J. Clark, Simon King |
| 2005 | Interspeech | Modelling pitch accent types for Polish speech synthesis. | Dominika Oliver, Robert A. J. Clark |
| 1999 | Interspeech | Objective methods for evaluating synthetic intonation. | Robert A. J. Clark, Kurt E. Dusterhoff |