| 2024 | Interspeech | Positional Description for Numerical Normalization. | Deepanshu Gupta, Javier Latorre |
| 2021 | NAACL | Proteno: Text Normalization with Limited Data for Fast Deployment in Text to Speech Systems. | Shubhi Tyagi, Antonio Bonafonte, Jaime Lorenzo-Trueba, Javier Latorre |
| 2019 | ICASSP | Effect of Data Reduction on Sequence-to-sequence Neural TTS. | Javier Latorre, Jakub Lachowicz, Jaime Lorenzo-Trueba, Thomas Merritt, Thomas Drugman, Srikanth Ronanki, Viacheslav Klimkov |
| 2019 | Interspeech | Towards Achieving Robust Universal Neural Vocoding. | Jaime Lorenzo-Trueba, Thomas Drugman, Javier Latorre, Thomas Merritt, Bartosz Putrycz, Roberto Barra-Chicote, Alexis Moinet, Vatsal Aggarwal |
| 2015 | ICASSP | Attributing modelling errors in HMM synthesis by stepping gradually from natural to modelled speech. | Thomas Merritt, Javier Latorre, Simon King |
| 2014 | ICASSP | A fixed dimension and perceptually based dynamic sinusoidal model of speech. | Qiong Hu, Yannis Stylianou, Korin Richmond, Ranniery Maia, Junichi Yamagishi, Javier Latorre |
| 2014 | ICASSP | Cluster adaptive training of average voice models. | Vincent Wan, Javier Latorre, Kayoko Yanagisawa, Mark J. F. Gales, Yannis Stylianou |
| 2014 | Interspeech | An investigation of the application of dynamic sinusoidal models to statistical parametric speech synthesis. | Qiong Hu, Yannis Stylianou, Ranniery Maia, Korin Richmond, Junichi Yamagishi, Javier Latorre |
| 2014 | Interspeech | Generating multiple-accent pronunciations for TTS using joint sequence model interpolation. | BalaKrishna Kolluru, Vincent Wan, Javier Latorre, Kayoko Yanagisawa, Mark J. F. Gales |
| 2014 | Interspeech | Voice expression conversion with factorised HMM-TTS models. | Javier Latorre, Vincent Wan, Kayoko Yanagisawa |
| 2014 | Interspeech | Speech intonation for TTS: study on evaluation methodology. | Javier Latorre, Kayoko Yanagisawa, Vincent Wan, BalaKrishna Kolluru, Mark J. F. Gales |
| 2013 | ICASSP | Training a supra-segmental parametric F0 model without interpolating F0. | Javier Latorre, Mark J. F. Gales, Kate M. Knill, Masami Akamine |
| 2013 | Interspeech | Photo-realistic expressive text to talking head synthesis. | Vincent Wan, Robert Anderson, Art Blokland, Norbert Braunschweiler, Langzhou Chen, BalaKrishna Kolluru, Javier Latorre, Ranniery Maia, Bjrn Stenger, Kayoko Yanagisawa, Yannis Stylianou, Masami Akamine, Mark J. F. Gales, Roberto Cipolla |
| 2012 | ICASSP | Unsupervised clustering of emotion and voice styles for expressive TTS. | Florian Eyben, Sabine Buchholz, Norbert Braunschweiler, Javier Latorre, Vincent Wan, Mark J. F. Gales, Kate M. Knill |
| 2012 | Interspeech | Exploring Rich Expressive Information from Audiobook Data Using Cluster Adaptive Training. | Langzhou Chen, Mark J. F. Gales, Vincent Wan, Javier Latorre, Masami Akamine |
| 2012 | Interspeech | Speech factorization for HMM-TTS based on cluster adaptive training. | Javier Latorre, Vincent Wan, Mark J. F. Gales, Langzhou Chen, K. K. Chin, Kate M. Knill, Masami Akamine |
| 2012 | Interspeech | C2H: A Computational Model of H&H-based Phonetic Contrast in Synthetic Speech. | Mauro Nicolao, Javier Latorre, Roger K. Moore |
| 2012 | Interspeech | Combining multiple high quality corpora for improving HMM-TTS. | Vincent Wan, Javier Latorre, K. K. Chin, Langzhou Chen, Mark J. F. Gales, Heiga Zen, Kate M. Knill, Masami Akamine |
| 2011 | ICASSP | Continuous F0 in the source-excitation generation for HMM-based TTS: Do we need voiced/unvoiced classification? | Javier Latorre, Mark J. F. Gales, Sabine Buchholz, Kate M. Knill, Masatsune Tamura, Yamato Ohtani, Masami Akamine |
| 2011 | Interspeech | Crowdsourcing Preference Tests, and How to Detect Cheating. | Sabine Buchholz, Javier Latorre |
| 2010 | Interspeech | Training a parametric-based logF0 model with the minimum generation error criterion. | Javier Latorre, Mark J. F. Gales, Heiga Zen |
| 2009 | Interspeech | Feedback loop for prosody prediction in concatenative speech synthesis. | Javier Latorre, Sergio Gracia, Masami Akamine |
| 2008 | Interspeech | Multilevel parametric-base F0 model for speech synthesis. | Javier Latorre, Masami Akamine |
| 2007 | ICASSP | Combining Gaussian Mixture Model with Global Variance Term to Improve the Quality of an HMM-Based Polyglot Speech Synthesizer. | Javier Latorre, Koji Iwano, Sadaoki Furui |
| 2005 | ICASSP | Polyglot Synthesis Using a Mixture of Monolingual Corpora. | Javier Latorre, Koji Iwano, Sadaoki Furui |
| 2005 | Interspeech | Cross-language synthesis with a polyglot synthesizer. | Javier Latorre, Koji Iwano, Sadaoki Furui |