Skip to content

Thomas Merritt

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

20

Venues

3

Active years

2014–2024

Best venue rank

A

Where they publish

Papers

20 indexed papers, newest first.

YearVenueTitleAuthors
2024ICASSPInvestigating Self-Supervised Features for Expressive, Multilingual Voice Conversion.lvaro Martn-Cortinas, Daniel Sez-Trigueros, Grzegorz Beringer, Ivn Valls-Prez, Roberto Barra-Chicote, Biel Tura Vecino, Adam Gabrys, Thomas Merritt, Piotr Bilinski, Jaime Lorenzo-Trueba
2023ICASSPAE-Flow: Autoencoder Normalizing Flow.Jakub Mosinski, Piotr Bilinski, Thomas Merritt, Abdelhamid Ezzerg, Daniel Korzekwa
2023InterspeechComparing normalizing flows and diffusion models for prosody and acoustic modelling in text-to-speech.Guangyan Zhang, Thomas Merritt, Manuel Sam Ribeiro, Biel Tura Vecino, Kayoko Yanagisawa, Kamil Pokora, Abdelhamid Ezzerg, Sebastian Cygert, Ammar Abbas, Piotr Bilinski, Roberto Barra-Chicote, Daniel Korzekwa, Jaime Lorenzo-Trueba
2022ICASSPText-Free Non-Parallel Many-To-Many Voice Conversion Using Normalising Flow.Thomas Merritt, Abdelhamid Ezzerg, Piotr Bilinski, Magdalena Proszewska, Kamil Pokora, Roberto Barra-Chicote, Daniel Korzekwa
2022InterspeechExpressive, Variable, and Controllable Duration Modelling in TTS.Syed Ammar Abbas, Thomas Merritt, Alexis Moinet, Sri Karlapati, Ewa Muszynska, Simon Slangen, Elia Gatti, Thomas Drugman
2022InterspeechCreating New Voices using Normalizing Flows.Piotr Bilinski, Thomas Merritt, Abdelhamid Ezzerg, Kamil Pokora, Sebastian Cygert, Kayoko Yanagisawa, Roberto Barra-Chicote, Daniel Korzekwa
2022InterspeechGlowVC: Mel-spectrogram space disentangling model for language-independent text-free voice conversion.Magdalena Proszewska, Grzegorz Beringer, Daniel Sez-Trigueros, Thomas Merritt, Abdelhamid Ezzerg, Roberto Barra-Chicote
2021ICASSPCamp: A Two-Stage Approach to Modelling Prosody in Context.Zack Hodari, Alexis Moinet, Sri Karlapati, Jaime Lorenzo-Trueba, Thomas Merritt, Arnaud Joly, Ammar Abbas, Penny Karanasou, Thomas Drugman
2021ICASSPLow-Resource Expressive Text-To-Speech Using Data Augmentation.Goeric Huybrechts, Thomas Merritt, Giulia Comini, Bartek Perz, Raahil Shah, Jaime Lorenzo-Trueba
2019ICASSPEffect of Data Reduction on Sequence-to-sequence Neural TTS.Javier Latorre, Jakub Lachowicz, Jaime Lorenzo-Trueba, Thomas Merritt, Thomas Drugman, Srikanth Ronanki, Viacheslav Klimkov
2019InterspeechTowards Achieving Robust Universal Neural Vocoding.Jaime Lorenzo-Trueba, Thomas Drugman, Javier Latorre, Thomas Merritt, Bartosz Putrycz, Roberto Barra-Chicote, Alexis Moinet, Vatsal Aggarwal
2019NAACLIn Other News: a Bi-style Text-to-speech Model for Synthesizing Newscaster Voice with Limited Data.Nishant Prateek, Mateusz Lajszczak, Roberto Barra-Chicote, Thomas Drugman, Jaime Lorenzo-Trueba, Thomas Merritt, Srikanth Ronanki, Trevor Wood
2017InterspeechPhrase Break Prediction for Long-Form Reading TTS: Exploiting Text Structure Information.Viacheslav Klimkov, Adam Nadolski, Alexis Moinet, Bartosz Putrycz, Roberto Barra-Chicote, Thomas Merritt, Thomas Drugman
2016ICASSPDeep neural network-guided unit selection synthesis.Thomas Merritt, Robert A. J. Clark, Zhizheng Wu, Junichi Yamagishi, Simon King
2016ICASSPFrom HMMS to DNNS: Where do the improvements come from?Oliver Watts, Gustav Eje Henter, Thomas Merritt, Zhizheng Wu, Simon King
2015ICASSPAttributing modelling errors in HMM synthesis by stepping gradually from natural to modelled speech.Thomas Merritt, Javier Latorre, Simon King
2015InterspeechDeep neural network context embeddings for model selection in rich-context HMM synthesis.Thomas Merritt, Junichi Yamagishi, Zhizheng Wu, Oliver Watts, Simon King
2014InterspeechA flexible front-end for HTS.Matthew P. Aylett, Rasmus Dall, Arnab Ghoshal, Gustav Eje Henter, Thomas Merritt
2014InterspeechMeasuring the perceptual effects of modelling assumptions in speech synthesis using stimuli constructed from repeated natural speech.Gustav Eje Henter, Thomas Merritt, Matt Shannon, Catherine Mayo, Simon King
2014InterspeechInvestigating source and filter contributions, and their interaction, to statistical parametric speech synthesis.Thomas Merritt, Tuomo Raitio, Simon King