Thomas Merritt
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
20
Venues
3
Active years
2014–2024
Best venue rank
A
Where they publish
Papers
20 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2024 | ICASSP | Investigating Self-Supervised Features for Expressive, Multilingual Voice Conversion. | lvaro Martn-Cortinas, Daniel Sez-Trigueros, Grzegorz Beringer, Ivn Valls-Prez, Roberto Barra-Chicote, Biel Tura Vecino, Adam Gabrys, Thomas Merritt, Piotr Bilinski, Jaime Lorenzo-Trueba |
| 2023 | ICASSP | AE-Flow: Autoencoder Normalizing Flow. | Jakub Mosinski, Piotr Bilinski, Thomas Merritt, Abdelhamid Ezzerg, Daniel Korzekwa |
| 2023 | Interspeech | Comparing normalizing flows and diffusion models for prosody and acoustic modelling in text-to-speech. | Guangyan Zhang, Thomas Merritt, Manuel Sam Ribeiro, Biel Tura Vecino, Kayoko Yanagisawa, Kamil Pokora, Abdelhamid Ezzerg, Sebastian Cygert, Ammar Abbas, Piotr Bilinski, Roberto Barra-Chicote, Daniel Korzekwa, Jaime Lorenzo-Trueba |
| 2022 | ICASSP | Text-Free Non-Parallel Many-To-Many Voice Conversion Using Normalising Flow. | Thomas Merritt, Abdelhamid Ezzerg, Piotr Bilinski, Magdalena Proszewska, Kamil Pokora, Roberto Barra-Chicote, Daniel Korzekwa |
| 2022 | Interspeech | Expressive, Variable, and Controllable Duration Modelling in TTS. | Syed Ammar Abbas, Thomas Merritt, Alexis Moinet, Sri Karlapati, Ewa Muszynska, Simon Slangen, Elia Gatti, Thomas Drugman |
| 2022 | Interspeech | Creating New Voices using Normalizing Flows. | Piotr Bilinski, Thomas Merritt, Abdelhamid Ezzerg, Kamil Pokora, Sebastian Cygert, Kayoko Yanagisawa, Roberto Barra-Chicote, Daniel Korzekwa |
| 2022 | Interspeech | GlowVC: Mel-spectrogram space disentangling model for language-independent text-free voice conversion. | Magdalena Proszewska, Grzegorz Beringer, Daniel Sez-Trigueros, Thomas Merritt, Abdelhamid Ezzerg, Roberto Barra-Chicote |
| 2021 | ICASSP | Camp: A Two-Stage Approach to Modelling Prosody in Context. | Zack Hodari, Alexis Moinet, Sri Karlapati, Jaime Lorenzo-Trueba, Thomas Merritt, Arnaud Joly, Ammar Abbas, Penny Karanasou, Thomas Drugman |
| 2021 | ICASSP | Low-Resource Expressive Text-To-Speech Using Data Augmentation. | Goeric Huybrechts, Thomas Merritt, Giulia Comini, Bartek Perz, Raahil Shah, Jaime Lorenzo-Trueba |
| 2019 | ICASSP | Effect of Data Reduction on Sequence-to-sequence Neural TTS. | Javier Latorre, Jakub Lachowicz, Jaime Lorenzo-Trueba, Thomas Merritt, Thomas Drugman, Srikanth Ronanki, Viacheslav Klimkov |
| 2019 | Interspeech | Towards Achieving Robust Universal Neural Vocoding. | Jaime Lorenzo-Trueba, Thomas Drugman, Javier Latorre, Thomas Merritt, Bartosz Putrycz, Roberto Barra-Chicote, Alexis Moinet, Vatsal Aggarwal |
| 2019 | NAACL | In Other News: a Bi-style Text-to-speech Model for Synthesizing Newscaster Voice with Limited Data. | Nishant Prateek, Mateusz Lajszczak, Roberto Barra-Chicote, Thomas Drugman, Jaime Lorenzo-Trueba, Thomas Merritt, Srikanth Ronanki, Trevor Wood |
| 2017 | Interspeech | Phrase Break Prediction for Long-Form Reading TTS: Exploiting Text Structure Information. | Viacheslav Klimkov, Adam Nadolski, Alexis Moinet, Bartosz Putrycz, Roberto Barra-Chicote, Thomas Merritt, Thomas Drugman |
| 2016 | ICASSP | Deep neural network-guided unit selection synthesis. | Thomas Merritt, Robert A. J. Clark, Zhizheng Wu, Junichi Yamagishi, Simon King |
| 2016 | ICASSP | From HMMS to DNNS: Where do the improvements come from? | Oliver Watts, Gustav Eje Henter, Thomas Merritt, Zhizheng Wu, Simon King |
| 2015 | ICASSP | Attributing modelling errors in HMM synthesis by stepping gradually from natural to modelled speech. | Thomas Merritt, Javier Latorre, Simon King |
| 2015 | Interspeech | Deep neural network context embeddings for model selection in rich-context HMM synthesis. | Thomas Merritt, Junichi Yamagishi, Zhizheng Wu, Oliver Watts, Simon King |
| 2014 | Interspeech | A flexible front-end for HTS. | Matthew P. Aylett, Rasmus Dall, Arnab Ghoshal, Gustav Eje Henter, Thomas Merritt |
| 2014 | Interspeech | Measuring the perceptual effects of modelling assumptions in speech synthesis using stimuli constructed from repeated natural speech. | Gustav Eje Henter, Thomas Merritt, Matt Shannon, Catherine Mayo, Simon King |
| 2014 | Interspeech | Investigating source and filter contributions, and their interaction, to statistical parametric speech synthesis. | Thomas Merritt, Tuomo Raitio, Simon King |