| 2017 | ICASSP | Expressive visual text to speech and expression adaptation using deep neural networks. | Jonathan Parker, Ranniery Maia, Yannis Stylianou, Roberto Cipolla |
| 2017 | Interspeech | Speaker Adaptation in DNN-Based Speech Synthesis Using d-Vectors. | Rama Doddipatla, Norbert Braunschweiler, Ranniery Maia |
| 2016 | ICASSP | Iterative estimation of phase using complex cepstrum representation. | Ranniery Maia, Yannis Stylianou |
| 2016 | ICASSP | Multi-stream spectral representation for statistical parametric speech synthesis. | Kayoko Yanagisawa, Ranniery Maia, Yannis Stylianou |
| 2016 | Interspeech | Pause Prediction from Text for Speech Synthesis with User-Definable Pause Insertion Likelihood Threshold. | Norbert Braunschweiler, Ranniery Maia |
| 2015 | ICASSP | Methods for applying dynamic sinusoidal models to statistical parametric speech synthesis. | Qiong Hu, Yannis Stylianou, Ranniery Maia, Korin Richmond, Junichi Yamagishi |
| 2015 | Interspeech | Fusion of multiple parameterisations for DNN-based sinusoidal speech synthesis with multi-task learning. | Qiong Hu, Zhizheng Wu, Korin Richmond, Junichi Yamagishi, Yannis Stylianou, Ranniery Maia |
| 2015 | Interspeech | A maximum likelihood approach to the detection of moments of maximum excitation and its application to high-quality speech parameterization. | Ranniery Maia, Yannis Stylianou, Masami Akamine |
| 2015 | Interspeech | Towards a linear dynamical model based speech synthesizer. | Vassilios Tsiaras, Ranniery Maia, Vassilios Diakoloukas, Yannis Stylianou, Vassilios Digalakis |
| 2014 | ICASSP | A fixed dimension and perceptually based dynamic sinusoidal model of speech. | Qiong Hu, Yannis Stylianou, Korin Richmond, Ranniery Maia, Junichi Yamagishi, Javier Latorre |
| 2014 | ICASSP | Complex cepstrum factorization for statistical parametric synthesis. | Ranniery Maia, Yannis Stylianou |
| 2014 | ICASSP | Linear dynamical models in speech synthesis. | Vassilios Tsiaras, Ranniery Maia, Vassilios Diakoloukas, Yannis Stylianou, Vassilios Digalakis |
| 2014 | Interspeech | An investigation of the application of dynamic sinusoidal models to statistical parametric speech synthesis. | Qiong Hu, Yannis Stylianou, Ranniery Maia, Korin Richmond, Junichi Yamagishi, Javier Latorre |
| 2013 | ICASSP | Complex cepstrum analysis based on the minimum mean squared error. | Ranniery Maia, Masami Akamine, Mark J. F. Gales |
| 2013 | Interspeech | Minimum mean squared error based warped complex cepstrum analysis for statistical parametric speech synthesis. | Ranniery Maia, Mark J. F. Gales, Yannis Stylianou, Masami Akamine |
| 2013 | Interspeech | Photo-realistic expressive text to talking head synthesis. | Vincent Wan, Robert Anderson, Art Blokland, Norbert Braunschweiler, Langzhou Chen, BalaKrishna Kolluru, Javier Latorre, Ranniery Maia, Bjrn Stenger, Kayoko Yanagisawa, Yannis Stylianou, Masami Akamine, Mark J. F. Gales, Roberto Cipolla |
| 2012 | ICASSP | Complex cepstrum as phase information in statistical parametric speech synthesis. | Ranniery Maia, Masami Akamine, Mark J. F. Gales |
| 2012 | ICASSP | Cepstral analysis based on the glimpse proportion measure for improving the intelligibility of HMM-based synthetic speech in noise. | Cassia Valentini-Botinhao, Ranniery Maia, Junichi Yamagishi, Simon King, Heiga Zen |
| 2012 | Interspeech | Analysis on the Importance of Short-Term Speech Parameterizations for Emotional Statistical Parametric Speech Synthesis. | Ranniery Maia |
| 2011 | Interspeech | Multipulse Sequences for Residual Signal Modeling. | Ranniery Maia, Heiga Zen, Kate M. Knill, Mark J. F. Gales, Sabine Buchholz |
| 2009 | Interspeech | A decision tree-based clustering approach to state definition in an excitation modeling framework for HMM-based speech synthesis. | Ranniery Maia, Tomoki Toda, Keiichi Tokuda, Shinsuke Sakai, Satoshi Nakamura |
| 2009 | Interspeech | A close look into the probabilistic concatenation model for corpus-based speech synthesis. | Shinsuke Sakai, Ranniery Maia, Hisashi Kawai, Satoshi Nakamura |
| 2008 | ICASSP | On the state definition for a trainable excitation model in HMM-based speech synthesis. | Ranniery Maia, Tomoki Toda, Keiichi Tokuda, Shinichi Sakai, Shun Nakamura |
| 2007 | Interspeech | A trainable excitation model for HMM-based speech synthesis. | Ranniery Maia, Tomoki Toda, Heiga Zen, Yoshihiko Nankaku, Keiichi Tokuda |
| 2005 | Interspeech | HMM-based european Portuguese TTS system. | Maria Joo Barros, Ranniery Maia, Keiichi Tokuda, Fernando Gil Resende, Diamantino Freitas |
| 2003 | ICASSP | Mixed-excited phonetic vocoding at 265 bps. | Ranniery Maia, Ricardo J. da R. Cirigliano, Daniel Rojtenberg, Fernando Gil Vianna Resende Jr. |
| 2003 | Interspeech | Towards the development of a brazilian portuguese text-to-speech system based on HMM. | Ranniery Maia, Heiga Zen, Keiichi Tokuda, Tadashi Kitamura, Fernando Gil Vianna Resende Jr. |