| 2015 | Interspeech | A maximum likelihood approach to the detection of moments of maximum excitation and its application to high-quality speech parameterization. | Ranniery Maia, Yannis Stylianou, Masami Akamine |
| 2015 | Interspeech | Emotional transplant in statistical speech synthesis based on emotion additive model. | Yamato Ohtani, Yu Nasu, Masahiro Morita, Masami Akamine |
| 2014 | Interspeech | GMM-based bandwidth extension using sub-band basis spectrum model. | Yamato Ohtani, Masatsune Tamura, Masahiro Morita, Masami Akamine |
| 2013 | ICASSP | Integrated automatic expression prediction and speech synthesis from text. | Langzhou Chen, Mark J. F. Gales, Norbert Braunschweiler, Masami Akamine, Kate M. Knill |
| 2013 | ICASSP | Training a supra-segmental parametric F0 model without interpolating F0. | Javier Latorre, Mark J. F. Gales, Kate M. Knill, Masami Akamine |
| 2013 | ICASSP | Complex cepstrum analysis based on the minimum mean squared error. | Ranniery Maia, Masami Akamine, Mark J. F. Gales |
| 2013 | Interspeech | Minimum mean squared error based warped complex cepstrum analysis for statistical parametric speech synthesis. | Ranniery Maia, Mark J. F. Gales, Yannis Stylianou, Masami Akamine |
| 2013 | Interspeech | Photo-realistic expressive text to talking head synthesis. | Vincent Wan, Robert Anderson, Art Blokland, Norbert Braunschweiler, Langzhou Chen, BalaKrishna Kolluru, Javier Latorre, Ranniery Maia, Bjrn Stenger, Kayoko Yanagisawa, Yannis Stylianou, Masami Akamine, Mark J. F. Gales, Roberto Cipolla |
| 2012 | ICASSP | Complex cepstrum as phase information in statistical parametric speech synthesis. | Ranniery Maia, Masami Akamine, Mark J. F. Gales |
| 2012 | Interspeech | Exploring Rich Expressive Information from Audiobook Data Using Cluster Adaptive Training. | Langzhou Chen, Mark J. F. Gales, Vincent Wan, Javier Latorre, Masami Akamine |
| 2012 | Interspeech | Speech factorization for HMM-TTS based on cluster adaptive training. | Javier Latorre, Vincent Wan, Mark J. F. Gales, Langzhou Chen, K. K. Chin, Kate M. Knill, Masami Akamine |
| 2012 | Interspeech | Histogram-based spectral equalization for HMM-based speech synthesis using mel-LSP. | Yamato Ohtani, Masatsune Tamura, Masahiro Morita, Takehiko Kagoshima, Masami Akamine |
| 2012 | Interspeech | HMM-based speech synthesis using sub-band basis spectrum model. | Yamato Ohtani, Masatsune Tamura, Masahiro Morita, Takehiko Kagoshima, Masami Akamine |
| 2012 | Interspeech | Combining multiple high quality corpora for improving HMM-TTS. | Vincent Wan, Javier Latorre, K. K. Chin, Langzhou Chen, Mark J. F. Gales, Heiga Zen, Kate M. Knill, Masami Akamine |
| 2011 | ICASSP | Continuous F0 in the source-excitation generation for HMM-based TTS: Do we need voiced/unvoiced classification? | Javier Latorre, Mark J. F. Gales, Sabine Buchholz, Kate M. Knill, Masatsune Tamura, Yamato Ohtani, Masami Akamine |
| 2011 | ICASSP | One sentence voice adaptation using GMM-based frequency-warping and shift with a sub-band basis spectrum model. | Masatsune Tamura, Masahiro Morita, Takehiko Kagoshima, Masami Akamine |
| 2010 | ICASSP | Covariance clustering on Riemannian manifolds for acoustic model compression. | Yusuke Shinohara, Takashi Masuko, Masami Akamine |
| 2010 | ICASSP | Unit selection speech synthesis using multiple speech units at non-adjacent segments for prosody and waveform generation. | Masatsune Tamura, Norbert Braunschweiler, Takehiko Kagoshima, Masami Akamine |
| 2010 | Interspeech | Sub-band basis spectrum model for pitch-synchronous log-spectrum and phase based on approximation of sparse coding. | Masatsune Tamura, Takehiko Kagoshima, Masami Akamine |
| 2009 | ICASSP | Bayesian feature enhancement using a mixture of unscented transformation for uncertainty decoding of noisy speech. | Yusuke Shinohara, Masami Akamine |
| 2009 | Interspeech | Decision tree acoustic models for ASR. | Jitendra Ajmera, Masami Akamine |
| 2009 | Interspeech | Feedback loop for prosody prediction in concatenative speech synthesis. | Javier Latorre, Sergio Gracia, Masami Akamine |
| 2008 | ICASSP | Feature enhancement by speaker-normalized splice for robust speech recognition. | Yusuke Shinohara, Takashi Masuko, Masami Akamine |
| 2008 | Interspeech | Speech recognition using soft decision trees. | Jitendra Ajmera, Masami Akamine |
| 2008 | Interspeech | Comparative evaluation of different methods for voice activity detection. | Hongfei Ding, Koichi Yamamoto, Masami Akamine |
| 2008 | Interspeech | Multilevel parametric-base F0 model for speech synthesis. | Javier Latorre, Masami Akamine |
| 2007 | Interspeech | HMM-based speech recognition using decision trees instead of GMMs. | Remco Teunen, Masami Akamine |
| 1999 | ICASSP | CELP speech coding based on an adaptive pulse position codebook. | Tadashi Amada, Kimio Miseki, Masami Akamine |
| 1999 | Interspeech | Toshiba English text-to-speech synthesizer (TESS). | Chang K. Suh, Takehiko Kagoshima, Masahiro Morita, Shigenobu Seto, Masami Akamine |
| 1998 | ICASSP | A 2.4 kbps variable bit rate ADP-CELP speech coder. | Masahiro Oshikiri, Masami Akamine |
| 1998 | Interspeech | Analytic generation of synthesis units by closed loop training for totally speaker driven text to speech system (TOS drive TTS). | Masami Akamine, Takehiko Kagoshima |
| 1998 | Interspeech | An F0 contour control model for totally speaker driven text to speech system. | Takehiko Kagoshima, Masahiro Morita, Shigenobu Seto, Masami Akamine |
| 1998 | Interspeech | Automatic rule generation for linguistic features analysis using inductive learning technique: linguistic features analysis in TOS drive TTS system. | Shigenobu Seto, Masahiro Morita, Takehiko Kagoshima, Masami Akamine |
| 1997 | ICASSP | Automatic generation of speech synthesis units based on closed loop training. | Takehiko Kagoshima, Masami Akamine |
| 1991 | ICASSP | Adaptive bit-allocation between the pole-zero synthesis filter and excitation in CELP. | Kimio Miseki, Masami Akamine |
| 1990 | ICASSP | CELP coding with an adaptive density pulse excitation model. | Masami Akamine, Kimio Miseki |