Skip to content

Masami Akamine

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

36

Venues

2

Active years

1990–2015

Best venue rank

A

Where they publish

Papers

36 indexed papers, newest first.

YearVenueTitleAuthors
2015InterspeechA maximum likelihood approach to the detection of moments of maximum excitation and its application to high-quality speech parameterization.Ranniery Maia, Yannis Stylianou, Masami Akamine
2015InterspeechEmotional transplant in statistical speech synthesis based on emotion additive model.Yamato Ohtani, Yu Nasu, Masahiro Morita, Masami Akamine
2014InterspeechGMM-based bandwidth extension using sub-band basis spectrum model.Yamato Ohtani, Masatsune Tamura, Masahiro Morita, Masami Akamine
2013ICASSPIntegrated automatic expression prediction and speech synthesis from text.Langzhou Chen, Mark J. F. Gales, Norbert Braunschweiler, Masami Akamine, Kate M. Knill
2013ICASSPTraining a supra-segmental parametric F0 model without interpolating F0.Javier Latorre, Mark J. F. Gales, Kate M. Knill, Masami Akamine
2013ICASSPComplex cepstrum analysis based on the minimum mean squared error.Ranniery Maia, Masami Akamine, Mark J. F. Gales
2013InterspeechMinimum mean squared error based warped complex cepstrum analysis for statistical parametric speech synthesis.Ranniery Maia, Mark J. F. Gales, Yannis Stylianou, Masami Akamine
2013InterspeechPhoto-realistic expressive text to talking head synthesis.Vincent Wan, Robert Anderson, Art Blokland, Norbert Braunschweiler, Langzhou Chen, BalaKrishna Kolluru, Javier Latorre, Ranniery Maia, Bjrn Stenger, Kayoko Yanagisawa, Yannis Stylianou, Masami Akamine, Mark J. F. Gales, Roberto Cipolla
2012ICASSPComplex cepstrum as phase information in statistical parametric speech synthesis.Ranniery Maia, Masami Akamine, Mark J. F. Gales
2012InterspeechExploring Rich Expressive Information from Audiobook Data Using Cluster Adaptive Training.Langzhou Chen, Mark J. F. Gales, Vincent Wan, Javier Latorre, Masami Akamine
2012InterspeechSpeech factorization for HMM-TTS based on cluster adaptive training.Javier Latorre, Vincent Wan, Mark J. F. Gales, Langzhou Chen, K. K. Chin, Kate M. Knill, Masami Akamine
2012InterspeechHistogram-based spectral equalization for HMM-based speech synthesis using mel-LSP.Yamato Ohtani, Masatsune Tamura, Masahiro Morita, Takehiko Kagoshima, Masami Akamine
2012InterspeechHMM-based speech synthesis using sub-band basis spectrum model.Yamato Ohtani, Masatsune Tamura, Masahiro Morita, Takehiko Kagoshima, Masami Akamine
2012InterspeechCombining multiple high quality corpora for improving HMM-TTS.Vincent Wan, Javier Latorre, K. K. Chin, Langzhou Chen, Mark J. F. Gales, Heiga Zen, Kate M. Knill, Masami Akamine
2011ICASSPContinuous F0 in the source-excitation generation for HMM-based TTS: Do we need voiced/unvoiced classification?Javier Latorre, Mark J. F. Gales, Sabine Buchholz, Kate M. Knill, Masatsune Tamura, Yamato Ohtani, Masami Akamine
2011ICASSPOne sentence voice adaptation using GMM-based frequency-warping and shift with a sub-band basis spectrum model.Masatsune Tamura, Masahiro Morita, Takehiko Kagoshima, Masami Akamine
2010ICASSPCovariance clustering on Riemannian manifolds for acoustic model compression.Yusuke Shinohara, Takashi Masuko, Masami Akamine
2010ICASSPUnit selection speech synthesis using multiple speech units at non-adjacent segments for prosody and waveform generation.Masatsune Tamura, Norbert Braunschweiler, Takehiko Kagoshima, Masami Akamine
2010InterspeechSub-band basis spectrum model for pitch-synchronous log-spectrum and phase based on approximation of sparse coding.Masatsune Tamura, Takehiko Kagoshima, Masami Akamine
2009ICASSPBayesian feature enhancement using a mixture of unscented transformation for uncertainty decoding of noisy speech.Yusuke Shinohara, Masami Akamine
2009InterspeechDecision tree acoustic models for ASR.Jitendra Ajmera, Masami Akamine
2009InterspeechFeedback loop for prosody prediction in concatenative speech synthesis.Javier Latorre, Sergio Gracia, Masami Akamine
2008ICASSPFeature enhancement by speaker-normalized splice for robust speech recognition.Yusuke Shinohara, Takashi Masuko, Masami Akamine
2008InterspeechSpeech recognition using soft decision trees.Jitendra Ajmera, Masami Akamine
2008InterspeechComparative evaluation of different methods for voice activity detection.Hongfei Ding, Koichi Yamamoto, Masami Akamine
2008InterspeechMultilevel parametric-base F0 model for speech synthesis.Javier Latorre, Masami Akamine
2007InterspeechHMM-based speech recognition using decision trees instead of GMMs.Remco Teunen, Masami Akamine
1999ICASSPCELP speech coding based on an adaptive pulse position codebook.Tadashi Amada, Kimio Miseki, Masami Akamine
1999InterspeechToshiba English text-to-speech synthesizer (TESS).Chang K. Suh, Takehiko Kagoshima, Masahiro Morita, Shigenobu Seto, Masami Akamine
1998ICASSPA 2.4 kbps variable bit rate ADP-CELP speech coder.Masahiro Oshikiri, Masami Akamine
1998InterspeechAnalytic generation of synthesis units by closed loop training for totally speaker driven text to speech system (TOS drive TTS).Masami Akamine, Takehiko Kagoshima
1998InterspeechAn F0 contour control model for totally speaker driven text to speech system.Takehiko Kagoshima, Masahiro Morita, Shigenobu Seto, Masami Akamine
1998InterspeechAutomatic rule generation for linguistic features analysis using inductive learning technique: linguistic features analysis in TOS drive TTS system.Shigenobu Seto, Masahiro Morita, Takehiko Kagoshima, Masami Akamine
1997ICASSPAutomatic generation of speech synthesis units based on closed loop training.Takehiko Kagoshima, Masami Akamine
1991ICASSPAdaptive bit-allocation between the pole-zero synthesis filter and excitation in CELP.Kimio Miseki, Masami Akamine
1990ICASSPCELP coding with an adaptive density pulse excitation model.Masami Akamine, Kimio Miseki