| 2025 | Interspeech | PeriodCodec: A Pitch-Controllable Neural Audio Codec Using Periodic Signals for Singing Voice Synthesis. | Masato Takagi, Miku Nishihara, Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2024 | ICASSP | PeriodGrad: Towards Pitch-Controllable Neural Vocoder Based on a Diffusion Probabilistic Model. | Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2024 | ICASSP | LIMMITS'24: Multi-Speaker, Multi-Lingual Indic TTS with Voice Cloning. | Abhayjeet Singh, Amala Nagireddi, Deekshitha G, Jesuraja Bandekar, Roopa R., Sandhya Badiger, Sathvik Udupa, Prasanta Kumar Ghosh, Hema A. Murthy, Pranaw Kumar, Keiichi Tokuda, Mark Hasegawa-Johnson, Philipp Olbrich |
| 2023 | ICASSP | Singing Voice Synthesis Based on a Musical Note Position-Aware Attention Mechanism. | Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2023 | ICASSP | Lightweight, Multi-Speaker, Multi-Lingual Indic Text-to-Speech. | Abhayjeet Singh, Amala Nagireddi, Deekshitha G, Jesuraja Bandekar, Roopa R., Sandhya Badiger, Sathvik Udupa, Prasanta Kumar Ghosh, Hema A. Murthy, Heiga Zen, Pranaw Kumar, Kamal Kant, Amol Bole, Bira Chandra Singh, Keiichi Tokuda, Mark Hasegawa-Johnson, Philipp Olbrich |
| 2023 | ICASSP | Embedding a Differentiable Mel-Cepstral Synthesis Filter to a Neural Speech Synthesis System. | Takenori Yoshimura, Shinji Takaki, Kazuhiro Nakamura, Keiichiro Oura, Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2022 | ICASSP | Autoregressive Variational Autoencoder with a Hidden Semi-Markov Model-Based Structured Attention for Speech Synthesis. | Takato Fujimoto, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2022 | Interspeech | End-to-End Text-to-Speech Based on Latent Representation of Speaking Styles Using Spontaneous Dialogue. | Kentaro Mitsui, Tianyu Zhao, Kei Sawada, Yukiya Hono, Yoshihiko Nankaku, Keiichi Tokuda |
| 2021 | ICASSP | Periodnet: A Non-Autoregressive Waveform Generation Model with a Structure Separating Periodic and Aperiodic Components. | Yukiya Hono, Shinji Takaki, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2020 | ICASSP | Semi-Supervised Learning Based on Hierarchical Generative Models for End-to-End Speech Synthesis. | Takato Fujimoto, Shinji Takaki, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2020 | ICASSP | Fast and High-Quality Singing Voice Synthesis System Based on Convolutional Neural Networks. | Kazuhiro Nakamura, Shinji Takaki, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2020 | Interspeech | Hierarchical Multi-Grained Generative Model for Expressive Speech Synthesis. | Yukiya Hono, Kazuna Tsuboi, Kei Sawada, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2019 | ICASSP | Singing Voice Synthesis Based on Generative Adversarial Networks. | Yukiya Hono, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2019 | ICASSP | Speaker-dependent Wavenet-based Delay-free Adpcm Speech Coding. | Takenori Yoshimura, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2019 | Interspeech | Statistical Approach to Speech Synthesis: Past, Present and Future. | Keiichi Tokuda |
| 2018 | ICASSP | Image Recognition Based on Separable Lattice Hmms Using a Deep Neural Network for Output Probability Distributions. | Eiji Ichikawa, Kei Sawada, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2018 | ICASSP | Statistical Voice Conversion Based on Wavenet. | Jumpei Niwa, Takenori Yoshimura, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2017 | ASRU | The blizzard machine learning challenge 2017. | Kei Sawada, Keiichi Tokuda, Simon King, Alan W. Black |
| 2017 | ICASSP | Image recognition based on discriminative models using features generated from separable lattice HMMS. | Yoshinari Tsuzuki, Kei Sawada, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2017 | Interspeech | Articulatory Text-to-Speech Synthesis Using the Digital Waveguide Mesh Driven by a Deep Neural Network. | Amelia Jane Gully, Takenori Yoshimura, Damian T. Murphy, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2016 | ICASSP | Trajectory training considering global variance for speech synthesis based on neural networks. | Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2016 | ICASSP | Directly modeling voiced and unvoiced components in speech waveforms by neural networks. | Keiichi Tokuda, Heiga Zen |
| 2016 | Interspeech | Redefining the Linguistic Context Feature Set for HMM and DNN TTS Through Position and Parsing. | Rasmus Dall, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2016 | Interspeech | Voice Conversion Based on Trajectory Model Training of Neural Networks Considering Global Variance. | Naoki Hosaka, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2016 | Interspeech | Singing Voice Synthesis Based on Deep Neural Networks. | Masanari Nishimura, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2016 | Interspeech | A Hierarchical Predictor of Synthetic Speech Naturalness Using Neural Networks. | Takenori Yoshimura, Gustav Eje Henter, Oliver Watts, Mirjam Wester, Junichi Yamagishi, Keiichi Tokuda |
| 2015 | ICASSP | The effect of neural networks in statistical parametric speech synthesis. | Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2015 | ICASSP | Directly modeling speech waveforms by neural networks for statistical parametric speech synthesis. | Keiichi Tokuda, Heiga Zen |
| 2015 | Interspeech | Simultaneous optimization of multiple tree structures for factor analyzed HMM-based speech synthesis. | Takenori Yoshimura, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2014 | HAI | Voice interaction system with 3D-CG virtual agent for stand-alone smartphones. | Daisuke Yamamoto, Keiichiro Oura, Ryota Nishimura, Takahiro Uchiya, Akinobu Lee, Ichi Takumi, Keiichi Tokuda |
| 2014 | ICASSP | HMM-Based singing voice synthesis and its application to Japanese and English. | Kazuhiro Nakamura, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2014 | ICASSP | Integration of speaker and pitch adaptive training for HMM-based singing voice synthesis. | Kanako Shirota, Kazuhiro Nakamura, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2014 | Interspeech | A mel-cepstral analysis technique restoring high frequency components from low-sampling-rate speech. | Kazuhiro Nakamura, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2013 | ICASSP | Mmdagent - A fully open-source toolkit for voice interaction systems. | Akinobu Lee, Keiichiro Oura, Keiichi Tokuda |
| 2013 | ICASSP | Separable lattice 2-D HMMS introducing state duration control for recognition of images with various variations. | Takaya Makino, Shinji Takaki, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2013 | ICASSP | Integration of acoustic modeling and mel-cepstral analysis for HMM-based speech synthesis. | Kazuhiro Nakamura, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2013 | ICASSP | Contextual partial additive structure for HMM-based speech synthesis. | Shinji Takaki, Yoshihiko Nankaku, Keiichi Tokuda |
| 2013 | ICASSP | Image recognition based on separable lattice trajectory 2-D HMMS. | Akira Tamamori, Yoshihiko Nankaku, Keiichi Tokuda |
| 2012 | ICASSP | Face recognition based on extended separable lattice 2-D HMMS. | Keisuke Kumaki, Yoshihiko Nankaku, Keiichi Tokuda |
| 2012 | ICASSP | Pitch adaptive training for hmm-based singing voice synthesis. | Keiichiro Oura, Ayami Mase, Yoshihiko Nankaku, Keiichi Tokuda |
| 2012 | ICASSP | Face recognition based on separable lattice 2-D HMMS using variational bayesian method. | Kei Sawada, Akira Tamamori, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2012 | ICASSP | A model structure integration based on a Bayesian framework for speech recognition. | Sayaka Shiota, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2012 | Interspeech | A Bayesian Approach to Speaker Recognition Based on GMMs Using Multiple Model Structures. | Takafumi Hattori, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2012 | Interspeech | Cross-lingual Speaker Adaptation for HMM-based Speech Synthesis based on Perceptual Characteristics and Speaker Interpolation. | Viviane de Franca Oliveira, Sayaka Shiota, Yoshihiko Nankaku, Keiichi Tokuda |
| 2011 | ICASSP | An analysis of machine translation and speech synthesis in speech-to-speech translation system. | Kei Hashimoto, Junichi Yamagishi, William J. Byrne, Simon King, Keiichi Tokuda |
| 2011 | ICASSP | Global variance modeling on frequency domain delta LSP for HMM-based speech synthesis. | Shifeng Pan, Yoshihiko Nankaku, Keiichi Tokuda, Jianhua Tao |
| 2011 | ICASSP | An optimization algorithm of independent mean and variance parameter tying structures for HMM-based speech synthesis. | Shinji Takaki, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2011 | Interspeech | Estimation of Window Coefficients for Dynamic Feature Extraction for HMM-Based Speech Synthesis. | Ling-Hui Chen, Yoshihiko Nankaku, Heiga Zen, Keiichi Tokuda, Zhen-Hua Ling, Li-Rong Dai |
| 2011 | Interspeech | Multi-Speaker Modeling with Shared Prior Distributions and Model Structures for Bayesian Speech Synthesis. | Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2011 | Interspeech | Large-Scale Subjective Evaluations of Speech Rate Control Methods for HMM-Based Speech Synthesizers. | Tsuneo Kato, Makoto Yamada, Nobuyuki Nishizawa, Keiichiro Oura, Keiichi Tokuda |
| 2011 | Interspeech | A Bayesian Approach to Voice Conversion Based on GMMs Using Multiple Model Structures. | Lei Li, Yoshihiko Nankaku, Keiichi Tokuda |
| 2011 | Interspeech | GMM-Based Missing-Feature Reconstruction on Multi-Frame Windows. | Ulpu Remes, Yoshihiko Nankaku, Keiichi Tokuda |
| 2011 | Interspeech | Estimation of Perceptual Spaces for Speaker Identities Based on the Cross-Lingual Discrimination Task. | Minoru Tsuzaki, Keiichi Tokuda, Hisashi Kawai, Jinfu Ni |
| 2010 | EAMT | A Deterministic Annealing-Based Training Algorithm For Statistical Machine Translation Models. | Pascual Martnez-Gmez, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda, Germn Sanchis-Trilles |
| 2010 | ICASSP | Factor analyzed voice models for HMM-based speech synthesis. | Kyosuke Kazumi, Yoshihiko Nankaku, Keiichi Tokuda |
| 2010 | ICASSP | Unsupervised cross-lingual speaker adaptation for HMM-based speech synthesis. | Keiichiro Oura, Keiichi Tokuda, Junichi Yamagishi, Simon King, Mirjam Wester |
| 2010 | ICASSP | Face recognition based on separable lattice 2-D HMM with state duration modeling. | Yoshiaki Takahashi, Akira Tamamori, Yoshihiko Nankaku, Keiichi Tokuda |
| 2010 | ICASSP | An extension of Separable Lattice 2-D HMMS for rotational data variations. | Akira Tamamori, Yoshihiko Nankaku, Keiichi Tokuda |
| 2010 | ICASSP | Statistical parametric speech synthesis based on product of experts. | Heiga Zen, Mark J. F. Gales, Yoshihiko Nankaku, Keiichi Tokuda |
| 2010 | Interspeech | Speaker adaptation based on nonlinear spectral transform for speech recognition. | Toyohiro Hayashi, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |
| 2010 | Interspeech | HMM-based singing voice synthesis system using pitch-shifted pseudo training data. | Ayami Mase, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2010 | Interspeech | Voice activity detection based on conditional random fields using multiple features. | Akira Saito, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |
| 2009 | ICASSP | A Bayesian approach to HMM-based speech synthesis. | Kei Hashimoto, Heiga Zen, Yoshihiko Nankaku, Takashi Masuko, Keiichi Tokuda |
| 2009 | ICASSP | Full covariance state duration modeling for HMM-based speech synthesis. | Heng Lu, Yi-Jian Wu, Keiichi Tokuda, Li-Rong Dai, Ren-Hua Wang |
| 2009 | ICASSP | Minimum generation error training by using original spectrum as reference for log spectral distortion measure. | Yi-Jian Wu, Keiichi Tokuda |
| 2009 | ICASSP | Voice conversion based on simultaneous modelling of spectrum and F0. | Kaori Yutani, Yosuke Uto, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |
| 2009 | ICASSP | Stereo-based stochastic noise compensation based on trajectory GMMS. | Heiga Zen, Yoshihiko Nankaku, Keiichi Tokuda |
| 2009 | Interspeech | A Bayesian approach to Hidden Semi-Markov Model based speech synthesis. | Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2009 | Interspeech | A decision tree-based clustering approach to state definition in an excitation modeling framework for HMM-based speech synthesis. | Ranniery Maia, Tomoki Toda, Keiichi Tokuda, Shinsuke Sakai, Satoshi Nakamura |
| 2009 | Interspeech | Tying covariance matrices to reduce the footprint of HMM-based speech synthesis systems. | Keiichiro Oura, Heiga Zen, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |
| 2009 | Interspeech | Deterministic annealing based training algorithm for Bayesian speech recognition. | Sayaka Shiota, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2009 | Interspeech | State mapping based method for cross-lingual speaker adaptation in HMM-based speech synthesis. | Yi-Jian Wu, Yoshihiko Nankaku, Keiichi Tokuda |
| 2009 | Interspeech | An improved minimum generation error based model adaptation for HMM-based speech synthesis. | Yi-Jian Wu, Long Qin, Keiichi Tokuda |
| 2009 | Interspeech | Thousands of voices for HMM-based speech synthesis. | Junichi Yamagishi, Bela Usabaev, Simon King, Oliver Watts, John Dines, Jilei Tian, Rile Hu, Yong Guan, Keiichiro Oura, Keiichi Tokuda, Reima Karhila, Mikko Kurimo |
| 2008 | ICASSP | On the state definition for a trainable excitation model in HMM-based speech synthesis. | Ranniery Maia, Tomoki Toda, Keiichi Tokuda, Shinichi Sakai, Shun Nakamura |
| 2008 | ICASSP | Acoustic modeling with contextual additive structure for HMM-based speech recognition. | Yoshihiko Nankaku, Kazuhiro Nakamura, Heiga Zen, Keiichi Tokuda |
| 2008 | ICASSP | Statistical approach to vocal tract transfer function estimation based on factor analyzed trajectory HMM. | Tomoki Toda, Keiichi Tokuda |
| 2008 | ICASSP | Performance evaluation of the speaker-independent HMM-based speech synthesis system "HTS 2007" for the Blizzard Challenge 2007. | Junichi Yamagishi, Takashi Nose, Heiga Zen, Tomoki Toda, Keiichi Tokuda |
| 2008 | Interspeech | Bayesian context clustering using cross valid prior distribution for HMM-based speech recognition. | Kei Hashimoto, Heiga Zen, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |
| 2008 | Interspeech | Speaker recognition based on variational Bayesian method. | Tatsuya Ito, Kei Hashimoto, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |
| 2008 | Interspeech | Unsupervised adaptation for HMM-based speech synthesis. | Simon King, Keiichi Tokuda, Heiga Zen, Junichi Yamagishi |
| 2008 | Interspeech | Acoustic modeling based on model structure annealing for speech recognition. | Sayaka Shiota, Kei Hashimoto, Heiga Zen, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |
| 2008 | Interspeech | Minimum generation error training with direct log spectral distortion on LSPs for HMM-based speech synthesis. | Yi-Jian Wu, Keiichi Tokuda |
| 2008 | Interspeech | Probabilistic answer selection based on conditional random fields for spoken dialog system. | Yoshitaka Yoshimi, Ryota Kakitsuba, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |
| 2008 | Interspeech | Simultaneous conversion of duration and spectrum based on statistical models including time-sequence matching. | Kaori Yutani, Yosuke Uto, Yoshihiko Nankaku, Tomoki Toda, Keiichi Tokuda |
| 2008 | Interspeech | Probabilistic feature mapping based on trajectory HMMs. | Heiga Zen, Yoshihiko Nankaku, Keiichi Tokuda |
| 2007 | ICASSP | Statistical Parametric Speech Synthesis. | Alan W. Black, Heiga Zen, Keiichi Tokuda |
| 2007 | ICASSP | Face Recognition using Hidden Markov Eigenface Models. | Yoshihiko Nankaku, Keiichi Tokuda |
| 2007 | Interspeech | A trainable excitation model for HMM-based speech synthesis. | Ranniery Maia, Tomoki Toda, Heiga Zen, Yoshihiko Nankaku, Keiichi Tokuda |
| 2007 | Interspeech | Model-space MLLR for trajectory HMMs. | Heiga Zen, Yoshihiko Nankaku, Keiichi Tokuda |
| 2006 | ICASSP | Face Recognition Based on Separable Lattice HMMS. | Daisuke Kurata, Yoshihiko Nankaku, Keiichi Tokuda, Tadashi Kitamura, Zoubin Ghahramani |
| 2006 | ICASSP | On the Use of Phonetic Information for Mapping from Articulatory Movements to Vocal Tract Spectrum. | Kenichi Nakamura, Tomoki Toda, Yoshihiko Nankaku, Keiichi Tokuda |
| 2006 | ICASSP | Hidden Semi-Markov Model Based Speech Recognition System using Weighted Finite-State Transducer. | Keiichiro Oura, Heiga Zen, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |
| 2006 | ICASSP | Estimating Trajectory Hmm Parameters Using Monte Carlo Em With Gibbs Sampler. | Heiga Zen, Yoshihiko Nankaku, Keiichi Tokuda, Tadashi Kitamura |
| 2006 | Interspeech | Reducing computation on parallel decoding using frame-wise confidence scores. | Tomohiro Hakamata, Akinobu Lee, Yoshihiko Nankaku, Keiichi Tokuda |
| 2006 | Interspeech | An HMM-based singing voice synthesis system. | Keijiro Saino, Heiga Zen, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |
| 2006 | Interspeech | Voice conversion based on mixtures of factor analyzers. | Yosuke Uto, Yoshihiko Nankaku, Tomoki Toda, Akinobu Lee, Keiichi Tokuda |
| 2006 | Interspeech | Speaker adaptation of trajectory HMMs using feature-space MLLR. | Heiga Zen, Yoshihiko Nankaku, Keiichi Tokuda, Tadashi Kitamura |
| 2005 | ICASSP | Minimum Classification Error Interactive Training for Speaker Identification. | Yusuke Kida, Hiroyoshi Yamamoto, Chiyomi Miyajima, Keiichi Tokuda, Tadashi Kitamura |
| 2005 | ICASSP | Sparse KPCA for Feature Extraction in Speech Recognition. | Amaro A. de Lima, Heiga Zen, Yoshihiko Nankaku, Keiichi Tokuda, Tadashi Kitamura, Fernando Gil Resende |
| 2005 | ICASSP | Spectral Conversion Based on Maximum Likelihood Estimation Considering Global Variance of Converted Parameter. | Tomoki Toda, Alan W. Black, Keiichi Tokuda |
| 2005 | Interspeech | HMM-based european Portuguese TTS system. | Maria Joo Barros, Ranniery Maia, Keiichi Tokuda, Fernando Gil Resende, Diamantino Freitas |
| 2005 | Interspeech | The blizzard challenge - 2005: evaluating corpus-based speech synthesis on common datasets. | Alan W. Black, Keiichi Tokuda |
| 2005 | Interspeech | Speech parameter generation algorithm considering global variance for HMM-based speech synthesis. | Tomoki Toda, Keiichi Tokuda |
| 2004 | ICASSP | Parameter sharing and minimum classification error training of mixtures of factor analyzers for speaker identification. | Hiroyoshi Yamamoto, Yoshihiko Nankaku, Chiyomi Miyajima, Keiichi Tokuda, Tadashi Kitamura |
| 2004 | ICASSP | A Viterbi algorithm for a trajectory model derived from HMM with explicit relationship between static and dynamic features. | Heiga Zen, Keiichi Tokuda, Tadashi Kitamura |
| 2004 | Interspeech | Deterministic annealing EM algorithm in parameter estimation for acoustic model. | Yohei Itaya, Heiga Zen, Yoshihiko Nankaku, Chiyomi Miyajima, Keiichi Tokuda, Tadashi Kitamura |
| 2004 | Interspeech | Decision-tree backing-off in HMM-based speech synthesis. | Shunsuke Kataoka, Nobuaki Mizutani, Keiichi Tokuda, Tadashi Kitamura |
| 2004 | Interspeech | Acoustic-to-articulatory inversion mapping with Gaussian mixture model. | Tomoki Toda, Alan W. Black, Keiichi Tokuda |
| 2004 | Interspeech | Constructing emotional speech synthesizers with limited speech database. | Heiga Zen, Tadashi Kitamura, Murtaza Bulut, Shrikanth S. Narayanan, Ryosuke Tsuzuki, Keiichi Tokuda |
| 2004 | Interspeech | Hidden semi-Markov model based speech synthesis. | Heiga Zen, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 2003 | ICASSP | Improving the performance of HMM-based very low bit rate speech coding. | Takahiro Hoshiya, Shinji Sako, Heiga Zen, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 2003 | ICASSP | Speech recognition using voice-characteristic-dependent acoustic models. | Hiroyuki Suzuki, Heiga Zen, Yoshihiko Nankaku, Chiyomi Miyajima, Keiichi Tokuda, Tadashi Kitamura |
| 2003 | ICASSP | A training method for average voice model based on shared decision tree context clustering and speaker adaptive training. | Junichi Yamagishi, Takashi Masuko, Keiichi Tokuda, Takao Kobayashi |
| 2003 | Interspeech | On the use of kernel PCA for feature extraction in speech recognition. | Amaro A. de Lima, Heiga Zen, Yoshihiko Nankaku, Chiyomi Miyajima, Keiichi Tokuda, Tadashi Kitamura |
| 2003 | Interspeech | Towards the development of a brazilian portuguese text-to-speech system based on HMM. | Ranniery Maia, Heiga Zen, Keiichi Tokuda, Tadashi Kitamura, Fernando Gil Vianna Resende Jr. |
| 2003 | Interspeech | Trajectory modeling based on HMMs with the explicit relationship between static and dynamic features. | Keiichi Tokuda, Heiga Zen, Tadashi Kitamura |
| 2003 | Interspeech | Decision tree-based simultaneous clustering of phonetic contexts, dimensions, and state positions for acoustic modeling. | Heiga Zen, Keiichi Tokuda, Tadashi Kitamura |
| 2002 | Interspeech | Eigenvoices for HMM-based speech synthesis. | Kengo Shichiri, Atsushi Sawabe, Takayoshi Yoshimura, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 2002 | Interspeech | A context clustering technique for average voice model in HMM-based speech synthesis. | Junichi Yamagishi, Masatsune Tamura, Takashi Masuko, Keiichi Tokuda, Takao Kobayashi |
| 2002 | Interspeech | Decision tree distribution tying based on a dimensional split technique. | Heiga Zen, Keiichi Tokuda, Tadashi Kitamura |
| 2001 | ICASSP | Speaker identification using Gaussian mixture models based on multi-space probability distribution. | Chiyomi Miyajima, Yosuke Hattori, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 2001 | ICASSP | Adaptation of pitch and spectrum for HMM-based speech synthesis using MLLR. | Masatsune Tamura, Takashi Masuko, Keiichi Tokuda, Takao Kobayashi |
| 2001 | Interspeech | Minimum classification error training for speaker identification using Gaussian mixture models based on multi-space probability distribution. | Chiyomi Miyajima, Keiichi Tokuda, Tadashi Kitamura |
| 2001 | Interspeech | A robust speaker verification system against imposture using an HMM-based speech synthesis system. | Takayuki Satoh, Takashi Masuko, Takao Kobayashi, Keiichi Tokuda |
| 2001 | Interspeech | Text-to-speech synthesis with arbitrary speaker's voice from average voice. | Masatsune Tamura, Takashi Masuko, Keiichi Tokuda, Takao Kobayashi |
| 2001 | Interspeech | Mixed excitation for HMM-based speech synthesis. | Takayoshi Yoshimura, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 2001 | ISCAS | Fast convergence transversal adaptive filtering algorithm for impulsive environment based on T distribution assumption. | Junibakti Sanubari, Keiichi Tokuda |
| 2000 | ICASSP | Speech parameter generation algorithms for HMM-based speech synthesis. | Keiichi Tokuda, Takayoshi Yoshimura, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 2000 | ICIP | Normalized Training for HMM-Based Visual Speech Recognition. | Yoshihiko Nankaku, Keiichi Tokuda, Tadashi Kitamura, Takao Kobayashi |
| 2000 | Interspeech | Imposture using synthetic speech against speaker verification based on spectrum and pitch. | Takashi Masuko, Keiichi Tokuda, Takao Kobayashi |
| 2000 | Interspeech | Audio-visual speech recognition using MCE-based hmms and model-dependent stream weights. | Chiyomi Miyajima, Keiichi Tokuda, Tadashi Kitamura |
| 2000 | Interspeech | HMM-based text-to-audio-visual speech synthesis. | Shinji Sako, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 1999 | ICASSP | Hidden Markov models based on multi-space probability distribution for pitch pattern modeling. | Keiichi Tokuda, Takashi Masuko, Noboru Miyazaki, Takao Kobayashi |
| 1999 | ICIP | Image Modeling Using Two Dimensional Exponential Systems. | Junibakti Sanubari, Keiichi Tokuda |
| 1999 | ICIP | Location Normalization of HMM-Based Lip Reading: Experiments for the M2VTS Database. | Oscar Vanegas, Keiichi Tokuda, Tadashi Kitamura |
| 1999 | Interspeech | On the security of HMM-based speaker verification systems against imposture using synthetic speech. | Takashi Masuko, Takafumi Hitotsumatsu, Keiichi Tokuda, Takao Kobayashi |
| 1999 | Interspeech | Intensity- and location-normalized training for HMM-based visual speech recognition. | Yoshihiko Nankaku, Keiichi Tokuda, Tadashi Kitamura |
| 1999 | Interspeech | Simultaneous modeling of spectrum, pitch and duration in HMM-based speech synthesis. | Takayoshi Yoshimura, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 1998 | ICASSP | A wideband CELP speech coder at 16 kbit/s based on mel-generalized cepstral analysis. | Kazuhito Koishida, Gou Hirabayashi, Keiichi Tokuda, Takao Kobayashi |
| 1998 | ICASSP | Text-to-visual speech synthesis based on parameter generation from HMM. | Takashi Masuko, Takao Kobayashi, Masatsune Tamura, Jun Masubuchi, Keiichi Tokuda |
| 1998 | ICASSP | A very low bit rate speech coder using HMM-based speech recognition/synthesis techniques. | Keiichi Tokuda, Takashi Masuko, Jun Hiroi, Takao Kobayashi, Tadashi Kitamura |
| 1998 | Interspeech | A 16 kbit/s wideband CELP coder using MEL-generalized cepstral analysis and its subjective evaluation. | Kazuhito Koishida, Gou Hirabayashi, Keiichi Tokuda, Takao Kobayashi |
| 1998 | Interspeech | A very low bit rate speech coder using HMM with speaker adaptation. | Takashi Masuko, Keiichi Tokuda, Takao Kobayashi |
| 1998 | Interspeech | HMM-based visual speech recognition using intensity and location normalization. | Oscar Vanegas, Akiji Tanaka, Keiichi Tokuda, Tadashi Kitamura |
| 1998 | Interspeech | Duration modeling for HMM-based speech synthesis. | Takayoshi Yoshimura, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 1997 | ICASSP | Efficient encoding of mel-generalized cepstrum for CELP coders. | Kazuhito Koishida, Keiichi Tokuda, Takao Kobayashi, Satoshi Imai |
| 1997 | ICASSP | Voice characteristics conversion for HMM-based speech synthesis system. | Takashi Masuko, Keiichi Tokuda, Takao Kobayashi, Satoshi Imai |
| 1997 | Interspeech | HMM compensation for noisy speech recognition based on cepstral parameter generation. | Takao Kobayashi, Takashi Masuko, Keiichi Tokuda |
| 1997 | Interspeech | Speaker interpolation in HMM-based speech synthesis system. | Takayoshi Yoshimura, Takashi Masuko, Keiichi Tokuda, Takao Kobayashi, Tadashi Kitamura |
| 1996 | ICASSP | Speech synthesis using HMMs with dynamic features. | Takashi Masuko, Keiichi Tokuda, Takao Kobayashi, Satoshi Imai |
| 1996 | ICASSP | Robust two dimensional spectral estimation based on AR model excited by a t-distribution process. | Junibakti Sanubari, Keiichi Tokuda, Mahoki Onoda |
| 1996 | Interspeech | CELP coding system based on mel-generalized cepstral analysis. | Kazuhito Koishida, Keiichi Tokuda, Takao Kobayashi, Satoshi Imai |
| 1995 | ICASSP | CELP coding based on mel-cepstral analysis. | Kazuhito Koishida, Keiichi Tokuda, Takao Kobayashi, Satoshi Imai |
| 1995 | ICASSP | Speech parameter generation from HMM using dynamic features. | Keiichi Tokuda, Takao Kobayashi, Satoshi Imai |
| 1995 | Interspeech | An algorithm for speech parameter generation from continuous mixture HMMs with dynamic features. | Keiichi Tokuda, Takashi Masuko, Tetsuya Yamada, Takao Kobayashi, Satoshi Imai |
| 1994 | ICASSP | Robust recursive spectral estimation based on AR model excited by a t-distribution process. | Junibakti Sanubari, Keiichi Tokuda, Mahoki Onoda |
| 1994 | ICASSP | Speech coding based on adaptive mel-cepstral analysis. | Keiichi Tokuda, Hidetoshi Matsumura, Takao Kobayashi, Satoshi Imai |
| 1994 | Interspeech | Speech coding based on adaptive MEL-cepstral analysis for noisy channels. | Kazuhito Koishida, Keiichi Tokuda, Takao Kobayashi, Satoshi Imai |
| 1994 | Interspeech | Mel-generalized cepstral analysis - a unified approach to speech spectral estimation. | Keiichi Tokuda, Takao Kobayashi, Takashi Masuko, Satoshi Imai |
| 1994 | ISCAS | AR Spectrum Estimation Based on Wavelet Representation. | Fernando Gil Resende, Keiichi Tokuda, Mineo Kaneko |
| 1992 | ICASSP | An adaptive algorithm for mel-cepstral analysis of speech. | Toshiaki Fukada, Keiichi Tokuda, Takao Kobayashi, Satoshi Imai |
| 1992 | ICASSP | Design of stable two-dimensional IIR digital filters with arbitrary magnitude function. | Takao Kobayashi, Kazuyoshi Fukushi, Keiichi Tokuda, Satoshi Imai |
| 1992 | ICASSP | Spectral estimation based on AR-model excited by t-distribution process. | Junibakti Sanubari, Keiichi Tokuda, Mahoki Onoda |
| 1990 | ICASSP | Adaptive filtering based on cepstral representation-adaptive cepstral analysis of speech. | Keiichi Tokuda, Takao Kobayashi, Shoji Shiomoto, Satoshi Imai |
| 1990 | Interspeech | Generalized cepstral analysis of speech - unified approach to LPC and cepstral method. | Keiichi Tokuda, Takao Kobayashi, Satoshi Imai |