| 2018 | ICASSP | Simultaneous Speech Recognition and Acoustic Event Detection Using an LSTM-CTC Acoustic Model and a WFST Decoder. | Hiroshi Fujimura, Manabu Nagao, Takashi Masuko |
| 2017 | ASRU | Computational cost reduction of long short-term memory based on simultaneous compression of input and hidden state. | Takashi Masuko |
| 2013 | Interspeech | N-best rescoring by phoneme classifiers using subclass adaboost algorithm. | Hiroshi Fujimura, Yusuke Shinohara, Takashi Masuko |
| 2011 | ASRU | N-Best rescoring by adaboost phoneme classifiers for isolated word recognition. | Hiroshi Fujimura, Masanobu Nakamura, Yusuke Shinohara, Takashi Masuko |
| 2010 | ICASSP | Covariance clustering on Riemannian manifolds for acoustic model compression. | Yusuke Shinohara, Takashi Masuko, Masami Akamine |
| 2010 | Interspeech | A duration modeling technique with incremental speech rate normalization. | Hiroshi Fujimura, Takashi Masuko, Mitsuyoshi Tachimori |
| 2009 | ICASSP | A Bayesian approach to HMM-based speech synthesis. | Kei Hashimoto, Heiga Zen, Yoshihiko Nankaku, Takashi Masuko, Keiichi Tokuda |
| 2009 | Interspeech | Robust F0 estimation based on log-time scale autocorrelation and its application to Mandarin tone recognition. | Yusuke Kida, Masaru Sakai, Takashi Masuko, Akinori Kawamura |
| 2008 | ICASSP | Feature enhancement by speaker-normalized splice for robust speech recognition. | Yusuke Shinohara, Takashi Masuko, Masami Akamine |
| 2006 | ICASSP | Speech Recognition Using Syllable Duration Ratio Model. | Masahide Ariu, Takashi Masuko, Shinichi Tanaka, Akinori Kawamura |
| 2005 | Interspeech | Performance evaluation of style adaptation for hidden semi-Markov model based speech synthesis. | Makoto Tachibana, Junichi Yamagishi, Takashi Masuko, Takao Kobayashi |
| 2004 | ICASSP | Speaking style adaptation using context clustering decision tree for HMM-based speech synthesis. | Junichi Yamagishi, Makoto Tachibana, Takashi Masuko, Takao Kobayashi |
| 2004 | Interspeech | A style control technique for HMM-based speech synthesis. | Takashi Masuko, Takao Kobayashi, Keisuke Miyanaga |
| 2004 | Interspeech | MLLR adaptation for hidden semi-Markov model based speech synthesis. | Junichi Yamagishi, Takashi Masuko, Takao Kobayashi |
| 2004 | Interspeech | Hidden semi-Markov model based speech synthesis. | Heiga Zen, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 2003 | ICASSP | Improving the performance of HMM-based very low bit rate speech coding. | Takahiro Hoshiya, Shinji Sako, Heiga Zen, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 2003 | ICASSP | A training method for average voice model based on shared decision tree context clustering and speaker adaptive training. | Junichi Yamagishi, Takashi Masuko, Keiichi Tokuda, Takao Kobayashi |
| 2003 | Interspeech | Modeling of various speaking styles and emotions for HMM-based speech synthesis. | Junichi Yamagishi, Koji Onishi, Takashi Masuko, Takao Kobayashi |
| 2002 | ICASSP | Fundamental frequency estimation based on instantaneous frequency amplitude spectrum. | Tomohiro Tanaka, Takao Kobayashi, Dhany Arifianto, Takashi Masuko |
| 2002 | Interspeech | Eigenvoices for HMM-based speech synthesis. | Kengo Shichiri, Atsushi Sawabe, Takayoshi Yoshimura, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 2002 | Interspeech | A context clustering technique for average voice model in HMM-based speech synthesis. | Junichi Yamagishi, Masatsune Tamura, Takashi Masuko, Keiichi Tokuda, Takao Kobayashi |
| 2001 | ICASSP | Speaker identification using Gaussian mixture models based on multi-space probability distribution. | Chiyomi Miyajima, Yosuke Hattori, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 2001 | ICASSP | Adaptation of pitch and spectrum for HMM-based speech synthesis using MLLR. | Masatsune Tamura, Takashi Masuko, Keiichi Tokuda, Takao Kobayashi |
| 2001 | Interspeech | A robust speaker verification system against imposture using an HMM-based speech synthesis system. | Takayuki Satoh, Takashi Masuko, Takao Kobayashi, Keiichi Tokuda |
| 2001 | Interspeech | Text-to-speech synthesis with arbitrary speaker's voice from average voice. | Masatsune Tamura, Takashi Masuko, Keiichi Tokuda, Takao Kobayashi |
| 2001 | Interspeech | Mixed excitation for HMM-based speech synthesis. | Takayoshi Yoshimura, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 2000 | ICASSP | Speech parameter generation algorithms for HMM-based speech synthesis. | Keiichi Tokuda, Takayoshi Yoshimura, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 2000 | Interspeech | Imposture using synthetic speech against speaker verification based on spectrum and pitch. | Takashi Masuko, Keiichi Tokuda, Takao Kobayashi |
| 2000 | Interspeech | HMM-based text-to-audio-visual speech synthesis. | Shinji Sako, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 1999 | ICASSP | Hidden Markov models based on multi-space probability distribution for pitch pattern modeling. | Keiichi Tokuda, Takashi Masuko, Noboru Miyazaki, Takao Kobayashi |
| 1999 | Interspeech | On the security of HMM-based speaker verification systems against imposture using synthetic speech. | Takashi Masuko, Takafumi Hitotsumatsu, Keiichi Tokuda, Takao Kobayashi |
| 1999 | Interspeech | Text-to-audio-visual speech synthesis based on parameter generation from HMM. | Masatsune Tamura, Shigekazu Kondo, Takashi Masuko, Takao Kobayashi |
| 1999 | Interspeech | Simultaneous modeling of spectrum, pitch and duration in HMM-based speech synthesis. | Takayoshi Yoshimura, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 1998 | ICASSP | Text-to-visual speech synthesis based on parameter generation from HMM. | Takashi Masuko, Takao Kobayashi, Masatsune Tamura, Jun Masubuchi, Keiichi Tokuda |
| 1998 | ICASSP | A very low bit rate speech coder using HMM-based speech recognition/synthesis techniques. | Keiichi Tokuda, Takashi Masuko, Jun Hiroi, Takao Kobayashi, Tadashi Kitamura |
| 1998 | Interspeech | A very low bit rate speech coder using HMM with speaker adaptation. | Takashi Masuko, Keiichi Tokuda, Takao Kobayashi |
| 1998 | Interspeech | Duration modeling for HMM-based speech synthesis. | Takayoshi Yoshimura, Keiichi Tokuda, Takashi Masuko, Takao Kobayashi, Tadashi Kitamura |
| 1997 | ICASSP | Voice characteristics conversion for HMM-based speech synthesis system. | Takashi Masuko, Keiichi Tokuda, Takao Kobayashi, Satoshi Imai |
| 1997 | Interspeech | HMM compensation for noisy speech recognition based on cepstral parameter generation. | Takao Kobayashi, Takashi Masuko, Keiichi Tokuda |
| 1997 | Interspeech | Speaker interpolation in HMM-based speech synthesis system. | Takayoshi Yoshimura, Takashi Masuko, Keiichi Tokuda, Takao Kobayashi, Tadashi Kitamura |
| 1996 | ICASSP | Speech synthesis using HMMs with dynamic features. | Takashi Masuko, Keiichi Tokuda, Takao Kobayashi, Satoshi Imai |
| 1995 | Interspeech | An algorithm for speech parameter generation from continuous mixture HMMs with dynamic features. | Keiichi Tokuda, Takashi Masuko, Tetsuya Yamada, Takao Kobayashi, Satoshi Imai |
| 1994 | Interspeech | Mel-generalized cepstral analysis - a unified approach to speech spectral estimation. | Keiichi Tokuda, Takao Kobayashi, Takashi Masuko, Satoshi Imai |