| 2025 | Interspeech | PeriodCodec: A Pitch-Controllable Neural Audio Codec Using Periodic Signals for Singing Voice Synthesis. | Masato Takagi, Miku Nishihara, Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2024 | ICASSP | PeriodGrad: Towards Pitch-Controllable Neural Vocoder Based on a Diffusion Probabilistic Model. | Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2023 | ICASSP | Singing Voice Synthesis Based on a Musical Note Position-Aware Attention Mechanism. | Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2023 | ICASSP | Embedding a Differentiable Mel-Cepstral Synthesis Filter to a Neural Speech Synthesis System. | Takenori Yoshimura, Shinji Takaki, Kazuhiro Nakamura, Keiichiro Oura, Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2022 | ICASSP | Autoregressive Variational Autoencoder with a Hidden Semi-Markov Model-Based Structured Attention for Speech Synthesis. | Takato Fujimoto, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2021 | ICASSP | Periodnet: A Non-Autoregressive Waveform Generation Model with a Structure Separating Periodic and Aperiodic Components. | Yukiya Hono, Shinji Takaki, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2020 | ICASSP | Semi-Supervised Learning Based on Hierarchical Generative Models for End-to-End Speech Synthesis. | Takato Fujimoto, Shinji Takaki, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2020 | ICASSP | Fast and High-Quality Singing Voice Synthesis System Based on Convolutional Neural Networks. | Kazuhiro Nakamura, Shinji Takaki, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2020 | Interspeech | Hierarchical Multi-Grained Generative Model for Expressive Speech Synthesis. | Yukiya Hono, Kazuna Tsuboi, Kei Sawada, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2019 | ICASSP | Singing Voice Synthesis Based on Generative Adversarial Networks. | Yukiya Hono, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2019 | ICASSP | Speaker-dependent Wavenet-based Delay-free Adpcm Speech Coding. | Takenori Yoshimura, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2018 | ICASSP | Image Recognition Based on Separable Lattice Hmms Using a Deep Neural Network for Output Probability Distributions. | Eiji Ichikawa, Kei Sawada, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2018 | ICASSP | Statistical Voice Conversion Based on Wavenet. | Jumpei Niwa, Takenori Yoshimura, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2017 | ICASSP | Image recognition based on discriminative models using features generated from separable lattice HMMS. | Yoshinari Tsuzuki, Kei Sawada, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2017 | Interspeech | Articulatory Text-to-Speech Synthesis Using the Digital Waveguide Mesh Driven by a Deep Neural Network. | Amelia Jane Gully, Takenori Yoshimura, Damian T. Murphy, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2016 | ICASSP | Trajectory training considering global variance for speech synthesis based on neural networks. | Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2016 | ICASSP | Privacy-preserving sound to degrade automatic speaker verification performance. | Kei Hashimoto, Junichi Yamagishi, Isao Echizen |
| 2016 | Interspeech | Redefining the Linguistic Context Feature Set for HMM and DNN TTS Through Position and Parsing. | Rasmus Dall, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2016 | Interspeech | Voice Conversion Based on Trajectory Model Training of Neural Networks Considering Global Variance. | Naoki Hosaka, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2016 | Interspeech | Singing Voice Synthesis Based on Deep Neural Networks. | Masanari Nishimura, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2015 | ICASSP | The effect of neural networks in statistical parametric speech synthesis. | Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2015 | Interspeech | Simultaneous optimization of multiple tree structures for factor analyzed HMM-based speech synthesis. | Takenori Yoshimura, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2014 | ICASSP | Integration of speaker and pitch adaptive training for HMM-based singing voice synthesis. | Kanako Shirota, Kazuhiro Nakamura, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2014 | Interspeech | A mel-cepstral analysis technique restoring high frequency components from low-sampling-rate speech. | Kazuhiro Nakamura, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2013 | ICASSP | Separable lattice 2-D HMMS introducing state duration control for recognition of images with various variations. | Takaya Makino, Shinji Takaki, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2013 | ICASSP | Integration of acoustic modeling and mel-cepstral analysis for HMM-based speech synthesis. | Kazuhiro Nakamura, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2012 | ICASSP | Face recognition based on separable lattice 2-D HMMS using variational bayesian method. | Kei Sawada, Akira Tamamori, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2012 | ICASSP | A model structure integration based on a Bayesian framework for speech recognition. | Sayaka Shiota, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2012 | Interspeech | A Bayesian Approach to Speaker Recognition Based on GMMs Using Multiple Model Structures. | Takafumi Hattori, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2011 | ICASSP | An analysis of machine translation and speech synthesis in speech-to-speech translation system. | Kei Hashimoto, Junichi Yamagishi, William J. Byrne, Simon King, Keiichi Tokuda |
| 2011 | Interspeech | Multi-Speaker Modeling with Shared Prior Distributions and Model Structures for Bayesian Speech Synthesis. | Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2010 | EAMT | A Deterministic Annealing-Based Training Algorithm For Statistical Machine Translation Models. | Pascual Martnez-Gmez, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda, Germn Sanchis-Trilles |
| 2009 | ICASSP | A Bayesian approach to HMM-based speech synthesis. | Kei Hashimoto, Heiga Zen, Yoshihiko Nankaku, Takashi Masuko, Keiichi Tokuda |
| 2009 | Interspeech | A Bayesian approach to Hidden Semi-Markov Model based speech synthesis. | Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2009 | Interspeech | Deterministic annealing based training algorithm for Bayesian speech recognition. | Sayaka Shiota, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2008 | Interspeech | Bayesian context clustering using cross valid prior distribution for HMM-based speech recognition. | Kei Hashimoto, Heiga Zen, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |
| 2008 | Interspeech | Speaker recognition based on variational Bayesian method. | Tatsuya Ito, Kei Hashimoto, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |
| 2008 | Interspeech | Acoustic modeling based on model structure annealing for speech recognition. | Sayaka Shiota, Kei Hashimoto, Heiga Zen, Yoshihiko Nankaku, Akinobu Lee, Keiichi Tokuda |