| 2023 | ICASSP | Embedding a Differentiable Mel-Cepstral Synthesis Filter to a Neural Speech Synthesis System. | Takenori Yoshimura, Shinji Takaki, Kazuhiro Nakamura, Keiichiro Oura, Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2021 | ICASSP | Periodnet: A Non-Autoregressive Waveform Generation Model with a Structure Separating Periodic and Aperiodic Components. | Yukiya Hono, Shinji Takaki, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2020 | ICASSP | Semi-Supervised Learning Based on Hierarchical Generative Models for End-to-End Speech Synthesis. | Takato Fujimoto, Shinji Takaki, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2020 | ICASSP | Fast and High-Quality Singing Voice Synthesis System Based on Convolutional Neural Networks. | Kazuhiro Nakamura, Shinji Takaki, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2019 | ICASSP | STFT Spectral Loss for Training a Neural Speech Waveform Model. | Shinji Takaki, Toru Nakashika, Xin Wang, Junichi Yamagishi |
| 2019 | ICASSP | Neural Source-filter-based Waveform Model for Statistical Parametric Speech Synthesis. | Xin Wang, Shinji Takaki, Junichi Yamagishi |
| 2019 | ICASSP | Investigation of Enhanced Tacotron Text-to-speech Synthesis Systems with Self-attention for Pitch Accent Language. | Yusuke Yasuda, Xin Wang, Shinji Takaki, Junichi Yamagishi |
| 2019 | Interspeech | Does the Lombard Effect Improve Emotional Communication in Noise? - Analysis of Emotional Speech Acted in Noise. | Yi Zhao, Atsushi Ando, Shinji Takaki, Junichi Yamagishi, Satoshi Kobashikawa |
| 2018 | ICASSP | A Comparison of Recent Waveform Generation and Acoustic Modeling Methods for Neural-Network-Based Speech Synthesis. | Xin Wang, Jaime Lorenzo-Trueba, Shinji Takaki, Lauri Juvela, Junichi Yamagishi |
| 2017 | ICASSP | Adapting and controlling DNN-based speech synthesis using input codes. | Hieu-Thi Luong, Shinji Takaki, Gustav Eje Henter, Junichi Yamagishi |
| 2017 | ICASSP | An autoregressive recurrent mixture density network for parametric speech synthesis. | Xin Wang, Shinji Takaki, Junichi Yamagishi |
| 2017 | Interspeech | Generative Adversarial Network-Based Postfilter for STFT Spectrograms. | Takuhiro Kaneko, Shinji Takaki, Hirokazu Kameoka, Junichi Yamagishi |
| 2017 | Interspeech | Complex-Valued Restricted Boltzmann Machine for Direct Learning of Frequency Spectra. | Toru Nakashika, Shinji Takaki, Junichi Yamagishi |
| 2017 | Interspeech | Direct Modeling of Frequency Spectra and Waveform Generation Based on Phase Recovery for DNN-Based Speech Synthesis. | Shinji Takaki, Hirokazu Kameoka, Junichi Yamagishi |
| 2017 | Interspeech | An RNN-Based Quantized F0 Model with Multi-Tier Feedback Links for Text-to-Speech Synthesis. | Xin Wang, Shinji Takaki, Junichi Yamagishi |
| 2016 | ICASSP | A deep auto-encoder based low-dimensional feature extraction from FFT spectral envelopes for statistical parametric speech synthesis. | Shinji Takaki, Junichi Yamagishi |
| 2016 | Interspeech | Using Text and Acoustic Features in Predicting Glottal Excitation Waveforms for Parametric Speech Synthesis with Recurrent Neural Networks. | Lauri Juvela, Xin Wang, Shinji Takaki, Manu Airaksinen, Junichi Yamagishi, Paavo Alku |
| 2016 | Interspeech | Speech Enhancement for a Noise-Robust Text-to-Speech Synthesis System Using Deep Recurrent Neural Networks. | Cassia Valentini-Botinhao, Xin Wang, Shinji Takaki, Junichi Yamagishi |
| 2016 | Interspeech | Enhance the Word Vector with Prosodic Information for the Recurrent Neural Network Based TTS System. | Xin Wang, Shinji Takaki, Junichi Yamagishi |
| 2015 | Interspeech | Multiple feed-forward deep neural networks for statistical parametric speech synthesis. | Shinji Takaki, Sangjin Kim, Junichi Yamagishi, JongJin Kim |
| 2013 | ICASSP | Separable lattice 2-D HMMS introducing state duration control for recognition of images with various variations. | Takaya Makino, Shinji Takaki, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2013 | ICASSP | Contextual partial additive structure for HMM-based speech synthesis. | Shinji Takaki, Yoshihiko Nankaku, Keiichi Tokuda |
| 2011 | ICASSP | An optimization algorithm of independent mean and variance parameter tying structures for HMM-based speech synthesis. | Shinji Takaki, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |