| 2021 | ICASSP | High-Intelligibility Speech Synthesis for Dysarthric Speakers with LPCNet-Based TTS and CycleVAE-Based VC. | Keisuke Matsubara, Takuma Okamoto, Ryoichi Takashima, Tetsuya Takiguchi, Tomoki Toda, Yoshinori Shiga, Hisashi Kawai |
| 2021 | ICASSP | Noise Level Limited Sub-Modeling for Diffusion Probabilistic Vocoders. | Takuma Okamoto, Tomoki Toda, Yoshinori Shiga, Hisashi Kawai |
| 2020 | ICASSP | Transformer-Based Text-to-Speech with Weighted Forced Attention. | Takuma Okamoto, Tomoki Toda, Yoshinori Shiga, Hisashi Kawai |
| 2019 | ASRU | Tacotron-Based Acoustic Model Using Phoneme Alignment for Practical Neural Text-to-Speech Systems. | Takuma Okamoto, Tomoki Toda, Yoshinori Shiga, Hisashi Kawai |
| 2019 | ICASSP | Investigations of Real-time Gaussian Fftnet and Parallel Wavenet Neural Vocoders with Simple Acoustic Features. | Takuma Okamoto, Tomoki Toda, Yoshinori Shiga, Hisashi Kawai |
| 2019 | Interspeech | Duration Modeling with Global Phoneme-Duration Vectors. | Jinfu Ni, Yoshinori Shiga, Hisashi Kawai |
| 2019 | Interspeech | Real-Time Neural Text-to-Speech with Sequence-to-Sequence Acoustic Model and WaveGlow or Single Gaussian WaveRNN Vocoders. | Takuma Okamoto, Tomoki Toda, Yoshinori Shiga, Hisashi Kawai |
| 2018 | ICASSP | An Investigation of Subband Wavenet Vocoder Covering Entire Audible Frequency Range with Limited Acoustic Features. | Takuma Okamoto, Kentaro Tachibana, Tomoki Toda, Yoshinori Shiga, Hisashi Kawai |
| 2018 | ICASSP | An Investigation of Noise Shaping with Perceptual Weighting for Wavenet-Based Speech Generation. | Kentaro Tachibana, Tomoki Toda, Yoshinori Shiga, Hisashi Kawai |
| 2018 | Interspeech | Multilingual Grapheme-to-Phoneme Conversion with Global Character Vectors. | Jinfu Ni, Yoshinori Shiga, Hisashi Kawai |
| 2017 | ASRU | Subband wavenet with overlapped single-sideband filterbanks. | Takuma Okamoto, Kentaro Tachibana, Tomoki Toda, Yoshinori Shiga, Hisashi Kawai |
| 2017 | Interspeech | Global Syllable Vectors for Building TTS Front-End with Deep Learning. | Jinfu Ni, Yoshinori Shiga, Hisashi Kawai |
| 2016 | Interspeech | Using Zero-Frequency Resonator to Extract Multilingual Intonation Structure. | Jinfu Ni, Yoshinori Shiga, Hisashi Kawai |
| 2016 | Interspeech | Model Integration for HMM- and DNN-Based Speech Synthesis Using Product-of-Experts Framework. | Kentaro Tachibana, Tomoki Toda, Yoshinori Shiga, Hisashi Kawai |
| 2015 | ICASSP | Extraction of pitch register from expressive speech in Japanese. | Jinfu Ni, Yoshinori Shiga, Chiori Hori |
| 2015 | Interspeech | Entropy-based sentence selection for speech synthesis using phonetic and prosodic contexts. | Takashi Nose, Yusuke Arao, Takao Kobayashi, Komei Sugiura, Yoshinori Shiga, Akinori Ito |
| 2015 | Interspeech | HMM based myanmar text to speech system. | Ye Kyaw Thu, Win Pa Pa, Jinfu Ni, Yoshinori Shiga, Andrew M. Finch, Chiori Hori, Hisashi Kawai, Eiichiro Sumita |
| 2014 | ICRA | Non-monologue HMM-based speech synthesis for service robots: A cloud robotics approach. | Komei Sugiura, Yoshinori Shiga, Hisashi Kawai, Teruhisa Misu, Chiori Hori |
| 2013 | Interspeech | A targets-based superpositional model of fundamental frequency contours applied to HMM-based speech synthesis. | Jinfu Ni, Yoshinori Shiga, Chiori Hori, Yutaka Kidawara |
| 2013 | Interspeech | Improvements to HMM-based speech synthesis based on parameter generation with rich context models. | Shinnosuke Takamichi, Tomoki Toda, Yoshinori Shiga, Sakriani Sakti, Graham Neubig, Satoshi Nakamura |
| 2013 | MDM | Multilingual Speech-to-Speech Translation System: VoiceTra. | Shigeki Matsuda, Xinhui Hu, Yoshinori Shiga, Hideki Kashioka, Chiori Hori, Keiji Yasuda, Hideo Okuma, Masao Uchiyama, Eiichiro Sumita, Hisashi Kawai, Satoshi Nakamura |
| 2012 | ICASSP | Effect of anti-aliasing filtering on the quality of speech from an HMM-based synthesizer. | Yoshinori Shiga |
| 2012 | Interspeech | An Evaluation of Parameter Generation Methods with Rich Context Models in HMM-Based Speech Synthesis. | Shinnosuke Takamichi, Tomoki Toda, Yoshinori Shiga, Hisashi Kawai, Sakriani Sakti, Satoshi Nakamura |
| 2011 | SIGdial | Toward Construction of Spoken Dialogue System that Evokes Users' Spontaneous Backchannels. | Teruhisa Misu, Etsuo Mizukami, Yoshinori Shiga, Shinichi Kawamoto, Hisashi Kawai, Satoshi Nakamura |
| 2010 | Interspeech | Improved training of excitation for HMM-based parametric speech synthesis. | Yoshinori Shiga, Tomoki Toda, Shinsuke Sakai, Hisashi Kawai |
| 2009 | Interspeech | Pulse density representation of spectrum for statistical speech processing. | Yoshinori Shiga |
| 2004 | Interspeech | Estimating detailed spectral envelopes using articulatory clustering. | Yoshinori Shiga, Simon King |
| 2004 | Interspeech | Source-filter separation for articulation-to-speech synthesis. | Yoshinori Shiga, Simon King |
| 2003 | Interspeech | Estimating the spectral envelope of voiced speech using multi-frame analysis. | Yoshinori Shiga, Simon King |
| 2003 | Interspeech | Estimation of voice source and vocal tract characteristics based on multi-frame analysis. | Yoshinori Shiga, Simon King |
| 1998 | Interspeech | Segmental duration control based on an articulatory model. | Yoshinori Shiga, Hiroshi Matsuura, Tsuneo Nitta |
| 1994 | Interspeech | A novel segment-concatenation algorithm for a cepstrum-based synthesizer. | Yoshinori Shiga, Yoshiyuki Hara, Tsuneo Nitta |