| 2024 | ICASSP | Electrolaryngeal Speech Intelligibility Enhancement through Robust Linguistic Encoders. | Lester Phillip Violeta, Wen-Chin Huang, Ding Ma, Ryuichi Yamamoto, Kazuhiro Kobayashi, Tomoki Toda |
| 2023 | ICASSP | Low-Latency Electrolaryngeal Speech Enhancement Based on Fastspeech2-Based Voice Conversion and Self-Supervised Speech Representation. | Kazuhiro Kobayashi, Tomoki Hayashi, Tomoki Toda |
| 2022 | ICASSP | An Investigation of Streaming Non-Autoregressive sequence-to-sequence Voice Conversion. | Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2021 | ASRU | Mandarin Electrolaryngeal Speech Voice Conversion with Sequence-to-Sequence Modeling. | Ming-Chi Yen, Wen-Chin Huang, Kazuhiro Kobayashi, Yu-Huai Peng, Shu-Wei Tsai, Yu Tsao, Tomoki Toda, Jyh-Shing Roger Jang, Hsin-Min Wang |
| 2021 | ICASSP | Non-Autoregressive Sequence-To-Sequence Voice Conversion. | Tomoki Hayashi, Wen-Chin Huang, Kazuhiro Kobayashi, Tomoki Toda |
| 2021 | ICASSP | Crank: An Open-Source Software for Nonparallel Voice Conversion Based on Vector-Quantized Variational Autoencoder. | Kazuhiro Kobayashi, Wen-Chin Huang, Yi-Chiao Wu, Patrick Lumban Tobing, Tomoki Hayashi, Tomoki Toda |
| 2021 | Interspeech | A Preliminary Study of a Two-Stage Paradigm for Preserving Speaker Identity in Dysarthric Voice Conversion. | Wen-Chin Huang, Kazuhiro Kobayashi, Yu-Huai Peng, Ching-Feng Liu, Yu Tsao, Hsin-Min Wang, Tomoki Toda |
| 2020 | ICASSP | Efficient Shallow Wavenet Vocoder Using Multiple Samples Output Based on Laplacian Distribution and Linear Prediction. | Patrick Lumban Tobing, Yi-Chiao Wu, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2020 | Interspeech | Intelligibility Enhancement Based on Speech Waveform Modification Using Hearing Impairment. | Shu Hikosaka, Shogo Seki, Tomoki Hayashi, Kazuhiro Kobayashi, Kazuya Takeda, Hideki Banno, Tomoki Toda |
| 2020 | Interspeech | Cyclic Spectral Modeling for Unsupervised Unit Discovery into Voice Conversion with Excitation and Waveform Modeling. | Patrick Lumban Tobing, Tomoki Hayashi, Yi-Chiao Wu, Kazuhiro Kobayashi, Tomoki Toda |
| 2019 | ASSETS | Development of a Real-time Bionic Voice Generation System based on Statistical Excitation Prediction. | Farzaneh Ahmadi, Kazuhiro Kobayashi, Tomoki Toda |
| 2019 | ICASSP | Voice Conversion with Cyclic Recurrent Neural Network and Fine-tuned Wavenet Vocoder. | Patrick Lumban Tobing, Yi-Chiao Wu, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2019 | Interspeech | Investigation of F0 Conditioning and Fully Convolutional Networks in Variational Autoencoder Based Voice Conversion. | Wen-Chin Huang, Yi-Chiao Wu, Chen-Chou Lo, Patrick Lumban Tobing, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda, Yu Tsao, Hsin-Min Wang |
| 2019 | Interspeech | Robustness of Statistical Voice Conversion Based on Direct Waveform Modification Against Background Sounds. | Yusuke Kurita, Kazuhiro Kobayashi, Kazuya Takeda, Tomoki Toda |
| 2019 | Interspeech | Non-Parallel Voice Conversion with Cyclic Variational Autoencoder. | Patrick Lumban Tobing, Yi-Chiao Wu, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2019 | Interspeech | Quasi-Periodic WaveNet Vocoder: A Pitch Dependent Dilated Convolution Model for Parametric Speech Generation. | Yi-Chiao Wu, Tomoki Hayashi, Patrick Lumban Tobing, Kazuhiro Kobayashi, Tomoki Toda |
| 2018 | Interspeech | Collapsed Speech Segment Detection and Suppression for WaveNet Vocoder. | Yi-Chiao Wu, Kazuhiro Kobayashi, Tomoki Hayashi, Patrick Lumban Tobing, Tomoki Toda |
| 2017 | ASRU | An investigation of multi-speaker training for wavenet vocoder. | Tomoki Hayashi, Akira Tamamori, Kazuhiro Kobayashi, Kazuya Takeda, Tomoki Toda |
| 2017 | Interspeech | Statistical Voice Conversion with WaveNet-Based Waveform Generation. | Kazuhiro Kobayashi, Tomoki Hayashi, Akira Tamamori, Tomoki Toda |
| 2017 | Interspeech | Speaker-Dependent WaveNet Vocoder. | Akira Tamamori, Tomoki Hayashi, Kazuhiro Kobayashi, Kazuya Takeda, Tomoki Toda |
| 2016 | ICASSP | Implementation of F0 transformation for statistical singing voice conversion based on direct waveform modification. | Kazuhiro Kobayashi, Tomoki Toda, Satoshi Nakamura |
| 2016 | ICASSP | An estimation method of voice timbre evaluation values using feature extraction with Gaussian mixture model based on reference singer. | Soichi Yamane, Kazuhiro Kobayashi, Tomoki Toda, Tomoyasu Nakano, Masataka Goto, Satoshi Nakamura |
| 2016 | Interspeech | The NU-NAIST Voice Conversion System for the Voice Conversion Challenge 2016. | Kazuhiro Kobayashi, Shinnosuke Takamichi, Satoshi Nakamura, Tomoki Toda |
| 2015 | Interspeech | Statistical singing voice conversion based on direct waveform modification with global variance. | Kazuhiro Kobayashi, Tomoki Toda, Graham Neubig, Sakriani Sakti, Satoshi Nakamura |
| 2015 | Interspeech | Articulatory controllable speech modification based on Gaussian mixture models with direct waveform modification using spectrum differential. | Patrick Lumban Tobing, Kazuhiro Kobayashi, Tomoki Toda, Graham Neubig, Sakriani Sakti, Satoshi Nakamura |
| 2014 | ICASSP | Regression approaches to perceptual age control in singing voice conversion. | Kazuhiro Kobayashi, Tomoki Toda, Tomoyasu Nakano, Masataka Goto, Graham Neubig, Sakriani Sakti, Satoshi Nakamura |
| 2014 | Interspeech | Statistical singing voice conversion with direct waveform modification based on the spectrum differential. | Kazuhiro Kobayashi, Tomoki Toda, Graham Neubig, Sakriani Sakti, Satoshi Nakamura |
| 2013 | Interspeech | An investigation of acoustic features for singing voice conversion based on perceptual age. | Kazuhiro Kobayashi, Hironori Doi, Tomoki Toda, Tomoyasu Nakano, Masataka Goto, Graham Neubig, Sakriani Sakti, Satoshi Nakamura |
| 2002 | LREC | The Valence Patterns of Japanese Verbs Extracted From The EDR Corpus. | Takano Ogino, Hitoshi Isahara, Kazuhiro Kobayashi |