| 2025 | Interspeech | PeriodCodec: A Pitch-Controllable Neural Audio Codec Using Periodic Signals for Singing Voice Synthesis. | Masato Takagi, Miku Nishihara, Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2024 | ACL | Integrating Pre-Trained Speech and Language Models for End-to-End Speech Recognition. | Yukiya Hono, Koh Mitsuda, Tianyu Zhao, Kentaro Mitsui, Toshiaki Wakatsuki, Kei Sawada |
| 2024 | COLING | Release of Pre-Trained Models for the Japanese Language. | Kei Sawada, Tianyu Zhao, Makoto Shing, Kentaro Mitsui, Akio Kaga, Yukiya Hono, Toshiaki Wakatsuki, Koh Mitsuda |
| 2024 | EMNLP | PSLM: Parallel Generation of Text and Speech with LLMs for Low-Latency Spoken Dialogue Systems. | Kentaro Mitsui, Koh Mitsuda, Toshiaki Wakatsuki, Yukiya Hono, Kei Sawada |
| 2024 | ICASSP | PeriodGrad: Towards Pitch-Controllable Neural Vocoder Based on a Diffusion Probabilistic Model. | Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2023 | ICASSP | Singing Voice Synthesis Based on a Musical Note Position-Aware Attention Mechanism. | Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2023 | ICASSP | Embedding a Differentiable Mel-Cepstral Synthesis Filter to a Neural Speech Synthesis System. | Takenori Yoshimura, Shinji Takaki, Kazuhiro Nakamura, Keiichiro Oura, Yukiya Hono, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2023 | Interspeech | UniFLG: Unified Facial Landmark Generator from Text or Speech. | Kentaro Mitsui, Yukiya Hono, Kei Sawada |
| 2022 | Interspeech | End-to-End Text-to-Speech Based on Latent Representation of Speaking Styles Using Spontaneous Dialogue. | Kentaro Mitsui, Tianyu Zhao, Kei Sawada, Yukiya Hono, Yoshihiko Nankaku, Keiichi Tokuda |
| 2021 | ICASSP | Periodnet: A Non-Autoregressive Waveform Generation Model with a Structure Separating Periodic and Aperiodic Components. | Yukiya Hono, Shinji Takaki, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2020 | Interspeech | Hierarchical Multi-Grained Generative Model for Expressive Speech Synthesis. | Yukiya Hono, Kazuna Tsuboi, Kei Sawada, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2019 | ICASSP | Singing Voice Synthesis Based on Generative Adversarial Networks. | Yukiya Hono, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |