| 2024 | ACL | Integrating Pre-Trained Speech and Language Models for End-to-End Speech Recognition. | Yukiya Hono, Koh Mitsuda, Tianyu Zhao, Kentaro Mitsui, Toshiaki Wakatsuki, Kei Sawada |
| 2024 | COLING | Release of Pre-Trained Models for the Japanese Language. | Kei Sawada, Tianyu Zhao, Makoto Shing, Kentaro Mitsui, Akio Kaga, Yukiya Hono, Toshiaki Wakatsuki, Koh Mitsuda |
| 2024 | EMNLP | PSLM: Parallel Generation of Text and Speech with LLMs for Low-Latency Spoken Dialogue Systems. | Kentaro Mitsui, Koh Mitsuda, Toshiaki Wakatsuki, Yukiya Hono, Kei Sawada |
| 2023 | ACL | Focused Prefix Tuning for Controllable Text Generation. | Congda Ma, Tianyu Zhao, Makoto Shing, Kei Sawada, Manabu Okumura |
| 2023 | Interspeech | UniFLG: Unified Facial Landmark Generator from Text or Speech. | Kentaro Mitsui, Yukiya Hono, Kei Sawada |
| 2022 | HAI | Backchannel Generation Model for a Third Party Listener Agent. | Divesh Lala, Koji Inoue, Tatsuya Kawahara, Kei Sawada |
| 2022 | Interspeech | MSR-NV: Neural Vocoder Using Multiple Sampling Rates. | Kentaro Mitsui, Kei Sawada |
| 2022 | Interspeech | End-to-End Text-to-Speech Based on Latent Representation of Speaking Styles Using Spontaneous Dialogue. | Kentaro Mitsui, Tianyu Zhao, Kei Sawada, Yukiya Hono, Yoshihiko Nankaku, Keiichi Tokuda |
| 2021 | ICLR | Dance Revolution: Long-Term Dance Generation with Music via Curriculum Learning. | Ruozi Huang, Huang Hu, Wei Wu, Kei Sawada, Mi Zhang, Daxin Jiang |
| 2020 | Interspeech | Hierarchical Multi-Grained Generative Model for Expressive Speech Synthesis. | Yukiya Hono, Kazuna Tsuboi, Kei Sawada, Kei Hashimoto, Keiichiro Oura, Yoshihiko Nankaku, Keiichi Tokuda |
| 2018 | ICASSP | Image Recognition Based on Separable Lattice Hmms Using a Deep Neural Network for Output Probability Distributions. | Eiji Ichikawa, Kei Sawada, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2017 | ASRU | The blizzard machine learning challenge 2017. | Kei Sawada, Keiichi Tokuda, Simon King, Alan W. Black |
| 2017 | ICASSP | Image recognition based on discriminative models using features generated from separable lattice HMMS. | Yoshinari Tsuzuki, Kei Sawada, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |
| 2012 | ICASSP | Face recognition based on separable lattice 2-D HMMS using variational bayesian method. | Kei Sawada, Akira Tamamori, Kei Hashimoto, Yoshihiko Nankaku, Keiichi Tokuda |