| 2025 | ASRU | Benchmarking Prosody Encoding in Discrete Speech Tokens. | Kentaro Onda, Satoru Fukayama, Daisuke Saito, Nobuaki Minematsu |
| 2025 | Interspeech | A Perception-Based L2 Speech Intelligibility Indicator: Leveraging a Rater's Shadowing and Sequence-to-sequence Voice Conversion. | Haopeng Geng, Daisuke Saito, Nobuaki Minematsu |
| 2025 | Interspeech | Discrete Tokens Exhibit Interlanguage Speech Intelligibility Benefit: an Analytical Study Towards Accent-robust ASR Only with Native Speech Data. | Kentaro Onda, Keisuke Imoto, Satoru Fukayama, Daisuke Saito, Nobuaki Minematsu |
| 2025 | Interspeech | Prosodically Enhanced Foreign Accent Simulation by Discrete Token-based Resynthesis Only with Native Speech Corpora. | Kentaro Onda, Keisuke Imoto, Satoru Fukayama, Daisuke Saito, Nobuaki Minematsu |
| 2024 | ICASSP | Do Learned Speech Symbols Follow Zipf's Law? | Shinnosuke Takamichi, Hiroki Maeda, Joonyong Park, Daisuke Saito, Hiroshi Saruwatari |
| 2024 | Interspeech | A ChatGPT-based oral Q&A practice system for first-time student participants in international conferences. | Mayuko Aiba, Daisuke Saito, Nobuaki Minematsu |
| 2024 | Interspeech | A Pilot Study of GSLM-based Simulation of Foreign Accentuation Only Using Native Speech Corpora. | Kentaro Onda, Joonyong Park, Nobuaki Minematsu, Daisuke Saito |
| 2024 | Interspeech | Acceleration of Posteriorgram-based DTW by Distilling the Class-to-class Distances Encoded in the Classifier Used to Calculate Posteriors. | Haitong Sun, Jaehyun Choi, Nobuaki Minematsu, Daisuke Saito |
| 2024 | Interspeech | Analysis and Visualization of Directional Diversity in Listening Fluency of World Englishes Speakers in the Framework of Mutual Shadowing. | Yu Tomita, Yingxiang Gao, Nobuaki Minematsu, Noriko Nakanishi, Daisuke Saito |
| 2024 | SIGCSE | Enhancing Programming Education through Game-Based Learning: Design and Implementation of a Puyo Puyo-Inspired Teaching Tool. | Ruochen Tian, Daisuke Saito, Hironori Washizaki, Yoshiaki Fukazawa, Hiroshi Kobayashi, Ayumi Tsuji |
| 2023 | EDUCON | Programming Education for Young People using the Falling-Puzzle Game, "Puyo Puyo". | Daisuke Saito, Ruochen Tian, Hironori Washizaki, Yoshiaki Fukazawa |
| 2023 | ICASSP | Multiple Acoustic Features Speech Emotion Recognition Using Cross-Attention Transformer. | Yurun He, Nobuaki Minematsu, Daisuke Saito |
| 2023 | Interspeech | Automatic Prediction of Language Learners' Listenability Using Speech and Text Features Extracted from Listening Drills. | Yingxiang Gao, Jaehyun Choi, Nobuaki Minematsu, Noriko Nakanishi, Daisuke Saito |
| 2023 | SIGCSE | Gender Characteristics and Computational Thinking in Scratch. | Rose Niousha, Daisuke Saito, Hironori Washizaki, Yoshiaki Fukazawa |
| 2022 | ICASSP | Quantifying Discriminability between NMF Bases. | Eisuke Konno, Daisuke Saito, Nobuaki Minematsu |
| 2022 | Interspeech | Text-to-speech synthesis using spectral modeling based on non-negative autoencoder. | Takeru Gorai, Daisuke Saito, Nobuaki Minematsu |
| 2022 | Interspeech | Detection of Learners' Listening Breakdown with Oral Dictation and Its Use to Model Listening Skill Improvement Exclusively Through Shadowing. | Takuya Kunihara, Chuanbo Zhu, Daisuke Saito, Nobuaki Minematsu, Noriko Nakanishi |
| 2021 | ASRU | Multi-Granularity Annotation of Instantaneous Intelligibility of Learners' Utterances Based on Shadowing Techniques. | Chuanbo Zhu, Ryo Hakoda, Daisuke Saito, Nobuaki Minematsu, Noriko Nakanishi, Tazuko Nishimura |
| 2021 | COMPSAC | Automated Educational Program Mapping on Learning Standards in Computer Science. | Koki Miura, Daisuke Saito, Hironori Washizaki, Yoshiaki Fukazawa |
| 2021 | COMPSAC | Preliminary Literature Review of Machine Learning System Development Practices. | Yasuhiro Watanabe, Hironori Washizaki, Kazunori Sakamoto, Daisuke Saito, Kiyoshi Honda, Naohiko Tsuda, Yoshiaki Fukazawa, Nobukazu Yoshioka |
| 2021 | DAFX | Quality Diversity for Synthesizer Sound Matching. | Naotake Masuda, Daisuke Saito |
| 2021 | EDUCON | Work-in-Progress: Analysis of the use of Mentoring with Online Mob Programming. | Shota Kaieda, Daisuke Saito, Hironori Washizaki, Yoshiaki Fukazawa |
| 2021 | Interspeech | Lexical Density Analysis of Word Productions in Japanese English Using Acoustic Word Embeddings. | Shintaro Ando, Nobuaki Minematsu, Daisuke Saito |
| 2020 | Interspeech | Attention-Based Speaker Embeddings for One-Shot Voice Conversion. | Tatsuma Ishihara, Daisuke Saito |
| 2020 | Interspeech | Shadowability Annotation with Fine Granularity on L2 Utterances and its Improvement with Native Listeners' Script-Shadowing. | Zhenchao Lin, Ryo Takashima, Daisuke Saito, Nobuaki Minematsu, Noriko Nakanishi |
| 2020 | Interspeech | Discriminative Method to Extract Coarse Prosodic Structure and its Application for Statistical Phrase/Accent Command Estimation. | Yuma Shirahata, Daisuke Saito, Nobuaki Minematsu |
| 2020 | Interspeech | Nonparallel Training of Exemplar-Based Voice Conversion System Using INCA-Based Alignment Technique. | Hitoshi Suda, Gaku Kotani, Daisuke Saito |
| 2019 | Interspeech | Analysis of Native Listeners' Facial Microexpressions While Shadowing Non-Native Speech - Potential of Shadowers' Facial Expressions for Comprehensibility Prediction. | Tasavat Trisitichoke, Shintaro Ando, Daisuke Saito, Nobuaki Minematsu |
| 2019 | SIGCSE | Rubric to Evaluate Programming Learning of Elementary School Students. | Daisuke Saito, Hironori Washizaki, Yoshiaki Fukazawa, Mariko Tamura, Yuki Sakuragi |
| 2018 | Interspeech | A Study of Objective Measurement of Comprehensibility through Native Speakers' Shadowing of Learners' Utterances. | Yusuke Inoue, Suguru Kabashima, Daisuke Saito, Nobuaki Minematsu, Kumi Kanamura, Yutaka Yamauchi |
| 2018 | Interspeech | A Comparative Study of Statistical Conversion of Face to Voice Based on Their Subjective Impressions. | Yasuhito Ohsugi, Daisuke Saito, Nobuaki Minematsu |
| 2018 | Tencon | Noise Reduction Method for Intra-Body Communication by Using Compensation Electrode. | Yutaro Toyoshima, Yoshiki Matsui, Ryota Kato, Kenta Nezu, Mitsuru Shinagawa, Daisuke Saito, Ken Seo, Kyoji Oohashi |
| 2017 | Interspeech | Parallel-Data-Free Many-to-Many Voice Conversion Based on DNN Integrated with Eigenspace Using a Non-Parallel Speech Corpus. | Tetsuya Hashimoto, Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu |
| 2017 | Interspeech | Use of Global and Acoustic Features Associated with Contextual Factors to Adapt Language Models for Spontaneous Speech Recognition. | Shohei Toyama, Daisuke Saito, Nobuaki Minematsu |
| 2017 | Interspeech | Acoustic-to-Articulatory Mapping Based on Mixture of Probabilistic Canonical Correlation Analysis. | Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu |
| 2017 | Interspeech | Automatic Scoring of Shadowing Speech Based on DNN Posteriors and Their DTW. | Junwei Yue, Fumiya Shiozawa, Shohei Toyama, Yutaka Yamauchi, Kayoko Ito, Daisuke Saito, Nobuaki Minematsu |
| 2016 | ICASSP | Divergence estimation based on deep neural networks and its use for language identification. | Yosuke Kashiwagi, Congying Zhang, Daisuke Saito, Nobuaki Minematsu |
| 2016 | Interspeech | Automatic Assessment and Error Detection of Shadowing Speech: Case of English Spoken by Japanese Learners. | Shuju Shi, Yosuke Kashiwagi, Shohei Toyama, Junwei Yue, Yutaka Yamauchi, Daisuke Saito, Nobuaki Minematsu |
| 2016 | Interspeech | The Voice Conversion Challenge 2016. | Tomoki Toda, Ling-Hui Chen, Daisuke Saito, Fernando Villavicencio, Mirjam Wester, Zhizheng Wu, Junichi Yamagishi |
| 2016 | Interspeech | Prediction of the Articulatory Movements of Unseen Phonemes of a Speaker Using the Speech Structure of Another Speaker. | Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu |
| 2016 | Interspeech | Voice Conversion Based on Matrix Variate Gaussian Mixture Model Using Multiple Frame Features. | Yi Yang, Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu |
| 2016 | Interspeech | Speaker Representations for Speaker Adaptation in Multiple Speakers' BLSTM-RNN-Based Speech Synthesis. | Yi Zhao, Daisuke Saito, Nobuaki Minematsu |
| 2016 | ITiCSE | Influence of the Programming Environment on Programming Education. | Daisuke Saito, Hironori Washizaki, Yoshiaki Fukazawa |
| 2015 | ICASSP | SAS: A speaker verification spoofing database containing diverse attacks. | Zhizheng Wu, Ali Khodabakhsh, Cenk Demiroglu, Junichi Yamagishi, Daisuke Saito, Tomoki Toda, Simon King |
| 2015 | Interspeech | Statistical acoustic-to-articulatory mapping unified with speaker normalization based on voice conversion. | Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose |
| 2014 | ICASSP | Improved and robust prediction of pronunciation distance for individual-basis clustering of World Englishes pronunciation. | Shun Kasahara, S. Kitahara, Nobuaki Minematsu, Han-Ping Shen, Takehiko Makino, Daisuke Saito, K. Hiorse |
| 2014 | ICASSP | Semi-supervised noise dictionary adaptation for exemplar-based noise robust speech recognition. | Yi Luan, Daisuke Saito, Yosuke Kashiwagi, Nobuaki Minematsu, Keikichi Hirose |
| 2014 | Interspeech | Application of matrix variate Gaussian mixture model to statistical voice conversion. | Daisuke Saito, Hidenobu Doi, Nobuaki Minematsu, Keikichi Hirose |
| 2013 | ASRU | Discriminative piecewise linear transformation based on deep learning for noise robust automatic speech recognition. | Yosuke Kashiwagi, Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose |
| 2013 | Interspeech | Probabilistic speech F | Tatsuma Ishihara, Hirokazu Kameoka, Kota Yoshizato, Daisuke Saito, Shigeki Sagayama |
| 2012 | ICASSP | A tandem connectionist model using combination of multi-scale spectro-temporal features for acoustic event detection. | Miquel Espi, Masakiyo Fujimoto, Daisuke Saito, Nobutaka Ono, Shigeki Sagayama |
| 2012 | Interspeech | Effects of Speaker Adaptive Training on Tensor-based Arbitrary Speaker Conversion. | Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose |
| 2012 | Interspeech | Hidden Markov Convolutive Mixture Model for Pitch Contour Analysis of Speech. | Kota Yoshizato, Hirokazu Kameoka, Daisuke Saito, Shigeki Sagayama |
| 2011 | ICASSP | High accurate model-integration-based voice conversion using dynamic features and model structure optimization. | Daisuke Saito, Shinji Watanabe, Atsushi Nakamura, Nobuaki Minematsu |
| 2011 | Interspeech | Adaptation of Prosody in Speech Synthesis by Changing Command Values of the Generation Process Model of Fundamental Frequency. | Keikichi Hirose, Keiko Ochi, Ryusuke Mihara, Hiroya Hashimoto, Daisuke Saito, Nobuaki Minematsu |
| 2011 | Interspeech | Gesture Design of Hand-to-Speech Converter Derived from Speech-to-Hand Converter Based on Probabilistic Integration Model. | Aki Kunikoshi, Yu Qiao, Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose |
| 2011 | Interspeech | One-to-Many Voice Conversion Based on Tensor Representation of Speaker Space. | Daisuke Saito, Keisuke Yamamoto, Nobuaki Minematsu, Keikichi Hirose |
| 2010 | ICASSP | HMM-based sequence-to-frame mapping for voice conversion. | Yu Qiao, Daisuke Saito, Nobuaki Minematsu |
| 2010 | Interspeech | Probabilistic integration of joint density model and speaker model for voice conversion. | Daisuke Saito, Shinji Watanabe, Atsushi Nakamura, Nobuaki Minematsu |
| 2009 | Interspeech | Optimal event search using a structural cost function - improvement of structure to speech conversion. | Daisuke Saito, Yu Qiao, Nobuaki Minematsu, Keikichi Hirose |
| 2008 | ICASSP | Directional dependency of cepstrum on vocal tract length. | Daisuke Saito, Ryo Matsuura, Satoshi Asakawa, Nobuaki Minematsu, Keikichi Hirose |
| 2008 | Interspeech | Structure to speech conversion - speech generation based on infant-like vocal imitation. | Daisuke Saito, Satoshi Asakawa, Nobuaki Minematsu, Keikichi Hirose |
| 2008 | Interspeech | Decomposition of rotational distortion caused by VTL difference using eigenvalues of its transformation matrix. | Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose |