Skip to content

Daisuke Saito

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

63

Venues

9

Active years

2008–2025

Best venue rank

A

Where they publish

Papers

63 indexed papers, newest first.

YearVenueTitleAuthors
2025ASRUBenchmarking Prosody Encoding in Discrete Speech Tokens.Kentaro Onda, Satoru Fukayama, Daisuke Saito, Nobuaki Minematsu
2025InterspeechA Perception-Based L2 Speech Intelligibility Indicator: Leveraging a Rater's Shadowing and Sequence-to-sequence Voice Conversion.Haopeng Geng, Daisuke Saito, Nobuaki Minematsu
2025InterspeechDiscrete Tokens Exhibit Interlanguage Speech Intelligibility Benefit: an Analytical Study Towards Accent-robust ASR Only with Native Speech Data.Kentaro Onda, Keisuke Imoto, Satoru Fukayama, Daisuke Saito, Nobuaki Minematsu
2025InterspeechProsodically Enhanced Foreign Accent Simulation by Discrete Token-based Resynthesis Only with Native Speech Corpora.Kentaro Onda, Keisuke Imoto, Satoru Fukayama, Daisuke Saito, Nobuaki Minematsu
2024ICASSPDo Learned Speech Symbols Follow Zipf's Law?Shinnosuke Takamichi, Hiroki Maeda, Joonyong Park, Daisuke Saito, Hiroshi Saruwatari
2024InterspeechA ChatGPT-based oral Q&A practice system for first-time student participants in international conferences.Mayuko Aiba, Daisuke Saito, Nobuaki Minematsu
2024InterspeechA Pilot Study of GSLM-based Simulation of Foreign Accentuation Only Using Native Speech Corpora.Kentaro Onda, Joonyong Park, Nobuaki Minematsu, Daisuke Saito
2024InterspeechAcceleration of Posteriorgram-based DTW by Distilling the Class-to-class Distances Encoded in the Classifier Used to Calculate Posteriors.Haitong Sun, Jaehyun Choi, Nobuaki Minematsu, Daisuke Saito
2024InterspeechAnalysis and Visualization of Directional Diversity in Listening Fluency of World Englishes Speakers in the Framework of Mutual Shadowing.Yu Tomita, Yingxiang Gao, Nobuaki Minematsu, Noriko Nakanishi, Daisuke Saito
2024SIGCSEEnhancing Programming Education through Game-Based Learning: Design and Implementation of a Puyo Puyo-Inspired Teaching Tool.Ruochen Tian, Daisuke Saito, Hironori Washizaki, Yoshiaki Fukazawa, Hiroshi Kobayashi, Ayumi Tsuji
2023EDUCONProgramming Education for Young People using the Falling-Puzzle Game, "Puyo Puyo".Daisuke Saito, Ruochen Tian, Hironori Washizaki, Yoshiaki Fukazawa
2023ICASSPMultiple Acoustic Features Speech Emotion Recognition Using Cross-Attention Transformer.Yurun He, Nobuaki Minematsu, Daisuke Saito
2023InterspeechAutomatic Prediction of Language Learners' Listenability Using Speech and Text Features Extracted from Listening Drills.Yingxiang Gao, Jaehyun Choi, Nobuaki Minematsu, Noriko Nakanishi, Daisuke Saito
2023SIGCSEGender Characteristics and Computational Thinking in Scratch.Rose Niousha, Daisuke Saito, Hironori Washizaki, Yoshiaki Fukazawa
2022ICASSPQuantifying Discriminability between NMF Bases.Eisuke Konno, Daisuke Saito, Nobuaki Minematsu
2022InterspeechText-to-speech synthesis using spectral modeling based on non-negative autoencoder.Takeru Gorai, Daisuke Saito, Nobuaki Minematsu
2022InterspeechDetection of Learners' Listening Breakdown with Oral Dictation and Its Use to Model Listening Skill Improvement Exclusively Through Shadowing.Takuya Kunihara, Chuanbo Zhu, Daisuke Saito, Nobuaki Minematsu, Noriko Nakanishi
2021ASRUMulti-Granularity Annotation of Instantaneous Intelligibility of Learners' Utterances Based on Shadowing Techniques.Chuanbo Zhu, Ryo Hakoda, Daisuke Saito, Nobuaki Minematsu, Noriko Nakanishi, Tazuko Nishimura
2021COMPSACAutomated Educational Program Mapping on Learning Standards in Computer Science.Koki Miura, Daisuke Saito, Hironori Washizaki, Yoshiaki Fukazawa
2021COMPSACPreliminary Literature Review of Machine Learning System Development Practices.Yasuhiro Watanabe, Hironori Washizaki, Kazunori Sakamoto, Daisuke Saito, Kiyoshi Honda, Naohiko Tsuda, Yoshiaki Fukazawa, Nobukazu Yoshioka
2021DAFXQuality Diversity for Synthesizer Sound Matching.Naotake Masuda, Daisuke Saito
2021EDUCONWork-in-Progress: Analysis of the use of Mentoring with Online Mob Programming.Shota Kaieda, Daisuke Saito, Hironori Washizaki, Yoshiaki Fukazawa
2021InterspeechLexical Density Analysis of Word Productions in Japanese English Using Acoustic Word Embeddings.Shintaro Ando, Nobuaki Minematsu, Daisuke Saito
2020InterspeechAttention-Based Speaker Embeddings for One-Shot Voice Conversion.Tatsuma Ishihara, Daisuke Saito
2020InterspeechShadowability Annotation with Fine Granularity on L2 Utterances and its Improvement with Native Listeners' Script-Shadowing.Zhenchao Lin, Ryo Takashima, Daisuke Saito, Nobuaki Minematsu, Noriko Nakanishi
2020InterspeechDiscriminative Method to Extract Coarse Prosodic Structure and its Application for Statistical Phrase/Accent Command Estimation.Yuma Shirahata, Daisuke Saito, Nobuaki Minematsu
2020InterspeechNonparallel Training of Exemplar-Based Voice Conversion System Using INCA-Based Alignment Technique.Hitoshi Suda, Gaku Kotani, Daisuke Saito
2019InterspeechAnalysis of Native Listeners' Facial Microexpressions While Shadowing Non-Native Speech - Potential of Shadowers' Facial Expressions for Comprehensibility Prediction.Tasavat Trisitichoke, Shintaro Ando, Daisuke Saito, Nobuaki Minematsu
2019SIGCSERubric to Evaluate Programming Learning of Elementary School Students.Daisuke Saito, Hironori Washizaki, Yoshiaki Fukazawa, Mariko Tamura, Yuki Sakuragi
2018InterspeechA Study of Objective Measurement of Comprehensibility through Native Speakers' Shadowing of Learners' Utterances.Yusuke Inoue, Suguru Kabashima, Daisuke Saito, Nobuaki Minematsu, Kumi Kanamura, Yutaka Yamauchi
2018InterspeechA Comparative Study of Statistical Conversion of Face to Voice Based on Their Subjective Impressions.Yasuhito Ohsugi, Daisuke Saito, Nobuaki Minematsu
2018TenconNoise Reduction Method for Intra-Body Communication by Using Compensation Electrode.Yutaro Toyoshima, Yoshiki Matsui, Ryota Kato, Kenta Nezu, Mitsuru Shinagawa, Daisuke Saito, Ken Seo, Kyoji Oohashi
2017InterspeechParallel-Data-Free Many-to-Many Voice Conversion Based on DNN Integrated with Eigenspace Using a Non-Parallel Speech Corpus.Tetsuya Hashimoto, Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu
2017InterspeechUse of Global and Acoustic Features Associated with Contextual Factors to Adapt Language Models for Spontaneous Speech Recognition.Shohei Toyama, Daisuke Saito, Nobuaki Minematsu
2017InterspeechAcoustic-to-Articulatory Mapping Based on Mixture of Probabilistic Canonical Correlation Analysis.Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu
2017InterspeechAutomatic Scoring of Shadowing Speech Based on DNN Posteriors and Their DTW.Junwei Yue, Fumiya Shiozawa, Shohei Toyama, Yutaka Yamauchi, Kayoko Ito, Daisuke Saito, Nobuaki Minematsu
2016ICASSPDivergence estimation based on deep neural networks and its use for language identification.Yosuke Kashiwagi, Congying Zhang, Daisuke Saito, Nobuaki Minematsu
2016InterspeechAutomatic Assessment and Error Detection of Shadowing Speech: Case of English Spoken by Japanese Learners.Shuju Shi, Yosuke Kashiwagi, Shohei Toyama, Junwei Yue, Yutaka Yamauchi, Daisuke Saito, Nobuaki Minematsu
2016InterspeechThe Voice Conversion Challenge 2016.Tomoki Toda, Ling-Hui Chen, Daisuke Saito, Fernando Villavicencio, Mirjam Wester, Zhizheng Wu, Junichi Yamagishi
2016InterspeechPrediction of the Articulatory Movements of Unseen Phonemes of a Speaker Using the Speech Structure of Another Speaker.Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu
2016InterspeechVoice Conversion Based on Matrix Variate Gaussian Mixture Model Using Multiple Frame Features.Yi Yang, Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu
2016InterspeechSpeaker Representations for Speaker Adaptation in Multiple Speakers' BLSTM-RNN-Based Speech Synthesis.Yi Zhao, Daisuke Saito, Nobuaki Minematsu
2016ITiCSEInfluence of the Programming Environment on Programming Education.Daisuke Saito, Hironori Washizaki, Yoshiaki Fukazawa
2015ICASSPSAS: A speaker verification spoofing database containing diverse attacks.Zhizheng Wu, Ali Khodabakhsh, Cenk Demiroglu, Junichi Yamagishi, Daisuke Saito, Tomoki Toda, Simon King
2015InterspeechStatistical acoustic-to-articulatory mapping unified with speaker normalization based on voice conversion.Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose
2014ICASSPImproved and robust prediction of pronunciation distance for individual-basis clustering of World Englishes pronunciation.Shun Kasahara, S. Kitahara, Nobuaki Minematsu, Han-Ping Shen, Takehiko Makino, Daisuke Saito, K. Hiorse
2014ICASSPSemi-supervised noise dictionary adaptation for exemplar-based noise robust speech recognition.Yi Luan, Daisuke Saito, Yosuke Kashiwagi, Nobuaki Minematsu, Keikichi Hirose
2014InterspeechApplication of matrix variate Gaussian mixture model to statistical voice conversion.Daisuke Saito, Hidenobu Doi, Nobuaki Minematsu, Keikichi Hirose
2013ASRUDiscriminative piecewise linear transformation based on deep learning for noise robust automatic speech recognition.Yosuke Kashiwagi, Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose
2013InterspeechProbabilistic speech FTatsuma Ishihara, Hirokazu Kameoka, Kota Yoshizato, Daisuke Saito, Shigeki Sagayama
2012ICASSPA tandem connectionist model using combination of multi-scale spectro-temporal features for acoustic event detection.Miquel Espi, Masakiyo Fujimoto, Daisuke Saito, Nobutaka Ono, Shigeki Sagayama
2012InterspeechEffects of Speaker Adaptive Training on Tensor-based Arbitrary Speaker Conversion.Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose
2012InterspeechHidden Markov Convolutive Mixture Model for Pitch Contour Analysis of Speech.Kota Yoshizato, Hirokazu Kameoka, Daisuke Saito, Shigeki Sagayama
2011ICASSPHigh accurate model-integration-based voice conversion using dynamic features and model structure optimization.Daisuke Saito, Shinji Watanabe, Atsushi Nakamura, Nobuaki Minematsu
2011InterspeechAdaptation of Prosody in Speech Synthesis by Changing Command Values of the Generation Process Model of Fundamental Frequency.Keikichi Hirose, Keiko Ochi, Ryusuke Mihara, Hiroya Hashimoto, Daisuke Saito, Nobuaki Minematsu
2011InterspeechGesture Design of Hand-to-Speech Converter Derived from Speech-to-Hand Converter Based on Probabilistic Integration Model.Aki Kunikoshi, Yu Qiao, Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose
2011InterspeechOne-to-Many Voice Conversion Based on Tensor Representation of Speaker Space.Daisuke Saito, Keisuke Yamamoto, Nobuaki Minematsu, Keikichi Hirose
2010ICASSPHMM-based sequence-to-frame mapping for voice conversion.Yu Qiao, Daisuke Saito, Nobuaki Minematsu
2010InterspeechProbabilistic integration of joint density model and speaker model for voice conversion.Daisuke Saito, Shinji Watanabe, Atsushi Nakamura, Nobuaki Minematsu
2009InterspeechOptimal event search using a structural cost function - improvement of structure to speech conversion.Daisuke Saito, Yu Qiao, Nobuaki Minematsu, Keikichi Hirose
2008ICASSPDirectional dependency of cepstrum on vocal tract length.Daisuke Saito, Ryo Matsuura, Satoshi Asakawa, Nobuaki Minematsu, Keikichi Hirose
2008InterspeechStructure to speech conversion - speech generation based on infant-like vocal imitation.Daisuke Saito, Satoshi Asakawa, Nobuaki Minematsu, Keikichi Hirose
2008InterspeechDecomposition of rotational distortion caused by VTL difference using eigenvalues of its transformation matrix.Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose