| 2025 | ASRU | Benchmarking Prosody Encoding in Discrete Speech Tokens. | Kentaro Onda, Satoru Fukayama, Daisuke Saito, Nobuaki Minematsu |
| 2025 | HAI | LangInLab: Augmenting Engineering Lab Instruction with Vision- and Voice-Enabled AI Agents for Language Learning. | Masanori Shigi, Zackary Rackauckas, Yuka Akiyama, Nobuaki Minematsu |
| 2025 | Interspeech | A Perception-Based L2 Speech Intelligibility Indicator: Leveraging a Rater's Shadowing and Sequence-to-sequence Voice Conversion. | Haopeng Geng, Daisuke Saito, Nobuaki Minematsu |
| 2025 | Interspeech | Discrete Tokens Exhibit Interlanguage Speech Intelligibility Benefit: an Analytical Study Towards Accent-robust ASR Only with Native Speech Data. | Kentaro Onda, Keisuke Imoto, Satoru Fukayama, Daisuke Saito, Nobuaki Minematsu |
| 2025 | Interspeech | Prosodically Enhanced Foreign Accent Simulation by Discrete Token-based Resynthesis Only with Native Speech Corpora. | Kentaro Onda, Keisuke Imoto, Satoru Fukayama, Daisuke Saito, Nobuaki Minematsu |
| 2024 | Interspeech | A ChatGPT-based oral Q&A practice system for first-time student participants in international conferences. | Mayuko Aiba, Daisuke Saito, Nobuaki Minematsu |
| 2024 | Interspeech | Exploring Pre-trained Speech Model for Articulatory Feature Extraction in Dysarthric Speech Using ASR. | Yuqin Lin, Longbiao Wang, Jianwu Dang, Nobuaki Minematsu |
| 2024 | Interspeech | A Pilot Study of GSLM-based Simulation of Foreign Accentuation Only Using Native Speech Corpora. | Kentaro Onda, Joonyong Park, Nobuaki Minematsu, Daisuke Saito |
| 2024 | Interspeech | Acceleration of Posteriorgram-based DTW by Distilling the Class-to-class Distances Encoded in the Classifier Used to Calculate Posteriors. | Haitong Sun, Jaehyun Choi, Nobuaki Minematsu, Daisuke Saito |
| 2024 | Interspeech | Analysis and Visualization of Directional Diversity in Listening Fluency of World Englishes Speakers in the Framework of Mutual Shadowing. | Yu Tomita, Yingxiang Gao, Nobuaki Minematsu, Noriko Nakanishi, Daisuke Saito |
| 2023 | ICASSP | Multiple Acoustic Features Speech Emotion Recognition Using Cross-Attention Transformer. | Yurun He, Nobuaki Minematsu, Daisuke Saito |
| 2023 | ICASSP | Hierarchical Softmax for End-To-End Low-Resource Multilingual Speech Recognition. | Qianying Liu, Zhuo Gong, Zhengdong Yang, Yuhang Yang, Sheng Li, Chenchen Ding, Nobuaki Minematsu, Hao Huang, Fei Cheng, Chenhui Chu, Sadao Kurohashi |
| 2023 | Interspeech | Automatic Prediction of Language Learners' Listenability Using Speech and Text Features Extracted from Listening Drills. | Yingxiang Gao, Jaehyun Choi, Nobuaki Minematsu, Noriko Nakanishi, Daisuke Saito |
| 2023 | Interspeech | A Unified Framework to Improve Learners' Skills of Perception and Production Based on Speech Shadowing and Overlapping. | Nobuaki Minematsu, Noriko Nakanishi, Yingxiang Gao, Haitong Sun |
| 2022 | ICASSP | Quantifying Discriminability between NMF Bases. | Eisuke Konno, Daisuke Saito, Nobuaki Minematsu |
| 2022 | Interspeech | Text-to-speech synthesis using spectral modeling based on non-negative autoencoder. | Takeru Gorai, Daisuke Saito, Nobuaki Minematsu |
| 2022 | Interspeech | Gradual Improvements Observed in Learners' Perception and Production of L2 Sounds Through Continuing Shadowing Practices on a Daily Basis. | Takuya Kunihara, Chuanbo Zhu, Nobuaki Minematsu, Noriko Nakanishi |
| 2022 | Interspeech | Detection of Learners' Listening Breakdown with Oral Dictation and Its Use to Model Listening Skill Improvement Exclusively Through Shadowing. | Takuya Kunihara, Chuanbo Zhu, Daisuke Saito, Nobuaki Minematsu, Noriko Nakanishi |
| 2021 | ASRU | Multi-Granularity Annotation of Instantaneous Intelligibility of Learners' Utterances Based on Shadowing Techniques. | Chuanbo Zhu, Ryo Hakoda, Daisuke Saito, Nobuaki Minematsu, Noriko Nakanishi, Tazuko Nishimura |
| 2021 | Interspeech | Lexical Density Analysis of Word Productions in Japanese English Using Acoustic Word Embeddings. | Shintaro Ando, Nobuaki Minematsu, Daisuke Saito |
| 2020 | ICASSP | Converting Written Language to Spoken Language with Neural Machine Translation for Language Modeling. | Shintaro Ando, Masayuki Suzuki, Nobuyasu Itoh, Gakuto Kurata, Nobuaki Minematsu |
| 2020 | Interspeech | Shadowability Annotation with Fine Granularity on L2 Utterances and its Improvement with Native Listeners' Script-Shadowing. | Zhenchao Lin, Ryo Takashima, Daisuke Saito, Nobuaki Minematsu, Noriko Nakanishi |
| 2020 | Interspeech | Discriminative Method to Extract Coarse Prosodic Structure and its Application for Statistical Phrase/Accent Command Estimation. | Yuma Shirahata, Daisuke Saito, Nobuaki Minematsu |
| 2019 | Interspeech | Analysis of Native Listeners' Facial Microexpressions While Shadowing Non-Native Speech - Potential of Shadowers' Facial Expressions for Comprehensibility Prediction. | Tasavat Trisitichoke, Shintaro Ando, Daisuke Saito, Nobuaki Minematsu |
| 2018 | Interspeech | A Study of Objective Measurement of Comprehensibility through Native Speakers' Shadowing of Learners' Utterances. | Yusuke Inoue, Suguru Kabashima, Daisuke Saito, Nobuaki Minematsu, Kumi Kanamura, Yutaka Yamauchi |
| 2018 | Interspeech | A Comparative Study of Statistical Conversion of Face to Voice Based on Their Subjective Impressions. | Yasuhito Ohsugi, Daisuke Saito, Nobuaki Minematsu |
| 2017 | Interspeech | Parallel-Data-Free Many-to-Many Voice Conversion Based on DNN Integrated with Eigenspace Using a Non-Parallel Speech Corpus. | Tetsuya Hashimoto, Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu |
| 2017 | Interspeech | Use of Global and Acoustic Features Associated with Contextual Factors to Adapt Language Models for Spontaneous Speech Recognition. | Shohei Toyama, Daisuke Saito, Nobuaki Minematsu |
| 2017 | Interspeech | Acoustic-to-Articulatory Mapping Based on Mixture of Probabilistic Canonical Correlation Analysis. | Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu |
| 2017 | Interspeech | Automatic Scoring of Shadowing Speech Based on DNN Posteriors and Their DTW. | Junwei Yue, Fumiya Shiozawa, Shohei Toyama, Yutaka Yamauchi, Kayoko Ito, Daisuke Saito, Nobuaki Minematsu |
| 2016 | ICASSP | Divergence estimation based on deep neural networks and its use for language identification. | Yosuke Kashiwagi, Congying Zhang, Daisuke Saito, Nobuaki Minematsu |
| 2016 | Interspeech | Automatic Assessment and Error Detection of Shadowing Speech: Case of English Spoken by Japanese Learners. | Shuju Shi, Yosuke Kashiwagi, Shohei Toyama, Junwei Yue, Yutaka Yamauchi, Daisuke Saito, Nobuaki Minematsu |
| 2016 | Interspeech | Prediction of the Articulatory Movements of Unseen Phonemes of a Speaker Using the Speech Structure of Another Speaker. | Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu |
| 2016 | Interspeech | Voice Conversion Based on Matrix Variate Gaussian Mixture Model Using Multiple Frame Features. | Yi Yang, Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu |
| 2016 | Interspeech | Speaker Representations for Speaker Adaptation in Multiple Speakers' BLSTM-RNN-Based Speech Synthesis. | Yi Zhao, Daisuke Saito, Nobuaki Minematsu |
| 2015 | Interspeech | Statistical acoustic-to-articulatory mapping unified with speaker normalization based on voice conversion. | Hidetsugu Uchida, Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose |
| 2014 | ICASSP | Improved and robust prediction of pronunciation distance for individual-basis clustering of World Englishes pronunciation. | Shun Kasahara, S. Kitahara, Nobuaki Minematsu, Han-Ping Shen, Takehiko Makino, Daisuke Saito, K. Hiorse |
| 2014 | ICASSP | Semi-supervised noise dictionary adaptation for exemplar-based noise robust speech recognition. | Yi Luan, Daisuke Saito, Yosuke Kashiwagi, Nobuaki Minematsu, Keikichi Hirose |
| 2014 | Interspeech | Application of matrix variate Gaussian mixture model to statistical voice conversion. | Daisuke Saito, Hidenobu Doi, Nobuaki Minematsu, Keikichi Hirose |
| 2013 | ASRU | Discriminative piecewise linear transformation based on deep learning for noise robust automatic speech recognition. | Yosuke Kashiwagi, Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose |
| 2013 | ASRU | Automatic pronunciation clustering using a World English archive and pronunciation structure analysis. | Han-Ping Shen, Nobuaki Minematsu, Takehiko Makino, Steven H. Weinberger, Teeraphon Pongkittiphan, Chung-Hsien Wu |
| 2013 | ICASSP | Improved estimation of femininity using GMM supervectors and SVR for voice therapy of Gender Identity Disorder Clients. | Chengshuo Wang, Masayuki Suzuki, Nobuaki Minematsu, Kyoko Sakuraba, Keikichi Hirose |
| 2013 | Interspeech | Artificial bandwidth extension based on regularized piecewise linear mapping with discriminative region weighting and long-Span features. | Nguyen Duc Duy, Masayuki Suzuki, Nobuaki Minematsu, Keikichi Hirose |
| 2013 | Interspeech | A free online accent and intonation dictionary for teachers and learners of Japanese. | Hiroko Hirano, Ibuki Nakamura, Nobuaki Minematsu, Masayuki Suzuki, Chieko Nakagawa, Noriko Nakamura, Yukinori Tagawa, Keikichi Hirose, Hiroya Hashimoto |
| 2013 | Interspeech | Generation of fundamental frequency contours for Thai speech synthesis using tone nucleus model. | Oraphan Krityakien, Keikichi Hirose, Nobuaki Minematsu |
| 2013 | Interspeech | Development of a web framework for teaching and learning Japanese prosody: OJAD (online Japanese accent dictionary). | Ibuki Nakamura, Nobuaki Minematsu, Masayuki Suzuki, Hiroko Hirano, Chieko Nakagawa, Noriko Nakamura, Yukinori Tagawa, Keikichi Hirose, Hiroya Hashimoto |
| 2013 | Interspeech | Failure transitions for joint n-gram models and G2p conversion. | Josef R. Novak, Nobuaki Minematsu, Keikichi Hirose |
| 2012 | ICASSP | Unseen noise robust speech recognition using adaptive piecewise linear transformation. | Keigo Chijiiwa, Masayuki Suzuki, Nobuaki Minematsu, Keikichi Hirose |
| 2012 | ICASSP | MFCC enhancement using joint corrupted and noise feature space for highly non-stationary noise environments. | Masayuki Suzuki, Takuya Yoshioka, Shinji Watanabe, Nobuaki Minematsu, Keikichi Hirose |
| 2012 | Interspeech | Improved Automatic Extraction of Generation Process Model Commands and Its use for Generating Fundamental Frequency Contours for Training HMM-based Speech Synthesis. | Hiroya Hashimoto, Keikichi Hirose, Nobuaki Minematsu |
| 2012 | Interspeech | Improved Prediction of Japanese Word Accent Sandhi Using CRF. | Nobuaki Minematsu, Shumpei Kobayashi, Shinya Shimizu, Keikichi Hirose |
| 2012 | Interspeech | Dynamic Grammars with Lookahead Composition for WFST-based Speech Recognition. | Josef R. Novak, Nobuaki Minematsu, Keikichi Hirose |
| 2012 | Interspeech | Improving WFST-based G2P Conversion with Alignment Constraints and RNNLM N-best Rescoring. | Josef R. Novak, Nobuaki Minematsu, Keikichi Hirose, Chiori Hori, Hideki Kashioka, Paul R. Dixon |
| 2012 | Interspeech | Effects of Speaker Adaptive Training on Tensor-based Arbitrary Speaker Conversion. | Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose |
| 2012 | Interspeech | Discriminative Reranking for LVCSR Leveraging Invariant Structure. | Masayuki Suzuki, Gakuto Kurata, Masafumi Nishimura, Nobuaki Minematsu |
| 2011 | ASRU | Decision of response timing for incremental speech recognition with reinforcement learning. | Di Lu, Takuya Nishimoto, Nobuaki Minematsu |
| 2011 | ICASSP | Improved F0 modeling and generation in voice conversion. | Aki Kunikoshi, Yao Qian, Frank K. Soong, Nobuaki Minematsu |
| 2011 | ICASSP | High accurate model-integration-based voice conversion using dynamic features and model structure optimization. | Daisuke Saito, Shinji Watanabe, Atsushi Nakamura, Nobuaki Minematsu |
| 2011 | Interspeech | Adaptation of Prosody in Speech Synthesis by Changing Command Values of the Generation Process Model of Fundamental Frequency. | Keikichi Hirose, Keiko Ochi, Ryusuke Mihara, Hiroya Hashimoto, Daisuke Saito, Nobuaki Minematsu |
| 2011 | Interspeech | Gesture Design of Hand-to-Speech Converter Derived from Speech-to-Hand Converter Based on Probabilistic Integration Model. | Aki Kunikoshi, Yu Qiao, Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose |
| 2011 | Interspeech | Measurement of Objective Intelligibility of Japanese Accented English Using ERJ (English Read by Japanese) Database. | Nobuaki Minematsu, Koji Okabe, Keisuke Ogaki, Keikichi Hirose |
| 2011 | Interspeech | Painless WFST Cascade Construction for LVCSR - Transducersaurus. | Josef R. Novak, Nobuaki Minematsu, Keikichi Hirose |
| 2011 | Interspeech | A Study on Bag of Gaussian Model with Application to Voice Conversion. | Yu Qiao, Tong Tong, Nobuaki Minematsu |
| 2011 | Interspeech | One-to-Many Voice Conversion Based on Tensor Representation of Speaker Space. | Daisuke Saito, Keisuke Yamamoto, Nobuaki Minematsu, Keikichi Hirose |
| 2011 | Interspeech | Continuous Digits Recognition Leveraging Invariant Structure. | Masayuki Suzuki, Gakuto Kurata, Masafumi Nishimura, Nobuaki Minematsu |
| 2011 | Interspeech | Prosody Conversion for Emotional Mandarin Speech Synthesis Using the Tone Nucleus Model. | Miaomiao Wen, Miaomiao Wang, Keikichi Hirose, Nobuaki Minematsu |
| 2010 | ICASSP | HMM-based sequence-to-frame mapping for voice conversion. | Yu Qiao, Daisuke Saito, Nobuaki Minematsu |
| 2010 | Interspeech | Regularized-MLLR speaker adaptation for computer-assisted language learning system. | Dean Luo, Yu Qiao, Nobuaki Minematsu, Yutaka Yamauchi, Keikichi Hirose |
| 2010 | Interspeech | Probabilistic integration of joint density model and speaker model for voice conversion. | Daisuke Saito, Shinji Watanabe, Atsushi Nakamura, Nobuaki Minematsu |
| 2010 | Interspeech | Integration of multilayer regression analysis with structure-based pronunciation assessment. | Masayuki Suzuki, Yu Qiao, Nobuaki Minematsu, Keikichi Hirose |
| 2010 | Interspeech | Improved generation of fundamental frequency in HMM-based speech synthesis using generation process model. | Miaomiao Wang, Miaomiao Wen, Keikichi Hirose, Nobuaki Minematsu |
| 2010 | Interspeech | Improving Mandarin segmental duration prediction with automatically extracted syntax features. | Miaomiao Wen, Miaomiao Wang, Keikichi Hirose, Nobuaki Minematsu |
| 2009 | ASRU | A study on Hidden Structural Model and its application to labeling sequences. | Yu Qiao, Masayuki Suzuki, Nobuaki Minematsu |
| 2009 | ASRU | Sub-structure-based estimation of pronunciation proficiency and classification of learners. | Masayuki Suzuki, Nobuaki Minematsu, Dean Luo, Keikichi Hirose |
| 2009 | ICASSP | Control of prosodic focus in corpus-based generation of fundamental frequency contours of Japanese based on the generation process model. | Keiko Ochi, Keikichi Hirose, Nobuaki Minematsu |
| 2009 | ICASSP | Mixture of Probabilistic Linear Regressions: A unified view of GMM-based mapping techiques. | Yu Qiao, Nobuaki Minematsu |
| 2009 | ICASSP | Affine invariant features and their application to speech recognition. | Yu Qiao, Masayuki Suzuki, Nobuaki Minematsu |
| 2009 | Interspeech | Speech generation from hand gestures based on space mapping. | Aki Kunikoshi, Yu Qiao, Nobuaki Minematsu, Keikichi Hirose |
| 2009 | Interspeech | Analysis and utilization of MLLR speaker adaptation technique for learners' pronunciation evaluation. | Dean Luo, Yu Qiao, Nobuaki Minematsu, Yutaka Yamauchi, Keikichi Hirose |
| 2009 | Interspeech | Structural analysis of dialects, sub-dialects and sub-sub-dialects of Chinese. | Xuebin Ma, Akira Nemoto, Nobuaki Minematsu, Yu Qiao, Keikichi Hirose |
| 2009 | Interspeech | On invariant structural representation for speech recognition: theoretical validation and experimental improvement. | Yu Qiao, Nobuaki Minematsu, Keikichi Hirose |
| 2009 | Interspeech | How to improve TTS systems for emotional expressivity. | Antonio Rui Ferreira Rebordo, Shaikh Mostafa Al Masum, Keikichi Hirose, Nobuaki Minematsu |
| 2009 | Interspeech | Optimal event search using a structural cost function - improvement of structure to speech conversion. | Daisuke Saito, Yu Qiao, Nobuaki Minematsu, Keikichi Hirose |
| 2008 | ICASSP | Multi-stream parameterization for structural speech recognition. | Satoshi Asakawa, Nobuaki Minematsu, Keikichi Hirose |
| 2008 | ICASSP | Unsupervised optimal phoneme segmentation: Objectives, algorithm and comparisons. | Yu Qiao, Naoya Shimomura, Nobuaki Minematsu |
| 2008 | ICASSP | Phase singularities for image representation and matching. | Yu Qiao, Wei Wang, Nobuaki Minematsu, Jianzhuang Liu, Xiaoou Tang |
| 2008 | ICASSP | Directional dependency of cepstrum on vocal tract length. | Daisuke Saito, Ryo Matsuura, Satoshi Asakawa, Nobuaki Minematsu, Keikichi Hirose |
| 2008 | Interspeech | Automatic pronunciation evaluation of language learners' utterances generated through shadowing. | Dean Luo, Naoya Shimomura, Nobuaki Minematsu, Yutaka Yamauchi, Keikichi Hirose |
| 2008 | Interspeech | Robust voiced/unvoiced speech classification using empirical mode decomposition and periodic correlation model. | Md. Khademul Islam Molla, Keikichi Hirose, Nobuaki Minematsu |
| 2008 | Interspeech | Control of prosodic focus in corpus-based generation of fundamental frequency based on the generation process model. | Keiko Ochi, Keikichi Hirose, Nobuaki Minematsu |
| 2008 | Interspeech | Metric learning for unsupervised phoneme segmentation. | Yu Qiao, Nobuaki Minematsu |
| 2008 | Interspeech | f-divergence is a generalized invariant measure between distributions. | Yu Qiao, Nobuaki Minematsu |
| 2008 | Interspeech | Structure to speech conversion - speech generation based on infant-like vocal imitation. | Daisuke Saito, Satoshi Asakawa, Nobuaki Minematsu, Keikichi Hirose |
| 2008 | Interspeech | Decomposition of rotational distortion caused by VTL difference using eigenvalues of its transformation matrix. | Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose |
| 2007 | ASRU | Random discriminant structure analysis for automatic recognition of connected vowels. | Yu Qiao, Satoshi Asakawa, Nobuaki Minematsu |
| 2007 | ICASSP | Development of a Femininity Estimator using Speaker Recognition Techniques for Voice Therapy of Gender Identity Disorder Clients. | Nobuaki Minematsu, Kazutaka Maruyama, Kyoko Sakuraba, Keikichi Hirose, Niro Tayama, Satoshi Imaizumi, Toshio Yamauchi |
| 2007 | Interspeech | Automatic recognition of connected vowels only using speaker-invariant representation of speech dynamics. | Satoshi Asakawa, Nobuaki Minematsu, Keikichi Hirose |
| 2007 | Interspeech | EMD based soft-thresholding for speech enhancement. | Erhan Deger, Md. Khademul Islam Molla, Keikichi Hirose, Nobuaki Minematsu, Md. Kamrul Hasan |
| 2007 | Interspeech | F | Hiroko Hirano, Keikichi Hirose, Goh Kawai, Wentao Gu, Nobuaki Minematsu |
| 2007 | Interspeech | Corpus-based generation of prosodic features from text based on generation process model. | Keikichi Hirose, Keiko Ochi, Nobuaki Minematsu |
| 2007 | Interspeech | Structural assessment of language learners' pronunciation. | Nobuaki Minematsu, K. Kamata, Satoshi Asakawa, Takehiko Makino, Tazuko Nishimura, Keikichi Hirose |
| 2007 | Interspeech | Pitch estimation of noisy speech signals using empirical mode decomposition. | Md. Khademul Islam Molla, Keikichi Hirose, Nobuaki Minematsu, Md. Kamrul Hasan |
| 2007 | Interspeech | A framework of reply speech generation for concept-to-speech conversion in spoken dialogue systems. | Seiya Takada, Yuji Yagi, Keikichi Hirose, Nobuaki Minematsu |
| 2007 | Interspeech | Features of pauses and conjunctions at syntactic and discourse boundaries in Japanese monologues. | Michiko Watanabe, Yasuharu Den, Keikichi Hirose, Shusaku Miwa, Nobuaki Minematsu |
| 2006 | ICASSP | Para-Linguistic Information Represented as Distortion of the Acoustic Universal Structure In Speech. | Nobuaki Minematsu, Satoshi Asakawa, Keikichi Hirose |
| 2006 | ICASSP | Localization Based Separation of Mixed Audio Signals with Binary Masking of Hilbert Spectrum. | Md. Khademul Islam Molla, Keikichi Hirose, Nobuaki Minematsu |
| 2006 | Interspeech | Unfilled pauses in Japanese sentences read aloud by non-native learners. | Hiroko Hirano, Goh Kawai, Keikichi Hirose, Nobuaki Minematsu |
| 2006 | Interspeech | Corpus-based generation of fundamental frequency contours using generation process model and considering emotional focuses. | Keikichi Hirose, Yasufumi Asano, Nobuaki Minematsu |
| 2006 | Interspeech | Tone recognition of continuous speech of standard Chinese using neural network and tone nucleus model. | Keikichi Hirose, Hui Hu, Xiaodong Wang, Nobuaki Minematsu |
| 2006 | Interspeech | Development of a program for self assessment of Japanese pronunciation by English learners. | Chiharu Tsurutani, Yutaka Yamauchi, Nobuaki Minematsu, Dean Luo, Kazutaka Maruyama, Keikichi Hirose |
| 2006 | Interspeech | Factors affecting speakers² choice of fillers in Japanese presentations. | Michiko Watanabe, Yasuharu Den, Keikichi Hirose, Shusaku Miwa, Nobuaki Minematsu |
| 2006 | ISCAS | Localization based audio source separation by sub-band beamforming. | M. Khademul Islam Molla, Keikichi Hirose, Nobuaki Minematsu |
| 2005 | CW | Improved concept-to-speech generation in a dialogue system on road guidance. | Yuji Yagi, Keikichi Hirose, Seiya Takada, Nobuaki Minematsu |
| 2005 | ICASSP | Mathematical Evidence of the Acoustic Universal Structure in Speech. | Nobuaki Minematsu |
| 2005 | Interspeech | Structural representation of the non-native pronunciations. | Satoshi Asakawa, Nobuaki Minematsu, Toshiko Isei-Jaakkola, Keikichi Hirose |
| 2005 | Interspeech | Corpus-based extraction of F0 contour generation process model parameters. | Keikichi Hirose, Yusuke Furuyama, Nobuaki Minematsu |
| 2005 | Interspeech | Multi-band approach of audio source discrimination with empirical mode decomposition. | Md. Khademul Islam Molla, Keikichi Hirose, Nobuaki Minematsu |
| 2005 | Interspeech | Japanese vowel recognition based on structural representation of speech. | Takao Murakami, Kazutaka Maruyama, Nobuaki Minematsu, Keikichi Hirose |
| 2005 | Interspeech | Generation of fundamental frequency contours for Mandarin speech synthesis based on tone nucleus model. | Qinghua Sun, Keikichi Hirose, Wentao Gu, Nobuaki Minematsu |
| 2005 | Interspeech | Filled pauses as cues to the complexity of following phrases. | Michiko Watanabe, Keikichi Hirose, Yasuharu Den, Nobuaki Minematsu |
| 2005 | ISCAS | Audio source separation by source localization with Hilbert spectrum. | M. Khademul Islam Molla, Keikichi Hirose, Nobuaki Minematsu |
| 2004 | ICASSP | Yet another acoustic representation of speech sounds. | Nobuaki Minematsu |
| 2004 | Interspeech | N-gram language modeling of Japanese using bunsetsu boundaries. | Sungyup Chung, Keikichi Hirose, Nobuaki Minematsu |
| 2004 | Interspeech | Use of prosodic features for speech recognition. | Keikichi Hirose, Nobuaki Minematsu |
| 2004 | Interspeech | Pronunciation assessment based upon the compatibility between a learner's pronunciation structure and the target language's lexical structure. | Nobuaki Minematsu |
| 2004 | Interspeech | Pronunciation assessment based upon the phonological distortions observed in language learners' utterances. | Nobuaki Minematsu |
| 2004 | Interspeech | Audio source separation from the mixture using empirical mode decomposition with independent subspace analysis. | Md. Khademul Islam Molla, Keikichi Hirose, Nobuaki Minematsu |
| 2004 | Interspeech | Clause types and filed pauses in Japanese spontaneous monologues. | Michiko Watanabe, Yasuharu Den, Keikichi Hirose, Nobuaki Minematsu |
| 2003 | Interspeech | Use of linguistic information for automatic extraction of f_0 contour generation process model parameters. | Keikichi Hirose, Yusuke Furuyama, Shuichi Narusawa, Nobuaki Minematsu, Hiroya Fujisaki |
| 2003 | Interspeech | A pronunciation training system for Japanese lexical accents with corrective feedback in learner's voice. | Keikichi Hirose, Frdric Gendrin, Nobuaki Minematsu |
| 2003 | Interspeech | Corpus-based synthesis of fundamental frequency contours of Japanese using automatically-generated prosodic corpus and generation process model. | Keikichi Hirose, Takayuki Ono, Nobuaki Minematsu |
| 2003 | Interspeech | Speech generation from concept for realizing conversation with an agent in a virtual room. | Keikichi Hirose, Junji Tago, Nobuaki Minematsu |
| 2003 | Interspeech | CART-based factor analysis of intelligibility reduction in Japanese English. | Nobuaki Minematsu, Changchen Guo, Keikichi Hirose |
| 2003 | Interspeech | Prosodic analysis and modeling of the NAGAUTA singing to synthesize its prosodic patterns from the standard notation. | Nobuaki Minematsu, Bungo Matsuoka, Keikichi Hirose |
| 2003 | Interspeech | Improvement of non-native speech recognition by effectively modeling frequently observed pronunciation habits. | Nobuaki Minematsu, Koichi Osaki, Keikichi Hirose |
| 2003 | Interspeech | Automatic estimation of perceptual age using speaker modeling techniques. | Nobuaki Minematsu, Keita Yamauchi, Keikichi Hirose |
| 2003 | Interspeech | Considerations on vowel durations for Japanese CALL system. | Taro Mouri, Keikichi Hirose, Nobuaki Minematsu |
| 2003 | Interspeech | Estimation of resonant characteristics based on AR-HMM modeling and spectral envelope conversion of vowel sounds. | Nobuyuki Nishizawa, Keikichi Hirose, Nobuaki Minematsu |
| 2002 | ICASSP | Automatic estimation of one's age with his/her speech based upon acoustic modeling techniques of speakers. | Nobuaki Minematsu, Mariko Sekiguchi, Keikichi Hirose |
| 2002 | ICASSP | A method for automatic extraction of model parameters from fundamental frequency contours of speech. | Shuichi Narusawa, Nobuaki Minematsu, Keikichi Hirose, Hiroya Fujisaki |
| 2002 | Interspeech | Improved corpus-based synthesis of fundamental frequency contours using generation process model. | Keikichi Hirose, Masaya Eto, Nobuaki Minematsu |
| 2002 | Interspeech | Statistical language modeling with prosodic boundaries and its use for continuous speech recognition. | Keikichi Hirose, Nobuaki Minematsu, Makoto Terao |
| 2002 | Interspeech | Robust speech recognition using inter-speaker and intra-speaker adaptation. | Baojie Li, Keikichi Hirose, Nobuaki Minematsu |
| 2002 | Interspeech | Integration of MLLR adaptation with pronunciation proficiency adaptation for non-native speech recognition. | Nobuaki Minematsu, Gakuto Kurata, Keikichi Hirose |
| 2002 | Interspeech | Corpus-based analysis of English spoken by Japanese students in view of the entire phonemic system of English. | Nobuaki Minematsu, Gakuto Kurata, Keikichi Hirose |
| 2002 | Interspeech | Acoustic modeling of sentence stress using differential features between syllables for English rhythm learning system development. | Nobuaki Minematsu, Satoshi Kobashikawa, Keikichi Hirose, Donna Erickson |
| 2002 | Interspeech | Automatic extraction of model parameters from fundamental frequency contours of English utterances. | Shuichi Narusawa, Nobuaki Minematsu, Keikichi Hirose, Hiroya Fujisaki |
| 2002 | Interspeech | Separation of voiced source characteristics and vocal tract transfer function characteristics for speech sounds by iterative analysis based on AR-HMM model. | Nobuyuki Nishizawa, Keikichi Hirose, Nobuaki Minematsu |
| 2002 | LREC | English Speech Database Read by Japanese Learners for CALL System Development. | Nobuaki Minematsu, Yoshihiro Tomiyama, Kei Yoshimoto, Katsumasa Shimizu, Seiichi Nakagawa, Masatake Dantsuji, Shozo Makino |
| 2001 | ICASSP | Generation of F | Atsuhiro Sakurai, Keikichi Hirose, Nobuaki Minematsu |
| 2001 | Interspeech | Corpus-based synthesis of fundamental frequency contours based on a generation process model. | Keikichi Hirose, Masaya Eto, Nobuaki Minematsu, Atsuhiro Sakurai |
| 2001 | Interspeech | Identification of accent and intonation in sentences for CALL systems. | Carlos Toshinori Ishi, Nobuaki Minematsu, Ryuji Nishide, Keikichi Hirose |
| 2001 | Interspeech | Use of topic knowledge in spoken dialogue information retrieval system for academic documents. | Shinya Kiriyama, Keikichi Hirose, Nobuaki Minematsu |
| 2001 | Interspeech | Instantaneous estimation of accentuation habits for Japanese students to learn English pronunciation. | Naoki Nakamura, Nobuaki Minematsu, Seiichi Nakagawa |
| 2000 | Interspeech | Analytical and perceptual study on the role of acoustic features in realizing emotional speech. | Keikichi Hirose, Nobuaki Minematsu, Hiromichi Kawanami |
| 2000 | Interspeech | Identification of Japanese double-mora phonemes considering speaking rate for the use in CALL systems. | Carlos Toshinori Ishi, Keikichi Hirose, Nobuaki Minematsu |
| 2000 | Interspeech | Free software toolkit for Japanese large vocabulary continuous speech recognition. | Tatsuya Kawahara, Akinobu Lee, Tetsunori Kobayashi, Kazuya Takeda, Nobuaki Minematsu, Shigeki Sagayama, Katsunobu Itou, Akinori Ito, Mikio Yamamoto, Atsushi Yamada, Takehito Utsuro, Kiyohiro Shikano |
| 2000 | Interspeech | Efficient search strategy in large vocabulary continuous speech recognition using prosodic boundary information. | Shi-wook Lee, Keikichi Hirose, Nobuaki Minematsu |
| 2000 | Interspeech | Modeling phone correlation for speaker adaptive speech recognition. | Baojie Li, Keikichi Hirose, Nobuaki Minematsu |
| 2000 | Interspeech | Performance comparison among HMM, DTW, and human abilities in terms of identifying stress patterns of word utterances. | Nobuaki Minematsu, Yukiko Fujisawa, Seiichi Nakagawa |
| 2000 | Interspeech | Quality improvement of PSOLA analysis-synthesis using partial zero-phase conversion. | Nobuaki Minematsu, Seiichi Nakagawa |
| 2000 | Interspeech | Instantaneous estimation of prosodic pronunciation habits for Japanese students to learn English pronunciation. | Nobuaki Minematsu, Seiichi Nakagawa |
| 2000 | Interspeech | Development of a formant-based analysis-synthesis system and generation of high quality liquid sounds of Japanese. | Nobuyuki Nishizawa, Nobuaki Minematsu, Keikichi Hirose |
| 2000 | Interspeech | Data-driven intonation modeling using a neural network and a command response model. | Atsuhiro Sakurai, Nobuaki Minematsu, Keikichi Hirose |
| 2000 | LREC | IPA Japanese Dictation Free Software Project. | Katsunobu Itou, Kiyohiro Shikano, Tatsuya Kawahara, Kazuya Takeda, Atsushi Yamada, Akinori Ito, Takehito Utsuro, Tetsunori Kobayashi, Nobuaki Minematsu, Mikio Yamamoto, Shigeki Sagayama, Akinobu Lee |
| 1998 | Interspeech | Evaluation of Japanese manners of generating word accent of English based on a stressed syllable detection technique. | Yukiko Fujisawa, Nobuaki Minematsu, Seiichi Nakagawa |
| 1998 | Interspeech | Continuous speech recognition using segmental unit input HMMs with a mixture of probability density functions and context dependency. | Kengo Hanai, Kazumasa Yamamoto, Nobuaki Minematsu, Seiichi Nakagawa |
| 1998 | Interspeech | Sharable software repository for Japanese large vocabulary continuous speech recognition. | Tatsuya Kawahara, Tetsunori Kobayashi, Kazuya Takeda, Nobuaki Minematsu, Katsunobu Itou, Mikio Yamamoto, Atsushi Yamada, Takehito Utsuro, Kiyohiro Shikano |
| 1998 | Interspeech | Modeling of variations in cepstral coefficients caused by F0 changes and its application to speech processing. | Nobuaki Minematsu, Seiichi Nakagawa |
| 1997 | Interspeech | Automatic detection of accent in English words spoken by Japanese students. | Nobuaki Minematsu, Nariaki Ohashi, Seiichi Nakagawa |
| 1996 | Interspeech | Automatic detection of accent nuclei at the head of words for speech recognition. | Nobuaki Minematsu, Seiichi Nakagawa |
| 1996 | Interspeech | Prosodic manipulation system of speech material for perceptual experiments. | Nobuaki Minematsu, Seiichi Nakagawa, Keikichi Hirose |
| 1994 | Interspeech | Speech recognition using HMM with decreased intra-group variation in the temporal structure. | Nobuaki Minematsu, Keikichi Hirose |
| 1994 | Interspeech | Role of prosodic features in the human process of speech perception. | Nobuaki Minematsu, Keikichi Hirose |
| 1992 | Interspeech | The influence of semantic and syntactic information on spoken sentence recognition. | Nobuaki Minematsu, Sumio Ohno, Keikichi Hirose, Hiroya Fujisaki |
| 1990 | Interspeech | Influence of context and knowledge on the perception of continuous speech. | Hiroya Fujisaki, Keikichi Hirose, Sumio Ohno, Nobuaki Minematsu |