| 2022 | Interspeech | Speak Like a Professional: Increasing Speech Intelligibility by Mimicking Professional Announcer Voice with Voice Conversion. | Tuan Vu Ho, Maori Kobayashi, Masato Akagi |
| 2022 | Interspeech | Vector-quantized Variational Autoencoder for Phase-aware Speech Enhancement. | Tuan Vu Ho, Quoc Huy Nguyen, Masato Akagi, Masashi Unoki |
| 2022 | Interspeech | Data Augmentation Using McAdams-Coefficient-Based Speaker Anonymization for Fake Audio Detection. | Kai Li, Sheng Li, Xugang Lu, Masato Akagi, Meng Liu, Lin Zhang, Chang Zeng, Longbiao Wang, Jianwu Dang, Masashi Unoki |
| 2020 | ICASSP | Multitask Learning and Multistage Fusion for Dimensional Audiovisual Emotion Recognition. | Bagus Tris Atmaja, Masato Akagi |
| 2020 | Interspeech | Segment-Level Effects of Gender, Nationality and Emotion Information on Text-Independent Speaker Verification. | Kai Li, Masato Akagi, Yibo Wu, Jianwu Dang |
| 2020 | Interspeech | Comparison of Glottal Source Parameter Values in Emotional Vowels. | Yongwei Li, Jianhua Tao, Bin Liu, Donna Erickson, Masato Akagi |
| 2020 | Tencon | On The Differences Between Song and Speech Emotion Recognition: Effect of Feature Sets, Feature Types, and Classifiers. | Bagus Tris Atmaja, Masato Akagi |
| 2020 | Tencon | Predicting Valence and Arousal by Aggregating Acoustic Features for Acoustic-Linguistic Information Fusion. | Bagus Tris Atmaja, Yasuhiro Hamada, Masato Akagi |
| 2019 | Interspeech | The Contribution of Acoustic Features Analysis to Model Emotion Perceptual Process for Language Diversity. | Xingfeng Li, Masato Akagi |
| 2018 | Interspeech | A Three-Layer Emotion Perception Model for Valence and Arousal-Based Detection from Multilingual Speech. | Xingfeng Li, Masato Akagi |
| 2017 | ICONIP | Weighted Robust Principal Component Analysis with Gammatone Auditory Filterbank for Singing Voice Separation. | Feng Li, Masato Akagi |
| 2016 | Interspeech | Multilingual Speech Emotion Recognition System Based on a Three-Layer Model. | Xingfeng Li, Masato Akagi |
| 2013 | Interspeech | Comparative investigation of objective speech intelligibility prediction measures for noise-reduced signals in Mandarin and Japanese. | Junfeng Li, Fei Chen, Masato Akagi, Yonghong Yan |
| 2012 | ICASSP | Evaluation of objective intelligibility prediction measures for noise-reduced signals in mandarin. | Risheng Xia, Junfeng Li, Masato Akagi, Yonghong Yan |
| 2011 | Interspeech | Voice Activity Detection in MTF-Based Power Envelope Restoration. | Masashi Unoki, Xugang Lu, Rico Petrick, Shota Morita, Masato Akagi, Rdiger Hoffmann |
| 2010 | Interspeech | A DOA estimation algorithm based on equalization-cancellation theory. | Duc Thanh Chau, Junfeng Li, Masato Akagi |
| 2009 | ICASSP | Psychoacoustically-motivated adaptive beta-order generalized spectral subtraction for cochlear implant patients. | Junfeng Li, Qian-Jie Fu, Hui Jiang, Masato Akagi |
| 2009 | Interspeech | Efficient modeling of temporal structure of speech for applications in voice transformation. | Binh Phu Nguyen, Masato Akagi |
| 2008 | Interspeech | Psychoacoustically-motivated adaptive β-order generalized spectral subtraction based on data-driven optimization. | Junfeng Li, Hui Jiang, Masato Akagi |
| 2008 | Interspeech | High-quality analysis/synthesis method based on temporal decomposition for speech modification. | Binh Phu Nguyen, Takeshi Shibata, Masato Akagi |
| 2008 | Interspeech | Robust front end processing for speech recognition in reverberant environments: utilization of speech characteristics. | Rico Petrick, Xugang Lu, Masashi Unoki, Masato Akagi, Rdiger Hoffmann |
| 2007 | Interspeech | A rule-based speech morphing for verifying a expressive speech perception model. | Chun-Fang Huang, Masato Akagi |
| 2007 | Interspeech | Noise reduction based on adaptive β-order generalized spectral subtraction for speech enhancement. | Junfeng Li, Shuichi Sakamoto, Satoshi Hongo, Masato Akagi, Yiti Suzuki |
| 2007 | Interspeech | A flexible spectral modification method based on temporal decomposition and Gaussian mixture model. | Binh Phu Nguyen, Masato Akagi |
| 2007 | Interspeech | Vocal conversion from speaking voice to singing voice using STRAIGHT. | Takeshi Saitou, Masataka Goto, Masashi Unoki, Masato Akagi |
| 2007 | Interspeech | Method of LP-based blind restoration for improving intelligibility of bone-conducted speech. | Thang Tat Vu, Germine Seide, Masashi Unoki, Masato Akagi |
| 2006 | Interspeech | Improved hybrid microphone array post-filter by integrating a robust speech absence probability estimator for speech enhancement. | Junfeng Li, Masato Akagi, Yiti Suzuki |
| 2006 | Interspeech | A robust feature extraction based on the MTF concept for speech recognition in reverberant environment. | Xugang Lu, Masashi Unoki, Masato Akagi |
| 2005 | ACII | Toward a Rule-Based Synthesis of Emotional Speech on Linguistic Descriptions of Perception. | Chun-Fang Huang, Masato Akagi |
| 2005 | ICASSP | A noise reduction system in arbitrary noise environments and its applications to speech enhancement and speech recognition. | Junfeng Li, Xugang Lu, Masato Akagi |
| 2005 | Interspeech | A multi-layer fuzzy logical model for emotional speech perception. | Chun-Fang Huang, Masato Akagi |
| 2005 | Interspeech | A hybrid microphone array post-filter in a diffuse noise field. | Junfeng Li, Masato Akagi |
| 2005 | Interspeech | A model for selective segregation of a target instrument sound from the mixed sound of various instruments. | Masashi Unoki, Masaaki Kubo, Atsushi Haniu, Masato Akagi |
| 2004 | Interspeech | Noise reduction using hybrid noise estimation technique and post-filtering. | Junfeng Li, Masato Akagi |
| 2004 | Interspeech | Analysis of acoustic features affecting "singing-ness" and its application to singing-voice synthesis from speaking-voice. | Takeshi Saitou, Naoya Tsuji, Masashi Unoki, Masato Akagi |
| 2003 | ICASSP | Temporal decomposition: a promising approach to VQ-based speaker identification. | Phu Chien Nguyen, Masato Akagi, Tu Bao Ho |
| 2003 | ICASSP | A method based on the MTF concept for dereverberating the power envelope from the reverberant signal. | Masashi Unoki, Masashi Furukawa, Keigo Sakata, Masato Akagi |
| 2003 | Interspeech | Efficient quantization of speech excitation parameters using temporal decomposition. | Phu Chien Nguyen, Masato Akagi |
| 2003 | Interspeech | A speech dereverberation method based on the MTF concept. | Masashi Unoki, Keigo Sakata, Masato Akagi |
| 2002 | ICASSP | Noise reduction using a small-scale microphone array in multi noise source environment. | Masato Akagi, Takashi Kago |
| 2002 | ICASSP | Improvement of the restricted temporal decomposition method for line spectral frequency parameters. | Phu Chien Nguyen, Masato Akagi |
| 2002 | Interspeech | Coding speech at very low rates using straight and temporal decomposition. | Phu Chien Nguyen, Takao Ochi, Masato Akagi |
| 2001 | Interspeech | A fundamental frequency estimation method for noisy speech based on instantaneous amplitude and frequency. | Yuichi Ishimoto, Masashi Unoki, Masato Akagi |
| 2000 | Interspeech | Perception of synthesized singing voices with fine fluctuations in their fundamental frequency contours. | Masato Akagi, Hironori Kitakaze |
| 2000 | Interspeech | Design of robust subtractive beamformer for noisy speech recognition. | Mitsunori Mizumachi, Masato Akagi, Satoshi Nakamura |
| 1999 | Interspeech | An objective distortion estimator for hearing aids and its application to noise reduction. | Mitsunori Mizumachi, Masato Akagi |
| 1999 | Interspeech | Segregation of vowel in background noise using the model of segregating two acoustic sources based on auditory scene analysis. | Masashi Unoki, Masato Akagi |
| 1998 | ICASSP | Noise reduction by paired-microphones using spectral subtraction. | Mitsunori Mizumachi, Masato Akagi |
| 1998 | ICASSP | Spectral stability based event localizing temporal decomposition. | A. C. R. Nandasena, Masato Akagi |
| 1998 | Interspeech | Fundamental frequency fluctuation in continuous vowel utterance and its perception. | Masato Akagi, Mamoru Iwaki, Tomoya Minakawa |
| 1998 | Interspeech | Spectral sequence compensation based on continuity of spectral sequence. | Masato Akagi, Mamoru Iwaki, Noriyoshi Sakaguchi |
| 1998 | Interspeech | Signal extraction from noisy signal based on auditory scene analysis. | Masashi Unoki, Masato Akagi |
| 1997 | Interspeech | Noise reduction by paired microphones. | Masato Akagi, Mitsunori Mizumachi |
| 1997 | Interspeech | A method of signal extraction from noisy signal. | Masashi Unoki, Masato Akagi |
| 1996 | Interspeech | Modeling of contextual effects and its application to word spotting. | Yuji Yonezawa, Masato Akagi |
| 1995 | Interspeech | Speaker individualities in fundamental frequency contours and its control. | Masato Akagi, Taw Ienaga |
| 1994 | Interspeech | Perception of central vowel with pre- and post-anchors. | Masato Akagi, Astrid van Wieringen, Louis C. W. Pols |
| 1994 | Interspeech | Speaker individualities in speech spectral envelopes. | Tatsuya Kitamura, Masato Akagi |
| 1990 | Interspeech | Contextual effect models and psycho acoustic evidence for the models. | Masato Akagi |
| 1988 | ICASSP | On the application of spectrum target prediction model to speech recognition. | Masato Akagi, Yoh'ichi Tohkura |