| 2025 | COLING | Indonesian Speech Content De-Identification in Low Resource Transcripts. | Rifqi Naufal Abdjul, Dessi Puji Lestari, Ayu Purwarianti, Candy Olivia Mawalim, Sakriani Sakti, Masashi Unoki |
| 2025 | ICASSP | Fine-tuning TitaNet-Large Model for Speaker Anonymization Attacker Systems. | Candy Olivia Mawalim, Aulia Adila, Masashi Unoki |
| 2025 | ICMI | Culture-Aware Multimodal Personality Prediction using Audio, Pose, and Cultural Embeddings. | Islam J. A. M. Samiul, Khalid Zaman, Marius Funk, Masashi Unoki, Yukiko Nakano, Shogo Okada |
| 2024 | Interspeech | Are Recent Deep Learning-Based Speech Enhancement Methods Ready to Confront Real-World Noisy Environments? | Candy Olivia Mawalim, Shogo Okada, Masashi Unoki |
| 2023 | ICASSP | An Improved Optimal Transport Kernel Embedding Method with Gating Mechanism for Singing Voice Separation and Speaker Identification. | Weitao Yuan, Yuren Bian, Shengbei Wang, Masashi Unoki, Wenwu Wang |
| 2023 | Interspeech | Consonant-emphasis Method Incorporating Robust Consonant-section Detection to Improve Intelligibility of Bone-conducted speech. | Yasufumi Uezu, Sicheng Wang, Teruki Toya, Masashi Unoki |
| 2022 | ICONIP | An Improved Stimulus Reconstruction Method for EEG-Based Short-Time Auditory Attention Detection. | Kai Yang, Zhuo Zhang, Gaoyan Zhang, Masashi Unoki, Jianwu Dang, Longbiao Wang |
| 2022 | Interspeech | Vector-quantized Variational Autoencoder for Phase-aware Speech Enhancement. | Tuan Vu Ho, Quoc Huy Nguyen, Masato Akagi, Masashi Unoki |
| 2022 | Interspeech | Data Augmentation Using McAdams-Coefficient-Based Speaker Anonymization for Fake Audio Detection. | Kai Li, Sheng Li, Xugang Lu, Masato Akagi, Meng Liu, Lin Zhang, Chang Zeng, Longbiao Wang, Jianwu Dang, Masashi Unoki |
| 2022 | Interspeech | Global Signal-to-noise Ratio Estimation Based on Multi-subband Processing Using Convolutional Neural Network. | Nan Li, Meng Ge, Longbiao Wang, Masashi Unoki, Sheng Li, Jianwu Dang |
| 2022 | Interspeech | Automatic Mean Opinion Score Estimation with Temporal Modulation Features on Gammatone Filterbank for Speech Assessment. | Huy Nguyen, Kai Li, Masashi Unoki |
| 2022 | Interspeech | Method for improving the word intelligibility of presented speech using bone-conduction headphones. | Teruki Toya, Wenyu Zhu, Maori Kobayashi, Kenichi Nakamura, Masashi Unoki |
| 2021 | ICASSP | Robust Voice Activity Detection Using a Masked Auditory Encoder Based Convolutional Neural Network. | Nan Li, Longbiao Wang, Masashi Unoki, Sheng Li, Rui Wang, Meng Ge, Jianwu Dang |
| 2021 | ICASSP | Synchronous Multi-Bit Audio Watermarking Based on Phase Shifting. | Shengbei Wang, Weitao Yuan, Zhen Zhang, Jianming Wang, Masashi Unoki |
| 2021 | Interspeech | Crossfire Conditional Generative Adversarial Networks for Singing Voice Extraction. | Weitao Yuan, Shengbei Wang, Xiangrui Li, Masashi Unoki, Wenwu Wang |
| 2020 | Interspeech | X-Vector Singular Value Modification and Statistical-Based Decomposition with Ensemble Regression Modeling for Speaker Anonymization System. | Candy Olivia Mawalim, Kasorn Galajit, Jessada Karnjana, Masashi Unoki |
| 2020 | Interspeech | Cortical Oscillatory Hierarchy for Natural Sentence Processing. | Bin Zhao, Jianwu Dang, Gaoyan Zhang, Masashi Unoki |
| 2019 | HCI | Multimodal BigFive Personality Trait Analysis Using Communication Skill Indices and Multiple Discussion Types Dataset. | Candy Olivia Mawalim, Shogo Okada, Yukiko I. Nakano, Masashi Unoki |
| 2019 | ICASSP | Inaudible Speech Watermarking Based on Self-compensated Echo-hiding and Sparse Subspace Clustering. | Shengbei Wang, Weitao Yuan, Jianming Wang, Masashi Unoki |
| 2019 | ICASSP | Proximal Deep Recurrent Neural Network for Monaural Singing Voice Separation. | Weitao Yuan, Shengbei Wang, Xiangrui Li, Masashi Unoki, Wenwu Wang |
| 2018 | ICASSP | Method of Estimating Direction of Arrival of Sound Source for Monaural Hearing Based on Temporal Modulation Perception. | Nguyen Khanh Bui, Daisuke Morikawa, Masashi Unoki |
| 2018 | ICASSP | Speech Watermarking Based on Robust Principal Component Analysis and Formant Manipulations. | Shengbei Wang, Weitao Yuan, Jianming Wang, Masashi Unoki |
| 2017 | Interspeech | Robust Method for Estimating F | Kenichiro Miwa, Masashi Unoki |
| 2016 | ICASSP | Investigations into vowel and consonant structures in articulatory and auditory spaces using Laplacian eigenmaps. | Jianwu Dang, Shengbei Wang, Masashi Unoki |
| 2016 | Interspeech | Modulation Spectral Features for Predicting Vocal Emotion Recognition by Simulated Cochlear Implants. | Zhi Zhu, Ryota Miyauchi, Yukiko Araki, Masashi Unoki |
| 2015 | ICASSP | Robust and reliable audio watermarking based on phase coding. | Nhut Minh Ngo, Masashi Unoki |
| 2015 | Interspeech | Complex tensor factorization in modulation frequency domain for single-channel speech enhancement. | Shogo Masaya, Masashi Unoki |
| 2014 | ICASSP | Restoration of instantaneous amplitude and phase using Kalman filter for speech enhancement. | Naushin Nower, Yang Liu, Masashi Unoki |
| 2014 | Interspeech | Formant enhancement based speech watermarking for tampering detection. | Shengbei Wang, Masashi Unoki, Nam Soo Kim |
| 2014 | IWDW | An Audio Watermarking Scheme Based on Singular-Spectrum Analysis. | Jessada Karnjana, Masashi Unoki, Pakinee Aimmanee, Chai Wutiwiwatchai |
| 2014 | IWDW | Watermarking for Digital Audio Based on Adaptive Phase Modulation. | Nhut Minh Ngo, Masashi Unoki |
| 2013 | Interspeech | Concurrent processing of voice activity detection and noise reduction using empirical mode decomposition and modulation spectrum analysis. | Yasuaki Kanai, Shota Morita, Masashi Unoki |
| 2011 | Interspeech | Adaptive Regularization Framework for Robust Voice Activity Detection. | Xugang Lu, Masashi Unoki, Ryosuke Isotani, Hisashi Kawai, Satoshi Nakamura |
| 2011 | Interspeech | Voice Activity Detection in MTF-Based Power Envelope Restoration. | Masashi Unoki, Xugang Lu, Rico Petrick, Shota Morita, Masato Akagi, Rdiger Hoffmann |
| 2010 | Interspeech | Voice activity detection in a reguarized reproducing kernel hilbert space. | Xugang Lu, Masashi Unoki, Ryosuke Isotani, Hisashi Kawai, Satoshi Nakamura |
| 2010 | Interspeech | Methods for robust speech recognition in reverberant environments: a comparison. | Rico Petrick, Thomas Fehr, Masashi Unoki, Rdiger Hoffmann |
| 2009 | ICASSP | Temporal contrast normalization and edge-preserved smoothing on temporal modulation structure for robust speech recognition. | Xugang Lu, Shigeki Matsuda, Masashi Unoki, Tohru Shimizu, Satoshi Nakamura |
| 2009 | Interspeech | Subband temporal modulation spectrum normalization for automatic speech recognition in reverberant environments. | Xugang Lu, Masashi Unoki, Satoshi Nakamura |
| 2008 | ICASSP | Comparative evaluations of robust and accurate F0 estimates in reverberant environments. | Masashi Unoki, Toshihiro Hosorogiya, Yuichi Ishimoto |
| 2008 | Interspeech | Robust front end processing for speech recognition in reverberant environments: utilization of speech characteristics. | Rico Petrick, Xugang Lu, Masashi Unoki, Masato Akagi, Rdiger Hoffmann |
| 2008 | Interspeech | A comprehensive study on the effects of room reverberation on fundamental frequency estimation. | Rico Petrick, Masashi Unoki, Anish Mittal, Carlos Segura, Rdiger Hoffmann |
| 2007 | Interspeech | Vocal conversion from speaking voice to singing voice using STRAIGHT. | Takeshi Saitou, Masataka Goto, Masashi Unoki, Masato Akagi |
| 2007 | Interspeech | Method of LP-based blind restoration for improving intelligibility of bone-conducted speech. | Thang Tat Vu, Germine Seide, Masashi Unoki, Masato Akagi |
| 2006 | Interspeech | A robust feature extraction based on the MTF concept for speech recognition in reverberant environment. | Xugang Lu, Masashi Unoki, Masato Akagi |
| 2005 | Interspeech | A model for selective segregation of a target instrument sound from the mixed sound of various instruments. | Masashi Unoki, Masaaki Kubo, Atsushi Haniu, Masato Akagi |
| 2004 | Interspeech | Analysis of acoustic features affecting "singing-ness" and its application to singing-voice synthesis from speaking-voice. | Takeshi Saitou, Naoya Tsuji, Masashi Unoki, Masato Akagi |
| 2003 | ICASSP | A method based on the MTF concept for dereverberating the power envelope from the reverberant signal. | Masashi Unoki, Masashi Furukawa, Keigo Sakata, Masato Akagi |
| 2003 | Interspeech | A speech dereverberation method based on the MTF concept. | Masashi Unoki, Keigo Sakata, Masato Akagi |
| 2001 | Interspeech | A fundamental frequency estimation method for noisy speech based on instantaneous amplitude and frequency. | Yuichi Ishimoto, Masashi Unoki, Masato Akagi |
| 1999 | Interspeech | Segregation of vowel in background noise using the model of segregating two acoustic sources based on auditory scene analysis. | Masashi Unoki, Masato Akagi |
| 1998 | ICASSP | A time-varying, analysis/synthesis auditory filterbank using the gammachirp. | Toshio Irino, Masashi Unoki |
| 1998 | Interspeech | Signal extraction from noisy signal based on auditory scene analysis. | Masashi Unoki, Masato Akagi |
| 1997 | Interspeech | A method of signal extraction from noisy signal. | Masashi Unoki, Masato Akagi |