Skip to content

Yoshiki Masuyama

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

23

Venues

5

Active years

2018–2025

Best venue rank

Multiconference

Where they publish

Papers

23 indexed papers, newest first.

YearVenueTitleAuthors
2025ASRURobot Confirmation Generation and Action Planning Using Long-context Q-Former Integrated with Multimodal LLM.Chiori Hori, Yoshiki Masuyama, Siddarth Jain, Radu Corcodel, Devesh K. Jha, Diego Romeres, Jonathan Le Roux
2025ICASSPMel-Spectrogram Inversion via Alternating Direction Method of Multipliers.Yoshiki Masuyama, Natsuki Ueno, Nobutaka Ono
2025ICASSPRetrieval-Augmented Neural Field for HRTF Upsampling and Personalization.Yoshiki Masuyama, Gordon Wichern, Franois G. Germain, Christopher Ick, Jonathan Le Roux
2025InterspeechDirection-Aware Neural Acoustic Fields for Few-Shot Interpolation of Ambisonic Impulse Responses.Christopher Ick, Gordon Wichern, Yoshiki Masuyama, Franois G. Germain, Jonathan Le Roux
2025InterspeechFactorized RVQ-GAN For Disentangled Speech Tokenization.Sameer Khurana, Dominik Klement, Antoine Laurent, Dominik Bobos, Juraj Novosad, Peter Gazdik, Ellen Zhang, Zili Huang, Amir Hussein, Ricard Marxer, Yoshiki Masuyama, Ryo Aihara, Chiori Hori, Franois G. Germain, Gordon Wichern, Jonathan Le Roux
2025InterspeechInvestigating continuous autoregressive generative speech enhancement.Haici Yang, Gordon Wichern, Ryo Aihara, Yoshiki Masuyama, Sameer Khurana, Franois G. Germain, Jonathan Le Roux
2025NAACLESPnet-SpeechLM: An Open Speech Language Model Toolkit.Jinchuan Tian, Jiatong Shi, William Chen, Siddhant Arora, Yoshiki Masuyama, Takashi Maekaku, Yihan Wu, Junyi Peng, Shikhar Bharadwaj, Yiwen Zhao, Samuele Cornell, Yifan Peng, Xiang Yue, Chao-Han Huck Yang, Graham Neubig, Shinji Watanabe
2024ICASSPNIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization.Yoshiki Masuyama, Gordon Wichern, Franois G. Germain, Zexu Pan, Sameer Khurana, Chiori Hori, Jonathan Le Roux
2024InterspeechExploring the Capability of Mamba in Speech Applications.Koichi Miyazaki, Yoshiki Masuyama, Masato Murata
2023ASRUScenario-Aware Audio-Visual TF-Gridnet for Target Speech Extraction.Zexu Pan, Gordon Wichern, Yoshiki Masuyama, Franois G. Germain, Sameer Khurana, Chiori Hori, Jonathan Le Roux
2023ICASSPMulti-Channel Speaker Extraction with Adversarial Training: The Wavlab Submission to The Clarity ICASSP 2023 Grand Challenge.Samuele Cornell, Zhong-Qiu Wang, Yoshiki Masuyama, Shinji Watanabe, Manuel Pariente, Nobutaka Ono, Stefano Squartini
2022InterspeechESPnet-SE++: Speech Enhancement for Robust Speech Recognition, Translation, and Understanding.Yen-Ju Lu, Xuankai Chang, Chenda Li, Wangyou Zhang, Samuele Cornell, Zhaoheng Ni, Yoshiki Masuyama, Brian Yan, Robin Scheibler, Zhong-Qiu Wang, Yu Tsao, Yanmin Qian, Shinji Watanabe
2022InterspeechJoint Optimization of Sampling Rate Offsets Based on Entire Signal Relationship Among Distributed Microphones.Yoshiki Masuyama, Kouei Yamaoka, Nobutaka Ono
2020ICASSPSpeech Enhancement Using Self-Adaptation and Multi-Head Self-Attention.Yuma Koizumi, Kohei Yatabe, Marc Delcroix, Yoshiki Masuyama, Daiki Takeuchi
2020ICASSPConsistency-Aware Multi-Channel Speech Enhancement Using Deep Neural Networks.Yoshiki Masuyama, Masahito Togami, Tatsuya Komatsu
2020ICASSPPhase Reconstruction Based On Recurrent Phase Unwrapping With Deep Neural Networks.Yoshiki Masuyama, Kohei Yatabe, Yuma Koizumi, Yasuhiro Oikawa, Noboru Harada
2020ICASSPUnsupervised Training for Deep Speech Source Separation with Kullback-Leibler Divergence Based Probabilistic Loss Function.Masahito Togami, Yoshiki Masuyama, Tatsuya Komatsu, Yu Nakagome
2020IROSSelf-supervised Neural Audio-Visual Sound Source Localization via Probabilistic Spatial Modeling.Yoshiki Masuyama, Yoshiaki Bando, Kohei Yatabe, Yoko Sasaki, Masaki Onishi, Yasuhiro Oikawa
2019ICASSPDeep Griffin-Lim Iteration.Yoshiki Masuyama, Kohei Yatabe, Yuma Koizumi, Yasuhiro Oikawa, Noboru Harada
2019ICASSPLow-rankness of Complex-valued Spectrogram and Its Application to Phase-aware Audio Processing.Yoshiki Masuyama, Kohei Yatabe, Yasuhiro Oikawa
2019ICASSPPhase-aware Harmonic/percussive Source Separation via Convex Optimization.Yoshiki Masuyama, Kohei Yatabe, Yasuhiro Oikawa
2019InterspeechMultichannel Loss Function for Supervised Speech Source Separation by Mask-Based Beamforming.Yoshiki Masuyama, Masahito Togami, Tatsuya Komatsu
2018ICASSPModal Decomposition of Musical Instrument Sound Via Alternating Direction Method of Multipliers.Yoshiki Masuyama, Tsubasa Kusano, Kohei Yatabe, Yasuhiro Oikawa