Yoshiki Masuyama
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
23
Venues
5
Active years
2018–2025
Best venue rank
Multiconference
Where they publish
Papers
23 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ASRU | Robot Confirmation Generation and Action Planning Using Long-context Q-Former Integrated with Multimodal LLM. | Chiori Hori, Yoshiki Masuyama, Siddarth Jain, Radu Corcodel, Devesh K. Jha, Diego Romeres, Jonathan Le Roux |
| 2025 | ICASSP | Mel-Spectrogram Inversion via Alternating Direction Method of Multipliers. | Yoshiki Masuyama, Natsuki Ueno, Nobutaka Ono |
| 2025 | ICASSP | Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization. | Yoshiki Masuyama, Gordon Wichern, Franois G. Germain, Christopher Ick, Jonathan Le Roux |
| 2025 | Interspeech | Direction-Aware Neural Acoustic Fields for Few-Shot Interpolation of Ambisonic Impulse Responses. | Christopher Ick, Gordon Wichern, Yoshiki Masuyama, Franois G. Germain, Jonathan Le Roux |
| 2025 | Interspeech | Factorized RVQ-GAN For Disentangled Speech Tokenization. | Sameer Khurana, Dominik Klement, Antoine Laurent, Dominik Bobos, Juraj Novosad, Peter Gazdik, Ellen Zhang, Zili Huang, Amir Hussein, Ricard Marxer, Yoshiki Masuyama, Ryo Aihara, Chiori Hori, Franois G. Germain, Gordon Wichern, Jonathan Le Roux |
| 2025 | Interspeech | Investigating continuous autoregressive generative speech enhancement. | Haici Yang, Gordon Wichern, Ryo Aihara, Yoshiki Masuyama, Sameer Khurana, Franois G. Germain, Jonathan Le Roux |
| 2025 | NAACL | ESPnet-SpeechLM: An Open Speech Language Model Toolkit. | Jinchuan Tian, Jiatong Shi, William Chen, Siddhant Arora, Yoshiki Masuyama, Takashi Maekaku, Yihan Wu, Junyi Peng, Shikhar Bharadwaj, Yiwen Zhao, Samuele Cornell, Yifan Peng, Xiang Yue, Chao-Han Huck Yang, Graham Neubig, Shinji Watanabe |
| 2024 | ICASSP | NIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization. | Yoshiki Masuyama, Gordon Wichern, Franois G. Germain, Zexu Pan, Sameer Khurana, Chiori Hori, Jonathan Le Roux |
| 2024 | Interspeech | Exploring the Capability of Mamba in Speech Applications. | Koichi Miyazaki, Yoshiki Masuyama, Masato Murata |
| 2023 | ASRU | Scenario-Aware Audio-Visual TF-Gridnet for Target Speech Extraction. | Zexu Pan, Gordon Wichern, Yoshiki Masuyama, Franois G. Germain, Sameer Khurana, Chiori Hori, Jonathan Le Roux |
| 2023 | ICASSP | Multi-Channel Speaker Extraction with Adversarial Training: The Wavlab Submission to The Clarity ICASSP 2023 Grand Challenge. | Samuele Cornell, Zhong-Qiu Wang, Yoshiki Masuyama, Shinji Watanabe, Manuel Pariente, Nobutaka Ono, Stefano Squartini |
| 2022 | Interspeech | ESPnet-SE++: Speech Enhancement for Robust Speech Recognition, Translation, and Understanding. | Yen-Ju Lu, Xuankai Chang, Chenda Li, Wangyou Zhang, Samuele Cornell, Zhaoheng Ni, Yoshiki Masuyama, Brian Yan, Robin Scheibler, Zhong-Qiu Wang, Yu Tsao, Yanmin Qian, Shinji Watanabe |
| 2022 | Interspeech | Joint Optimization of Sampling Rate Offsets Based on Entire Signal Relationship Among Distributed Microphones. | Yoshiki Masuyama, Kouei Yamaoka, Nobutaka Ono |
| 2020 | ICASSP | Speech Enhancement Using Self-Adaptation and Multi-Head Self-Attention. | Yuma Koizumi, Kohei Yatabe, Marc Delcroix, Yoshiki Masuyama, Daiki Takeuchi |
| 2020 | ICASSP | Consistency-Aware Multi-Channel Speech Enhancement Using Deep Neural Networks. | Yoshiki Masuyama, Masahito Togami, Tatsuya Komatsu |
| 2020 | ICASSP | Phase Reconstruction Based On Recurrent Phase Unwrapping With Deep Neural Networks. | Yoshiki Masuyama, Kohei Yatabe, Yuma Koizumi, Yasuhiro Oikawa, Noboru Harada |
| 2020 | ICASSP | Unsupervised Training for Deep Speech Source Separation with Kullback-Leibler Divergence Based Probabilistic Loss Function. | Masahito Togami, Yoshiki Masuyama, Tatsuya Komatsu, Yu Nakagome |
| 2020 | IROS | Self-supervised Neural Audio-Visual Sound Source Localization via Probabilistic Spatial Modeling. | Yoshiki Masuyama, Yoshiaki Bando, Kohei Yatabe, Yoko Sasaki, Masaki Onishi, Yasuhiro Oikawa |
| 2019 | ICASSP | Deep Griffin-Lim Iteration. | Yoshiki Masuyama, Kohei Yatabe, Yuma Koizumi, Yasuhiro Oikawa, Noboru Harada |
| 2019 | ICASSP | Low-rankness of Complex-valued Spectrogram and Its Application to Phase-aware Audio Processing. | Yoshiki Masuyama, Kohei Yatabe, Yasuhiro Oikawa |
| 2019 | ICASSP | Phase-aware Harmonic/percussive Source Separation via Convex Optimization. | Yoshiki Masuyama, Kohei Yatabe, Yasuhiro Oikawa |
| 2019 | Interspeech | Multichannel Loss Function for Supervised Speech Source Separation by Mask-Based Beamforming. | Yoshiki Masuyama, Masahito Togami, Tatsuya Komatsu |
| 2018 | ICASSP | Modal Decomposition of Musical Instrument Sound Via Alternating Direction Method of Multipliers. | Yoshiki Masuyama, Tsubasa Kusano, Kohei Yatabe, Yasuhiro Oikawa |