| 2025 | ASRU | KyotoMOS2: MOS Prediction for Speech Across Multiple Sampling Rates. | Wangjin Zhou, Yizhou Zhang, Keisuke Imoto, Tatsuya Kawahara |
| 2025 | ICASSP | Formula-Supervised Sound Event Detection: Pre-Training Without Real Data. | Yuto Shibata, Keitaro Tanaka, Yoshiaki Bando, Keisuke Imoto, Hirokatsu Kataoka, Yoshimitsu Aoki |
| 2025 | ICASSP | Trainingless Adaptation of Pretrained Models for Environmental Sound Classification. | Noriyuki Tonami, Wataru Kohno, Keisuke Imoto, Yoshiyuki Yajima, Sakiko Mishima, Reishi Kondo, Tomoyuki Hino |
| 2025 | Interspeech | Discrete Tokens Exhibit Interlanguage Speech Intelligibility Benefit: an Analytical Study Towards Accent-robust ASR Only with Native Speech Data. | Kentaro Onda, Keisuke Imoto, Satoru Fukayama, Daisuke Saito, Nobuaki Minematsu |
| 2025 | Interspeech | Prosodically Enhanced Foreign Accent Simulation by Discrete Token-based Resynthesis Only with Native Speech Corpora. | Kentaro Onda, Keisuke Imoto, Satoru Fukayama, Daisuke Saito, Nobuaki Minematsu |
| 2025 | Interspeech | Training Onset-and-Offset-Aware Sound Event Detection on a Heterogeneous Dataset via Probabilistic Sequential Modeling. | Tomoya Yoshinaga, Yoshiaki Bando, Keitaro Tanaka, Keisuke Imoto, Masaki Onishi, Shigeo Morishima |
| 2024 | ICASSP | Environmental Sound Synthesis from Vocal Imitations and Sound Event Labels. | Yuki Okamoto, Keisuke Imoto, Shinnosuke Takamichi, Ryotaro Nagase, Takahiro Fukumori, Yoichi Yamashita |
| 2024 | ICASSP | F1-EV score: Measuring The Likelihood of Estimating a Good Decision Threshold for Semi-Supervised Anomaly Detection. | Kevin Wilkinghoff, Keisuke Imoto |
| 2024 | Interspeech | M2D-CLAP: Masked Modeling Duo Meets CLAP for Learning General-purpose Audio-Language Representation. | Daisuke Niizumi, Daiki Takeuchi, Yasunori Ohishi, Noboru Harada, Masahiro Yasuda, Shunsuke Tsubaki, Keisuke Imoto |
| 2023 | ICASSP | Visual Onoma-to-Wave: Environmental Sound Synthesis from Visual Onomatopoeias and Sound-Source Images. | Hien Ohnaka, Shinnosuke Takamichi, Keisuke Imoto, Yuki Okamoto, Kazuki Fujii, Hiroshi Saruwatari |
| 2023 | Interspeech | CAPTDURE: Captioned Sound Dataset of Single Sources. | Yuki Okamoto, Kanta Shimonishi, Keisuke Imoto, Kota Dohi, Shota Horiguchi, Yohei Kawaguchi |
| 2022 | ICASSP | Environmental Sound Extraction Using Onomatopoeic Words. | Yuki Okamoto, Shota Horiguchi, Masaaki Yamamoto, Keisuke Imoto, Yohei Kawaguchi |
| 2022 | ICASSP | Sound Event Detection Guided by Semantic Contexts of Scenes. | Noriyuki Tonami, Keisuke Imoto, Ryotaro Nagase, Yuki Okamoto, Takahiro Fukumori, Yoichi Yamashita |
| 2021 | ICASSP | Impact of Sound Duration and Inactive Frames on Sound Event Detection Performance. | Keisuke Imoto, Sakiko Mishima, Yumi Arai, Reishi Kondo |
| 2021 | ICASSP | Sound Event Detection Based on Curriculum Learning Considering Learning Difficulty of Events. | Noriyuki Tonami, Keisuke Imoto, Yuki Okamoto, Takahiro Fukumori, Yoichi Yamashita |
| 2020 | ICASSP | Sound Event Detection by Multitask Learning of Sound Events and Scenes with Soft Scene Labels. | Keisuke Imoto, Noriyuki Tonami, Yuma Koizumi, Masahiro Yasuda, Ryosuke Yamanishi, Yoichi Yamashita |
| 2020 | ICASSP | Scene-Dependent Acoustic Event Detection with Scene Conditioning and Fake-Scene-Conditioned Loss. | Tatsuya Komatsu, Keisuke Imoto, Masahito Togami |
| 2020 | ICASSP | Sound Event Localization Based on Sound Intensity Vector Refined by Dnn-Based Denoising and Source Separation. | Masahiro Yasuda, Yuma Koizumi, Shoichiro Saito, Hisashi Uematsu, Keisuke Imoto |
| 2019 | ICASSP | Sound Event Detection Using Graph Laplacian Regularization Based on Event Co-occurrence. | Keisuke Imoto, Seisuke Kyochi |
| 2019 | ICASSP | Joint Acoustic and Class Inference for Weakly Supervised Sound Event Detection. | Sandeep Kothinti, Keisuke Imoto, Debmalya Chakrabarty, Gregory Sell, Shinji Watanabe, Mounya Elhilali |
| 2017 | MMSP | Acoustic scene classification using asynchronous multichannel observations with different lengths. | Keisuke Imoto, Nobutaka Ono |
| 2015 | ICASSP | Acoustic scene analysis from acoustic event sequence with intermittent missing event. | Keisuke Imoto, Nobutaka Ono |
| 2013 | Interspeech | User activity estimation method based on probabilistic generative model of acoustic event sequence with user activity and its subordinate categories. | Keisuke Imoto, Suehiro Shimauchi, Hisashi Uematsu, Hitoshi Ohmuro |