| 2024 | ICPR | UAV-Enhanced Combination to Application: Comprehensive Analysis and Benchmarking of a Human Detection Dataset for Disaster Scenarios. | Ragib Amin Nihal, Benjamin Yen, Katsutoshi Itoyama, Kazuhiro Nakadai |
| 2024 | RO-MAN | Improving Impressions of Response Delay in AI-based Spoken Dialogue Systems. | Shuhei Asaka, Katsutoshi Itoyama, Kazuhiro Nakadai |
| 2023 | Interspeech | miniStreamer: Enhancing Small Conformer with Chunked-Context Masking for Streaming ASR Applications on the Edge. | Haris Gulzar, Monikka Roslianna Busto, Takeharu Eda, Katsutoshi Itoyama, Kazuhiro Nakadai |
| 2023 | RO-MAN | Improving Sign Language Understanding Introducing Label Smoothing. | Sihan Tan, Khan Nabeela Khanum, Katsutoshi Itoyama, Kazuhiro Nakadai |
| 2022 | Interspeech | Weakly-Supervised Neural Full-Rank Spatial Covariance Analysis for a Front-End System of Distant Speech Recognition. | Yoshiaki Bando, Takahiro Aizawa, Katsutoshi Itoyama, Kazuhiro Nakadai |
| 2022 | IROS | Spotforming by NMF Using Multiple Microphone Arrays. | Yasuhiro Kagimoto, Katsutoshi Itoyama, Kenji Nishida, Kazuhiro Nakadai |
| 2022 | IROS | Outdoor evaluation of sound source localization for drone groups using microphone arrays. | Taiki Yamada, Katsutoshi Itoyama, Kenji Nishida, Kazuhiro Nakadai |
| 2021 | Interspeech | Assessment of von Mises-Bernoulli Deep Neural Network in Sound Source Localization. | Katsutoshi Itoyama, Yoshiya Morimoto, Shungo Masaki, Ryosuke Kojima, Kenji Nishida, Kazuhiro Nakadai |
| 2020 | IROS | Synchronization of Microphones Based on Rank Minimization of Warped Spectrum for Asynchronous Distributed Recording. | Katsutoshi Itoyama, Kazuhiro Nakadai |
| 2019 | ICASSP | Joint Transcription of Lead, Bass, and Rhythm Guitars Based on a Factorial Hidden Semi-Markov Model. | Kentaro Shibata, Ryo Nishikimi, Satoru Fukayama, Masataka Goto, Eita Nakamura, Katsutoshi Itoyama, Kazuyoshi Yoshii |
| 2019 | IROS | Environmental sound segmentation utilizing Mask U-Net. | Yui Sudo, Katsutoshi Itoyama, Kenji Nishida, Kazuhiro Nakadai |
| 2018 | ICASSP | Statistical Speech Enhancement Based on Probabilistic Integration of Variational Autoencoder and Non-Negative Matrix Factorization. | Yoshiaki Bando, Masato Mimura, Katsutoshi Itoyama, Kazuyoshi Yoshii, Tatsuya Kawahara |
| 2018 | ICASSP | Unsupervised Beamforming Based on Multichannel Nonnegative Matrix Factorization for Noisy Speech Recognition. | Kazuki Shimada, Yoshiaki Bando, Masato Mimura, Katsutoshi Itoyama, Kazuyoshi Yoshii, Tatsuya Kawahara |
| 2018 | RO-MAN | Signal Restoration based on Bi-directional LSTM with Spectral Filtering for Robot Audition. | Ryosuke Taniguchi, Kotaro Hoshiba, Katsutoshi Itoyama, Kenji Nishida, Kazuhiro Nakadai |
| 2017 | ICASSP | Bayesian multichannel nonnegative matrix factorization for audio source separation and localization. | Kousuke Itakura, Yoshiaki Bando, Eita Nakamura, Katsutoshi Itoyama, Kazuyoshi Yoshii, Tatsuya Kawahara |
| 2016 | ICASSP | Student's T nonnegative matrix factorization and positive semidefinite tensor factorization for single-channel audio source separation. | Kazuyoshi Yoshii, Katsutoshi Itoyama, Masataka Goto |
| 2016 | IROS | Online simultaneous localization and mapping of multiple sound sources and asynchronous microphone arrays. | Kouhei Sekiguchi, Yoshiaki Bando, Keisuke Nakamura, Kazuhiro Nakadai, Katsutoshi Itoyama, Kazuyoshi Yoshii |
| 2016 | LREC | Parallel Speech Corpora of Japanese Dialects. | Koichiro Yoshino, Naoki Hirayama, Shinsuke Mori, Fumihiko Takahashi, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2015 | AAAI | Recognition of In-Field Frog Chorusing Using Bayesian Nonparametric Microphone Array Processing. | Yoshiaki Bando, Takuma Otsuka, Ikkyu Aihara, Hiromitsu Awano, Katsutoshi Itoyama, Kazuyoshi Yoshii, Hiroshi Gitchang Okuno |
| 2015 | ICASSP | Challenges in deploying a microphone array to localize and separate sound sources in real auditory scenes. | Yoshiaki Bando, Takuma Otsuka, Katsutoshi Itoyama, Kazuyoshi Yoshii, Yoko Sasaki, Satoshi Kagami, Hiroshi G. Okuno |
| 2015 | ICASSP | Singing voice analysis and editing based on mutually dependent F0 estimation and source separation. | Yukara Ikemiya, Kazuyoshi Yoshii, Katsutoshi Itoyama |
| 2015 | ICASSP | A feedback framework for improved chord recognition based on NMF-based approximate note transcription. | Satoshi Maruo, Kazuyoshi Yoshii, Katsutoshi Itoyama, Matthias Mauch, Masataka Goto |
| 2015 | Interspeech | Bayesian integration of sound source separation and speech recognition: a new approach to simultaneous speech recognition. | Kousuke Itakura, Izaya Nishimuta, Yoshiaki Bando, Katsutoshi Itoyama, Kazuyoshi Yoshii |
| 2015 | IROS | Microphone-accelerometer based 3D posture estimation for a hose-shaped rescue robot. | Yoshiaki Bando, Katsutoshi Itoyama, Masashi Konyo, Satoshi Tadokoro, Kazuhiro Nakadai, Kazuyoshi Yoshii, Hiroshi G. Okuno |
| 2015 | IROS | Audio-visual beat tracking based on a state-space model for a music robot dancing with humans. | Misato Ohkita, Yoshiaki Bando, Yukara Ikemiya, Katsutoshi Itoyama, Kazuyoshi Yoshii |
| 2015 | IROS | Optimizing the layout of multiple mobile robots for cooperative sound source separation. | Kouhei Sekiguchi, Yoshiaki Bando, Katsutoshi Itoyama, Kazuyoshi Yoshii |
| 2015 | SMC | Identification and Localization of One or Two Concurrent Speakers in a Binaural Robotic Context. | Karim Youssef, Katsutoshi Itoyama, Kazuyoshi Yoshii |
| 2014 | ICASSP | Transcribing vocal expression from polyphonic music. | Yukara Ikemiya, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2014 | ICASSP | Automatic transcription of guitar tablature from audio signals in accordance with player's proficiency. | Kazuki Yazawa, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2014 | IROS | Visualization of auditory awareness based on sound source positions estimated by depth sensor and microphone array. | Takahiro Iyama, Osamu Sugiyama, Takuma Otsuka, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2014 | SMC | Sound annotation tool for multidirectional sounds based on spatial information extracted by HARK robot audition software. | Osamu Sugiyama, Katsutoshi Itoyama, Kazuhiro Nakadai, Hiroshi G. Okuno |
| 2013 | ICASSP | Multiple index combination for Japanese spoken term detection with optimum index selection based on OOV-region classifier. | Naoyuki Kanda, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2013 | ICASSP | Initialization-robust Bayesian multipitch analyzer based on psychoacoustical and musical criteria. | Daichi Sakaue, Takuma Otsuka, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2013 | ICASSP | Audio-based guitar tablature transcription using multipitch analysis and playability constraints. | Kazuki Yazawa, Daichi Sakaue, Kohei Nagira, Katsutoshi Itoyama, Hiroshi G. Okuno |
| 2013 | Interspeech | Automatic estimation of dialect mixing ratio for dialect speech recognition. | Naoki Hirayama, Koichiro Yoshino, Katsutoshi Itoyama, Shinsuke Mori, Hiroshi G. Okuno |
| 2013 | IROS | Posture estimation of hose-shaped robot using microphone array localization. | Yoshiaki Bando, Takeshi Mizumoto, Katsutoshi Itoyama, Kazuhiro Nakadai, Hiroshi G. Okuno |
| 2013 | IROS | Noise correlation matrix estimation for improving sound source localization by multirotor UAV. | Koutarou Furukawa, Keita Okutani, Kohei Nagira, Takuma Otsuka, Katsutoshi Itoyama, Kazuhiro Nakadai, Hiroshi G. Okuno |
| 2012 | ICASSP | Initialization-robust multipitch estimation based on latent harmonic allocation using overtone corpus. | Daichi Sakaue, Katsutoshi Itoyama, Tetsuya Ogata, Hiroshi G. Okuno |
| 2011 | ICASSP | Simultaneous processing of sound source separation and musical instrument identification using Bayesian spectral modeling. | Katsutoshi Itoyama, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2009 | ISM | Bowed String Sequence Estimation of a Violin Based on Adaptive Audio Signal Classification and Context-Dependent Error Correction. | Akira Maezawa, Katsutoshi Itoyama, Toru Takahashi, Tetsuya Ogata, Hiroshi G. Okuno |
| 2007 | ICASSP | Integration and Adaptation of Harmonic and Inharmonic Models for Separating Polyphonic Musical Signals. | Katsutoshi Itoyama, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |