| 2026 | COMPSAC | Selective Privacy Protection in Speech Data: Suppressing Age Inference While Preserving Speaker Utility. | Aoi Ito, Katunobu Itou, Ryuichi Nisimura |
| 2025 | ISM | Dialogue-Pseudo: A Speaker Pseudonymization Framework for Privacy Protection in Dialogue Speech Data. | Aoi Ito, Katunobu Itou |
| 2024 | ISM | Speaker Pseudonymization for Japanese Speech Using Duration Embeddings. | Aoi Ito, Katunobu Itou |
| 2024 | ISM | Homophonic Music Composition Using a GAN and LSTM Pipeline for Melody and Harmony Generation. | Clment Saint-Marc, Katunobu Itou |
| 2024 | ISM | Instrumentality Classification Evaluation System for Natural Sounds | Yuhuan Wang, Katunobu Itou |
| 2022 | Interspeech | Cross-Lingual Transfer Learning Approach to Phoneme Error Detection via Latent Phonetic Representation. | Jovan M. Dalhouse, Katunobu Itou |
| 2018 | ISM | Automatic Electronic Organ Reduction System Based on Melody Clustering Considering Melodic and Instrumental Characteristics. | Katunobu Itou, Daiki Tanaka |
| 2014 | ICASSP | Intra-note segmentation via sticky HMM with DP emission. | Yuma Koizumi, Katunobu Itou |
| 2010 | IIWAS | Speaker model updating by the conversational sounds in speaker verification. | Keita Yamamuro, Katunobu Itou |
| 2009 | ICASSP | The use of acoustically detected filled and silent pauses in spontaneous speech recognition. | Jun Ogata, Masataka Goto, Katunobu Itou |
| 2008 | LREC | Test Collections for Spoken Document Retrieval from Lecture Audio Data. | Tomoyosi Akiba, Kiyoaki Aikawa, Yoshiaki Itoh, Tatsuya Kawahara, Hiroaki Nanjo, Hiromitsu Nishizaki, Norihito Yasuda, Yoichi Yamashita, Katunobu Itou |
| 2008 | LREC | In-car Speech Data Collection along with Various Multimodal Signals. | Akira Ozaki, Sunao Hara, Takashi Kusakawa, Chiyomi Miyajima, Takanori Nishino, Norihide Kitaoka, Katunobu Itou, Kazuya Takeda |
| 2007 | ICMI | Statistical segmentation and recognition of fingertip trajectories for a gesture interface. | Kazuhiro Morimoto, Chiyomi Miyajima, Norihide Kitaoka, Katunobu Itou, Kazuya Takeda |
| 2006 | LREC | Statistical Analysis for Thesaurus Construction using an Encyclopedic Corpus. | Yasunori Ohishi, Katunobu Itou, Kazuya Takeda, Atsushi Fujii |
| 2005 | ICASSP | Two-stage Noise Spectra Estimation and Regression based In-car Speech Recognition using Single Distant Microphone. | Weifeng Li, Katunobu Itou, Kazuya Takeda, Fumitada Itakura |
| 2005 | ICDE | Improved Noise Spectra Estimation and Log-spectral Regression for In-car Speech Recognition. | Weifeng Li, Katunobu Itou, Kazuya Takeda, Fumitada Itakura |
| 2005 | Interspeech | Subjective and objective quality assessment of regression-enhanced speech in real car environments. | Weifeng Li, Katunobu Itou, Kazuya Takeda, Fumitada Itakura |
| 2005 | Interspeech | Discrimination between singing and speaking voices. | Yasunori Ohishi, Masataka Goto, Katunobu Itou, Kazuya Takeda |
| 2005 | Interspeech | Data collection and evaluation of speech recognition for motorbike riders. | H. Tanaka, Hiroshi Fujimura, Chiyomi Miyajima, Takanori Nishino, Katunobu Itou, Kazuya Takeda |
| 2005 | WWW | Cyclone: an encyclopedic web search site. | Atsushi Fujii, Katunobu Itou, Tetsuya Ishikawa |
| 2004 | LREC | Collecting Spontaneously Spoken Queries for Information Retrieval. | Tomoyosi Akiba, Atsushi Fujii, Katunobu Itou |
| 2003 | Interspeech | Adapting language models for frequent fixed phrases by emphasizing n-gram subsets. | Tomoyosi Akiba, Katunobu Itou, Atsushi Fujii |
| 2003 | Interspeech | Building a test collection for speech-driven web retrieval. | Atsushi Fujii, Katunobu Itou |
| 2003 | Interspeech | A cross-media retrieval system for lecture videos. | Atsushi Fujii, Katunobu Itou, Tomoyosi Akiba, Tetsuya Ishikawa |
| 2003 | Interspeech | Speech shift: direct speech-input-mode switching through intentional control of voice pitch. | Masataka Goto, Yukihiro Omoto, Katunobu Itou, Tetsunori Kobayashi |
| 2003 | Interspeech | Speech starter: noise-robust endpoint detection by using filled pauses. | Koji Kitayama, Masataka Goto, Katunobu Itou, Tetsunori Kobayashi |
| 2002 | EMNLP | A Method for Open-Vocabulary Speech-Driven Text Retrieval. | Atsushi Fujii, Katunobu Itou, Tetsuya Ishikawa |
| 2002 | Interspeech | Selective back-off smoothing for incorporating grammatical constraints into the n-gram language model. | Tomoyosi Akiba, Katunobu Itou, Atsushi Fujii, Tetsuya Ishikawa |
| 2002 | Interspeech | Speech completion: on-demand completion assistance using filled pauses for speech input interfaces. | Masataka Goto, Katunobu Itou, Satoru Hayamizu |
| 2002 | LREC | Producing a Large-scale Encyclopedic Corpus over the Web. | Atsushi Fujii, Katunobu Itou, Tetsuya Ishikawa |
| 2001 | Interspeech | A structured statistical language model conditioned by arbitrarily abstracted grammatical categories based on GLR parsing. | Tomoyosi Akiba, Katunobu Itou |
| 2001 | Interspeech | Real-time sound source localization and separation system and its application to automatic speech recognition. | Futoshi Asano, Masataka Goto, Katunobu Itou, Hideki Asoh |
| 2001 | SIGIR | Speech-Driven Text Retrieval: Using Target IR Collections for Statistical Language Model Adaptation in Speech Recognition. | Atsushi Fujii, Katunobu Itou, Tetsuya Ishikawa |
| 1999 | Interspeech | A real-time filled pause detection system for spontaneous speech recognition. | Masataka Goto, Katunobu Itou, Satoru Hayamizu |
| 1998 | Interspeech | The design of the newspaper-based Japanese large vocabulary continuous speech recognition corpus. | Katunobu Itou, Mikio Yamamoto, Kazuya Takeda, Toshiyuki Takezawa, Tatsuo Matsuoka, Tetsunori Kobayashi, Kiyohiro Shikano, Shuichi Itahashi |
| 1996 | Interspeech | RWC multimodal database for interactions by integration of spoken language and visual information. | Satoru Hayamizu, Osamu Hasegawa, Katunobu Itou, Katsuhiko Sakaue, Kazuyo Tanaka, Shigeki Nagaya, Masayuki Nakazawa, T. Endoh, Fumio Togawa, Kenji Sakamoto, Kazuhiko Yamamoto |
| 1994 | Interspeech | Collecting and analyzing nonverbal elements for maintenance of dialog using a wizard of oz simulation. | Katunobu Itou, Tomoyosi Akiba, Osamu Hasegawa, Satoru Hayamizu, Kazuyo Tanaka |
| 1993 | Interspeech | Detection of unknown words in large vocabulary speech recognition. | Satoru Hayamizu, Katunobu Itou, Kazuyo Tanaka |
| 1992 | ICASSP | Continuous speech recognition by context-dependent phonetic HMM and an efficient algorithm for finding N-Best sentence hypotheses. | Katunobu Itou, Satoru Hayamizu, Hozumi Tanaka |
| 1992 | Interspeech | A spoken language dialogue system for automatic collection of spontaneous speech. | Satoru Hayamizu, Katunobu Itou, Masafumi Tamoto, Kazuyo Tanaka |
| 1992 | Interspeech | Detection of unknown words and automatic estimation of their transcriptions in continuous speech recognition. | Katunobu Itou, Satoru Hayamizu, Hozumi Tanaka |