| 2009 | ICMI | A speaker diarization method based on the probabilistic fusion of audio-visual location information. | Kentaro Ishizuka, Shoko Araki, Kazuhiro Otsuka, Tomohiro Nakatani, Masakiyo Fujimoto |
| 2009 | ICMI | Realtime meeting analysis and 3D meeting viewer based on omnidirectional multimodal sensors. | Kazuhiro Otsuka, Shoko Araki, Dan Mikami, Kentaro Ishizuka, Masakiyo Fujimoto, Junji Yamato |
| 2009 | Interspeech | A study of mutual front-end processing method based on statistical model for noise robust speech recognition. | Masakiyo Fujimoto, Kentaro Ishizuka, Tomohiro Nakatani |
| 2008 | ICASSP | Speaker indexing and speech enhancement in real meetings / conversations. | Shoko Araki, Masakiyo Fujimoto, Kentaro Ishizuka, Hiroshi Sawada, Shoji Makino |
| 2008 | ICASSP | A voice activity detection based on the adaptive integration of multiple speech features and a signal decision scheme. | Masakiyo Fujimoto, Kentaro Ishizuka, Tomohiro Nakatani |
| 2008 | ICMI | A realtime multimodal system for analyzing group meetings by combining face pose tracking and speaker diarization. | Kazuhiro Otsuka, Shoko Araki, Kentaro Ishizuka, Masakiyo Fujimoto, Martin Heinrich, Junji Yamato |
| 2008 | Interspeech | Study of integration of statistical model-based voice activity detection and noise suppression. | Masakiyo Fujimoto, Kentaro Ishizuka, Tomohiro Nakatani |
| 2008 | Interspeech | Statistical speech activity detection based on spatial power distribution for analyses of poster presentations. | Kentaro Ishizuka, Shoko Araki, Tatsuya Kawahara |
| 2008 | Interspeech | Multi-modal recording, analysis and indexing of poster sessions. | Tatsuya Kawahara, Hisao Setoguchi, Katsuya Takanashi, Kentaro Ishizuka, Shoko Araki |
| 2007 | ICASSP | Noise Robust Voice Activity Detection Based on Statistical Model and Parallel Non-Linear Kalman Filtering. | Masakiyo Fujimoto, Kentaro Ishizuka, Hiroko Kato Solvang |
| 2007 | ICASSP | Two-Microphone Voice Activity Detection Based on the Homogeneity of the Direction of Arrival Estimates. | Juan E. Rubio, Kentaro Ishizuka, Hiroshi Sawada, Shoko Araki, Tomohiro Nakatani, Masakiyo Fujimoto |
| 2007 | ICMI | The world of mushrooms: human-computer interaction prototype systems for ambient intelligence. | Yasuhiro Minami, Minako Sawaki, Kohji Dohsaka, Ryuichiro Higashinaka, Kentaro Ishizuka, Hideki Isozaki, Tatsushi Matsubayashi, Masato Miyoshi, Atsushi Nakamura, Takanobu Oba, Hiroshi Sawada, Takeshi Yamada, Eisaku Maeda |
| 2007 | Interspeech | Noise robust voice activity detection based on switching kalman filter. | Masakiyo Fujimoto, Kentaro Ishizuka |
| 2007 | Interspeech | Noise robust front-end processing with voice activity detection based on periodic to aperiodic component ratio. | Kentaro Ishizuka, Tomohiro Nakatani, Masakiyo Fujimoto, Noboru Miyazaki |
| 2006 | ICASSP | A Feature for Voice Activity Detection Derived from Speech Analysis with the Exponential Autoregressive Model. | Kentaro Ishizuka, Hiroko Kato Solvang |
| 2006 | Interspeech | Study of noise robust voice activity detection based on periodic component to aperiodic component ratio. | Kentaro Ishizuka, Tomohiro Nakatani |
| 2005 | ICASSP | Speech Signal Analysis with Exponential Autoregressive Model. | Kentaro Ishizuka, Hiroko Kato Solvang, Tomohiro Nakatani |
| 2005 | Interspeech | A longitudinal analysis of the spectral peaks of vowels for a Japanese infant. | Kentaro Ishizuka, Ryoko Mugitani, Hiroko Kato Solvang, Shigeaki Amano |
| 2004 | ICASSP | Speech feature extraction method representing periodicity and aperiodicity in sub bands for robust speech recognition. | Kentaro Ishizuka, Noboru Miyazaki |
| 2004 | Interspeech | Improvement in robustness of speech feature extraction method using sub-band based periodicity and aperiodicity decomposition. | Kentaro Ishizuka, Noboru Miyazaki, Tomohiro Nakatani, Yasuhiro Minami |
| 2002 | ICASSP | Noise-robust speech recognition using a new spectral estimation method "PHASOR". | Kiyoaki Aikawa, Kentaro Ishizuka |
| 2002 | Interspeech | Effect of F0 fluctuation and amplitude modulation of natural vowels on vowel identification in noisy environments. | Kentaro Ishizuka, Kiyoaki Aikawa |
| 1998 | Interspeech | Speaking-style dependent lexicalized filler model for key-phrase detection and verification. | Tatsuya Kawahara, Kentaro Ishizuka, Shuji Doshita, Chin-Hui Lee |