| 2012 | LREC | Method for Collection of Acted Speech Using Various Situation Scripts. | Takahiro Miyajima, Hideaki Kikuchi, Katsuhiko Shirai, Shigeki Okawa |
| 2009 | CISIS | Decision Model for a Robot to Start Communicating with a Human. | Satoshi Ushiyama, Kazunori Matsui, Motoi Yamagiwa, Makoto Murakami, Minoru Uehara, Katsuhiko Shirai |
| 2008 | ICASSP | Noisy speech recognition using temporal AM-FM combination. | Yotaro Kubo, Akira Kurematsu, Katsuhiko Shirai, Shigeki Okawa |
| 2008 | Interspeech | A comparative study on AM and FM features. | Yotaro Kubo, Shigeki Okawa, Akira Kurematsu, Katsuhiko Shirai |
| 2007 | Interspeech | A study on temporal features derived by analytic signal. | Yotaro Kubo, Shigeki Okawa, Akira Kurematsu, Katsuhiko Shirai |
| 2005 | Interspeech | Discrimination of speech, musical instruments and singing voices using the temporal patterns of sinusoidal segments in audio signals. | Toru Taniguchi, Akishige Adachi, Shigeki Okawa, Masaaki Honda, Katsuhiko Shirai |
| 2004 | Interspeech | Analysis of the phone level contributions to objective evaluation of English speech by non-natives. | Yasuo Suzuki, Yoshinori Sagisaka, Katsuhiko Shirai, Makiko Muto |
| 2004 | MMSP | Approach of feature with confident weight for robust speech recognition. | Yubo Ge, Jun Song, Lingnan Ge, Katsuhiko Shirai |
| 2003 | Interspeech | Corpus-based modeling of naturalness estimation in timing control for non-native speech. | Makiko Muto, Yoshinori Sagisaka, Takuro Naito, Daiju Maeki, Aki Kondo, Katsuhiko Shirai |
| 2003 | Interspeech | Statistical estimation of phoneme's most stable point based on universal constraint. | Shigeki Okawa, Katsuhiko Shirai |
| 2002 | HIS | Accurate Human Face Extraction using Genetic Algorithm and Subspace Method. | Makoto Murakami, Masahide Yoneyama, Katsuhiko Shirai |
| 2001 | Interspeech | Multi-class composite n-gram language model using multiple word clusters and word successions. | Shuntaro Isogai, Katsuhiko Shirai, Hirofumi Yamamoto, Yoshinori Sagisaka |
| 2001 | Interspeech | Speech enhancement based on IMM with NPHMM. | Yunjung Lee, Joohun Lee, Ki Yong Lee, Katsuhiko Shirai |
| 2001 | Interspeech | Pronunciation variant analysis using speaking style parallel corpus. | Hideharu Nakajima, Izumi Hirano, Yoshinori Sagisaka, Katsuhiko Shirai |
| 2001 | SMC | The multi-lateral security framework for the ubiquitous audiovisual services. | Itaru Kaneko, Katsuhiko Shirai |
| 2000 | ICASSP | Visual approach for automatic pitch period estimation. | Zhang Sen, Katsuhiko Shirai |
| 2000 | Interspeech | Designing a domain independent platform of spoken dialogue system. | Kazumi Aoyama, Izumi Hirano, Hideaki Kikuchi, Katsuhiko Shirai |
| 2000 | Interspeech | Overview of an intelligent system for information retrieval based on human-machine dialogue through spoken language. | Hiroya Fujisaki, Katsuhiko Shirai, Shuji Doshita, Seiichi Nakagawa, Keikichi Hirose, Shuichi Itahashi, Tatsuya Kawahara, Sumio Ohno, Hideaki Kikuchi, Kenji Abe, Shinya Kiriyama |
| 2000 | Interspeech | Improvement of dialogue efficiency by dialogue control model according to performance of processes. | Hideaki Kikuchi, Katsuhiko Shirai |
| 2000 | Interspeech | An automatic timing detection method for superimposing closed captions of TV programs. | Ichiro Maruyama, Yoshiharu Abe, Terumasa Ehara, Katsuhiko Shirai |
| 2000 | Interspeech | Using machine learning method and subword unit representations for spoken document categorization. | Weidong Qu, Katsuhiko Shirai |
| 2000 | Interspeech | Re-estimation of LPC coefficients in the sense of l∞ criterion. | Zhang Sen, Katsuhiko Shirai |
| 2000 | SMC | Controlling non-verbal information in speaker-change for spoken dialogue. | Kazumi Aoyama, Masao Yokoyama, Hideaki Kikuchi, Katsuhiko Shirai |
| 2000 | SMC | Modeling of spoken dialogue control for improvement of dialogue efficiency. | Hideaki Kikuchi, Katsuhiko Shirai |
| 1999 | Interspeech | A post-processing of speech for hearing impaired integrate into standard digital audio decoders. | Shinichi Hoshino, Itaru Kaneko, Hideaki Kikuchi, Katsuhiko Shirai |
| 1999 | Interspeech | Cognitive experiments on timing lag for superimposing closed captions. | Ichiro Maruyama, Yoshiharu Abe, Eiji Sawamura, Tetsuo Mitsuhashi, Terumasa Ehara, Katsuhiko Shirai |
| 1999 | Interspeech | A recombination strategy for multi-band speech recognition based on mutual information criterion. | Shigeki Okawa, Takehiro Nakajima, Katsuhiko Shirai |
| 1999 | Interspeech | Improving recognition correct rate of important words in large vocabulary speech recognition. | Yasuo Shirosaki, Hideaki Kikuchi, Katsuhiko Shirai |
| 1998 | ACL | Project for Production of Closed-Caption TV Programs for the Hearing Impaired. | Takahiro Wakao, Eiji Sawamura, Terumasa Ehara, Ichiro Maruyama, Katsuhiko Shirai |
| 1998 | Interspeech | Word sequence pair spotting for synchronization of speech and text in production of closed-caption TV programs for the hearing impaired. | Ichiro Maruyama, Yoshiharu Abe, Takahiro Wakao, Eiji Sawamura, Terumasa Ehara, Katsuhiko Shirai |
| 1998 | Interspeech | Use of non-verbal information in communication between human and robot. | Masao Yokoyama, Kazumi Aoyama, Hideaki Kikuchi, Katsuhiko Shirai |
| 1998 | IROS | Controlling gaze of humanoid in communication with human. | Hideaki Kikuchi, Masao Yokoyama, Keiichiro Hoashi, Yasuaki Hidaki, Tetsunori Kobayashi, Katsuhiko Shirai |
| 1997 | ICASSP | Difference in visual information between face to face and telephone dialogues. | Yuri Iwano, Yosuke Sugita, Yusuke Kasahara, Shu Nakazato, Katsuhiko Shirai |
| 1997 | ICASSP | Japanese large-vocabulary continuous-speech recognition using a business-newspaper corpus. | Tatsuo Matsuoka, Katsutoshi Ohtsuki, Takeshi Mori, Kotaro Yoshida, Sadaoki Furui, Katsuhiko Shirai |
| 1997 | ICIP | Facial Expressions Recognition Using Discrete Hopfield Neural Network. | Masahide Yoneyama, Yuri Iwano, Akihiro Ohtake, Katsuhiko Shirai |
| 1997 | Interspeech | Toward automatic transcription of Japanese broadcast news. | Tatsuo Matsuoka, Yuichi Taguchi, Katsutoshi Ohtsuki, Sadaoki Furui, Katsuhiko Shirai |
| 1996 | Interspeech | Analysis of head movements and its role in spoken dialogue. | Yuri Iwano, Shioya Kageyama, Emi Morikawa, Shu Nakazato, Katsuhiko Shirai |
| 1996 | Interspeech | Japanese large-vocabulary continuous-speech recognition using a business-newspaper corpus. | Tatsuo Matsuoka, Katsutoshi Ohtsuki, Takeshi Mori, Sadaoki Furui, Katsuhiko Shirai |
| 1996 | Interspeech | Estimation of statistical phoneme center considering phonemic environments. | Shigeki Okawa, Katsuhiko Shirai |
| 1996 | Interspeech | Modeling of spoken dialogue with and without visual information. | Katsuhiko Shirai |
| 1996 | Interspeech | Spoken dialogue interface in a dual task situation. | Shuichi Tanaka, Shu Nakazato, Keiichiro Hoashi, Katsuhiko Shirai |
| 1995 | Interspeech | Estimation of statistical phoneme center and its application to accurate phoneme modelling. | Shigeki Okawa, Katsuhiko Shirai |
| 1994 | ICASSP | Markov model based noise modeling and its application to noisy speech recognition using dynamical features of speech. | Tetsunori Kobayashi, Ryuji Mine, Katsuhiko Shirai |
| 1994 | ICASSP | Automatic training of phoneme dictionary based on mutual information criterion. | Shigeki Okawa, Tetsunori Kobayashi, Katsuhiko Shirai |
| 1994 | Interspeech | Effects on utterances caused by knowledge on the hearer. | Shu Nakazato, Katsuhiko Shirai |
| 1994 | Interspeech | Multimodal drawing tool using speech, mouse and key-board. | Takuya Nishirnoto, Nobutoshi Shida, Tetsunori Kobayashi, Katsuhiko Shirai |
| 1994 | Interspeech | Phoneme recognition in various styles of utterance based on mutual information criterion. | Shigeki Okawa, Tetsunori Kobayashi, Katsuhiko Shirai |
| 1994 | Interspeech | Evaluation of phonetic feature recognition with a time-delay neural network. | Shigeki Okawa, Christoph Windheuser, Frdric Bimbot, Katsuhiko Shirai |
| 1994 | Interspeech | Generation of prosody in speech synthesis using large speech data-base. | Naohiro Sakurai, Takerni Mochida, Tetsunori Kobayashi, Katsuhiko Shirai |
| 1993 | Interspeech | Speech recognition under the unstationary noise based on the noise Markov model and spectral-subtraction. | Tetsunori Kobayashi, Ryuji Mine, Katsuhiko Shirai |
| 1993 | Interspeech | Word spotting in conversational speech based on phonemic unit likelihood by mutual information criterion. | Shigeki Okawa, Tetsunori Kobayashi, Katsuhiko Shirai |
| 1992 | ICASSP | Speaker adaptive phoneme recognition based on feature mapping from spectral domain to probabilistic domain. | Tetsunori Kobayashi, Y. Uchiyama, J. Osada, Katsuhiko Shirai |
| 1992 | Interspeech | Spectral mapping onto probabilistic domain using neural networks and its application to speaker adaptive phoneme recognition. | Tetsunori Kobayashi, Katsuhiko Shirai |
| 1992 | Interspeech | Phoneme recognition in continuous speech based on mutual information considering phonemic duration and connectivity. | Katsuhiko Shirai, Shigeki Okawa, Tetsunori Kobayashi |
| 1991 | ICASSP | Application of neural networks to articulatory motion estimation. | Tetsunori Kobayashi, Masayuki Yagyu, Katsuhiko Shirai |
| 1991 | Interspeech | Text-to-speech synthesizer using superposition of sinusoidal waves generated by synchronized oscillators. | Katsuhiko Shirai, Kazuo Hashimoto, Tetsunori Kobayashi |
| 1991 | Interspeech | Optimal construction of context sensitive quantizer for phoneme recognition in continuous speech. | Katsuhiko Shirai, Eiichiro Kitagawa, T. Endo |
| 1990 | ICASSP | Speaker adaptive phoneme recognition by multi-level clustering based on mutual information criterion. | Katsuhiko Shirai, Naoki Hosaka, Eiichiro Kitagawa |
| 1990 | ICASSP | Interactive design environment of VLSI architecture for digital signal processing. | Toshiyuki Takezawa, Katsuhiko Shirai |
| 1990 | Interspeech | Speaker adaptable phoneme recognition selecting reliable acoustic features based on mutual information. | Katsuhiko Shirai, Naoki Hosaka, Eiichiro Kitagawa, T. Endo |
| 1990 | Interspeech | Speech synthesis using superposition of sinusoidal waves generated by synchronized oscillators. | Katsuhiko Shirai, Y. Sato, Kazuo Hashimoto |
| 1989 | ICASSP | Multi-level clustering of acoustic features for phoneme recognition based on mutual information. | Katsuhiko Shirai, Noriyuki Aoki, Naoki Hosaka |
| 1989 | Interspeech | Phoneme recognition in continuous speech using feature selection based on mutual information. | Katsuhiko Shirai, Noriyuki Aoki, Naoki Hosaka |
| 1987 | Interspeech | Description of task dependent knowledge for speech understanding system. | Tetsunori Kobayashi, Katsuhiko Shirai |
| 1987 | Interspeech | Speaker adaptive phoneme recognition in continuous speech based on vector quantization. | Katsuhiko Shirai, Kazunori Mano, Koji Sano |
| 1986 | COLING | Linguistic Knowledge Extraction from Real Language Behavior. | Katsuhiko Shirai, T. Hamada |
| 1986 | ICASSP | A network model dealing with focus of conversation for speech understanding system. | Tetsunori Kobayashi, Katsuhiko Shirai |
| 1986 | ICASSP | Phoneme recognition in connected speech using both static and dynamic properties of spectrum described by vector quantization. | Kazunori Mano, Shunichi Ishige, Katsuhiko Shirai |
| 1986 | ICASSP | Effects of tempo and context on jaw openings for vowels in vowel sequence words. | Shinobu Masaki, Katsuhiko Shirai, Shigeru Kiritani |
| 1986 | ICASSP | Pitch contour control in Japanese conversational speech. | Katsuhiko Shirai, Kazuhiko Iwata, Takeshi Ohno |
| 1986 | ICASSP | Estimation of articulatory parameters by table look-up method and its application for speaker independent phoneme recognition. | Katsuhiko Shirai, Tetsunori Kobayashi, J. Yazawa |
| 1984 | ICASSP | Phrase speech recognition of large vocabulary using feature in articulatory domain. | Katsuhiko Shirai, Tetsunori Kobayashi |
| 1983 | ICASSP | Considerations on articulatory dynamics for continuous speech recognition. | Katsuhiko Shirai, Tetsunori Kobayashi |
| 1982 | COLING | Japanese Sentence Analysis System Essay - Evaluation Of Dictionary Derived From Real Text Data. | Katsuhiko Shirai, J. Kubota, Yoshihiko Hayashi |
| 1982 | ICASSP | Recognition of semivowels and consonants in continuous speech using articulatory parameters. | Katsuhiko Shirai, Tetsunori Kobayashi |
| 1981 | ICASSP | Vowel identification in continuous speech using articulatory parameters. | Katsuhiko Shirai |
| 1980 | COLING | A Trial Of Japanese Text Input System Using Speech Recognition. | Katsuhiko Shirai, Y. Fukazawa, T. Matzui, H. Matzuura |