| 2000 | IV | Creation of a Noh Mask Using a 3D Computer Graphics Technique. | Akiko Tohma, Katsunori Shimohara, Yoh'ichi Tohkura, Naokazu Yokoya |
| 1996 | Interspeech | Does training in speech perception modify speech production? | Reiko Akahane-Yamada, Yoh'ichi Tohkura, Ann R. Bradlow, David B. Pisoni |
| 1996 | Interspeech | A few factors which affect the degree of incorporating lip-read information into speech perception. | Kaoru Sekiyama, Yoh'ichi Tohkura, Michio Umeda |
| 1994 | Interspeech | Human processing of auditory-visual information in speech perception: potential for multimodal human-machine interfaces. | Patricia K. Kuhl, Minoru Tsuzaki, Yoh'ichi Tohkura, Andrew N. Meltzoff |
| 1993 | ICASSP | A dynamic cepstrum incorporating time-frequency masking and its application to continuous speech recognition. | Kiyoaki Aikawa, Harald Singer, Hideki Kawahara, Yoh'ichi Tohkura |
| 1990 | ICASSP | A hybrid speech recognition system using HMMs with an LVQ-trained codebook. | Hitoshi Iwamida, Shigeru Katagiri, Erik McDermott, Yoh'ichi Tohkura |
| 1990 | Interspeech | Perception and production of syllable-initial English /r/ and /l/ by native speakers of Japanese. | Reiko Akahane-Yamada, Yoh'ichi Tohkura |
| 1990 | Interspeech | The role of temporal structure of speech in word perception and spoken language understanding. | Yoshinori Kitahara, Yoh'ichi Tohkura |
| 1988 | ICASSP | On the application of spectrum target prediction model to speech recognition. | Masato Akagi, Yoh'ichi Tohkura |
| 1986 | ICASSP | A weighted cepstral distance measure for speech recognition. | Yoh'ichi Tohkura |