| 2014 | ICASSP | Role play dialogue topic model for language model adaptation in multi-party conversation speech recognition. | Ryo Masumura, Takanobu Oba, Hirokazu Masataki, Osamu Yoshioka, Satoshi Takahashi |
| 2013 | ICASSP | HMM-based expressive speech synthesis based on phrase-level F0 context labeling. | Yu Maeno, Takashi Nose, Takao Kobayashi, Tomoki Koriyama, Yusuke Ijima, Hideharu Nakajima, Hideyuki Mizuno, Osamu Yoshioka |
| 2013 | ICASSP | Use of latent words language models in ASR: A sampling-based implementation. | Ryo Masumura, Hirokazu Masataki, Takanobu Oba, Osamu Yoshioka, Satoshi Takahashi |
| 2013 | Interspeech | Unsupervised confidence calibration using examples of recognized words and their contexts. | Taichi Asami, Satoshi Kobashikawa, Hirokazu Masataki, Osamu Yoshioka, Satoshi Takahashi |
| 2013 | Interspeech | Viterbi decoding for latent words language models using gibbs sampling. | Ryo Masumura, Hirokazu Masataki, Takanobu Oba, Osamu Yoshioka, Satoshi Takahashi |
| 2013 | Interspeech | Which resemblance is useful to predict phrase boundary rise labels for Japanese expressive text-to-speech synthesis, numerically-expressed stylistic or distribution-based semantic? | Hideharu Nakajima, Hideyuki Mizuno, Osamu Yoshioka, Satoshi Takahashi |
| 2012 | Interspeech | Speech Data Clustering Based on Phoneme Error Trend for Unsupervised Acoustic Model Adaptation. | Taichi Asami, Satoshi Kobashikawa, Hirokazu Masataki, Osamu Yoshioka, Satoshi Takahashi |
| 2012 | Interspeech | Automatic Vocabulary Adaptation Based on Semantic Similarity and Speech Recognition Confidence Measure. | Shoko Yamahata, Yoshikazu Yamaguchi, Atsunori Ogawa, Hirokazu Masataki, Osamu Yoshioka, Satoshi Takahashi |
| 2011 | Interspeech | HMM-Based Emphatic Speech Synthesis Using Unsupervised Context Labeling. | Yu Maeno, Takashi Nose, Takao Kobayashi, Yusuke Ijima, Hideharu Nakajima, Hideyuki Mizuno, Osamu Yoshioka |
| 2011 | Interspeech | Anger Recognition in Spoken Dialog Using Linguistic and Para-Linguistic Information. | Narichika Nomoto, Masafumi Tamoto, Hirokazu Masataki, Osamu Yoshioka, Satoshi Takahashi |
| 2010 | Interspeech | Detection of anger emotion in dialog speech using prosody feature and temporal relation of utterances. | Narichika Nomoto, Hirokazu Masataki, Osamu Yoshioka, Satoshi Takahashi |
| 1994 | Interspeech | A multi-modal dialogue system for telephone directory assistance. | Osamu Yoshioka, Yasuhiro Minami, Kiyohiro Shikano |
| 1994 | NAACL | A Large-Vocabulary Continuous Speech Recognition Algorithm and its Application to a Multi-Modal Telephone Directory Assistance System. | Yasuhiro Minami, Kiyohiro Shikano, Osamu Yoshioka, Satoshi Takahashi, Tomokazu Yamada, Sadaoki Furui |