| 2025 | Interspeech | Improving User Impression of Spoken Dialogue Systems by Controlling Para-linguistic Expression Based on Intimacy. | Shoki Kawanishi, Akinori Ito, Yuya Chiba, Takashi Nose |
| 2023 | HCI | Multimodal Expressive Embodied Conversational Agent Design. | Simon Jolibois, Akinori Ito, Takashi Nose |
| 2021 | Interspeech | Improvement of Automatic English Pronunciation Assessment with Small Number of Utterances Using Sentence Speakability. | Satsuki Naijo, Akinori Ito, Takashi Nose |
| 2021 | Interspeech | Neural Spoken-Response Generation Using Prosodic and Linguistic Context for Conversational Systems. | Yoshihiro Yamazaki, Yuya Chiba, Takashi Nose, Akinori Ito |
| 2020 | Interspeech | Multi-Stream Attention-Based BLSTM with Feature Segmentation for Speech Emotion Recognition. | Yuya Chiba, Takashi Nose, Akinori Ito |
| 2020 | LREC | Construction and Analysis of a Multimodal Chat-talk Corpus for Dialog Systems Considering Interpersonal Closeness. | Yoshihiro Yamazaki, Yuya Chiba, Takashi Nose, Akinori Ito |
| 2018 | Interspeech | Analyzing Effect of Physical Expression on English Proficiency for Multimodal Computer-Assisted Language Learning. | Haoran Wu, Yuya Chiba, Takashi Nose, Akinori Ito |
| 2018 | SIGdial | An Analysis of the Effect of Emotional Speech Synthesis on Non-Task-Oriented Dialogue System. | Yuya Chiba, Takashi Nose, Taketo Kase, Mai Yamanaka, Akinori Ito |
| 2018 | SIGdial | Improving User Impression in Spoken Dialog System with Gradual Speech Form Control. | Yukiko Kageyama, Yuya Chiba, Takashi Nose, Akinori Ito |
| 2017 | HCI | Collection of Example Sentences for Non-task-Oriented Dialog Using a Spoken Dialog System and Comparison with Hand-Crafted DB. | Yukiko Kageyama, Yuya Chiba, Takashi Nose, Akinori Ito |
| 2015 | EMNLP | Hierarchical Latent Words Language Models for Robust Modeling to Out-Of Domain Tasks. | Ryo Masumura, Taichi Asami, Takanobu Oba, Hirokazu Masataki, Sumitaka Sakauchi, Akinori Ito |
| 2015 | HCI | On Appropriateness and Estimation of the Emotion of Synthesized Response Speech in a Spoken Dialogue System. | Taketo Kase, Takashi Nose, Akinori Ito |
| 2015 | Interspeech | Combinations of various language model technologies including data expansion and adaptation in spontaneous speech recognition. | Ryo Masumura, Taichi Asami, Takanobu Oba, Hirokazu Masataki, Sumitaka Sakauchi, Akinori Ito |
| 2015 | Interspeech | Latent words recurrent neural network language models. | Ryo Masumura, Taichi Asami, Takanobu Oba, Hirokazu Masataki, Sumitaka Sakauchi, Akinori Ito |
| 2015 | Interspeech | Entropy-based sentence selection for speech synthesis using phonetic and prosodic contexts. | Takashi Nose, Yusuke Arao, Takao Kobayashi, Komei Sugiura, Yoshinori Shiga, Akinori Ito |
| 2015 | SIGGRAPH | Game jam based iterative curriculum for game production in Japan. | Koji Mikami, Yosuke Nakamura, Akinori Ito, Motonobu Kawashima, Taichi Watanabe, Yoshihiro Kishimoto, Kunio Kondo |
| 2015 | RO-MAN | Development of a mobile robot moving on a handrail - Control for preceding a person keeping a distance. | Yuma Fujiwara, Yutaka Hiroi, Yuki Tanaka, Akinori Ito |
| 2014 | HCI | Controlling Switching Pause Using an AR Agent for Interactive CALL System. | Naoto Suzuki, Takashi Nose, Yutaka Hiroi, Akinori Ito |
| 2014 | Interspeech | Analysis of spectral enhancement using global variance in HMM-based speech synthesis. | Takashi Nose, Akinori Ito |
| 2014 | SIGdial | User Modeling by Using Bag-of-Behaviors for Building a Dialog System Sensitive to the Interlocutor's Internal State. | Yuya Chiba, Masashi Ito, Takashi Nose, Akinori Ito |
| 2013 | HCI | Estimation of User's State during a Dialog Turn with Sequential Multi-modal Features. | Yuya Chiba, Masashi Ito, Akinori Ito |
| 2013 | HRI | ASAHI: OK for failure: a robot for supporting daily life, equipped with a robot avatar. | Yutaka Hiroi, Akinori Ito |
| 2012 | HSI | Estimation of User's Internal State before the User's First Utterance Using Acoustic Features and Face Orientation. | Yuya Chiba, Masashi Ito, Akinori Ito |
| 2012 | HSI | Effect of Robot Height on Comfortableness of Spoken Dialog. | Yutaka Hiroi, Takayuki Nakayama, Hisanori Kuroda, Shinji Miyake, Akinori Ito |
| 2012 | ICASSP | Spoken document retrieval by discriminative modeling in a high dimensional feature space. | Takanobu Oba, Takaaki Hori, Atsushi Nakamura, Akinori Ito |
| 2011 | ICASSP | Bit rate reduction of the MELP coder using Lempel-Ziv segment quantization. | Minoru Kohata, Motoyuki Suzuki, Akinori Ito, Shozo Makino |
| 2011 | ICASSP | Round-robin duel discriminative language models in one-pass decoding with on-the-fly error correction. | Takanobu Oba, Takaaki Hori, Akinori Ito, Atsushi Nakamura |
| 2011 | Interspeech | Evaluation of Abnormal Sound Detection using Multi-Stage GMM in Various Environments. | Akinori Ito, Akihito Aiba, Masashi Ito, Shozo Makino |
| 2011 | Interspeech | Training a Language Model Using Webdata for Large Vocabulary Japanese Spontaneous Speech Recognition. | Ryo Masumura, Seongjun Hahm, Akinori Ito |
| 2011 | Interspeech | Language Model Expansion Using Webdata for Spoken Document Retrieval. | Ryo Masumura, Seongjun Hahm, Akinori Ito |
| 2010 | ICASSP | Aspect-model-based reference speaker weighting. | Seongjun Hahm, Yuichi Ohkawa, Masashi Ito, Motoyuki Suzuki, Akinori Ito, Shozo Makino |
| 2010 | Interspeech | An effect of formant amplitude in vowel perception. | Masashi Ito, Keiji Ohara, Akinori Ito, Masafumi Yano |
| 2009 | ICASSP | Information hiding for G.711 speech based on substitution of least significant bits and estimation of tolerable distortion. | Akinori Ito, Shun'ichiro Abe, Yiti Suzuki |
| 2009 | Interspeech | Evaluation of English intonation based on combination of multiple evaluation scores. | Akinori Ito, Tomoaki Konno, Masashi Ito, Shozo Makino |
| 2009 | Interspeech | Relative importance of formant and whole-spectral cues for vowel perception. | Masashi Ito, Keiji Ohara, Akinori Ito, Masafumi Yano |
| 2009 | Interspeech | Detailed description of triphone model using SSS-free algorithm. | Motoyuki Suzuki, Daisuke Honma, Akinori Ito, Shozo Makino |
| 2009 | SIGGRAPH | Construction trial of a practical education curriculum for game development by industry/university collaboration. | Koji Mikami, Taichi Watanabe, Katsunori Yamaji, Kenji Ozawa, Akinori Ito, Motonobu Kawashima, Ryota Takeuchi, Kunio Kondo, Mitsuru Kaneko |
| 2008 | Interspeech | A fast speaker adaptation method using aspect model. | Seongjun Hahm, Akinori Ito, Shozo Makino, Motoyuki Suzuki |
| 2008 | Interspeech | Discrimination of task-related words for vocabulary design of spoken dialog systems. | Akinori Ito, Toyomi Meguro, Shozo Makino, Motoyuki Suzuki |
| 2008 | Interspeech | Recognition of English utterances with grammatical and lexical mistakes for dialogue-based CALL system. | Akinori Ito, Ryohei Tsutsui, Shozo Makino, Motoyuki Suzuki |
| 2006 | Interspeech | A user simulator based on voiceXML for evaluation of spoken dialog systems. | Akinori Ito, Keisuke Shimada, Motoyuki Suzuki, Shozo Makino |
| 2006 | Interspeech | Unsupervised language model adaptation based on automatic text collection from WWW. | Motoyuki Suzuki, Yasutomo Kajiura, Akinori Ito, Shozo Makino |
| 2005 | CW | Smile and Laughter Recognition using Speech Processing and Face Recognition from Conversation Video. | Akinori Ito, Xinyue Wang, Motoyuki Suzuki, Shozo Makino |
| 2005 | Interspeech | Internal noise suppression for speech recognition by small robots. | Akinori Ito, Takashi Kanayama, Motoyuki Suzuki, Shozo Makino |
| 2005 | Interspeech | Pronunciation error detection method based on error rule clustering using a decision tree. | Akinori Ito, Yen-Ling Lim, Motoyuki Suzuki, Shozo Makino |
| 2005 | Interspeech | Construction method of acoustic models dealing with various background noises based on combination of HMMs. | Motoyuki Suzuki, Yusuke Kato, Akinori Ito, Shozo Makino |
| 2004 | Interspeech | Noise adaptive spoken dialog system based on selection of multiple dialog strategies. | Akinori Ito, Takanobu Oba, Takashi Konashi, Motoyuki Suzuki, Shozo Makino |
| 2004 | Interspeech | A spoken dialog system based on automatic grammar generation and template-based weighting for autonomous mobile robots. | Takashi Konashi, Motoyuki Suzuki, Akinori Ito, Shozo Makino |
| 2004 | Interspeech | A Japanese dialogue-based CALL system with mispronunciation and grammar error detection. | Oh Pyo Kweon, Akinori Ito, Motoyuki Suzuki, Shozo Makino |
| 2004 | Interspeech | Speaker adaptation method for CALL system using bilingual speakers' utterances. | Motoyuki Suzuki, Hirokazu Ogasawara, Akinori Ito, Yuichi Ohkawa, Shozo Makino |
| 2003 | Interspeech | An optimized multi-duration HMM for spontaneous speech recognition. | Yuichi Ohkawa, Akihiro Yoshida, Motoyuki Suzuki, Akinori Ito, Shozo Makino |
| 2002 | LREC | Continuous Speech Recognition Consortium an Open Repository for CSR Tools and Models. | Akinobu Lee, Tatsuya Kawahara, Kazuya Takeda, Masato Mimura, Atsushi Yamada, Akinori Ito, Katsunobu Itou, Kiyohiro Shikano |
| 2001 | MMSP | New state clustering of hidden Markov network with Korean phonological rules for speech recognition. | Se-Jin Oh, Hyun-Yeol Chung, Cheol-Jun Hwang, Bum-Koog Kim, Akinori Ito |
| 2000 | Interspeech | Language modeling by stochastic dependency grammar for Japanese speech recognition. | Akinori Ito, Chiori Hori, Masaharu Katoh, Masaki Kohda |
| 2000 | Interspeech | Free software toolkit for Japanese large vocabulary continuous speech recognition. | Tatsuya Kawahara, Akinobu Lee, Tetsunori Kobayashi, Kazuya Takeda, Nobuaki Minematsu, Shigeki Sagayama, Katsunobu Itou, Akinori Ito, Mikio Yamamoto, Atsushi Yamada, Takehito Utsuro, Kiyohiro Shikano |
| 2000 | LREC | IPA Japanese Dictation Free Software Project. | Katsunobu Itou, Kiyohiro Shikano, Tatsuya Kawahara, Kazuya Takeda, Atsushi Yamada, Akinori Ito, Takehito Utsuro, Tetsunori Kobayashi, Nobuaki Minematsu, Mikio Yamamoto, Shigeki Sagayama, Akinobu Lee |
| 1999 | Interspeech | A new metric for stochastic language model evaluation. | Akinori Ito, Masaki Kohda, Mari Ostendorf |
| 1997 | Interspeech | N-gram language model adaptation using small corpus for spoken dialog recognition. | Akinori Ito, Hideyuki Saitoh, Masaharu Katoh, Masaki Kohda |
| 1996 | Interspeech | Language modeling by string pattern n-gram for Japanese speech recognition. | Akinori Ito, Masaki Kohda |
| 1994 | ICASSP | The performance prediction method on sentence recognition system using a finite state automaton. | Takashi Otsuki, Akinori Ito, Shozo Makino, Teruhiko Otomo |
| 1993 | ICASSP | A new word pre-selection method based on an extended redundant hash addressing for continuous speech recognition. | Akinori Ito, Shozo Makino |
| 1992 | Interspeech | Word pre-selection using a redundant hash addressing method for continuous speech recognition. | Akinori Ito, Shozo Makino |
| 1991 | ICASSP | A Japanese text dictation system based on phoneme recognition and a dependency grammar. | Shozo Makino, Akinori Ito, Mitsuru Endo, Ken'iti Kido |
| 1990 | Interspeech | A Japanese text dictation system based on phoneme recognition using a modified LVQ2 method. | Shozo Makino, Akinori Ito, Mitsuru Endo, Ken'iti Kido |