| 2026 | ACL | MMAC: A Multilingual, Multimodal Alignment Framework for Cultural Grounding Evaluation. | Weihua Zheng, Zhengyuan Liu, Tanmoy Chakraborty, Weiwen Xu, Xiaoxue Gao, Bryan Chen Zhengyu Tan, Bowei Zou, Chang Liu, Yujia Hu, Xing Xie, Xiaoyuan Yi, Jing Yao, Chaojun Wang, Long Li, Rui Liu, Huiyao Liu, Koji Inoue, Ryuichi Sumida, Tatsuya Kawahara, Fan Xu, Lingyu Ye, Wei Tian, Dongjun Kim, Jimin Jung, Jaehyung Seo, Nadya Yuki Wangsajaya, Pham Minh Duc, Ojasva Saxena, Palash Nandi, Xiyan Tao, Wiwik Karlina, Tuan Luong, Keertana Arun Vasan, Roy Ka-Wei Lee, Nancy F. Chen |
| 2026 | LREC | Adapting Pretrained Models to Endangered Languages in Japan: A Comparative Study on Ryukyuan and Ainu Speech Recognition. | Kohei Matsuura, Takanori Ashihara, Tatsuya Kawahara |
| 2026 | SIGdial | On the Structure of Address in Multi-Party Dialogue: From Discrete Labels to Continuous Levels. | Taiga Mori, Koji Inoue, Divesh Lala, Tatsuya Kawahara |
| 2026 | SIGdial | I Understand How You Feel: Enhancing Deeper Emotional Support Through Multilingual Emotional Validation in Dialogue System. | Zi Haur Pang, Yahui Fu, Koji Inoue, Tatsuya Kawahara |
| 2025 | ASRU | KyotoMOS2: MOS Prediction for Speech Across Multiple Sampling Rates. | Wangjin Zhou, Yizhou Zhang, Keisuke Imoto, Tatsuya Kawahara |
| 2025 | CHI | Does the Appearance of Autonomous Conversational Robots Affect User Spoken Behaviors in Real-World Conference Interactions? | Zi Haur Pang, Yahui Fu, Divesh Lala, Mikey Elmers, Koji Inoue, Tatsuya Kawahara |
| 2025 | COLING | Human-Like Embodied AI Interviewer: Employing Android ERICA in Real International Conference. | Zi Haur Pang, Yahui Fu, Divesh Lala, Mikey Elmers, Koji Inoue, Tatsuya Kawahara |
| 2025 | HAI | Can LLMs be Surprised? Evaluation and Analysis of Surprise Expression of LLMs. | Motoori Takeuchi, Koji Inoue, Keiko Ochi, Tatsuya Kawahara |
| 2025 | ICASSP | Leveraging IPA and Articulatory Features as Effective Inductive Biases for Multilingual ASR Training. | Jaeyoung Lee, Masato Mimura, Tatsuya Kawahara |
| 2025 | ICASSP | Extending Whisper for Emotion Prediction Using Word-level Pseudo Labels. | Kwok Chin Yuen, Sheng Li, Jia Qi Yip, Chenhui Chu, Tatsuya Kawahara, Eng Siong Chng |
| 2025 | ICMI | CCMI 2025: Cross-Cultural Multimodal Interaction. | Koji Inoue, Shogo Okada, Divesh Lala, Sahba Zojaji, Nancy F. Chen, Tatsuya Kawahara |
| 2025 | ICMI | Real-time Generation of Various Types of Nodding for Avatar Attentive Listening System. | Kazushi Kato, Koji Inoue, Divesh Lala, Keiko Ochi, Tatsuya Kawahara |
| 2025 | IJCNLP | Minority-Aware Satisfaction Estimation in Dialogue Systems via Preference-Adaptive Reinforcement Learning. | Yahui Fu, Zi Haur Pang, Tatsuya Kawahara |
| 2025 | Interspeech | Triadic Multi-party Voice Activity Projection for Turn-taking in Spoken Dialogue Systems. | Mikey Elmers, Koji Inoue, Divesh Lala, Tatsuya Kawahara |
| 2025 | Interspeech | Multi-lingual and Zero-Shot Speech Recognition by Incorporating Classification of Language-Independent Articulatory Features. | Ryo Magoshi, Shinsuke Sakai, Jaeyoung Lee, Tatsuya Kawahara |
| 2025 | Interspeech | Switch Conformer with Universal Phonetic Experts for Multilingual ASR. | Masato Mimura, Jaeyoung Lee, Tatsuya Kawahara |
| 2025 | Interspeech | Simple and Effective Content Encoder for Singing Voice Conversion via SSL-Embedding Dimension Reduction. | Wangjin Zhou, Tianjiao Du, Chenglin Xu, Sheng Li, Yi Zhao, Tatsuya Kawahara |
| 2025 | IROS | A Noise-Robust Turn-Taking System for Real-World Dialogue Robots: A Field Experiment. | Koji Inoue, Yuki Okafuji, Jun Baba, Yoshiki Ohira, Katsuya Hyodo, Tatsuya Kawahara |
| 2025 | NAACL | Yeah, Un, Oh: Continuous and Real-time Backchannel Prediction with Fine-tuning of Voice Activity Projection. | Koji Inoue, Divesh Lala, Gabriel Skantze, Tatsuya Kawahara |
| 2025 | SIGdial | Prompt-Guided Turn-Taking Prediction. | Koji Inoue, Mikey Elmers, Yahui Fu, Zi Haur Pang, Divesh Lala, Keiko Ochi, Tatsuya Kawahara |
| 2024 | COLING | Multilingual Turn-taking Prediction Using Voice Activity Projection. | Koji Inoue, Bing'er Jiang, Erik Ekstedt, Tatsuya Kawahara, Gabriel Skantze |
| 2024 | ICASSP | Enhancing Two-Stage Finetuning for Speech Emotion Recognition Using Adapters. | Yuan Gao, Hao Shi, Chenhui Chu, Tatsuya Kawahara |
| 2024 | ICASSP | Zero- and Few-Shot Sound Event Localization and Detection. | Kazuki Shimada, Kengo Uchida, Yuichiro Koyama, Takashi Shibuya, Shusuke Takahashi, Yuki Mitsufuji, Tatsuya Kawahara |
| 2024 | ICASSP | Diffusion-Based Speech Enhancement with Joint Generative and Predictive Decoders. | Hao Shi, Kazuki Shimada, Masato Hirano, Takashi Shibuya, Yuichiro Koyama, Zhi Zhong, Shusuke Takahashi, Tatsuya Kawahara, Yuki Mitsufuji |
| 2024 | ICASSP | MOS-FAD: Improving Fake Audio Detection Via Automatic Mean Opinion Score Prediction. | Wangjin Zhou, Zhengdong Yang, Chenhui Chu, Sheng Li, Raj Dabre, Yi Zhao, Tatsuya Kawahara |
| 2024 | Interspeech | Speech Emotion Recognition with Multi-level Acoustic and Semantic Information Extraction and Interaction. | Yuan Gao, Hao Shi, Chenhui Chu, Tatsuya Kawahara |
| 2024 | Interspeech | Efficient and Robust Long-Form Speech Recognition with Hybrid H3-Conformer. | Tomoki Honda, Shinsuke Sakai, Tatsuya Kawahara |
| 2024 | Interspeech | Entrainment Analysis and Prosody Prediction of Subsequent Interlocutor's Backchannels in Dialogue. | Keiko Ochi, Koji Inoue, Divesh Lala, Tatsuya Kawahara |
| 2024 | Interspeech | Dual-path Adaptation of Pretrained Feature Extraction Module for Robust Automatic Speech Recognition. | Hao Shi, Tatsuya Kawahara |
| 2024 | SIGdial | StyEmp: Stylizing Empathetic Response Generation via Multi-Grained Prefix Encoder and Personality Reinforcement. | Yahui Fu, Chenhui Chu, Tatsuya Kawahara |
| 2023 | CHI | I Know Your Feelings Before You Do: Predicting Future Affective Reactions in Human-Computer Dialogue. | Yuanchao Li, Koji Inoue, Leimin Tian, Changzeng Fu, Carlos Toshinori Ishi, Hiroshi Ishiguro, Tatsuya Kawahara, Catherine Lai |
| 2023 | ICASSP | Domain and Language Adaptation Using Heterogeneous Datasets for Wav2vec2.0-Based Speech Recognition of Low-Resource Language. | Soky Kak, Sheng Li, Chenhui Chu, Tatsuya Kawahara |
| 2023 | ICASSP | Time-Domain Speech Enhancement Assisted by Multi-Resolution Frequency Encoder and Decoder. | Hao Shi, Masato Mimura, Longbiao Wang, Jianwu Dang, Tatsuya Kawahara |
| 2023 | ICMI | Towards Objective Evaluation of Socially-Situated Conversational Robots: Assessing Human-Likeness through Multimodal User Behaviors. | Koji Inoue, Divesh Lala, Keiko Ochi, Tatsuya Kawahara, Gabriel Skantze |
| 2023 | Interspeech | Two-stage Finetuning of Wav2vec 2.0 for Speech Emotion Recognition with ASR and Gender Pretraining. | Yuan Gao, Chenhui Chu, Tatsuya Kawahara |
| 2023 | Interspeech | Embedding Articulatory Constraints for Low-resource Speech Recognition Based on Large Pre-trained Model. | Jaeyoung Lee, Masato Mimura, Tatsuya Kawahara |
| 2023 | PACLIC | RealPersonaChat: A Realistic Persona Chat Corpus with Interlocutors' Own Personalities. | Sanae Yamashita, Koji Inoue, Ao Guo, Shota Mochizuki, Tatsuya Kawahara, Ryuichiro Higashinaka |
| 2023 | RO-MAN | Robotic Backchanneling in Online Conversation Facilitation: A Cross-Generational Study. | Sota Kobuki, Katie Seaborn, Seiki Tokunaga, Kosuke Fukumori, Shun Hidaka, Kazuhiro Tamura, Koji Inoue, Tatsuya Kawahara, Mihoko Otake-Matsuura |
| 2023 | SIGdial | Reasoning before Responding: Integrating Commonsense-based Causality Explanation for Empathetic Response Generation. | Yahui Fu, Koji Inoue, Chenhui Chu, Tatsuya Kawahara |
| 2022 | HAI | Backchannel Generation Model for a Third Party Listener Agent. | Divesh Lala, Koji Inoue, Tatsuya Kawahara, Kei Sawada |
| 2022 | HRI | Alzheimer's Dementia Detection through Spontaneous Dialogue with Proactive Robotic Listeners. | Yuanchao Li, Catherine Lai, Divesh Lala, Koji Inoue, Tatsuya Kawahara |
| 2022 | ICASSP | Phone-Informed Refinement of Synthesized Mel Spectrogram for Data Augmentation in Speech Recognition. | Sei Ueno, Tatsuya Kawahara |
| 2022 | ICASSP | Selective Multi-Task Learning For Speech Emotion Recognition Using Corpora Of Different Styles. | Heran Zhang, Masato Mimura, Tatsuya Kawahara, Kenkichi Ishizuka |
| 2022 | Interspeech | Non-autoregressive Error Correction for CTC-based ASR with Phone-conditioned Masked LM. | Hayato Futami, Hirofumi Inaguma, Sei Ueno, Masato Mimura, Shinsuke Sakai, Tatsuya Kawahara |
| 2022 | Interspeech | Leveraging Simultaneous Translation for Enhancing Transcription of Low-resource Language via Cross Attention Mechanism. | Soky Kak, Sheng Li, Masato Mimura, Chenhui Chu, Tatsuya Kawahara |
| 2022 | Interspeech | Multimodal Persuasive Dialogue Corpus using Teleoperated Android. | Seiya Kawano, Muteki Arioka, Akishige Yuguchi, Kenta Yamamoto, Koji Inoue, Tatsuya Kawahara, Satoshi Nakamura, Koichiro Yoshino |
| 2022 | Interspeech | End-to-end Speech-to-Punctuated-Text Recognition. | Jumon Nozaki, Tatsuya Kawahara, Kenkichi Ishizuka, Taiichi Hashimoto |
| 2022 | Interspeech | Monaural Speech Enhancement Based on Spectrogram Decomposition for Convolutional Neural Network-sensitive Feature Extraction. | Hao Shi, Longbiao Wang, Sheng Li, Jianwu Dang, Tatsuya Kawahara |
| 2022 | SIGdial | Simultaneous Job Interview System Using Multiple Semi-autonomous Agents. | Haruki Kawai, Yusuke Muraki, Kenta Yamamoto, Divesh Lala, Koji Inoue, Tatsuya Kawahara |
| 2021 | ASRU | ASR Rescoring and Confidence Estimation with Electra. | Hayato Futami, Hirofumi Inaguma, Masato Mimura, Shinsuke Sakai, Tatsuya Kawahara |
| 2021 | ASRU | Data Augmentation for ASR Using TTS Via a Discrete Representation. | Sei Ueno, Masato Mimura, Shinsuke Sakai, Tatsuya Kawahara |
| 2021 | ICASSP | ORTHROS: non-autoregressive end-to-end speech translation With dual-decoder. | Hirofumi Inaguma, Yosuke Higuchi, Kevin Duh, Tatsuya Kawahara, Shinji Watanabe |
| 2021 | Interspeech | StableEmit: Selection Probability Discount for Reducing Emission Latency of Streaming Monotonic Attention ASR. | Hirofumi Inaguma, Tatsuya Kawahara |
| 2021 | Interspeech | VAD-Free Streaming Hybrid CTC/Attention ASR for Unsegmented Recording. | Hirofumi Inaguma, Tatsuya Kawahara |
| 2021 | NAACL | Source and Target Bidirectional Knowledge Distillation for End-to-end Speech Translation. | Hirofumi Inaguma, Tatsuya Kawahara, Shinji Watanabe |
| 2021 | SIGdial | A multi-party attentive listening robot which stimulates involvement from side participants. | Koji Inoue, Hiromi Sakamoto, Kenta Yamamoto, Divesh Lala, Tatsuya Kawahara |
| 2021 | SIGdial | ERICA: An Empathetic Android Companion for Covid-19 Quarantine. | Etsuko Ishii, Genta Indra Winata, Samuel Cahyawijaya, Divesh Lala, Tatsuya Kawahara, Pascale Fung |
| 2021 | SIGdial | Multi-Referenced Training for Dialogue Response Generation. | Tianyu Zhao, Tatsuya Kawahara |
| 2020 | ACL | Designing Precise and Robust Dialogue Response Evaluators. | Tianyu Zhao, Divesh Lala, Tatsuya Kawahara |
| 2020 | COLING | Topic-relevant Response Generation using Optimal Transport for an Open-domain Dialog System. | Shuying Zhang, Tianyu Zhao, Tatsuya Kawahara |
| 2020 | HRI | Autonomous Dialogue Technologies in Symbiotic Human-robot Interaction. | Hiroshi Ishiguro, Tatsuya Kawahara, Yutaka Nakamura |
| 2020 | ICMI | Job Interviewer Android with Elaborate Follow-up Question Generation. | Koji Inoue, Kohei Hara, Divesh Lala, Kenta Yamamoto, Shizuka Nakamura, Katsuya Takanashi, Tatsuya Kawahara |
| 2020 | ICMI | Prediction of Shared Laughter for Human-Robot Dialogue. | Divesh Lala, Koji Inoue, Tatsuya Kawahara |
| 2020 | Interspeech | End-to-End Speech-to-Dialog-Act Recognition. | Viet-Trung Dang, Tianyu Zhao, Sei Ueno, Hirofumi Inaguma, Tatsuya Kawahara |
| 2020 | Interspeech | End-to-End Speech Emotion Recognition Combined with Acoustic-to-Word ASR Model. | Han Feng, Sei Ueno, Tatsuya Kawahara |
| 2020 | Interspeech | Distilling the Knowledge of BERT for Sequence-to-Sequence ASR. | Hayato Futami, Hirofumi Inaguma, Sei Ueno, Masato Mimura, Shinsuke Sakai, Tatsuya Kawahara |
| 2020 | Interspeech | CTC-Synchronous Training for Monotonic Attention Model. | Hirofumi Inaguma, Masato Mimura, Tatsuya Kawahara |
| 2020 | Interspeech | Enhancing Monotonic Multihead Attention for Streaming ASR. | Hirofumi Inaguma, Masato Mimura, Tatsuya Kawahara |
| 2020 | Interspeech | Generative Adversarial Training Data Adaptation for Very Low-Resource Automatic Speech Recognition. | Kohei Matsuura, Masato Mimura, Shinsuke Sakai, Tatsuya Kawahara |
| 2020 | Interspeech | Semi-Supervised Learning for Character Expression of Spoken Dialogue Systems. | Kenta Yamamoto, Koji Inoue, Tatsuya Kawahara |
| 2020 | LREC | Speech Corpus of Ainu Folklore and End-to-end Speech Recognition for Ainu Language. | Kohei Matsuura, Sei Ueno, Masato Mimura, Shinsuke Sakai, Tatsuya Kawahara |
| 2020 | SIGdial | An Attentive Listening System with Android ERICA: Comparison of Autonomous and WOZ Interactions. | Koji Inoue, Divesh Lala, Kenta Yamamoto, Shizuka Nakamura, Katsuya Takanashi, Tatsuya Kawahara |
| 2019 | ASRU | Multilingual End-to-End Speech Translation. | Hirofumi Inaguma, Kevin Duh, Tatsuya Kawahara, Shinji Watanabe |
| 2019 | ICASSP | Transfer Learning of Language-independent End-to-end ASR with Language Model Fusion. | Hirofumi Inaguma, Jaejin Cho, Murali Karthick Baskar, Tatsuya Kawahara, Shinji Watanabe |
| 2019 | ICASSP | Multi-speaker Sequence-to-sequence Speech Synthesis for Data Augmentation in Acoustic-to-word Speech Recognition. | Sei Ueno, Masato Mimura, Shinsuke Sakai, Tatsuya Kawahara |
| 2019 | ICMI | Smooth Turn-taking by a Robot Using an Online Continuous Model to Generate Turn-taking Cues. | Divesh Lala, Koji Inoue, Tatsuya Kawahara |
| 2019 | IJCAI | ERICA and WikiTalk. | Divesh Lala, Graham Wilcock, Kristiina Jokinen, Tatsuya Kawahara |
| 2019 | Interspeech | End-to-End Articulatory Attribute Modeling for Low-Resource Multilingual Speech Recognition. | Sheng Li, Chenchen Ding, Xugang Lu, Peng Shen, Tatsuya Kawahara, Hisashi Kawai |
| 2019 | Interspeech | Turn-Taking Prediction Based on Detection of Transition Relevance Place. | Kohei Hara, Koji Inoue, Katsuya Takanashi, Tatsuya Kawahara |
| 2019 | Interspeech | Analysis of Effect and Timing of Fillers in Natural Turn-Taking. | Divesh Lala, Shizuka Nakamura, Tatsuya Kawahara |
| 2019 | Interspeech | Investigating Radical-Based End-to-End Speech Recognition Systems for Chinese Dialects and Japanese. | Sheng Li, Xugang Lu, Chenchen Ding, Peng Shen, Tatsuya Kawahara, Hisashi Kawai |
| 2019 | Interspeech | Improving Transformer-Based Speech Recognition Systems with Compressed Structure and Speech Attributes Augmentation. | Sheng Li, Raj Dabre, Xugang Lu, Peng Shen, Tatsuya Kawahara, Hisashi Kawai |
| 2019 | Interspeech | Improved End-to-End Speech Emotion Recognition Using Self Attention Mechanism and Multitask Learning. | Yuanchao Li, Tianyu Zhao, Tatsuya Kawahara |
| 2018 | ICASSP | Statistical Speech Enhancement Based on Probabilistic Integration of Variational Autoencoder and Non-Negative Matrix Factorization. | Yoshiaki Bando, Masato Mimura, Katsutoshi Itoyama, Kazuyoshi Yoshii, Tatsuya Kawahara |
| 2018 | ICASSP | Efficient Learning of Articulatory Models Based on Multi-Label Training and Label Correction for Pronunciation Learning. | Richeng Duan, Tatsuya Kawahara, Masatake Dantsuji, Hiroaki Nanjo |
| 2018 | ICASSP | An End-to-End Approach to Joint Social Signal Detection and Automatic Speech Recognition. | Hirofumi Inaguma, Masato Mimura, Koji Inoue, Kazuyoshi Yoshii, Tatsuya Kawahara |
| 2018 | ICASSP | Audio-Visual Conversation Analysis by Smart Posterboard and Humanoid Robot. | Tatsuya Kawahara, Koji Inoue, Divesh Lala, Katsuya Takanashi |
| 2018 | ICASSP | Unsupervised Beamforming Based on Multichannel Nonnegative Matrix Factorization for Noisy Speech Recognition. | Kazuki Shimada, Yoshiaki Bando, Masato Mimura, Katsutoshi Itoyama, Kazuyoshi Yoshii, Tatsuya Kawahara |
| 2018 | ICASSP | Acoustic-to-Word Attention-Based Model Complemented with Character-Level CTC-Based Model. | Sei Ueno, Hirofumi Inaguma, Masato Mimura, Tatsuya Kawahara |
| 2018 | ICMI | Evaluation of Real-time Deep Learning Turn-taking Models for Multiple Dialogue Scenarios. | Divesh Lala, Koji Inoue, Tatsuya Kawahara |
| 2018 | Interspeech | Prediction of Turn-taking Using Multitask Learning with Prediction of Backchannels and Fillers. | Kohei Hara, Koji Inoue, Katsuya Takanashi, Tatsuya Kawahara |
| 2018 | Interspeech | Engagement Recognition in Spoken Dialogue via Neural Network by Aggregating Different Annotators' Models. | Koji Inoue, Divesh Lala, Katsuya Takanashi, Tatsuya Kawahara |
| 2018 | Interspeech | Improving CTC-based Acoustic Model with Very Deep Residual Time-delay Neural Networks. | Sheng Li, Xugang Lu, Ryoichi Takashima, Peng Shen, Tatsuya Kawahara, Hisashi Kawai |
| 2018 | Interspeech | Forward-Backward Attention Decoder. | Masato Mimura, Shinsuke Sakai, Tatsuya Kawahara |
| 2018 | Interspeech | Encoder Transfer for Attention-based Acoustic-to-word Speech Recognition. | Sei Ueno, Takafumi Moriya, Masato Mimura, Shinsuke Sakai, Yusuke Shinohara, Yoshikazu Yamaguchi, Yushi Aono, Tatsuya Kawahara |
| 2018 | IUI | Voice Input Tutoring System for Older Adults using Input Stumble Detection. | Toshiyuki Hagiya, Keiichiro Hoashi, Tatsuya Kawahara |
| 2018 | SIGdial | A Unified Neural Architecture for Joint Dialog Act Segmentation and Recognition in Spoken Dialog System. | Tianyu Zhao, Tatsuya Kawahara |
| 2017 | ASRU | Incremental training and constructing the very deep convolutional residual network acoustic models. | Sheng Li, Xugang Lu, Peng Shen, Ryoichi Takashima, Tatsuya Kawahara, Hisashi Kawai |
| 2017 | ASRU | Cross-domain speech recognition using nonparallel corpora with cycle-consistent adversarial networks. | Masato Mimura, Shinsuke Sakai, Tatsuya Kawahara |
| 2017 | ICAART | Utterance Behavior of Users While Playing Basketball with a Virtual Teammate. | Divesh Lala, Yuanchao Li, Tatsuya Kawahara |
| 2017 | ICASSP | Effective articulatory modeling for pronunciation error detection of L2 learner without non-native training data. | Richeng Duan, Tatsuya Kawahara, Masatake Dantsuji, Jinsong Zhang |
| 2017 | ICASSP | Bayesian multichannel nonnegative matrix factorization for audio source separation and localization. | Kousuke Itakura, Yoshiaki Bando, Eita Nakamura, Katsutoshi Itoyama, Kazuyoshi Yoshii, Tatsuya Kawahara |
| 2017 | ICASSP | Semi-supervised ensemble DNN acoustic model training. | Sheng Li, Xugang Lu, Shinsuke Sakai, Masato Mimura, Tatsuya Kawahara |
| 2017 | IJCNLP | Joint Learning of Dialog Act Segmentation and Recognition in Spoken Dialog Using Neural Networks. | Tianyu Zhao, Tatsuya Kawahara |
| 2017 | Interspeech | Social Signal Detection in Spontaneous Dialogue Using Bidirectional LSTM-CTC. | Hirofumi Inaguma, Koji Inoue, Masato Mimura, Tatsuya Kawahara |
| 2017 | Interspeech | Combined Multi-Channel NMF-Based Robust Beamforming for Noisy Speech Recognition. | Masato Mimura, Yoshiaki Bando, Kazuki Shimada, Shinsuke Sakai, Kazuyoshi Yoshii, Tatsuya Kawahara |
| 2017 | Interspeech | Analysis of the Relationship Between Prosodic Features of Fillers and its Forms or Occurrence Positions. | Shizuka Nakamura, Ryosuke Nakanishi, Katsuya Takanashi, Tatsuya Kawahara |
| 2017 | SIGdial | Attentive listening system with backchanneling, response generation and flexible turn-taking. | Divesh Lala, Pierrick Milhorat, Koji Inoue, Masanari Ishida, Katsuya Takanashi, Tatsuya Kawahara |
| 2016 | ICASSP | Data selection from multiple ASR systems' hypotheses for unsupervised acoustic model training. | Sheng Li, Yuya Akita, Tatsuya Kawahara |
| 2016 | ICMI | Prediction of ice-breaking between participants using prosodic features in the first meeting dialogue. | Hirofumi Inaguma, Koji Inoue, Shizuka Nakamura, Katsuya Takanashi, Tatsuya Kawahara |
| 2016 | ICMI | Annotation and analysis of listener's engagement based on multi-modal behaviors. | Koji Inoue, Divesh Lala, Shizuka Nakamura, Katsuya Takanashi, Tatsuya Kawahara |
| 2016 | ICMI | Multimodal interaction with the autonomous Android ERICA. | Divesh Lala, Pierrick Milhorat, Koji Inoue, Tianyu Zhao, Tatsuya Kawahara |
| 2016 | Interspeech | Prediction and Generation of Backchannel Form for Attentive Listening Systems. | Tatsuya Kawahara, Takashi Yamaguchi, Koji Inoue, Katsuya Takanashi, Nigel G. Ward |
| 2016 | Interspeech | Joint Optimization of Denoising Autoencoder and DNN Acoustic Model Based on Multi-Target Learning for Noisy Speech Recognition. | Masato Mimura, Shinsuke Sakai, Tatsuya Kawahara |
| 2016 | IVA | Managing Dialog and Joint Actions for Virtual Basketball Teammates. | Divesh Lala, Tatsuya Kawahara |
| 2016 | RO-MAN | ERICA: The ERATO Intelligent Conversational Android. | Dylan F. Glas, Takashi Minato, Carlos Toshinori Ishi, Tatsuya Kawahara, Hiroshi Ishiguro |
| 2016 | SIGdial | Talking with ERICA, an autonomous android. | Koji Inoue, Pierrick Milhorat, Divesh Lala, Tianyu Zhao, Tatsuya Kawahara |
| 2015 | ICASSP | Language model adaptation for academic lectures using character recognition result of presentation slides. | Yuya Akita, Yizheng Tong, Tatsuya Kawahara |
| 2015 | ICASSP | Deep autoencoders augmented with phone-class feature for reverberant speech recognition. | Masato Mimura, Shinsuke Sakai, Tatsuya Kawahara |
| 2015 | Interspeech | Enhanced speaker diarization with detection of backchannels using eye-gaze information in poster conversations. | Koji Inoue, Yukoh Wakabayashi, Hiromasa Yoshimoto, Katsuya Takanashi, Tatsuya Kawahara |
| 2015 | Interspeech | Discriminative data selection for lightly supervised training of acoustic model using closed caption texts. | Sheng Li, Yuya Akita, Tatsuya Kawahara |
| 2015 | Interspeech | Ensemble speaker modeling using speaker adaptive training deep neural network for speaker adaptation. | Sheng Li, Xugang Lu, Yuya Akita, Tatsuya Kawahara |
| 2015 | Interspeech | Speech dereverberation using long short-term memory. | Masato Mimura, Shinsuke Sakai, Tatsuya Kawahara |
| 2015 | SP | News Navigation System Based on Proactive Dialogue Strategy. | Koichiro Yoshino, Tatsuya Kawahara |
| 2014 | AMTA | Japanese-to-English patent translation system based on domain-adapted word segmentation and post-ordering. | Katsuhito Sudoh, Masaaki Nagata, Shinsuke Mori, Tatsuya Kawahara |
| 2014 | ICCE | Partial and Synchronized Caption Generation to Develop Second Language Listening Skill. | Maryam Sadat Mirzaei, Yuya Akita, Tatsuya Kawahara |
| 2014 | Interspeech | Speaker diarization using eye-gaze information in multi-party conversations. | Koji Inoue, Yukoh Wakabayashi, Hiromasa Yoshimoto, Tatsuya Kawahara |
| 2014 | SIGdial | Information Navigation System Based on POMDP that Tracks User Focus. | Koichiro Yoshino, Tatsuya Kawahara |
| 2013 | HCI | Multi-party Human-Machine Interaction Using a Smart Multimodal Digital Signage. | Tony Tung, Randy Gomez, Tatsuya Kawahara, Takashi Matsuyama |
| 2013 | ICASSP | Incorporating semantic information to selection of web texts for language model of spoken dialogue system. | Koichiro Yoshino, Shinsuke Mori, Tatsuya Kawahara |
| 2013 | IJCNLP | Predicate Argument Structure Analysis using Partially Annotated Corpora. | Koichiro Yoshino, Shinsuke Mori, Tatsuya Kawahara |
| 2013 | ICRA | Hands-free human-robot communication robust to speaker's radial position. | Randy Gomez, Keisuke Nakamura, Kazuhiro Nakadai, Ui-Hyun Kim, Hiroshi G. Okuno, Tatsuya Kawahara |
| 2013 | Interspeech | Estimation of interest and comprehension level of audience through multi-modal behaviors in poster conversations. | Tatsuya Kawahara, Soichiro Hayashi, Katsuya Takanashi |
| 2012 | ACL | Machine Translation without Words through Substring Alignment. | Graham Neubig, Taro Watanabe, Shinsuke Mori, Tatsuya Kawahara |
| 2012 | COLING | Language Modeling for Spoken Dialogue System based on Filtering using Predicate-Argument Structures. | Koichiro Yoshino, Shinsuke Mori, Tatsuya Kawahara |
| 2012 | ECCV | Group Dynamics and Multimodal Interaction Modeling Using a Smart Digital Signage. | Tony Tung, Randy Gomez, Tatsuya Kawahara, Takashi Matsuyama |
| 2012 | HRI | Multi-party human-robot interaction with distant-talking speech recognition. | Randy Gomez, Tatsuya Kawahara, Keisuke Nakamura, Kazuhiro Nakadai |
| 2012 | IAAI | Transcription System Using Automatic Speech Recognition for the Japanese Parliament (Diet). | Tatsuya Kawahara |
| 2012 | ICASSP | Discriminative approach to lexical entry selection for automatic speech recognition of agglutinative language. | Mijit Ablimit, Tatsuya Kawahara, Askar Hamdulla |
| 2012 | Interspeech | Automatic Transcription of Lecture Speech using Language Model Based on Speaking-Style Transformation of Proceeding Texts. | Yuya Akita, Makoto Watanabe, Tatsuya Kawahara |
| 2012 | Interspeech | Dereverberation based on Wavelet Packet Filtering for Robust Automatic Speech Recognition. | Tatsuya Kawahara, Randy Gomez |
| 2012 | Interspeech | Prediction of Turn-Taking by Combining Prosodic and Eye-Gaze Information in Poster Conversations. | Tatsuya Kawahara, Takuma Iwatate, Katsuya Takanashi |
| 2012 | Interspeech | Comparative Analysis of Intensity between Native Speakers and Japanese Speakers of English. | Tomoko Nariai, Kazuyo Tanaka, Tatsuya Kawahara |
| 2012 | LREC | Designing an Evaluation Framework for Spoken Term Detection and Spoken Document Retrieval at the NTCIR-9 SpokenDoc Task. | Tomoyosi Akiba, Hiromitsu Nishizaki, Kiyoaki Aikawa, Tatsuya Kawahara, Tomoko Matsui |
| 2012 | SIGdial | Multi-modal Sensing and Analysis of Poster Conversations: Toward Smart Posterboard. | Tatsuya Kawahara |
| 2011 | ACL | An Unsupervised Model for Joint Phrase Alignment and Extraction. | Graham Neubig, Taro Watanabe, Eiichiro Sumita, Shinsuke Mori, Tatsuya Kawahara |
| 2011 | Interspeech | Automatic Comma Insertion of Lecture Transcripts Based on Multiple Annotations. | Yuya Akita, Tatsuya Kawahara |
| 2011 | Interspeech | Denoising Using Optimized Wavelet Filtering for Automatic Speech Recognition. | Randy Gomez, Tatsuya Kawahara |
| 2011 | SIGdial | Spoken Dialogue System based on Information Extraction using Similarity of Predicate Argument Structures. | Koichiro Yoshino, Shinsuke Mori, Tatsuya Kawahara |
| 2010 | ICASSP | Using online model comparison in the Variational Bayes framework for online unsupervised Voice Activity Detection. | David Cournapeau, Shinji Watanabe, Atsushi Nakamura, Tatsuya Kawahara |
| 2010 | ICASSP | Optimizing spectral subtraction and wiener filtering for robust speech recognition in reverberant and noisy conditions. | Randy Gomez, Tatsuya Kawahara |
| 2010 | ICASSP | Improved statistical models for SMT-based speaking style transformation. | Graham Neubig, Yuya Akita, Shinsuke Mori, Tatsuya Kawahara |
| 2010 | Interspeech | Semi-automated update of automatic transcription system for the Japanese national congress. | Yuya Akita, Masato Mimura, Graham Neubig, Tatsuya Kawahara |
| 2010 | Interspeech | An improved wavelet-based dereverberation for robust automatic speech recognition. | Randy Gomez, Tatsuya Kawahara |
| 2010 | Interspeech | Constructing Japanese test collections for spoken term detection. | Yoshiaki Itoh, Hiromitsu Nishizaki, Xinhui Hu, Hiroaki Nanjo, Tomoyosi Akiba, Tatsuya Kawahara, Seiichi Nakagawa, Tomoko Matsui, Yoichi Yamashita, Kiyoaki Aikawa |
| 2010 | Interspeech | Classroom note-taking system for hearing impaired students using automatic speech recognition adapted to lectures. | Tatsuya Kawahara, Norihiro Katsumaru, Yuya Akita, Shinsuke Mori |
| 2010 | Interspeech | Detection of hot spots in poster conversations based on reactive tokens of audience. | Tatsuya Kawahara, Kouhei Sumi, Zhi-Qiang Chang, Katsuya Takanashi |
| 2010 | Interspeech | Learning a language model from continuous speech. | Graham Neubig, Masato Mimura, Shinsuke Mori, Tatsuya Kawahara |
| 2009 | ASRU | New perspectives on spoken language understanding: Does machine need to fully understand speech? | Tatsuya Kawahara |
| 2009 | ICASSP | Language model transformation applied to lightly supervised training of acoustic model for congress meetings. | Tatsuya Kawahara, Masato Mimura, Yuya Akita |
| 2009 | ICASSP | Optimal learning of P-Layer additive F0 models with cross-validation. | Shinsuke Sakai, Tatsuya Kawahara, Tohru Shimizu, Satoshi Nakamura |
| 2009 | Interspeech | Automatic transcription system for meetings of the Japanese national congress. | Yuya Akita, Masato Mimura, Tatsuya Kawahara |
| 2009 | Interspeech | Optimization of dereverberation parameters based on likelihood of speech recognizer. | Randy Gomez, Tatsuya Kawahara |
| 2009 | Interspeech | A WFST-based log-linear framework for speaking-style transformation. | Graham Neubig, Shinsuke Mori, Tatsuya Kawahara |
| 2009 | Interspeech | Acoustic event detection for spotting "hot spots" in podcasts. | Kouhei Sumi, Tatsuya Kawahara, Jun Ogata, Masataka Goto |
| 2008 | COLING | Bayes Risk-based Dialogue Management for Document Retrieval System with Speech Interface. | Teruhisa Misu, Tatsuya Kawahara |
| 2008 | ICASSP | Using variational bayes free energy for unsupervised voice activity detection. | David Cournapeau, Tatsuya Kawahara |
| 2008 | ICASSP | Automatic lecture transcription by exploiting presentation slide information for language model adaptation. | Tatsuya Kawahara, Yusuke Nemoto, Yuya Akita |
| 2008 | ICASSP | Admissible stopping in viterbi beam search for unit selection in concatenative speech synthesis. | Shinichi Sakai, Tatsuya Kawahara, Shun Nakamura |
| 2008 | ICASSP | GMM and HMM training by aggregated EM algorithm with increased ensemble sizes for robust parameter estimation. | Takahiro Shinozaki, Tatsuya Kawahara |
| 2008 | ICASSP | Effective error prediction using decision tree for ASR grammar network in call system. | Hongcui Wang, Tatsuya Kawahara |
| 2008 | Interspeech | Statistical speech activity detection based on spatial power distribution for analyses of poster presentations. | Kentaro Ishizuka, Shoko Araki, Tatsuya Kawahara |
| 2008 | Interspeech | Multi-modal recording, analysis and indexing of poster sessions. | Tatsuya Kawahara, Hisao Setoguchi, Katsuya Takanashi, Kentaro Ishizuka, Shoko Araki |
| 2008 | Interspeech | Detection of feeling through back-channels in spoken dialogue. | Tatsuya Kawahara, Masayoshi Toyokura, Teruhisa Misu, Chiori Hori |
| 2008 | Interspeech | Predicting ASR errors by exploiting barge-in rate of individual users for spoken dialogue systems. | Kazunori Komatani, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2008 | Interspeech | Extracting word-pronunciation pairs from comparable set of text and speech. | Tetsuro Sasada, Shinsuke Mori, Tatsuya Kawahara |
| 2008 | Interspeech | Aggregated cross-validation and its efficient application to Gaussian mixture optimization. | Takahiro Shinozaki, Sadaoki Furui, Tatsuya Kawahara |
| 2008 | Interspeech | A Japanese CALL system based on dynamic question generation and error prediction for ASR. | Hongcui Wang, Tatsuya Kawahara |
| 2008 | LREC | Test Collections for Spoken Document Retrieval from Lecture Audio Data. | Tomoyosi Akiba, Kiyoaki Aikawa, Yoshiaki Itoh, Tatsuya Kawahara, Hiroaki Nanjo, Hiromitsu Nishizaki, Norihito Yasuda, Yoichi Yamashita, Katunobu Itou |
| 2007 | ASRU | HMM training based on CV-EM and CV Gaussian mixture optimization. | Takahiro Shinozaki, Tatsuya Kawahara |
| 2007 | ICASSP | Topic-Independent Speaking-Style Transformation of Language Model for Spontaneous Speech Recognition. | Yuya Akita, Tatsuya Kawahara |
| 2007 | ICASSP | Automatic Detection of Sentence and Clause Units using Local Syntactic Dependency. | Tatsuya Kawahara, Masahiro Saikou, Katsuya Takanashi |
| 2007 | ICASSP | Speech-Based Interactive Information Guidance System using Question-Answering Technique. | Teruhisa Misu, Tatsuya Kawahara |
| 2007 | ICMI | Multi-modal conversational analysis of poster presentations using multiple sensors. | Hisao Setoguchi, Katsuya Takanashi, Tatsuya Kawahara |
| 2007 | Interspeech | PLSA-based topic detection in meetings for adaptation of lexicon and language model. | Yuya Akita, Yusuke Nemoto, Tatsuya Kawahara |
| 2007 | Interspeech | Evaluation of real-time voice activity detection based on high order statistics. | David Cournapeau, Tatsuya Kawahara |
| 2007 | Interspeech | Analyzing temporal transition of real user's behaviors in a spoken dialogue system. | Kazunori Komatani, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2007 | Interspeech | Bayes risk-based optimization of dialogue management for document retrieval system with speech interface. | Teruhisa Misu, Tatsuya Kawahara |
| 2007 | Interspeech | Gaussian mixture optimization for HMM based on efficient cross-validation. | Takahiro Shinozaki, Tatsuya Kawahara |
| 2007 | Interspeech | Evaluating and optimizing Japanese tutor system featuring dynamic question generation and interactive guidance. | Christopher J. Waple, Hongcui Wang, Tatsuya Kawahara, Yasushi Tsubota, Masatake Dantsuji |
| 2007 | MMSP | Real-Time Continuous Speech Recognition System on SH-4A Microprocessor. | Hiroaki Kokubo, Nobuo Hataoka, Akinobu Lee, Tatsuya Kawahara, Kiyohiro Shikano |
| 2006 | ACL | Detection of Quotations and Inserted Clauses and Its Application to Dependency Structure Analysis in Spontaneous Japanese. | Ryoji Hamabe, Kiyotaka Uchimoto, Tatsuya Kawahara, Hitoshi Isahara |
| 2006 | ICASSP | Efficient Estimation of Language Model Statistics of Spontaneous Speech Via Statistical Transformation Model. | Yuya Akita, Tatsuya Kawahara |
| 2006 | Interspeech | Sentence boundary detection of spontaneous Japanese using statistical language model and support vector machines. | Yuya Akita, Masahiro Saikou, Hiroaki Nanjo, Tatsuya Kawahara |
| 2006 | Interspeech | Voice activity detector based on enhanced cumulant of LPC residual and on-line EM algorithm. | David Cournapeau, Tatsuya Kawahara, Kenji Mase, Tomoji Toriyama |
| 2006 | Interspeech | Detection of quotations and inserted clauses and its application to dependency structure analysis in spontaneous Japanese. | Ryoji Hamabe, Kiyotaka Uchimoto, Tatsuya Kawahara, Hitoshi Isahara |
| 2006 | Interspeech | Evaluation of voice activity detection by combining multiple features with weight adaptation. | Yusuke Kida, Tatsuya Kawahara |
| 2006 | Interspeech | A bootstrapping approach for developing language model of new spoken dialogue systems by selecting web texts. | Teruhisa Misu, Tatsuya Kawahara |
| 2006 | Interspeech | Decision tree-based training of probabilistic concatenation models for corpus-based speech synthesis. | Shinsuke Sakai, Tatsuya Kawahara |
| 2006 | Interspeech | Prototyping a call system for students of Japanese using dynamic diagram generation and interactive hints. | Christopher J. Waple, Yasushi Tsubota, Masatake Dantsuji, Tatsuya Kawahara |
| 2006 | LREC | Dependency-structure Annotation to Corpus of Spontaneous Japanese. | Kiyotaka Uchimoto, Ryoji Hamabe, Takehiko Maruyama, Katsuya Takanashi, Tatsuya Kawahara, Hitoshi Isahara |
| 2006 | MMSP | Embedded Julius: Continuous Speech Recognition Software for Microprocessor. | Hiroaki Kokubo, Hiroaki Hataoka, Akinobu Lee, Tatsuya Kawahara, Kiyohiro Shikano |
| 2005 | ICASSP | Generalized Statistical Modeling of Pronunciation Variations using Variable-length Phone Context. | Yuya Akita, Tatsuya Kawahara |
| 2005 | ICASSP | Incorporating Dialogue Context and Topic Clustering in Out-of-Domain Detection. | Ian R. Lane, Tatsuya Kawahara |
| 2005 | ICASSP | A New ASR Evaluation Measure and Minimum Bayes-Risk Decoding for Open-domain Speech Understanding. | Hiroaki Nanjo, Tatsuya Kawahara |
| 2005 | Interspeech | Voice activity detection based on optimally weighted combination of multiple features. | Yusuke Kida, Tatsuya Kawahara |
| 2005 | Interspeech | Utterance verification incorporating in-domain confidence and discourse coherence measures. | Ian R. Lane, Tatsuya Kawahara |
| 2005 | Interspeech | Dialogue strategy to clarify user's queries for document retrieval system with speech interface. | Teruhisa Misu, Tatsuya Kawahara |
| 2005 | Interspeech | Minimum Bayes-risk decoding considering word significance for information retrieval system. | Hiroaki Nanjo, Teruhisa Misu, Tatsuya Kawahara |
| 2005 | Interspeech | Trigger-based language model adaptation for automatic meeting transcription. | Carlos Troncoso, Tatsuya Kawahara |
| 2005 | NAACL | Speech-based Information Retrieval System with Clarification Dialogue Strategy. | Teruhisa Misu, Tatsuya Kawahara |
| 2004 | COLING | Efficient Confirmation Strategy for Large-scale Text Retrieval Systems with Spoken Dialogue Interface. | Kazunori Komatani, Teruhisa Misu, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2004 | COLING | Dependency Structure Analysis and Sentence Boundary Detection in Spontaneous Japanese. | Kazuya Shitaoka, Kiyotaka Uchimoto, Tatsuya Kawahara, Hitoshi Isahara |
| 2004 | ICASSP | Out-of-domain detection based on confidence measures from multiple topic classification. | Ian R. Lane, Tatsuya Kawahara, Tomoko Matsui, Satoshi Nakamura |
| 2004 | ICASSP | Real-time word confidence scoring using local posterior probabilities on tree trellis search. | Akinobu Lee, Kiyohiro Shikano, Tatsuya Kawahara |
| 2004 | ICASSP | Automatic indexing of key sentences for lecture archives using statistics of presumed discourse markers. | Hiroaki Nanjo, Tasuku Kitade, Tatsuya Kawahara |
| 2004 | ICASSP | Speaker indexing and adaptation using speaker clustering based on statistical model selection. | Masafumi Nishida, Tatsuya Kawahara |
| 2004 | Interspeech | Language model adaptation based on PLSA of topics and speakers. | Yuya Akita, Tatsuya Kawahara |
| 2004 | Interspeech | Practical use of English pronunciation system for Japanese students in the CALL classroom. | Tatsuya Kawahara, Masatake Dantsuji, Yasushi Tsubota |
| 2004 | Interspeech | Topic classification and verification modeling for out-of-domain utterance detection. | Tatsuya Kawahara, Ian Richard Lane, Tomoko Matsui, Satoshi Nakamura |
| 2004 | Interspeech | Recent progress of open-source LVCSR engine julius and Japanese model repository. | Tatsuya Kawahara, Akinobu Lee, Kazuya Takeda, Katsunobu Itou, Kiyohiro Shikano |
| 2004 | Interspeech | Automatic transformation of lecture transcription into document style using statistical framework. | Tatsuya Kawahara, Kazuya Shitaoka, Hiroaki Nanjo |
| 2004 | Interspeech | Dependency structure analysis and sentence boundary detection in spontaneous Japanese. | Tatsuya Kawahara, Kiyotaka Uchimoto, Hitoshi Isahara, Kazuya Shitaoka |
| 2004 | Interspeech | Automatic extraction of key sentences from oral presentations using statistical measure based on discourse markers. | Tasuku Kitade, Tatsuya Kawahara, Hiroaki Nanjo |
| 2004 | Interspeech | Example-based training of dialogue planning incorporating user and situation models. | Ian Richard Lane, Tatsuya Kawahara, Shinichi Ueno |
| 2004 | Interspeech | Confirmation strategy for document retrieval systems with spoken dialog interface. | Teruhisa Misu, Tatsuya Kawahara, Kazunori Komatani |
| 2003 | ACL | Dialog Navigator : A Spoken Dialog Q-A System based on Large Text Knowledge Base. | Yoji Kiyota, Sadao Kurohashi, Teruhisa Misu, Kazunori Komatani, Tatsuya Kawahara, Fuyuko Kido |
| 2003 | ACL | Flexible Guidance Generation Using User Model in Spoken Dialogue Systems. | Kazunori Komatani, Shinichi Ueno, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2003 | ICASSP | Language model switching based on topic detection for dialog speech recognition. | Ian R. Lane, Tatsuya Kawahara, Tomoko Matsui |
| 2003 | ICASSP | Unsupervised speaker indexing using speaker model selection based on Bayesian information criterion. | Masafumi Nishida, Tatsuya Kawahara |
| 2003 | Interspeech | Unsupervised speaker indexing using anchor models and automatic transcription of discussions. | Yuya Akita, Tatsuya Kawahara |
| 2003 | Interspeech | Spoken dialogue system for queries on appliance manuals using hierarchical confirmation strategy. | Tatsuya Kawahara, Ryosuke Ito, Kazunori Komatani |
| 2003 | Interspeech | User modeling in spoken dialogue systems for flexible guidance generation. | Kazunori Komatani, Shinichi Ueno, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2003 | Interspeech | Hierarchical topic classification for dialog speech recognition based on language model switching. | Ian R. Lane, Tatsuya Kawahara, Tomoko Matsui, Satoshi Nakamura |
| 2003 | Interspeech | Speaker model selection using Bayesian information criterion for speaker indexing and speaker adaptation. | Masafumi Nishida, Tatsuya Kawahara |
| 2003 | SIGdial | Flexible Spoken Dialogue System based on User Models and Dynamic Generation of VoiceXML Scripts. | Kazunori Komatani, Fumihiro Adachi, Shinichi Ueno, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2002 | COLING | Efficient Dialogue Strategy to Find Users' Intended Items from Information Query Results. | Kazunori Komatani, Tatsuya Kawahara, Ryosuke Ito, Hiroshi G. Okuno |
| 2002 | ICASSP | Automatic indexing of lecture speech by extracting topic-independent discourse markers. | Tatsuya Kawahara, Masahiro Hasegawa |
| 2002 | ICASSP | Speaking-rate dependent decoding and adaptation for spontaneous lecture speech recognition. | Hiroaki Nanjo, Tatsuya Kawahara |
| 2002 | Interspeech | Modeling and automatic detection of English sentence stress for computer-assisted English prosody learning system. | Kazunori Imoto, Yasushi Tsubota, Antoine Raux, Tatsuya Kawahara, Masatake Dantsuji |
| 2002 | Interspeech | Speaking rate compensation based on likelihood criterion in acoustic model training and decoding. | Kozo Okuda, Tatsuya Kawahara, Satoshi Nakamura |
| 2002 | Interspeech | Automatic intelligibility assessment and diagnosis of critical pronunciation errors for computer-assisted pronunciation learning. | Antoine Raux, Tatsuya Kawahara |
| 2002 | Interspeech | Recognition and verification of English by Japanese students for computer-assisted language learning system. | Yasushi Tsubota, Tatsuya Kawahara, Masatake Dantsuji |
| 2002 | Interspeech | Belief network based disambiguation of object reference in spoken dialogue system for robot. | Yoko Yamakata, Tatsuya Kawahara, Hiroshi G. Okuno |
| 2002 | LREC | Continuous Speech Recognition Consortium an Open Repository for CSR Tools and Models. | Akinobu Lee, Tatsuya Kawahara, Kazuya Takeda, Masato Mimura, Atsushi Yamada, Akinori Ito, Katsunobu Itou, Kiyohiro Shikano |
| 2001 | ICASSP | Gaussian mixture selection using context-independent HMM. | Akinobu Lee, Tatsuya Kawahara, Kiyohiro Shikano |
| 2001 | Interspeech | Domain-independent spoken dialogue platform using key-phrase spotting based on combined language model. | Kazunori Komatani, Katsuaki Tanaka, Hiroaki Kashima, Tatsuya Kawahara |
| 2001 | Interspeech | Julius - an open source real-time large vocabulary recognition engine. | Akinobu Lee, Tatsuya Kawahara, Kiyohiro Shikano |
| 2001 | Interspeech | Speaking rate dependent acoustic modeling for spontaneous lecture speech recognition. | Hiroaki Nanjo, Kazuomi Kato, Tatsuya Kawahara |
| 2000 | COLING | Flexible Mixed-Initiative Dialogue Management using Concept-Level Confidence Measures of Speech Recognizer Output. | Kazunori Komatani, Tatsuya Kawahara |
| 2000 | ICASSP | A new phonetic tied-mixture model for efficient decoding. | Akinobu Lee, Tatsuya Kawahara, Kazuya Takeda, Kiyohiro Shikano |
| 2000 | Interspeech | Overview of an intelligent system for information retrieval based on human-machine dialogue through spoken language. | Hiroya Fujisaki, Katsuhiko Shirai, Shuji Doshita, Seiichi Nakagawa, Keikichi Hirose, Shuichi Itahashi, Tatsuya Kawahara, Sumio Ohno, Hideaki Kikuchi, Kenji Abe, Shinya Kiriyama |
| 2000 | Interspeech | Modelling of the perception of English sentence stress for computer-assisted language learning. | Kazunori Imoto, Masatake Dantsuji, Tatsuya Kawahara |
| 2000 | Interspeech | Automatic transcription of lecture speech using topic-independent language modeling. | Kazuomi Kato, Hiroaki Nanjo, Tatsuya Kawahara |
| 2000 | Interspeech | Free software toolkit for Japanese large vocabulary continuous speech recognition. | Tatsuya Kawahara, Akinobu Lee, Tetsunori Kobayashi, Kazuya Takeda, Nobuaki Minematsu, Shigeki Sagayama, Katsunobu Itou, Akinori Ito, Mikio Yamamoto, Atsushi Yamada, Takehito Utsuro, Kiyohiro Shikano |
| 2000 | Interspeech | Generating effective confirmation and guidance using two-level confidence measures for dialogue systems. | Kazunori Komatani, Tatsuya Kawahara |
| 2000 | Interspeech | Automatic diagnosis of recognition errors in large vocabulary continuous speech recognition systems. | Hiroaki Nanjo, Akinobu Lee, Tatsuya Kawahara |
| 2000 | Interspeech | Computer-assisted English vowel learning system for Japanese speakers using cross language formant structures. | Yasushi Tsubota, Masatake Dantsuji, Tatsuya Kawahara |
| 2000 | LREC | IPA Japanese Dictation Free Software Project. | Katsunobu Itou, Kiyohiro Shikano, Tatsuya Kawahara, Kazuya Takeda, Atsushi Yamada, Akinori Ito, Takehito Utsuro, Tetsunori Kobayashi, Nobuaki Minematsu, Mikio Yamamoto, Shigeki Sagayama, Akinobu Lee |
| 1999 | ICASSP | Topic independent language model for key-phrase detection and verification. | Tatsuya Kawahara, Shuji Doshita |
| 1998 | Interspeech | Automatic pronunciation error detection and guidance for foreign language learning. | Chul-Ho Jo, Tatsuya Kawahara, Shuji Doshita, Masatake Dantsuji |
| 1998 | Interspeech | Speaking-style dependent lexicalized filler model for key-phrase detection and verification. | Tatsuya Kawahara, Kentaro Ishizuka, Shuji Doshita, Chin-Hui Lee |
| 1998 | Interspeech | Sharable software repository for Japanese large vocabulary continuous speech recognition. | Tatsuya Kawahara, Tetsunori Kobayashi, Kazuya Takeda, Nobuaki Minematsu, Katsunobu Itou, Mikio Yamamoto, Atsushi Yamada, Takehito Utsuro, Kiyohiro Shikano |
| 1998 | Interspeech | An efficient two-pass search algorithm using word trellis index. | Akinobu Lee, Tatsuya Kawahara, Shuji Doshita |
| 1998 | Interspeech | Prosodic analysis of fillers and self-repair in Japanese speech. | Felix C. M. Quimbo, Tatsuya Kawahara, Shuji Doshita |
| 1997 | ICASSP | Combining key-phrase detection and subword-based verification for flexible speech understanding. | Tatsuya Kawahara, Chin-Hui Lee, Biing-Hwang Juang |
| 1997 | ICASSP | Task adaptation using MAP estimation in N-gram language modeling. | Hirokazu Masataki, Yoshinori Sagisaka, Kazuya Hisaki, Tatsuya Kawahara |
| 1996 | ICASSP | Concept-based phrase spotting approach for spontaneous speech understanding. | Tatsuya Kawahara, Norihide Kitaoka, Shuji Doshita |
| 1996 | Interspeech | Key-phrase detection and verification for flexible speech understanding. | Tatsuya Kawahara, Chin-Hui Lee, Biing-Hwang Juang |
| 1994 | ICASSP | Heuristic search integrating syntactic, semantic and dialog-level constraints. | Tatsuya Kawahara, Masahiro Araki, Shuji Doshita |
| 1994 | Interspeech | Keyword and phrase spotting with heuristic language model. | Tatsuya Kawahara, Toshihiko Munetsugu, Norihide Kitaoka, Shuji Doshita |
| 1992 | ICASSP | HMM based on pair-wise Bayes classifiers. | Tatsuya Kawahara, Shuji Doshita |
| 1991 | ICASSP | Phoneme recognition by combining discriminant analysis and HMM. | Tatsuya Kawahara, Shuji Doshita |
| 1991 | Interspeech | Unsupervised speaker normalization by speaker Markov model converter for speaker-independent speech recognition. | Pascale Fung, Tatsuya Kawahara, Shuji Doshita |
| 1990 | Interspeech | Phoneme recognition by combining Bayesian linear discriminations of selected pairs of classes. | Tatsuya Kawahara, Toru Ogawa, Shigeyoshi Kitazawa, Shuji Doshita |