| 2025 | ICASSP | Deep Generic Representations for Domain-Generalized Anomalous Sound Detection. | Phurich Saengthong, Takahiro Shinozaki |
| 2024 | ICASSP | Self-Supervised Speaker Verification with Adaptive Threshold and Hierarchical Training. | Zehua Zhou, Haoyuan Yang, Takahiro Shinozaki |
| 2023 | ACL | Multi-Domain Dialogue State Tracking with Disentangled Domain-Slot Attention. | Longfei Yang, Jiyi Li, Sheng Li, Takahiro Shinozaki |
| 2023 | ICASSP | Continuous Action Space-Based Spoken Language Acquisition Agent Using Residual Sentence Embedding and Transformer Decoder. | Ryota Komatsu, Yusuke Kimura, Takuma Okamoto, Takahiro Shinozaki |
| 2023 | ICLR | FreeMatch: Self-adaptive Thresholding for Semi-supervised Learning. | Yidong Wang, Hao Chen, Qiang Heng, Wenxin Hou, Yue Fan, Zhen Wu, Jindong Wang, Marios Savvides, Takahiro Shinozaki, Bhiksha Raj, Bernt Schiele, Xing Xie |
| 2023 | Interspeech | Memory Network-Based End-To-End Neural ES-KMeans for Improved Word Segmentation. | Yu Iwamoto, Takahiro Shinozaki |
| 2022 | ACML | Margin Calibration for Long-Tailed Visual Recognition. | Yidong Wang, Bowen Zhang, Wenxin Hou, Zhen Wu, Jindong Wang, Takahiro Shinozaki |
| 2022 | COLING | Exploiting Unlabeled Data for Target-Oriented Opinion Words Extraction. | Yidong Wang, Hao Wu, Ao Liu, Wenxin Hou, Zhen Wu, Jindong Wang, Takahiro Shinozaki, Manabu Okumura, Yue Zhang |
| 2022 | ICASSP | Hybrid RNN-T/Attention-Based Streaming ASR with Triggered Chunkwise Attention and Dual Internal Language Model Integration. | Takafumi Moriya, Takanori Ashihara, Atsushi Ando, Hiroshi Sato, Tomohiro Tanaka, Kohei Matsuura, Ryo Masumura, Marc Delcroix, Takahiro Shinozaki |
| 2022 | Interspeech | Streaming Target-Speaker ASR with Neural Transducer. | Takafumi Moriya, Hiroshi Sato, Tsubasa Ochiai, Marc Delcroix, Takahiro Shinozaki |
| 2022 | Interspeech | Self-Supervised Learning with Multi-Target Contrastive Coding for Non-Native Acoustic Modeling of Mispronunciation Verification. | Longfei Yang, Jinsong Zhang, Takahiro Shinozaki |
| 2022 | Interspeech | Augmented Adversarial Self-Supervised Learning for Early-Stage Alzheimer's Speech Detection. | Longfei Yang, Wenqing Wei, Sheng Li, Jiyi Li, Takahiro Shinozaki |
| 2022 | Interspeech | Censer: Curriculum Semi-supervised Learning for Speech Recognition Based on Self-supervised Pre-training. | Bowen Zhang, Songjun Cao, Xiaoming Zhang, Yike Zhang, Long Ma, Takahiro Shinozaki |
| 2022 | SIGdial | Multi-Domain Dialogue State Tracking with Top-K Slot Self Attention. | Longfei Yang, Jiyi Li, Sheng Li, Takahiro Shinozaki |
| 2021 | ICASSP | Meta-Adapter: Efficient Cross-Lingual Adaptation With Meta-Learning. | Wenxin Hou, Yidong Wang, Shengzhou Gao, Takahiro Shinozaki |
| 2021 | Interspeech | Cross-Domain Speech Recognition with Unsupervised Character-Level Distribution Matching. | Wenxin Hou, Jindong Wang, Xu Tan, Tao Qin, Takahiro Shinozaki |
| 2020 | CEC | Dual Inheritance Evolution Strategy for Deep Neural Network Optimization. | Kent Hino, Yusuke Kimura, Yue Dong, Takahiro Shinozaki |
| 2020 | ICASSP | Spoken Language Acquisition Based on Reinforcement Learning and Word Unit Segmentation. | Shengzhou Gao, Wenxin Hou, Tomohiro Tanaka, Takahiro Shinozaki |
| 2020 | ICPR | Unsupervised Sound Source Localization From Audio-Image Pairs Using Input Gradient Map. | Tomohiro Tanaka, Takahiro Shinozaki |
| 2020 | Interspeech | Large-Scale End-to-End Multilingual Speech Recognition and Language Identification with Multi-Task Learning. | Wenxin Hou, Yue Dong, Bairong Zhuang, Longfei Yang, Jiatong Shi, Takahiro Shinozaki |
| 2020 | Interspeech | Pronunciation Erroneous Tendency Detection with Language Adversarial Represent Learning. | Longfei Yang, Kaiqi Fu, Jinsong Zhang, Takahiro Shinozaki |
| 2020 | Interspeech | Sound-Image Grounding Based Focusing Mechanism for Efficient Automatic Spoken Language Acquisition. | Mingxin Zhang, Tomohiro Tanaka, Wenxin Hou, Shengzhou Gao, Takahiro Shinozaki |
| 2020 | Interspeech | Time-Domain Target-Speaker Speech Separation with Waveform-Based Speaker Embedding. | Jianshu Zhao, Shengzhou Gao, Takahiro Shinozaki |
| 2019 | ASRU | Efficient Free Keyword Detection Based on CNN and End-to-End Continuous DP-Matching. | Tomohiro Tanaka, Takahiro Shinozaki |
| 2019 | ICASSP | Effective and Stable Neuron Model Optimization Based on Aggregated CMA-ES. | Han Xu, Takahiro Shinozaki, Ryota Kobayashi |
| 2018 | ICASSP | Reinforcement Learning of Speech Recognition System Based on Policy Gradient and Hypothesis Selection. | Taku Kato, Takahiro Shinozaki |
| 2017 | ASRU | Composite embedding systems for ZeroSpeech2017 Track1. | Hayato Shibata, Taku Kato, Takahiro Shinozaki, Shinji Watanabe |
| 2017 | Interspeech | Semi-Supervised Learning of a Pronunciation Dictionary from Disjoint Phonemic Transcripts and Text. | Takahiro Shinozaki, Shinji Watanabe, Daichi Mochihashi, Graham Neubig |
| 2015 | ASRU | Automation of system building for state-of-the-art large vocabulary speech recognition using evolution strategy. | Takafumi Moriya, Tomohiro Tanaka, Takahiro Shinozaki, Shinji Watanabe, Kevin Duh |
| 2015 | ICASSP | Structure discovery of deep neural network based on evolutionary algorithms. | Takahiro Shinozaki, Shinji Watanabe |
| 2014 | Interspeech | Accent type and phrase boundary estimation using acoustic and language models for automatic prosodic labeling. | Tomoki Koriyama, Hiroshi Suzuki, Takashi Nose, Takahiro Shinozaki, Takao Kobayashi |
| 2013 | ICIAP | Statistical Person Verification Using Behavioral Patterns from Complex Human Motion. | Felipe Gmez-Caballero, Takahiro Shinozaki, Sadaoki Furui, Koichi Shinoda |
| 2013 | Interspeech | Reverberant speech recognition based on denoising autoencoder. | Takaaki Ishii, Hiroki Komiyama, Takahiro Shinozaki, Yasuo Horiuchi, Shingo Kuroiwa |
| 2012 | ICASSP | Unsupervised CV language model adaptation based on direct likelihood maximization sentence selection. | Takahiro Shinozaki, Yasuo Horiuchi, Shingo Kuroiwa |
| 2012 | Interspeech | HMM Based Continuous EOG Recognition for Eye-input Speech Interface. | Fuming Fang, Takahiro Shinozaki, Yasuo Horiuchi, Shingo Kuroiwa, Sadaoki Furui, Toshimitsu Musha |
| 2011 | Interspeech | Sentence Selection by Direct Likelihood Maximization for Language Model Adaptation. | Takahiro Shinozaki, Yu Kubota, Sadaoki Furui, Eiji Utsunomiya, Yasutaka Shindoh |
| 2010 | ICASSP | Investigations on ensemble based unsupervised adaptation methods. | Yu Kubota, Takahiro Shinozaki, Sadaoki Furui |
| 2009 | ICASSP | Unsupervisec cross-validation adaptation algorithms for improved adaptation performance. | Takahiro Shinozaki, Yu Kubota, Sadaoki Furui |
| 2009 | Interspeech | Target speech GMM-based spectral compensation for noise robust speech recognition. | Takahiro Shinozaki, Sadaoki Furui |
| 2008 | ICASSP | GMM and HMM training by aggregated EM algorithm with increased ensemble sizes for robust parameter estimation. | Takahiro Shinozaki, Tatsuya Kawahara |
| 2008 | Interspeech | Aggregated cross-validation and its efficient application to Gaussian mixture optimization. | Takahiro Shinozaki, Sadaoki Furui, Tatsuya Kawahara |
| 2007 | ASRU | HMM training based on CV-EM and CV Gaussian mixture optimization. | Takahiro Shinozaki, Tatsuya Kawahara |
| 2007 | ICASSP | Model Complexity Selection and Cross-Validation EM Training for Robust Speaker Diarization. | Xavier Anguera Mir, Takahiro Shinozaki, Chuck Wooters, Javier Hernando |
| 2007 | ICASSP | Cross-Validation EM Training for Robust Parameter Estimation. | Takahiro Shinozaki, Mari Ostendorf |
| 2007 | Interspeech | Gaussian mixture optimization for HMM based on efficient cross-validation. | Takahiro Shinozaki, Tatsuya Kawahara |
| 2006 | ICASSP | Hmm State Clustering Based on Efficient Cross-Validation. | Takahiro Shinozaki |
| 2006 | Interspeech | Investigation on Mandarin broadcast news speech recognition. | Mei-Yuh Hwang, Xin Lei, Wen Wang, Takahiro Shinozaki |
| 2005 | Interspeech | Cluster-based modeling for ubiquitous speech recognition. | Sadaoki Furui, Tomohisa Ichiba, Takahiro Shinozaki, Edward W. D. Whittaker, Koji Iwano |
| 2005 | Interspeech | Data sampling for improved speech recognizer training. | Takahiro Shinozaki, Mari Ostendorf, Les E. Atlas |
| 2004 | Interspeech | Spontaneous speech recognition using a massively parallel decoder. | Takahiro Shinozaki, Sadaoki Furui |
| 2003 | ICASSP | Unsupervised class-based language model adaptation for spontaneous speech recognition. | Tadasuke Yokoyama, Takahiro Shinozaki, Koji Iwano, Sadaoki Furui |
| 2003 | Interspeech | Time adjustable mixture weights for speaking rate fluctuation. | Takahiro Shinozaki, Sadaoki Furui |
| 2002 | ICASSP | Analysis on individual differences in automatic transcription of spontaneous presentations. | Takahiro Shinozaki, Sadaoki Furui |
| 2002 | Interspeech | A new lexicon optimization method for LVCSR based on linguistic and acoustic characteristics of words. | Takahiro Shinozaki, Sadaoki Furui |
| 2001 | ICASSP | Ubiquitous speech processing. | Sadaoki Furui, Koji Iwano, Chiori Hori, Takahiro Shinozaki, Yohei Saito, Satoshi Tamura |
| 2001 | Interspeech | Towards automatic transcription of spontaneous presentations. | Takahiro Shinozaki, Chiori Hori, Sadaoki Furui |
| 2000 | Interspeech | Toward the realization of spontaneous speech recognition - introduction of a Japanese priority program and preliminary results -. | Sadaoki Furui, Kikuo Maekawa, Hitoshi Isahara, Takahiro Shinozaki, Takashi Ohdaira |