| 2025 | ICIP | Diffusion Pretraining for Gait Recognition in the Wild. | Wei Ming Neo, Koichi Shinoda, Tat-Jen Cham |
| 2025 | Interspeech | SepVAC: Multitask Learning of Speaker Separation, Speaker Localization, Microphone Array Localization, and Room Acoustic Parameter Estimation in Various Acoustic Conditions. | Roland Hartanto, Sakriani Sakti, Koichi Shinoda |
| 2025 | WACV | Diffusion-Based Generative Regularization for Supervised Discriminative Learning. | Takuya Asakura, Nakamasa Inoue, Koichi Shinoda |
| 2024 | ICIP | Pyramid Coder: Hierarchical Code Generator for Compositional Visual Question Answering. | Ruoyue Shen, Nakamasa Inoue, Koichi Shinoda |
| 2024 | Interspeech | MSDET: Multitask Speaker Separation and Direction-of-Arrival Estimation Training. | Roland Hartanto, Sakriani Sakti, Koichi Shinoda |
| 2024 | MMM | Co-speech Gesture Generation with Variational Auto Encoder. | Shinichi Ka, Koichi Shinoda |
| 2024 | WACV | CAMOT: Camera Angle-aware Multi-Object Tracking. | Felix Limanta, Kuniaki Uto, Koichi Shinoda |
| 2023 | ICASSP | Synthesizing Speech from ECoG with a Combination of Transformer-Based Encoder and Neural Vocoder. | Kai Shigemi, Shuji Komeiji, Takumi Mitsuhashi, Yasushi Iimura, Hiroharu Suzuki, Hidenori Sugano, Koichi Shinoda, Kohei Yatabe, Toshihisa Tanaka |
| 2023 | MMM | EvIs-Kitchen: Egocentric Human Activities Recognition with Video and Inertial Sensor Data. | Yuzhe Hao, Kuniaki Uto, Asako Kanezaki, Ikuro Sato, Rei Kawakami, Koichi Shinoda |
| 2023 | WACV | Text-Guided Object Detector for Multi-modal Video Question Answering. | Ruoyue Shen, Nakamasa Inoue, Koichi Shinoda |
| 2022 | ECCV | Implicit Neural Representations for Variable Length Human Motion Generation. | Pablo Cervantes, Yusuke Sekikawa, Ikuro Sato, Koichi Shinoda |
| 2022 | ICASSP | Transformer-Based Estimation of Spoken Sentences Using Electrocorticography. | Shuji Komeiji, Kai Shigemi, Takumi Mitsuhashi, Yasushi Iimura, Hiroharu Suzuki, Hidenori Sugano, Koichi Shinoda, Toshihisa Tanaka |
| 2022 | IJCNN | MSR-DARTS: Minimum Stable Rank of Differentiable Architecture Search. | Kengo Machida, Kuniaki Uto, Koichi Shinoda, Taiji Suzuki |
| 2022 | IGARSS | RI-DC: Rotation-Invariant Detection and Classification for Wheat Head Detection. | Takeru Ito, Kuniaki Uto, Koichi Shinoda |
| 2021 | ASRU | Multimodal Emotion Recognition with High-Level Speech and Text Features. | Mariana Rodrigues Makiuchi, Kuniaki Uto, Koichi Shinoda |
| 2020 | ICMI | Deep Video Understanding of Character Relationships in Movies. | Yang Lu, Asri Rizki Yuliani, Keisuke Ishikawa, Ronaldo Prata Amorim, Roland Hartanto, Nakamasa Inoue, Kuniaki Uto, Koichi Shinoda |
| 2020 | IGARSS | Estimation of Leaf Angle Distribution Based on Statistical Properties of Leaf Shading Distribution. | Kuniaki Uto, Mauro Dalla Mura, Yuka Sasaki, Koichi Shinoda |
| 2020 | Interspeech | NEC-TT Speaker Verification System for SRE'19 CTS Challenge. | Kong Aik Lee, Koji Okabe, Hitoshi Yamamoto, Qiongqiong Wang, Ling Guo, Takafumi Koshinaka, Jiacen Zhang, Keisuke Ishikawa, Koichi Shinoda |
| 2019 | ICASSP | Sequence-level Knowledge Distillation for Model Compression of Attention-based Sequence-to-sequence Speech Recognition. | Raden Mu'az Mun'im, Nakamasa Inoue, Koichi Shinoda |
| 2019 | IGARSS | Estimation of Diffuse Component of Global Radiation Based on Leaf-Scale Crop Images. | Kuniaki Uto, Mauro Dalla Mura, Jocelyn Chanussot, Koichi Shinoda |
| 2019 | Interspeech | The NEC-TT 2018 Speaker Verification System. | Kong Aik Lee, Hitoshi Yamamoto, Koji Okabe, Qiongqiong Wang, Ling Guo, Takafumi Koshinaka, Jiacen Zhang, Koichi Shinoda |
| 2019 | Interspeech | A Modified Algorithm for Multiple Input Spectrogram Inversion. | Dongxiao Wang, Hirokazu Kameoka, Koichi Shinoda |
| 2018 | BMVC | A Fine-to-Coarse Convolutional Neural Network for 3D Human Action Recognition. | Thao Le Minh, Nakamasa Inoue, Koichi Shinoda |
| 2018 | ICASSP | Multi-Task Autoencoder for Noise-Robust Speech Recognition. | Haoyi Zhang, Conggui Liu, Nakamasa Inoue, Koichi Shinoda |
| 2018 | IJCAI | Deep Learning Based Multi-modal Addressee Recognition in Visual Scenes with Utterances. | Thao Le Minh, Nobuyuki Shimizu, Takashi Miyazaki, Koichi Shinoda |
| 2018 | Interspeech | Attentive Statistics Pooling for Deep Speaker Embedding. | Koji Okabe, Takafumi Koshinaka, Koichi Shinoda |
| 2018 | Interspeech | Detecting Alzheimer's Disease Using Gated Convolutional Neural Network from Audio Data. | Tifani Warnita, Nakamasa Inoue, Koichi Shinoda |
| 2018 | Interspeech | I-vector Transformation Using Conditional Generative Adversarial Networks for Short Utterance Speaker Verification. | Jiacen Zhang, Nakamasa Inoue, Koichi Shinoda |
| 2017 | MMM | Boredom Recognition Based on Users' Spontaneous Behaviors in Multiparty Human-Robot Interactions. | Yasuhiro Shibasaki, Kotaro Funakoshi, Koichi Shinoda |
| 2016 | Interspeech | Recurrent Out-of-Vocabulary Word Detection Using Distribution of Features. | Taichi Asami, Ryo Masumura, Yushi Aono, Koichi Shinoda |
| 2014 | ACCV | Spectral Graph Skeletons for 3D Action Recognition. | Tommi Kerola, Nakamasa Inoue, Koichi Shinoda |
| 2014 | ACL | Semantics for Large-Scale Multimedia: New Challenges for NLP. | Florian Metze, Koichi Shinoda |
| 2014 | ICASSP | Constrained discriminative PLDA training for speaker verification. | Johan Rohdin, Sangeeta Biswas, Koichi Shinoda |
| 2014 | Interspeech | Simple gesture-based error correction interface for smartphone speech recognition. | Yuan Liang, Koji Iwano, Koichi Shinoda |
| 2014 | MMM | Event Detection by Velocity Pyramid. | Zhuolin Liang, Nakamasa Inoue, Koichi Shinoda |
| 2013 | ICCV | Neighbor-to-Neighbor Search for Fast Coding of Feature Vectors. | Nakamasa Inoue, Koichi Shinoda |
| 2013 | ICIAP | Statistical Person Verification Using Behavioral Patterns from Complex Human Motion. | Felipe Gmez-Caballero, Takahiro Shinozaki, Sadaoki Furui, Koichi Shinoda |
| 2013 | Interspeech | Combining deep speaker specific representations with GMM-SVM for speaker verification. | Ryan Price, Sangeeta Biswas, Koichi Shinoda |
| 2012 | ACCV | q-Gaussian Mixture Models Based on Non-extensive Statistics for Image and Video Semantic Indexing. | Nakamasa Inoue, Koichi Shinoda |
| 2012 | ICIP | Multimedia event detection using GMM supervectors and SVMS. | Yusuke Kamishima, Nakamasa Inoue, Koichi Shinoda, Shunsuke Sato |
| 2012 | Interspeech | Q-Gaussian based spectral subtraction for robust speech recognition. | Hilman Ferdinandus Pardede, Koichi Shinoda, Koji Iwano |
| 2012 | Interspeech | Overlapped Speech Detection in Meeting Using Cross-Channel Spectral Subtraction and Spectrum Similarity. | Ryo Yokoyama, Yu Nasu, Koichi Shinoda, Koji Iwano |
| 2011 | ASRU | Designing text corpus using phone-error distribution for acoustic modeling. | Hiroko Murakami, Koichi Shinoda, Sadaoki Furui |
| 2011 | ICASSP | Structural MAP adaptation in GMM-supervector based speaker recognition. | Marc Ferras, Koichi Shinoda, Sadaoki Furui |
| 2011 | ICASSP | Cross-Channel Spectral Subtraction for meeting speech recognition. | Yu Nasu, Koichi Shinoda, Sadaoki Furui |
| 2011 | Interspeech | Acoustic Forest for SMAP-Based Speaker Verification. | Sangeeta Biswas, Marc Ferras, Koichi Shinoda, Sadaoki Furui |
| 2011 | Interspeech | Structural Joint Factor Analysis for Speaker Recognition. | Marc Ferras, Koichi Shinoda, Sadaoki Furui |
| 2011 | Interspeech | Generalized-Log Spectral Mean Normalization for Speech Recognition. | Hilman Ferdinandus Pardede, Koichi Shinoda |
| 2010 | ICASSP | Speech modeling based on committee-based active learning. | Yuzo Hamanaka, Koichi Shinoda, Sadaoki Furui, Tadashi Emori, Takafumi Koshinaka |
| 2010 | ICPR | Robust Gait Recognition Against Speed Variation. | Muhammad Rasyid Aqmar, Koichi Shinoda, Sadaoki Furui |
| 2010 | ICPR | High-Level Feature Extraction Using SIFT GMMs and Audio Models. | Nakamasa Inoue, Tatsuhiko Saito, Koichi Shinoda, Sadaoki Furui |
| 2010 | Interspeech | Dynamic language model adaptation using keyword category classification. | Hitoshi Yamamoto, Ken Hanazawa, Kiyokazu Miki, Koichi Shinoda |
| 2009 | ICASSP | Independent component analysis for noisy speech recognition. | Hsin-Lung Hsieh, Jen-Tzung Chien, Koichi Shinoda, Sadaoki Furui |
| 2009 | ICASSP | Online speaker clustering using incremental learning of an ergodic hidden Markov model. | Takafumi Koshinaka, Kentaro Nagatomo, Koichi Shinoda |
| 2009 | Interspeech | Speaker adaptation based on two-step active learning. | Koichi Shinoda, Hiroko Murakami, Sadaoki Furui |
| 2008 | Interspeech | Robust spoken term detection using combination of phone-based and word-based recognition. | Kenji Iwata, Koichi Shinoda, Sadaoki Furui |
| 2008 | Interspeech | Improvement of eigenvoice-based speaker adaptation by parameter space clustering. | Shutaro Tanji, Koichi Shinoda, Sadaoki Furui, Antonio Ortega |
| 2008 | Interspeech | Time-lag adaptation for semi-synchronous speech and pen input. | Yasushi Watanabe, Koichi Shinoda, Sadaoki Furui |
| 2007 | ICASSP | Speech Recognition using FHMMS Robust Against Nonstationary Noise. | Agnieszka Betkowska, Koichi Shinoda, Sadaoki Furui |
| 2007 | ICASSP | Semi-Synchronous Speech and Pen Input. | Yasushi Watanabe, Kenji Iwata, Ryuta Nakagawa, Koichi Shinoda, Sadaoki Furui |
| 2007 | Interspeech | Predictive minimum Bayes risk classification for robust speech recognition. | Jen-Tzung Chien, Koichi Shinoda, Sadaoki Furui |
| 2007 | Interspeech | Automatic estimation of scaling factors among probabilistic models in speech recognition. | Tadashi Emori, Yoshifumi Onishi, Koichi Shinoda |
| 2007 | Interspeech | Dynamic language model adaptation using presentation slides for lecture speech recognition. | Hiroki Yamazaki, Koji Iwano, Koichi Shinoda, Sadaoki Furui, Haruo Yokota |
| 2006 | ICASSP | Towards Optimal Bayes Decision for Speech Recognition. | Jen-Tzung Chien, Chih-Hsien Huang, Koichi Shinoda, Sadaoki Furui |
| 2005 | ICIP | Robust highlight extraction using multi-stream hidden Markov models for baseball video. | Nguyen Huu Bach, Koichi Shinoda, Sadaoki Furui |
| 2002 | ICASSP | Efficient reduction of Gaussian components using MDL criterion for HMM-based speech recognition. | Koichi Shinoda, Ken-ichi Iso |
| 2001 | Interspeech | Rapid vocal tract length normalization using maximum likelihood estimation. | Tadashi Emori, Koichi Shinoda |
| 1998 | ICASSP | Unsupervised adaptation using structural Bayes approach. | Koichi Shinoda, Chin-Hui Lee |
| 1997 | Interspeech | Acoustic modeling based on the MDL principle for speech recognition. | Koichi Shinoda, Takao Watanabe |
| 1996 | ICASSP | Speaker adaptation with autonomous model complexity control by MDL principle. | Koichi Shinoda, Takao Watanabe |
| 1996 | Interspeech | Unsupervised and incremental speaker adaptation under adverse environmental conditions. | Keizaburo Takagi, Koichi Shinoda, Hiroaki Hattori, Takao Watanabe |
| 1995 | ICASSP | High speed speech recognition using tree-structured probability density function. | Takao Watanabe, Koichi Shinoda, Keizaburo Takagi, Ken-ichi Iso |
| 1995 | Interspeech | Speaker adaptation with autonomous control using tree structure. | Koichi Shinoda, Takao Watanabe |
| 1994 | Interspeech | Unsupervised speaker adaptation for speech recognition using demi-syllable HMM. | Koichi Shinoda, Takao Watanabe |
| 1994 | Interspeech | Speech recognition using tree-structured probability density function. | Takao Watanabe, Koichi Shinoda, Keizaburo Takagi, Eiko Yamada |
| 1991 | ICASSP | Speaker adaptation for demi-syllable based continuous density HMM. | Koichi Shinoda, Ken-ichi Iso, Takao Watanabe |
| 1990 | Interspeech | Speaker adaptation for demi-syllable based speech recognition using continuous HMM. | Koichi Shinoda, Ken-ichi Iso, Takao Watanabe |