| 2025 | ICASSP | Harnessing the Zero-Shot Power of Instruction-Tuned Large Language Model for Guiding End-to-End Speech Recognition. | Yosuke Higuchi, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2025 | ICINCO | Video-Based Vibration Analysis for Predictive Maintenance: A Motion Magnification and Random Forest Approach. | Walid Gomaa, Abdelrahman Wael Ammar, Ismael Abbo, Mohamed Galal Nassef, Tetsuji Ogawa, Mohab Hossam |
| 2025 | Interspeech | End-to-End Speech Translation Guided by Robust Translation Capability of Large Language Model. | Yosuke Higuchi, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2025 | Interspeech | Speaker-Distinguishable CTC: Learning Speaker Distinction Using CTC for Multi-Talker Speech Recognition. | Asahi Sakuma, Hiroaki Sato, Ryuga Sugano, Tadashi Kumano, Yoshihiko Kawai, Tetsuji Ogawa |
| 2025 | PACLIC | Analysis of the Correlation Between Theory of Mind and Dialogue Ability to Identify Essential ToM for Dialogue Systems. | Haruhisa Iseno, Atsumoto Ohashi, Tetsuji Ogawa, Shinnosuke Takamichi, Ryuichiro Higashinaka |
| 2024 | ICASSP | Parody Detection Using Source-Target Attention with Teacher-Forced Lyrics. | Tomoki Ariga, Yosuke Higuchi, Kazutoshi Hayasaka, Naoki Okamoto, Tetsuji Ogawa |
| 2024 | ICMLA | WindVibraTransformer: A Foundational Model for Precise and Robust Wind Turbine Condition Monitoring via Vibration Signals. | Takuya Wakayama, Taiki Inoue, Jun Ogata, Makoto Iida, Tetsuji Ogawa |
| 2024 | ICPR | Leveraging Data from Vast Unexplored Seas: Positive Unlabeled Learning for Refining Prediction Area in Good Fishing Ground Prediction. | Haruki Konii, Teppei Nakano, Yasumasa Miyazawa, Tetsuji Ogawa |
| 2024 | Interspeech | Hierarchical Multi-Task Learning with CTC and Recursive Operation. | Nahomi Kusunoki, Yosuke Higuchi, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2023 | ASRU | A Single Speech Enhancement Model Unifying Dereverberation, Denoising, Speaker Counting, Separation, And Extraction. | Kohei Saijo, Wangyou Zhang, Zhong-Qiu Wang, Shinji Watanabe, Tetsunori Kobayashi, Tetsuji Ogawa |
| 2023 | ICASSP | Neural Diarization with Non-Autoregressive Intermediate Attractors. | Yusuke Fujita, Tatsuya Komatsu, Robin Scheibler, Yusuke Kida, Tetsuji Ogawa |
| 2023 | ICASSP | Intermpl: Momentum Pseudo-Labeling With Intermediate CTC Loss. | Yosuke Higuchi, Tetsuji Ogawa, Tetsunori Kobayashi, Shinji Watanabe |
| 2023 | ICASSP | BECTRA: Transducer-Based End-To-End ASR with Bert-Enhanced Encoder. | Yosuke Higuchi, Tetsuji Ogawa, Tetsunori Kobayashi, Shinji Watanabe |
| 2023 | ICASSP | Self-Remixing: Unsupervised Speech Separation VIA Separation and Remixing. | Kohei Saijo, Tetsuji Ogawa |
| 2023 | ICASSP | Conversation-Oriented ASR with Multi-Look-Ahead CBS Architecture. | Huaibo Zhao, Shinya Fujie, Tetsuji Ogawa, Jin Sakuma, Yusuke Kida, Tetsunori Kobayashi |
| 2023 | ICINCO | Masry: A Text-to-Speech System for the Egyptian Arabic. | Ahmed Hammad Azab, Ahmed Bayoumy Zaki, Tetsuji Ogawa, Walid Gomaa |
| 2023 | ICMLA | Learning Discriminative Feature Representations via Metric Learning for Early Operation of Wind Turbine Anomaly Detection Systems. | Taiki Inoue, Jun Ogata, Makoto Iida, Tetsuji Ogawa |
| 2023 | IJCNN | Thermal Gait Dataset for Deep Learning-Oriented Gait Recognition. | Fatma Youssef, Ahmed El-Mahdy, Tetsuji Ogawa, Walid Gomaa |
| 2023 | Interspeech | Remixing-based Unsupervised Source Separation from Scratch. | Kohei Saijo, Tetsuji Ogawa |
| 2022 | EMNLP | BERT Meets CTC: New Formulation of End-to-End Speech Recognition with Pre-trained Masked Language Model. | Yosuke Higuchi, Brian Yan, Siddhant Arora, Tetsuji Ogawa, Tetsunori Kobayashi, Shinji Watanabe |
| 2022 | ICASSP | Hierarchical Conditional End-to-End ASR with CTC and Multi-Granular Subword Units. | Yosuke Higuchi, Keita Karube, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2022 | ICASSP | Remix-Cycle-Consistent Learning on Adversarially Learned Separator for Accurate and Stable Unsupervised Speech Separation. | Kohei Saijo, Tetsuji Ogawa |
| 2022 | Interspeech | Can Humans Correct Errors From System? Investigating Error Tendencies in Speaker Identification Using Crowdsourcing. | Yuta Ide, Susumu Saito, Teppei Nakano, Tetsuji Ogawa |
| 2022 | Interspeech | Confusion Detection for Adaptive Conversational Strategies of An Oral Proficiency Assessment Interview Agent. | Mao Saeki, Kotoka Miyagi, Shinya Fujie, Shungo Suzuki, Tetsuji Ogawa, Tetsunori Kobayashi, Yoichi Matsuyama |
| 2022 | Interspeech | Unsupervised Training of Sequential Neural Beamformer Using Coarsely-separated and Non-separated Signals. | Kohei Saijo, Tetsuji Ogawa |
| 2022 | Interspeech | Text-Only Domain Adaptation Based on Intermediate CTC. | Hiroaki Sato, Tomoyasu Komori, Takeshi Mishima, Yoshihiko Kawai, Takahiro Mochizuki, Shoei Sato, Tetsuji Ogawa |
| 2021 | ICASSP | Improved Mask-CTC for Non-Autoregressive End-to-End ASR. | Yosuke Higuchi, Hirofumi Inaguma, Shinji Watanabe, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2021 | Interspeech | Efficient and Stable Adversarial Learning Using Unpaired Data for Unsupervised Multichannel Speech Separation. | Yu Nakagome, Masahito Togami, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2021 | Interspeech | VocalTurk: Exploring Feasibility of Crowdsourced Speaker Identification. | Susumu Saito, Yuta Ide, Teppei Nakano, Tetsuji Ogawa |
| 2020 | COLING | Exploiting Narrative Context and A Priori Knowledge of Categories in Textual Emotion Classification. | Hikari Tanabe, Tetsuji Ogawa, Tetsunori Kobayashi, Yoshihiko Hayashi |
| 2020 | ICASSP | Deep Speech Extraction with Time-Varying Spatial Filtering Guided By Desired Direction Attractor. | Yu Nakagome, Masahito Togami, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2020 | ICASSP | Frame-Level Phoneme-Invariant Speaker Embedding for Text-Independent Speaker Recognition on Extremely Short Utterances. | Naohiro Tawara, Atsunori Ogawa, Tomoharu Iwata, Marc Delcroix, Tetsuji Ogawa |
| 2020 | ICPR | Feature Representation Learning for Calving Detection of Cows Using Video Frames. | Ryosuke Hyodo, Teppei Nakano, Tetsuji Ogawa |
| 2020 | ICPR | Toward Building a Data-Driven System For Detecting Mounting Actions of Black Beef Cattle. | Yuriko Kawano, Susumu Saito, Teppei Nakano, Ikumi Kondo, Ryota Yamazaki, Hiromi Kusaka, Minoru Sakaguchi, Tetsuji Ogawa |
| 2020 | ICPR | Crowdsourced Verification for Operating Calving Surveillance Systems at an Early Stage. | Yusuke Okimoto, Soshi Kawata, Susumu Saito, Teppei Nakano, Tetsuji Ogawa |
| 2020 | Interspeech | Mask CTC: Non-Autoregressive End-to-End ASR with CTC and Mask Predict. | Yosuke Higuchi, Shinji Watanabe, Nanxin Chen, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2020 | Interspeech | Mentoring-Reverse Mentoring for Unsupervised Multi-Channel Speech Source Separation. | Yu Nakagome, Masahito Togami, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2019 | ICASSP | Postfiltering Using an Adversarial Denoising Autoencoder with Noise-aware Training. | Naohiro Tawara, Hikari Tanabe, Tetsunori Kobayashi, Masaru Fujieda, Kazuhiro Katagiri, Takashi Yazu, Tetsuji Ogawa |
| 2019 | Interspeech | Speaker Adversarial Training of DPGMM-Based Feature Extractor for Zero-Resource Languages. | Yosuke Higuchi, Naohiro Tawara, Tetsunori Kobayashi, Tetsuji Ogawa |
| 2019 | Interspeech | Multi-Channel Speech Enhancement Using Time-Domain Convolutional Denoising Autoencoder. | Naohiro Tawara, Tetsunori Kobayashi, Tetsuji Ogawa |
| 2018 | ICASSP | Language Model Domain Adaptation Via Recurrent Neural Networks with Domain-Shared and Domain-Specific Representations. | Tsuyoshi Morioka, Naohiro Tawara, Tetsuji Ogawa, Atsunori Ogawa, Tomoharu Iwata, Tetsunori Kobayashi |
| 2018 | ICASSP | Speaker Invariant Feature Extraction for Zero-Resource Languages with Adversarial Learning. | Taira Tsuchiya, Naohiro Tawara, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2018 | ICPR | Sequential Fish Catch Forecasting Using Bayesian State Space Models. | Yuya Kokaki, Naohiro Tawara, Tetsunori Kobayashi, Kazuo Hashimoto, Tetsuji Ogawa |
| 2016 | ICPR | A new efficient measure for accuracy prediction and its application to multistream-based unsupervised adaptation. | Tetsuji Ogawa, Sri Harish Reddy Mallidi, Emmanuel Dupoux, Jordan Cohen, Naomi H. Feldman, Hynek Hermansky |
| 2015 | ASRU | Uncertainty estimation of DNN classifiers. | Sri Harish Reddy Mallidi, Tetsuji Ogawa, Hynek Hermansky |
| 2015 | ICASSP | Towards machines that know when they do not know: Summary of work done at 2014 Frederick Jelinek Memorial Workshop. | Hynek Hermansky, Luks Burget, Jordan Cohen, Emmanuel Dupoux, Naomi Feldman, John Godfrey, Sanjeev Khudanpur, Matthew Maciejewski, Sri Harish Reddy Mallidi, Anjali Menon, Tetsuji Ogawa, Vijayaditya Peddinti, Richard C. Rose, Richard M. Stern, Matthew Wiesner, Karel Vesel |
| 2015 | ICASSP | A comparative study of spectral clustering for i-vector-based speaker clustering under noisy conditions. | Naohiro Tawara, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2015 | Interspeech | Autoencoder based multi-stream combination for noise robust speech recognition. | Sri Harish Reddy Mallidi, Tetsuji Ogawa, Karel Vesel, Phani S. Nidadavolu, Hynek Hermansky |
| 2015 | Interspeech | Bilinear map of filter-bank outputs for DNN-based speech recognition. | Tetsuji Ogawa, Kenshiro Ueda, Kouichi Katsurada, Tetsunori Kobayashi, Tsuneo Nitta |
| 2014 | Interspeech | Effect of frequency weighting on MLP-based speaker canonicalization. | Yuichi Kubota, Motoi Omachi, Tetsuji Ogawa, Tetsunori Kobayashi, Tsuneo Nitta |
| 2013 | Interspeech | Stream selection and integration in multistream ASR using GMM-based performance monitoring. | Tetsuji Ogawa, Feipeng Li, Hynek Hermansky |
| 2012 | ICASSP | Fully Bayesian inference of multi-mixture Gaussian model and its evaluation using speaker clustering. | Naohiro Tawara, Tetsuji Ogawa, Shinji Watanabe, Tetsunori Kobayashi |
| 2012 | ICPR | An improved entropy-based multiple kernel learning. | Hideitsu Hino, Tetsuji Ogawa |
| 2012 | Interspeech | Fully Bayesian speaker clustering based on hierarchically structured utterance-oriented Dirichlet process mixture model. | Naohiro Tawara, Tetsuji Ogawa, Shinji Watanabe, Atsushi Nakamura, Tetsunori Kobayashi |
| 2011 | ICASSP | Speaker recognition using multiple kernel learning based on conditional entropy minimization. | Tetsuji Ogawa, Hideitsu Hino, Nima Reyhani, Noboru Murata, Tetsunori Kobayashi |
| 2011 | Interspeech | Speaker Verification Robust to Talking Style Variation Using Multiple Kernel Learning Based on Conditional Entropy Minimization. | Tetsuji Ogawa, Hideitsu Hino, Noboru Murata, Tetsunori Kobayashi |
| 2011 | Interspeech | Spatial Filter Calibration Based on Minimization of Modified LSD. | Nobuaki Tanaka, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2011 | Interspeech | Speaker Clustering Based on Utterance-Oriented Dirichlet Process Mixture Model. | Naohiro Tawara, Shinji Watanabe, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2009 | IROS | Robot auditory system using head-mounted square microphone array. | Kosuke Hosoya, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2008 | ICASSP | A Kalman filter based restoration method for in-vehicle camera images in foggy conditions. | Tomoki Hiramatsu, Tetsuji Ogawa, Miki Haseyama |
| 2008 | ICASSP | Kernel PCA-based resolution enhancement approach of still images using different levels of pyramid structure. | Tetsuji Ogawa, Miki Haseyama |
| 2008 | ICASSP | Speech enhancement using square microphone array for mobile devices. | Shintaro Takada, Tetsuji Ogawa, Kenzo Akagiri, Tetsunori Kobayashi |
| 2008 | Interspeech | CENSREC-4: development of evaluation framework for distant-talking speech recognition under reverberant environments. | Masato Nakayama, Takanobu Nishiura, Yuki Denda, Norihide Kitaoka, Kazumasa Yamamoto, Takeshi Yamada, Satoru Tsuge, Chiyomi Miyajima, Masakiyo Fujimoto, Tetsuya Takiguchi, Satoshi Tamura, Tetsuji Ogawa, Shigeki Matsuda, Shingo Kuroiwa, Kazuya Takeda, Satoshi Nakamura |
| 2007 | ICASSP | Adequacy Analysis of Simulation-Based Assessment of Speech Recognition System. | Tetsuji Ogawa, Satoshi Kanba, Tetsunori Kobayashi |
| 2006 | Interspeech | Manifold HLDA and its application to robust speech recognition. | Toshiaki Kubo, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2005 | Interspeech | Optimizing the structure of partly-hidden Markov models using weighted likelihood-ratio maximization criterion. | Tetsuji Ogawa, Tetsunori Kobayashi |
| 2004 | Interspeech | Recognition of three simultaneous utterance of speech by four-line directivity microphone mounted on head of robot. | Naoya Mochiki, Tetsunori Kobayashi, Toshiyuki Sekiya, Tetsuji Ogawa |
| 2003 | ICASSP | Hybrid modeling of PHMM and HMM for speech recognition. | Tetsuji Ogawa, Tetsunori Kobayashi |
| 2003 | Interspeech | Speech recognition of double talk using SAFIA-based audio segregation. | Toshiyuki Sekiya, Tetsuji Ogawa, Tetsunori Kobayashi |
| 2002 | Interspeech | Generalization of state-observation-dependency in partly hidden Markov models. | Tetsuji Ogawa, Tetsunori Kobayashi |