| 2022 | ICASSP | Knowledge Transfer from Large-Scale Pretrained Language Models to End-To-End Speech Recognizers. | Yotaro Kubo, Shigeki Karita, Michiel Bacchiani |
| 2021 | Interspeech | A Comparative Study on Neural Architectures and Training Methods for Japanese Speech Recognition. | Shigeki Karita, Yotaro Kubo, Michiel Adriaan Unico Bacchiani, Llion Jones |
| 2020 | ICASSP | Joint Phoneme-Grapheme Model for End-To-End Speech Recognition. | Yotaro Kubo, Michiel Bacchiani |
| 2014 | AISTATS | Online Passive-Aggressive Algorithms for Non-Negative Matrix Factorization and Completion. | Mathieu Blondel, Yotaro Kubo, Naonori Ueda |
| 2014 | ICASSP | Unsupervised non-parametric Bayesian modeling of non-stationary noise for model-based noise suppression. | Masakiyo Fujimoto, Yotaro Kubo, Tomohiro Nakatani |
| 2014 | ICASSP | Real-time one-pass decoding with recurrent neural network language model for speech recognition. | Takaaki Hori, Yotaro Kubo, Atsushi Nakamura |
| 2014 | Interspeech | Restructuring output layers of deep neural networks using minimum risk parameter clustering. | Yotaro Kubo, Jun Suzuki, Takaaki Hori, Atsushi Nakamura |
| 2013 | ICASSP | Large vocabulary continuous speech recognition based on WFST structured classifiers and deep bottleneck features. | Yotaro Kubo, Takaaki Hori, Atsushi Nakamura |
| 2013 | Interspeech | Is speech enhancement pre-processing still relevant when using deep neural networks for acoustic modeling? | Marc Delcroix, Yotaro Kubo, Tomohiro Nakatani, Atsushi Nakamura |
| 2013 | Interspeech | A method for structure estimation of weighted finite-state transducers and its application to grapheme-to-phoneme conversion. | Yotaro Kubo, Takaaki Hori, Atsushi Nakamura |
| 2012 | ICASSP | Decoding network optimization using minimum transition error training. | Yotaro Kubo, Shinji Watanabe, Atsushi Nakamura |
| 2012 | ICASSP | Basis vector orthogonalization for an improved kernel gradient matching pursuit method. | Yotaro Kubo, Shinji Watanabe, Atsushi Nakamura, Simon Wiesler, Ralf Schlter, Hermann Ney |
| 2012 | ICASSP | Bag Of ARCS: New representation of speech segment features based on finite state machines. | Shinji Watanabe, Yotaro Kubo, Takanobu Oba, Takaaki Hori, Atsushi Nakamura |
| 2012 | Interspeech | Integrating Deep Neural Networks into Structural Classification Approach based on Weighted Finite-State Transducers. | Yotaro Kubo, Takaaki Hori, Atsushi Nakamura |
| 2011 | ICASSP | Subspace pursuit method for kernel-log-linear models. | Yotaro Kubo, Simon Wiesler, Ralf Schlter, Hermann Ney, Shinji Watanabe, Atsushi Nakamura, Tetsunori Kobayashi |
| 2011 | ICASSP | Feature selection for log-linear acoustic models. | Simon Wiesler, Alexander Richard, Yotaro Kubo, Ralf Schlter, Hermann Ney |
| 2010 | Interspeech | A regularized discriminative training method of acoustic models derived by minimum relative entropy discrimination. | Yotaro Kubo, Shinji Watanabe, Atsushi Nakamura, Tetsunori Kobayashi |
| 2008 | ICASSP | Noisy speech recognition using temporal AM-FM combination. | Yotaro Kubo, Akira Kurematsu, Katsuhiko Shirai, Shigeki Okawa |
| 2008 | Interspeech | A comparative study on AM and FM features. | Yotaro Kubo, Shigeki Okawa, Akira Kurematsu, Katsuhiko Shirai |
| 2007 | Interspeech | A study on temporal features derived by analytic signal. | Yotaro Kubo, Shigeki Okawa, Akira Kurematsu, Katsuhiko Shirai |