Yosuke Kashiwagi
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
31
Venues
5
Active years
2013–2026
Best venue rank
A*
Where they publish
Papers
31 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | ACL | Optimizing Conversational Quality in Spoken Dialogue Systems with Reinforcement Learning from AI Feedback. | Siddhant Arora, Jinchuan Tian, Jiatong Shi, Hayato Futami, Yosuke Kashiwagi, Emiru Tsunoo, Shinji Watanabe |
| 2025 | ASRU | Spiralformer: Low Latency Encoder for Streaming Speech Recognition with Circular Layer Skipping and Early Exiting. | Emiru Tsunoo, Hayato Futami, Yosuke Kashiwagi, Siddhant Arora, Shinji Watanabe |
| 2025 | ICASSP | Hypothesis Clustering and Merging: Novel MultiTalker Speech Recognition with Speaker Tokens. | Yosuke Kashiwagi, Hayato Futami, Emiru Tsunoo, Siddhant Arora, Shinji Watanabe |
| 2025 | Interspeech | Chain-of-Thought Training for Open E2E Spoken Dialogue Systems. | Siddhant Arora, Jinchuan Tian, Hayato Futami, Jee-weon Jung, Jiatong Shi, Yosuke Kashiwagi, Emiru Tsunoo, Shinji Watanabe |
| 2025 | Interspeech | Scheduled Interleaved Speech-Text Training for Speech-to-Speech Translation with LLMs. | Hayato Futami, Emiru Tsunoo, Yosuke Kashiwagi, Yuki Ito, Hassan Shahmohammadi, Siddhant Arora, Shinji Watanabe |
| 2025 | Interspeech | Differentiable K-means for Fully-optimized Discrete Token-based ASR. | Kentaro Onda, Yosuke Kashiwagi, Emiru Tsunoo, Hayato Futami, Shinji Watanabe |
| 2025 | NAACL | ESPnet-SDS: Unified Toolkit and Demo for Spoken Dialogue Systems. | Siddhant Arora, Yifan Peng, Jiatong Shi, Jinchuan Tian, William Chen, Shikhar Bharadwaj, Hayato Futami, Yosuke Kashiwagi, Emiru Tsunoo, Shuichiro Shimizu, Vaibhav Srivastav, Shinji Watanabe |
| 2024 | ICASSP | Phoneme-Aware Encoding for Prefix-Tree-Based Contextual ASR. | Hayato Futami, Emiru Tsunoo, Yosuke Kashiwagi, Hiroaki Ogawa, Siddhant Arora, Shinji Watanabe |
| 2024 | Interspeech | Finding Task-specific Subnetworks in Multi-task Spoken Language Understanding Model. | Hayato Futami, Siddhant Arora, Yosuke Kashiwagi, Emiru Tsunoo, Shinji Watanabe |
| 2024 | Interspeech | Rapid Language Adaptation for Multilingual E2E Speech Recognition Using Encoder Prompting. | Yosuke Kashiwagi, Hayato Futami, Emiru Tsunoo, Siddhant Arora, Shinji Watanabe |
| 2024 | Interspeech | Decoder-only Architecture for Streaming End-to-end Speech Recognition. | Emiru Tsunoo, Hayato Futami, Yosuke Kashiwagi, Siddhant Arora, Shinji Watanabe |
| 2024 | NAACL | UniverSLU: Universal Spoken Language Understanding for Diverse Tasks with Natural Language Instructions. | Siddhant Arora, Hayato Futami, Jee-weon Jung, Yifan Peng, Roshan S. Sharma, Yosuke Kashiwagi, Emiru Tsunoo, Karen Livescu, Shinji Watanabe |
| 2023 | ICASSP | A Study on the Integration of Pipeline and E2E SLU Systems for Spoken Semantic Parsing Toward Stop Quality Challenge. | Siddhant Arora, Hayato Futami, Shih-Lun Wu, Jessica Huynh, Yifan Peng, Yosuke Kashiwagi, Emiru Tsunoo, Brian Yan, Shinji Watanabe |
| 2023 | ICASSP | The Pipeline System of ASR and NLU with MLM-based data Augmentation Toward Stop Low-Resource Challenge. | Hayato Futami, Jessica Huynh, Siddhant Arora, Shih-Lun Wu, Yosuke Kashiwagi, Yifan Peng, Brian Yan, Emiru Tsunoo, Shinji Watanabe |
| 2023 | ICASSP | Streaming Joint Speech Recognition and Disfluency Detection. | Hayato Futami, Emiru Tsunoo, Kentaro Shibata, Yosuke Kashiwagi, Takao Okuda, Siddhant Arora, Shinji Watanabe |
| 2023 | ICASSP | E-Branchformer-Based E2E SLU Toward Stop on-Device Challenge. | Yosuke Kashiwagi, Siddhant Arora, Hayato Futami, Jessica Huynh, Shih-Lun Wu, Yifan Peng, Brian Yan, Emiru Tsunoo, Shinji Watanabe |
| 2023 | Interspeech | Integrating Pretrained ASR and LM to Perform Sequence Generation for Spoken Language Understanding. | Siddhant Arora, Hayato Futami, Yosuke Kashiwagi, Emiru Tsunoo, Brian Yan, Shinji Watanabe |
| 2023 | Interspeech | Tensor decomposition for minimization of E2E SLU model toward on-device processing. | Yosuke Kashiwagi, Siddhant Arora, Hayato Futami, Jessica Huynh, Shih-Lun Wu, Yifan Peng, Brian Yan, Emiru Tsunoo, Shinji Watanabe |
| 2023 | Interspeech | Integration of Frame- and Label-synchronous Beam Search for Streaming Encoder-decoder Speech Recognition. | Emiru Tsunoo, Hayato Futami, Yosuke Kashiwagi, Siddhant Arora, Shinji Watanabe |
| 2022 | ICASSP | Joint Speech Recognition and Audio Captioning. | Chaitanya Narisetty, Emiru Tsunoo, Xuankai Chang, Yosuke Kashiwagi, Michael Hentschel, Shinji Watanabe |
| 2022 | ICASSP | Improving Character Error Rate is Not Equal to Having Clean Speech: Speech Enhancement for ASR Systems with Black-Box Acoustic Models. | Ryosuke Sawata, Yosuke Kashiwagi, Shusuke Takahashi |
| 2022 | ICASSP | Run-and-Back Stitch Search: Novel Block Synchronous Decoding For Streaming Encoder-Decoder ASR. | Emiru Tsunoo, Chaitanya Narisetty, Michael Hentschel, Yosuke Kashiwagi, Shinji Watanabe |
| 2022 | Interspeech | Residual Language Model for End-to-end Speech Recognition. | Emiru Tsunoo, Yosuke Kashiwagi, Chaitanya Prasad Narisetty, Shinji Watanabe |
| 2021 | ICASSP | Gaussian Kernelized Self-Attention for Long Sequence Data and its Application to CTC-Based Speech Recognition. | Yosuke Kashiwagi, Emiru Tsunoo, Shinji Watanabe |
| 2021 | Interspeech | Data Augmentation Methods for End-to-End Speech Recognition on Distant-Talk Scenarios. | Emiru Tsunoo, Kentaro Shibata, Chaitanya Narisetty, Yosuke Kashiwagi, Shinji Watanabe |
| 2019 | ASRU | Transformer ASR with Contextual Block Processing. | Emiru Tsunoo, Yosuke Kashiwagi, Toshiyuki Kumakura, Shinji Watanabe |
| 2019 | Interspeech | End-to-End Adaptation with Backpropagation Through WFST for On-Device Speech Recognition System. | Emiru Tsunoo, Yosuke Kashiwagi, Satoshi Asakawa, Toshiyuki Kumakura |
| 2016 | ICASSP | Divergence estimation based on deep neural networks and its use for language identification. | Yosuke Kashiwagi, Congying Zhang, Daisuke Saito, Nobuaki Minematsu |
| 2016 | Interspeech | Automatic Assessment and Error Detection of Shadowing Speech: Case of English Spoken by Japanese Learners. | Shuju Shi, Yosuke Kashiwagi, Shohei Toyama, Junwei Yue, Yutaka Yamauchi, Daisuke Saito, Nobuaki Minematsu |
| 2014 | ICASSP | Semi-supervised noise dictionary adaptation for exemplar-based noise robust speech recognition. | Yi Luan, Daisuke Saito, Yosuke Kashiwagi, Nobuaki Minematsu, Keikichi Hirose |
| 2013 | ASRU | Discriminative piecewise linear transformation based on deep learning for noise robust automatic speech recognition. | Yosuke Kashiwagi, Daisuke Saito, Nobuaki Minematsu, Keikichi Hirose |