Tomoki Hayashi
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
43
Venues
8
Active years
2010–2023
Best venue rank
A*
Where they publish
Papers
43 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2023 | ACL | ESPnet-ST-v2: Multipurpose Spoken Language Translation Toolkit. | Brian Yan, Jiatong Shi, Yun Tang, Hirofumi Inaguma, Yifan Peng, Siddharth Dalmia, Peter Polak, Patrick Fernandes, Dan Berrebbi, Tomoki Hayashi, Xiaohui Zhang, Zhaoheng Ni, Moto Hira, Soumi Maiti, Juan Pino, Shinji Watanabe |
| 2023 | ICASSP | Low-Latency Electrolaryngeal Speech Enhancement Based on Fastspeech2-Based Voice Conversion and Self-Supervised Speech Representation. | Kazuhiro Kobayashi, Tomoki Hayashi, Tomoki Toda |
| 2022 | BMVC | Improving Dense Representation Learning by Superpixelization and Contrasting Cluster Assignment. | Robin Karlsson, Tomoki Hayashi, Keisuke Fujii, Alexander Carballo, Kento Ohtani, Kazuya Takeda |
| 2022 | ICASSP | An Investigation of Streaming Non-Autoregressive sequence-to-sequence Voice Conversion. | Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2022 | ICASSP | S3PRL-VC: Open-Source Voice Conversion Framework with Self-Supervised Speech Representations. | Wen-Chin Huang, Shu-Wen Yang, Tomoki Hayashi, Hung-Yi Lee, Shinji Watanabe, Tomoki Toda |
| 2022 | Interspeech | Muskits: an End-to-end Music Processing Toolkit for Singing Voice Synthesis. | Jiatong Shi, Shuai Guo, Tao Qian, Tomoki Hayashi, Yuning Wu, Fangzheng Xu, Xuankai Chang, Huazhe Li, Peter Wu, Shinji Watanabe, Qin Jin |
| 2021 | ASRU | On Prosody Modeling for ASR+TTS Based Voice Conversion. | Wen-Chin Huang, Tomoki Hayashi, Xinjian Li, Shinji Watanabe, Tomoki Toda |
| 2021 | ICASSP | Recent Developments on Espnet Toolkit Boosted By Conformer. | Pengcheng Guo, Florian Boyer, Xuankai Chang, Tomoki Hayashi, Yosuke Higuchi, Hirofumi Inaguma, Naoyuki Kamo, Chenda Li, Daniel Garcia-Romero, Jiatong Shi, Jing Shi, Shinji Watanabe, Kun Wei, Wangyou Zhang, Yuekai Zhang |
| 2021 | ICASSP | Non-Autoregressive Sequence-To-Sequence Voice Conversion. | Tomoki Hayashi, Wen-Chin Huang, Kazuhiro Kobayashi, Tomoki Toda |
| 2021 | ICASSP | Any-to-One Sequence-to-Sequence Voice Conversion Using Self-Supervised Discrete Speech Representations. | Wen-Chin Huang, Yi-Chiao Wu, Tomoki Hayashi |
| 2021 | ICASSP | Crank: An Open-Source Software for Nonparallel Voice Conversion Based on Vector-Quantized Variational Autoencoder. | Kazuhiro Kobayashi, Wen-Chin Huang, Yi-Chiao Wu, Patrick Lumban Tobing, Tomoki Hayashi, Tomoki Toda |
| 2021 | Interspeech | Acoustic Event Detection with Classifier Chains. | Tatsuya Komatsu, Shinji Watanabe, Koichi Miyazaki, Tomoki Hayashi |
| 2020 | ACL | ESPnet-ST: All-in-One Speech Translation Toolkit. | Hirofumi Inaguma, Shun Kiyono, Kevin Duh, Shigeki Karita, Nelson Yalta, Tomoki Hayashi, Shinji Watanabe |
| 2020 | ICASSP | Espnet-TTS: Unified, Reproducible, and Integratable Open Source End-to-End Text-to-Speech Toolkit. | Tomoki Hayashi, Ryuichi Yamamoto, Katsuki Inoue, Takenori Yoshimura, Shinji Watanabe, Tomoki Toda, Kazuya Takeda, Yu Zhang, Xu Tan |
| 2020 | ICASSP | Semi-Supervised Speaker Adaptation for End-to-End Speech Synthesis with Pretrained Models. | Katsuki Inoue, Sunao Hara, Masanobu Abe, Tomoki Hayashi, Ryuichi Yamamoto, Shinji Watanabe |
| 2020 | ICASSP | Weakly-Supervised Sound Event Detection with Self-Attention. | Koichi Miyazaki, Tatsuya Komatsu, Tomoki Hayashi, Shinji Watanabe, Tomoki Toda, Kazuya Takeda |
| 2020 | ICASSP | Efficient Shallow Wavenet Vocoder Using Multiple Samples Output Based on Laplacian Distribution and Linear Prediction. | Patrick Lumban Tobing, Yi-Chiao Wu, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2020 | ICASSP | End-to-End Automatic Speech Recognition Integrated with CTC-Based Voice Activity Detection. | Takenori Yoshimura, Tomoki Hayashi, Kazuya Takeda, Shinji Watanabe |
| 2020 | Interspeech | Intelligibility Enhancement Based on Speech Waveform Modification Using Hearing Impairment. | Shu Hikosaka, Shogo Seki, Tomoki Hayashi, Kazuhiro Kobayashi, Kazuya Takeda, Hideki Banno, Tomoki Toda |
| 2020 | Interspeech | Voice Transformer Network: Sequence-to-Sequence Voice Conversion Using Transformer with Text-to-Speech Pretraining. | Wen-Chin Huang, Tomoki Hayashi, Yi-Chiao Wu, Hirokazu Kameoka, Tomoki Toda |
| 2020 | Interspeech | Cyclic Spectral Modeling for Unsupervised Unit Discovery into Voice Conversion with Excitation and Waveform Modeling. | Patrick Lumban Tobing, Tomoki Hayashi, Yi-Chiao Wu, Kazuhiro Kobayashi, Tomoki Toda |
| 2020 | Interspeech | Quasi-Periodic Parallel WaveGAN Vocoder: A Non-Autoregressive Pitch-Dependent Dilated Convolution Model for Parametric Speech Generation. | Yi-Chiao Wu, Tomoki Hayashi, Takuma Okamoto, Hisashi Kawai, Tomoki Toda |
| 2019 | ASRU | A Comparative Study on Transformer vs RNN in Speech Applications. | Shigeki Karita, Xiaofei Wang, Shinji Watanabe, Takenori Yoshimura, Wangyou Zhang, Nanxin Chen, Tomoki Hayashi, Takaaki Hori, Hirofumi Inaguma, Ziyan Jiang, Masao Someki, Nelson Enrique Yalta Soplin, Ryuichi Yamamoto |
| 2019 | ASRU | Attention-Based Speech Recognition Using Gaze Information. | Osamu Segawa, Tomoki Hayashi, Kazuya Takeda |
| 2019 | ASRU | Investigation of Shallow Wavenet Vocoder with Laplacian Distribution Output. | Patrick Lumban Tobing, Tomoki Hayashi, Tomoki Toda |
| 2019 | ICASSP | Cycle-consistency Training for End-to-end Speech Recognition. | Takaaki Hori, Ramn Fernandez Astudillo, Tomoki Hayashi, Yu Zhang, Shinji Watanabe, Jonathan Le Roux |
| 2019 | ICASSP | Scene-dependent Anomalous Acoustic-event Detection Based on Conditional Wavenet and I-vector. | Tatsuya Komatsu, Tomoki Hayashi, Reishi Kondo, Tomoki Toda, Kazuya Takeda |
| 2019 | ICASSP | Voice Conversion with Cyclic Recurrent Neural Network and Fine-tuned Wavenet Vocoder. | Patrick Lumban Tobing, Yi-Chiao Wu, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2019 | Interspeech | Pre-Trained Text Embeddings for Enhanced Text-to-Speech Synthesis. | Tomoki Hayashi, Shinji Watanabe, Tomoki Toda, Kazuya Takeda, Shubham Toshniwal, Karen Livescu |
| 2019 | Interspeech | Investigation of F0 Conditioning and Fully Convolutional Networks in Variational Autoencoder Based Voice Conversion. | Wen-Chin Huang, Yi-Chiao Wu, Chen-Chou Lo, Patrick Lumban Tobing, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda, Yu Tsao, Hsin-Min Wang |
| 2019 | Interspeech | Non-Parallel Voice Conversion with Cyclic Variational Autoencoder. | Patrick Lumban Tobing, Yi-Chiao Wu, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2019 | Interspeech | Quasi-Periodic WaveNet Vocoder: A Pitch Dependent Dilated Convolution Model for Parametric Speech Generation. | Yi-Chiao Wu, Tomoki Hayashi, Patrick Lumban Tobing, Kazuhiro Kobayashi, Tomoki Toda |
| 2018 | Interspeech | Multi-Head Decoder for End-to-End Speech Recognition. | Tomoki Hayashi, Shinji Watanabe, Tomoki Toda, Kazuya Takeda |
| 2018 | Interspeech | ESPnet: End-to-End Speech Processing Toolkit. | Shinji Watanabe, Takaaki Hori, Shigeki Karita, Tomoki Hayashi, Jiro Nishitoba, Yuya Unno, Nelson Enrique Yalta Soplin, Jahn Heymann, Matthew Wiesner, Nanxin Chen, Adithya Renduchintala, Tsubasa Ochiai |
| 2018 | Interspeech | Collapsed Speech Segment Detection and Suppression for WaveNet Vocoder. | Yi-Chiao Wu, Kazuhiro Kobayashi, Tomoki Hayashi, Patrick Lumban Tobing, Tomoki Toda |
| 2017 | ASRU | An investigation of multi-speaker training for wavenet vocoder. | Tomoki Hayashi, Akira Tamamori, Kazuhiro Kobayashi, Kazuya Takeda, Tomoki Toda |
| 2017 | ICASSP | BLSTM-HMM hybrid system combined with sound activity detection network for polyphonic Sound Event Detection. | Tomoki Hayashi, Shinji Watanabe, Tomoki Toda, Takaaki Hori, Jonathan Le Roux, Kazuya Takeda |
| 2017 | Interspeech | Statistical Voice Conversion with WaveNet-Based Waveform Generation. | Kazuhiro Kobayashi, Tomoki Hayashi, Akira Tamamori, Tomoki Toda |
| 2017 | Interspeech | Speaker-Dependent WaveNet Vocoder. | Akira Tamamori, Tomoki Hayashi, Kazuhiro Kobayashi, Kazuya Takeda, Tomoki Toda |
| 2015 | ICASSP | Exploring multi-channel features for denoising-autoencoder-based speech enhancement. | Shoko Araki, Tomoki Hayashi, Marc Delcroix, Masakiyo Fujimoto, Kazuya Takeda, Tomohiro Nakatani |
| 2013 | SIGGRAPH | Dream board: a visualization system by handwriting recognition. | Tomomi Hatanaka, Tomoki Hayashi, Keita Suzuki, Hiroaki Sawano, Takeshi Tsuchiya, Kei'ichi Koyanagi |
| 2011 | MVA | Skeleton Features Distribution for 3D Object Retrieval. | Tomoki Hayashi, Benjamin Raynal, Vincent Nozick, Hideo Saito |
| 2010 | ICPR | An Augmented Reality Setup with an Omnidirectional Camera Based on Multiple Object Detection. | Tomoki Hayashi, Hideaki Uchiyama, Julien Pilet, Hideo Saito |