Skip to content

Tomoki Koriyama

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

38

Venues

3

Active years

2010–2025

Best venue rank

A

Where they publish

Papers

38 indexed papers, newest first.

YearVenueTitleAuthors
2025InterspeechSpeaker-agnostic Emotion Vector for Cross-speaker Emotion Intensity Control.Masato Murata, Koichi Miyazaki, Tomoki Koriyama
2025InterspeechEigenvoice Synthesis based on Model Editing for Speaker Generation.Masato Murata, Koichi Miyazaki, Tomoki Koriyama, Tomoki Toda
2024InterspeechVAE-based Phoneme Alignment Using Gradient Annealing and SSL Acoustic Features.Tomoki Koriyama
2024InterspeechAn Attribute Interpolation Method in Speech Synthesis by Model Merging.Masato Murata, Koichi Miyazaki, Tomoki Koriyama
2024InterspeechFrame-Wise Breath Detection with Self-Training: An Exploration of Enhancing Breath Naturalness in Text-to-Speech.Dong Yang, Tomoki Koriyama, Yuki Saito
2023ICASSPStructured State Space Decoder for Speech Recognition and Synthesis.Koichi Miyazaki, Masato Murata, Tomoki Koriyama
2023ICASSPDuration-Aware Pause Insertion Using Pre-Trained Language Model for Multi-Speaker Text-To-Speech.Dong Yang, Tomoki Koriyama, Yuki Saito, Takaaki Saeki, Detai Xin, Hiroshi Saruwatari
2022InterspeechPredicting VQVAE-based Character Acting Style from Quotation-Annotated Text for Audiobook Speech Synthesis.Wataru Nakata, Tomoki Koriyama, Shinnosuke Takamichi, Yuki Saito, Yusuke Ijima, Ryo Masumura, Hiroshi Saruwatari
2022InterspeechUTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022.Takaaki Saeki, Detai Xin, Wataru Nakata, Tomoki Koriyama, Shinnosuke Takamichi, Hiroshi Saruwatari
2021InterspeechHarmonic WaveGAN: GAN-Based Speech Waveform Generation Model with Harmonic Structure Discriminator.Kazuki Mizuta, Tomoki Koriyama, Hiroshi Saruwatari
2021InterspeechSequence-to-Sequence Learning for Deep Gaussian Process Based Speech Synthesis Using Self-Attention GP Layer.Taiki Nakamura, Tomoki Koriyama, Hiroshi Saruwatari
2021InterspeechCross-Lingual Speaker Adaptation Using Domain Adaptation and Speaker Consistency Loss for Text-To-Speech Synthesis.Detai Xin, Yuki Saito, Shinnosuke Takamichi, Tomoki Koriyama, Hiroshi Saruwatari
2020ICASSPUtterance-Level Sequential Modeling for Deep Gaussian Process Based Speech Synthesis Using Simple Recurrent Unit.Tomoki Koriyama, Hiroshi Saruwatari
2020InterspeechMulti-Speaker Text-to-Speech Synthesis Using Deep Gaussian Processes.Kentaro Mitsui, Tomoki Koriyama, Hiroshi Saruwatari
2020InterspeechCross-Lingual Text-To-Speech Synthesis via Domain Adaptation and Perceptual Similarity Regression in Speaker Space.Detai Xin, Yuki Saito, Shinnosuke Takamichi, Tomoki Koriyama, Hiroshi Saruwatari
2020InterspeechInvestigating Effective Additional Contextual Factors in DNN-Based Spontaneous Speech Synthesis.Yuki Yamashita, Tomoki Koriyama, Yuki Saito, Shinnosuke Takamichi, Yusuke Ijima, Ryo Masumura, Hiroshi Saruwatari
2020LRECDNN-based Speech Synthesis Using Abundant Tags of Spontaneous Speech Corpus.Yuki Yamashita, Tomoki Koriyama, Yuki Saito, Shinnosuke Takamichi, Yusuke Ijima, Ryo Masumura, Hiroshi Saruwatari
2019ICASSPA Training Method Using DNN-guided Layerwise Pretraining for Deep Gaussian Processes.Tomoki Koriyama, Takao Kobayashi
2019ICASSPGenerative Moment Matching Network-based Random Modulation Post-filter for DNN-based Singing Voice Synthesis and Neural Double-tracking.Hiroki Tamaru, Yuki Saito, Shinnosuke Takamichi, Tomoki Koriyama, Hiroshi Saruwatari
2019InterspeechSemi-Supervised Prosody Modeling Using Deep Gaussian Process Latent Variable Model.Tomoki Koriyama, Takao Kobayashi
2017ICASSPDuration prediction using multiple Gaussian process experts for GPR-based speech synthesis.Decha Moungsri, Tomoki Koriyama, Takao Kobayashi
2017InterspeechSampling-Based Speech Parameter Generation Using Moment-Matching Networks.Shinnosuke Takamichi, Tomoki Koriyama, Hiroshi Saruwatari
2016ICASSPA speaker adaptation technique for Gaussian process regression based speech synthesis using feature space transform.Tomoki Koriyama, Syohei Oshio, Takao Kobayashi
2016InterspeechUnsupervised Stress Information Labeling Using Gaussian Process Latent Variable Model for Statistical Speech Synthesis.Decha Moungsri, Tomoki Koriyama, Takao Kobayashi
2015ICASSPProsody generation using frame-based Gaussian process regression and classification for statistical parametric speech synthesis.Tomoki Koriyama, Takao Kobayashi
2015InterspeechA comparison of speech synthesis systems based on GPR, HMM, and DNN with a small amount of training data.Tomoki Koriyama, Takao Kobayashi
2015InterspeechDuration prediction using multi-level model for GPR-based speech synthesis.Decha Moungsri, Tomoki Koriyama, Takao Kobayashi
2014ICASSPParametric speech synthesis based on Gaussian process regression using global variance and hyperparameter optimization.Tomoki Koriyama, Takashi Nose, Takao Kobayashi
2014InterspeechAccent type and phrase boundary estimation using acoustic and language models for automatic prosodic labeling.Tomoki Koriyama, Hiroshi Suzuki, Takashi Nose, Takahiro Shinozaki, Takao Kobayashi
2014InterspeechTransform mapping using shared decision tree context clustering for HMM-based cross-lingual speech synthesis.Daiki Nagahama, Takashi Nose, Tomoki Koriyama, Takao Kobayashi
2013ICASSPFrame-level acoustic modeling based on Gaussian process regression for statistical nonparametric speech synthesis.Tomoki Koriyama, Takashi Nose, Takao Kobayashi
2013ICASSPHMM-based expressive speech synthesis based on phrase-level F0 context labeling.Yu Maeno, Takashi Nose, Takao Kobayashi, Tomoki Koriyama, Yusuke Ijima, Hideharu Nakajima, Hideyuki Mizuno, Osamu Yoshioka
2013InterspeechStatistical nonparametric speech synthesis using sparse Gaussian processes.Tomoki Koriyama, Takashi Nose, Takao Kobayashi
2013InterspeechA style control technique for singing voice synthesis based on multiple-regression HSMM.Takashi Nose, Misa Kanemoto, Tomoki Koriyama, Takao Kobayashi
2012ICASSPAn F0 modeling technique based on prosodic events for spontaneous speech synthesis.Tomoki Koriyama, Takashi Nose, Takao Kobayashi
2012InterspeechDiscontinuous Observation HMM for Prosodic-Event-Based F0 Generation.Tomoki Koriyama, Takashi Nose, Takao Kobayashi
2011InterspeechOn the Use of Extended Context for HMM-Based Spontaneous Conversational Speech Synthesis.Tomoki Koriyama, Takashi Nose, Takao Kobayashi
2010InterspeechConversational spontaneous speech synthesis using average voice model.Tomoki Koriyama, Takashi Nose, Takao Kobayashi