Skip to content

R. J. Skerry-Ryan

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

16

Venues

5

Active years

2017–2025

Best venue rank

A*

Where they publish

Papers

16 indexed papers, newest first.

YearVenueTitleAuthors
2025ICMLLong-Form Speech Generation with Spoken Language Models.Se Jin Park, Julian Salazar, Aren Jansen, Keisuke Kinoshita, Yong Man Ro, R. J. Skerry-Ryan
2025InterspeechZero-Shot Mono-to-Binaural Speech Synthesis.Alon Levkovitch, Julian Salazar, Soroosh Mariooryad, R. J. Skerry-Ryan, Nadav Bar, W. Bastiaan Kleijn, Eliya Nachmani
2025NAACLRobust and Unbounded Length Generalization in Autoregressive Transformer-Based Text-to-Speech.Eric Battenberg, R. J. Skerry-Ryan, Daisy Stanton, Soroosh Mariooryad, Matt Shannon, Julian Salazar, David Kao
2024ICLRSpoken Question Answering and Speech Continuation Using Spectrogram-Powered LLM.Eliya Nachmani, Alon Levkovitch, Roy Hirsch, Julian Salazar, Chulayuth Asawaroengchai, Soroosh Mariooryad, Ehud Rivlin, R. J. Skerry-Ryan, Michelle Tadmor Ramanovich
2022ICASSPSpeaker Generation.Daisy Stanton, Matt Shannon, Soroosh Mariooryad, R. J. Skerry-Ryan, Eric Battenberg, Tom Bagby, David Kao
2021ICASSPWave-Tacotron: Spectrogram-Free End-to-End Text-to-Speech Synthesis.Ron J. Weiss, R. J. Skerry-Ryan, Eric Battenberg, Soroosh Mariooryad, Diederik P. Kingma
2021InterspeechParallel Tacotron 2: A Non-Autoregressive Neural TTS Model with Differentiable Duration Modeling.Isaac Elias, Heiga Zen, Jonathan Shen, Yu Zhang, Ye Jia, R. J. Skerry-Ryan, Yonghui Wu
2020ICASSPLocation-Relative Attention Mechanisms for Robust Long-Form Speech Synthesis.Eric Battenberg, R. J. Skerry-Ryan, Soroosh Mariooryad, Daisy Stanton, David Kao, Matt Shannon, Tom Bagby
2020ICLRSemi-Supervised Generative Modeling for Controllable Speech Synthesis.Raza Habib, Soroosh Mariooryad, Matt Shannon, Eric Battenberg, R. J. Skerry-Ryan, Daisy Stanton, David Kao, Tom Bagby
2019ICASSPSemi-supervised Training for Improving Data Efficiency in End-to-end Speech Synthesis.Yu-An Chung, Yuxuan Wang, Wei-Ning Hsu, Yu Zhang, R. J. Skerry-Ryan
2019InterspeechLearning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning.Yu Zhang, Ron J. Weiss, Heiga Zen, Yonghui Wu, Zhifeng Chen, R. J. Skerry-Ryan, Ye Jia, Andrew Rosenberg, Bhuvana Ramabhadran
2018ICASSPComplex Evolution Recurrent Neural Networks (ceRNNs).Izhak Shafran, Tom Bagby, R. J. Skerry-Ryan
2018ICASSPNatural TTS Synthesis by Conditioning Wavenet on MEL Spectrogram Predictions.Jonathan Shen, Ruoming Pang, Ron J. Weiss, Mike Schuster, Navdeep Jaitly, Zongheng Yang, Zhifeng Chen, Yu Zhang, Yuxuan Wang, R. J. Skerry-Ryan, Rif A. Saurous, Yannis Agiomyrgiannakis, Yonghui Wu
2018ICMLTowards End-to-End Prosody Transfer for Expressive Speech Synthesis with Tacotron.R. J. Skerry-Ryan, Eric Battenberg, Ying Xiao, Yuxuan Wang, Daisy Stanton, Joel Shor, Ron J. Weiss, Rob Clark, Rif A. Saurous
2018ICMLStyle Tokens: Unsupervised Style Modeling, Control and Transfer in End-to-End Speech Synthesis.Yuxuan Wang, Daisy Stanton, Yu Zhang, R. J. Skerry-Ryan, Eric Battenberg, Joel Shor, Ying Xiao, Ye Jia, Fei Ren, Rif A. Saurous
2017InterspeechTacotron: Towards End-to-End Speech Synthesis.Yuxuan Wang, R. J. Skerry-Ryan, Daisy Stanton, Yonghui Wu, Ron J. Weiss, Navdeep Jaitly, Zongheng Yang, Ying Xiao, Zhifeng Chen, Samy Bengio, Quoc V. Le, Yannis Agiomyrgiannakis, Rob Clark, Rif A. Saurous