R. J. Skerry-Ryan
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
16
Venues
5
Active years
2017–2025
Best venue rank
A*
Where they publish
Papers
16 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICML | Long-Form Speech Generation with Spoken Language Models. | Se Jin Park, Julian Salazar, Aren Jansen, Keisuke Kinoshita, Yong Man Ro, R. J. Skerry-Ryan |
| 2025 | Interspeech | Zero-Shot Mono-to-Binaural Speech Synthesis. | Alon Levkovitch, Julian Salazar, Soroosh Mariooryad, R. J. Skerry-Ryan, Nadav Bar, W. Bastiaan Kleijn, Eliya Nachmani |
| 2025 | NAACL | Robust and Unbounded Length Generalization in Autoregressive Transformer-Based Text-to-Speech. | Eric Battenberg, R. J. Skerry-Ryan, Daisy Stanton, Soroosh Mariooryad, Matt Shannon, Julian Salazar, David Kao |
| 2024 | ICLR | Spoken Question Answering and Speech Continuation Using Spectrogram-Powered LLM. | Eliya Nachmani, Alon Levkovitch, Roy Hirsch, Julian Salazar, Chulayuth Asawaroengchai, Soroosh Mariooryad, Ehud Rivlin, R. J. Skerry-Ryan, Michelle Tadmor Ramanovich |
| 2022 | ICASSP | Speaker Generation. | Daisy Stanton, Matt Shannon, Soroosh Mariooryad, R. J. Skerry-Ryan, Eric Battenberg, Tom Bagby, David Kao |
| 2021 | ICASSP | Wave-Tacotron: Spectrogram-Free End-to-End Text-to-Speech Synthesis. | Ron J. Weiss, R. J. Skerry-Ryan, Eric Battenberg, Soroosh Mariooryad, Diederik P. Kingma |
| 2021 | Interspeech | Parallel Tacotron 2: A Non-Autoregressive Neural TTS Model with Differentiable Duration Modeling. | Isaac Elias, Heiga Zen, Jonathan Shen, Yu Zhang, Ye Jia, R. J. Skerry-Ryan, Yonghui Wu |
| 2020 | ICASSP | Location-Relative Attention Mechanisms for Robust Long-Form Speech Synthesis. | Eric Battenberg, R. J. Skerry-Ryan, Soroosh Mariooryad, Daisy Stanton, David Kao, Matt Shannon, Tom Bagby |
| 2020 | ICLR | Semi-Supervised Generative Modeling for Controllable Speech Synthesis. | Raza Habib, Soroosh Mariooryad, Matt Shannon, Eric Battenberg, R. J. Skerry-Ryan, Daisy Stanton, David Kao, Tom Bagby |
| 2019 | ICASSP | Semi-supervised Training for Improving Data Efficiency in End-to-end Speech Synthesis. | Yu-An Chung, Yuxuan Wang, Wei-Ning Hsu, Yu Zhang, R. J. Skerry-Ryan |
| 2019 | Interspeech | Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning. | Yu Zhang, Ron J. Weiss, Heiga Zen, Yonghui Wu, Zhifeng Chen, R. J. Skerry-Ryan, Ye Jia, Andrew Rosenberg, Bhuvana Ramabhadran |
| 2018 | ICASSP | Complex Evolution Recurrent Neural Networks (ceRNNs). | Izhak Shafran, Tom Bagby, R. J. Skerry-Ryan |
| 2018 | ICASSP | Natural TTS Synthesis by Conditioning Wavenet on MEL Spectrogram Predictions. | Jonathan Shen, Ruoming Pang, Ron J. Weiss, Mike Schuster, Navdeep Jaitly, Zongheng Yang, Zhifeng Chen, Yu Zhang, Yuxuan Wang, R. J. Skerry-Ryan, Rif A. Saurous, Yannis Agiomyrgiannakis, Yonghui Wu |
| 2018 | ICML | Towards End-to-End Prosody Transfer for Expressive Speech Synthesis with Tacotron. | R. J. Skerry-Ryan, Eric Battenberg, Ying Xiao, Yuxuan Wang, Daisy Stanton, Joel Shor, Ron J. Weiss, Rob Clark, Rif A. Saurous |
| 2018 | ICML | Style Tokens: Unsupervised Style Modeling, Control and Transfer in End-to-End Speech Synthesis. | Yuxuan Wang, Daisy Stanton, Yu Zhang, R. J. Skerry-Ryan, Eric Battenberg, Joel Shor, Ying Xiao, Ye Jia, Fei Ren, Rif A. Saurous |
| 2017 | Interspeech | Tacotron: Towards End-to-End Speech Synthesis. | Yuxuan Wang, R. J. Skerry-Ryan, Daisy Stanton, Yonghui Wu, Ron J. Weiss, Navdeep Jaitly, Zongheng Yang, Ying Xiao, Zhifeng Chen, Samy Bengio, Quoc V. Le, Yannis Agiomyrgiannakis, Rob Clark, Rif A. Saurous |