Vincent Wan
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
30
Venues
5
Active years
2002–2022
Best venue rank
A*
Where they publish
Papers
30 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2022 | Interspeech | Training Text-To-Speech Systems From Synthetic Data: A Practical Approach For Accent Transfer Tasks. | Lev Finkelstein, Heiga Zen, Norman Casagrande, Chun-an Chan, Ye Jia, Tom Kenter, Alexey Petelin, Jonathan Shen, Vincent Wan, Yu Zhang, Yonghui Wu, Rob Clark |
| 2019 | ICML | CHiVE: Varying Prosody in Speech Synthesis with a Linguistically Driven Dynamic Hierarchical Conditional Variational Network. | Tom Kenter, Vincent Wan, Chun-an Chan, Rob Clark, Jakub Vit |
| 2017 | Interspeech | Google's Next-Generation Real-Time Unit-Selection Synthesizer Using Sequence-to-Sequence LSTM-Based Autoencoders. | Vincent Wan, Yannis Agiomyrgiannakis, Hanna Siln, Jakub Vt |
| 2014 | ICASSP | Cluster adaptive training of average voice models. | Vincent Wan, Javier Latorre, Kayoko Yanagisawa, Mark J. F. Gales, Yannis Stylianou |
| 2014 | Interspeech | An initial investigation of long-term adaptation for meeting transcription. | Xie Chen, Mark J. F. Gales, Kate M. Knill, Catherine Breslin, Langzhou Chen, K. K. Chin, Vincent Wan |
| 2014 | Interspeech | Generating multiple-accent pronunciations for TTS using joint sequence model interpolation. | BalaKrishna Kolluru, Vincent Wan, Javier Latorre, Kayoko Yanagisawa, Mark J. F. Gales |
| 2014 | Interspeech | Voice expression conversion with factorised HMM-TTS models. | Javier Latorre, Vincent Wan, Kayoko Yanagisawa |
| 2014 | Interspeech | Speech intonation for TTS: study on evaluation methodology. | Javier Latorre, Kayoko Yanagisawa, Vincent Wan, BalaKrishna Kolluru, Mark J. F. Gales |
| 2013 | CVPR | Expressive Visual Text-to-Speech Using Active Appearance Models. | Robert Anderson, Bjrn Stenger, Vincent Wan, Roberto Cipolla |
| 2013 | Interspeech | Photo-realistic expressive text to talking head synthesis. | Vincent Wan, Robert Anderson, Art Blokland, Norbert Braunschweiler, Langzhou Chen, BalaKrishna Kolluru, Javier Latorre, Ranniery Maia, Bjrn Stenger, Kayoko Yanagisawa, Yannis Stylianou, Masami Akamine, Mark J. F. Gales, Roberto Cipolla |
| 2013 | SIGGRAPH | An expressive text-driven 3D talking head. | Robert Anderson, Bjrn Stenger, Vincent Wan, Roberto Cipolla |
| 2012 | ICASSP | Unsupervised clustering of emotion and voice styles for expressive TTS. | Florian Eyben, Sabine Buchholz, Norbert Braunschweiler, Javier Latorre, Vincent Wan, Mark J. F. Gales, Kate M. Knill |
| 2012 | Interspeech | Exploring Rich Expressive Information from Audiobook Data Using Cluster Adaptive Training. | Langzhou Chen, Mark J. F. Gales, Vincent Wan, Javier Latorre, Masami Akamine |
| 2012 | Interspeech | Speech factorization for HMM-TTS based on cluster adaptive training. | Javier Latorre, Vincent Wan, Mark J. F. Gales, Langzhou Chen, K. K. Chin, Kate M. Knill, Masami Akamine |
| 2012 | Interspeech | Combining multiple high quality corpora for improving HMM-TTS. | Vincent Wan, Javier Latorre, K. K. Chin, Langzhou Chen, Mark J. F. Gales, Heiga Zen, Kate M. Knill, Masami Akamine |
| 2011 | Interspeech | Extending Audio Notetaker to Browse WebASR Transcriptions. | Roger C. F. Tucker, Dan Fry, Vincent Wan, Stuart N. Wrigley, Thomas Hain |
| 2010 | Interspeech | The AMIDA 2009 meeting transcription system. | Thomas Hain, Luks Burget, John Dines, Philip N. Garner, Asmaa El Hannani, Marijn Huijbregts, Martin Karafit, Mike Lincoln, Vincent Wan |
| 2009 | Interspeech | Real-time ASR from meetings. | Philip N. Garner, John Dines, Thomas Hain, Asmaa El Hannani, Martin Karafit, Danil Korchagin, Mike Lincoln, Vincent Wan, Le Zhang |
| 2008 | Interspeech | Combining neural network and rule-based systems for dysarthria diagnosis. | James Carmichael, Vincent Wan, Phil D. Green |
| 2008 | Interspeech | Automatic speech recognition for scientific purposes - webASR. | Thomas Hain, Asmaa El Hannani, Stuart N. Wrigley, Vincent Wan |
| 2007 | ICASSP | The AMI System for the Transcription of Speech in Meetings. | Thomas Hain, Vincent Wan, Luks Burget, Martin Karafit, John Dines, Jithendra Vepa, Giulia Garau, Mike Lincoln |
| 2007 | ICASSP | Finding Maximum Margin Segments in Speech. | Yago Pereiro-Estevan, Vincent Wan, Odette Scharenborg |
| 2007 | Interspeech | Segmentation of speech: child's play? | Odette Scharenborg, Mirjam Ernestus, Vincent Wan |
| 2007 | Interspeech | Can unquantised articulatory feature continuums be modelled? | Odette Scharenborg, Vincent Wan |
| 2006 | ICASSP | Strategies for Language Model Web-Data Collection. | Vincent Wan, Thomas Hain |
| 2005 | Interspeech | Transcription of conference room meetings: an investigation. | Thomas Hain, John Dines, Giulia Garau, Martin Karafit, Darren Moore, Vincent Wan, Roeland Ordelman, Steve Renals |
| 2005 | Interspeech | Polynomial dynamic time warping kernel support vector machines for dysarthric speech recognition with sparse training data. | Vincent Wan, James Carmichael |
| 2003 | ICASSP | SVMSVM: support vector machine speaker verification methodology. | Vincent Wan, Steve Renals |
| 2003 | Interspeech | Feature selection for the classification of crosstalk in multi-channel audio. | Stuart N. Wrigley, Guy J. Brown, Vincent Wan, Steve Renals |
| 2002 | ICASSP | Evaluation of kernel methods for speaker verification and identification. | Vincent Wan, Steve Renals |