| 2025 | ASRU | The AudioMOS Challenge 2025. | Wen-Chin Huang, Hui Wang, Cheng Liu, Yi-Chiao Wu, Andros Tjandra, Wei-Ning Hsu, Erica Cooper, Yong Qin, Tomoki Toda |
| 2025 | ASRU | Meta Audiobox Aesthetics: Unified Automatic Assessment for Speech, Music and Sound. | Andros Tjandra, Yi-Chiao Wu, Baishan Guo, John Hoffman, Brian Ellis, Apoorv Vyas, Bowen Shi, Sanyuan Chen, Matt Le, Nick Zacharov, Carleigh Wood, Ann Lee, Wei-Ning Hsu |
| 2025 | ICASSP | ComplexDec: A Domain-robust High-fidelity Neural Audio Codec with Complex Spectrum Modeling. | Yi-Chiao Wu, Dejan Markovic, Steven Krenn, Israel D. Gebru, Alexander Richard |
| 2025 | ICLR | FlowDec: A flow-based full-band general audio codec with high perceptual quality. | Simon Welker, Matthew Le, Ricky T. Q. Chen, Wei-Ning Hsu, Timo Gerkmann, Alexander Richard, Yi-Chiao Wu |
| 2024 | ICASSP | ScoreDec: A Phase-Preserving High-Fidelity Audio Codec with a Generalized Score-Based Diffusion Post-Filter. | Yi-Chiao Wu, Dejan Markovic, Steven Krenn, Israel D. Gebru, Alexander Richard |
| 2024 | Interspeech | EARS: An Anechoic Fullband Speech Dataset Benchmarked for Speech Enhancement and Dereverberation. | Julius Richter, Yi-Chiao Wu, Steven Krenn, Simon Welker, Bunlong Lay, Shinji Watanabe, Alexander Richard, Timo Gerkmann |
| 2023 | ICASSP | Audiodec: An Open-Source Streaming High-Fidelity Neural Audio Codec. | Yi-Chiao Wu, Israel D. Gebru, Dejan Markovic, Alexander Richard |
| 2023 | ICASSP | Source-Filter HiFi-GAN: Fast and Pitch Controllable High-Fidelity Neural Vocoder. | Reo Yoneyama, Yi-Chiao Wu, Tomoki Toda |
| 2022 | CVPR | Contactless Blood Pressure Measurement via Remote Photoplethysmography with Synthetic Data Generation Using Generative Adversarial Network. | Bing-Fei Wu, Li-Wen Chiu, Yi-Chiao Wu, Chun-Chih Lai, Pao-Hsien Chu |
| 2022 | ICASSP | Direct Noisy Speech Modeling for Noisy-To-Noisy Voice Conversion. | Chao Xie, Yi-Chiao Wu, Patrick Lumban Tobing, Wen-Chin Huang, Tomoki Toda |
| 2022 | Interspeech | Unified Source-Filter GAN with Harmonic-plus-Noise Source Excitation Generation. | Reo Yoneyama, Yi-Chiao Wu, Tomoki Toda |
| 2021 | ASRU | HASA-Net: A Non-Intrusive Hearing-Aid Speech Assessment Network. | Hsin-Tien Chiang, Yi-Chiao Wu, Cheng Yu, Tomoki Toda, Hsin-Min Wang, Yih-Chun Hu, Yu Tsao |
| 2021 | ICASSP | Any-to-One Sequence-to-Sequence Voice Conversion Using Self-Supervised Discrete Speech Representations. | Wen-Chin Huang, Yi-Chiao Wu, Tomoki Hayashi |
| 2021 | ICASSP | Crank: An Open-Source Software for Nonparallel Voice Conversion Based on Vector-Quantized Variational Autoencoder. | Kazuhiro Kobayashi, Wen-Chin Huang, Yi-Chiao Wu, Patrick Lumban Tobing, Tomoki Hayashi, Tomoki Toda |
| 2021 | Interspeech | Relational Data Selection for Data Augmentation of Speaker-Dependent Multi-Band MelGAN Vocoder. | Yi-Chiao Wu, Cheng-Hung Hu, Hung-Shin Lee, Yu-Huai Peng, Wen-Chin Huang, Yu Tsao, Hsin-Min Wang, Tomoki Toda |
| 2021 | Interspeech | Unified Source-Filter GAN: Unified Source-Filter Network Based On Factorization of Quasi-Periodic Parallel WaveGAN. | Reo Yoneyama, Yi-Chiao Wu, Tomoki Toda |
| 2020 | ICASSP | Efficient Shallow Wavenet Vocoder Using Multiple Samples Output Based on Laplacian Distribution and Linear Prediction. | Patrick Lumban Tobing, Yi-Chiao Wu, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2020 | Interspeech | Voice Transformer Network: Sequence-to-Sequence Voice Conversion Using Transformer with Text-to-Speech Pretraining. | Wen-Chin Huang, Tomoki Hayashi, Yi-Chiao Wu, Hirokazu Kameoka, Tomoki Toda |
| 2020 | Interspeech | Cyclic Spectral Modeling for Unsupervised Unit Discovery into Voice Conversion with Excitation and Waveform Modeling. | Patrick Lumban Tobing, Tomoki Hayashi, Yi-Chiao Wu, Kazuhiro Kobayashi, Tomoki Toda |
| 2020 | Interspeech | Quasi-Periodic Parallel WaveGAN Vocoder: A Non-Autoregressive Pitch-Dependent Dilated Convolution Model for Parametric Speech Generation. | Yi-Chiao Wu, Tomoki Hayashi, Takuma Okamoto, Hisashi Kawai, Tomoki Toda |
| 2020 | Interspeech | A Cyclical Post-Filtering Approach to Mismatch Refinement of Neural Vocoder for Text-to-Speech Systems. | Yi-Chiao Wu, Patrick Lumban Tobing, Kazuki Yasuhara, Noriyuki Matsunaga, Yamato Ohtani, Tomoki Toda |
| 2020 | SMC | Masked Neural Sparse Encoder for Face Occlusion Detection. | Bing-Fei Wu, Yi-Chiao Wu |
| 2019 | ICASSP | Voice Conversion with Cyclic Recurrent Neural Network and Fine-tuned Wavenet Vocoder. | Patrick Lumban Tobing, Yi-Chiao Wu, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2019 | Interspeech | Investigation of F0 Conditioning and Fully Convolutional Networks in Variational Autoencoder Based Voice Conversion. | Wen-Chin Huang, Yi-Chiao Wu, Chen-Chou Lo, Patrick Lumban Tobing, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda, Yu Tsao, Hsin-Min Wang |
| 2019 | Interspeech | Non-Parallel Voice Conversion with Cyclic Variational Autoencoder. | Patrick Lumban Tobing, Yi-Chiao Wu, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2019 | Interspeech | Quasi-Periodic WaveNet Vocoder: A Pitch Dependent Dilated Convolution Model for Parametric Speech Generation. | Yi-Chiao Wu, Tomoki Hayashi, Patrick Lumban Tobing, Kazuhiro Kobayashi, Tomoki Toda |
| 2018 | Interspeech | Exemplar-Based Spectral Detail Compensation for Voice Conversion. | Yu-Huai Peng, Hsin-Te Hwang, Yi-Chiao Wu, Yu Tsao, Hsin-Min Wang |
| 2018 | Interspeech | Collapsed Speech Segment Detection and Suppression for WaveNet Vocoder. | Yi-Chiao Wu, Kazuhiro Kobayashi, Tomoki Hayashi, Patrick Lumban Tobing, Tomoki Toda |
| 2017 | ICASSP | A locally linear embbeding based postfiltering approach for speech enhancement. | Yi-Chiao Wu, Hsin-Te Hwang, Syu-Siang Wang, Chin-Cheng Hsu, Ying-Hui Lai, Yu Tsao, Hsin-Min Wang |
| 2017 | Interspeech | Voice Conversion from Unaligned Corpora Using Variational Autoencoding Wasserstein Generative Adversarial Networks. | Chin-Cheng Hsu, Hsin-Te Hwang, Yi-Chiao Wu, Yu Tsao, Hsin-Min Wang |
| 2017 | Interspeech | A Post-Filtering Approach Based on Locally Linear Embedding Difference Compensation for Speech Enhancement. | Yi-Chiao Wu, Hsin-Te Hwang, Syu-Siang Wang, Chin-Cheng Hsu, Yu Tsao, Hsin-Min Wang |
| 2016 | Interspeech | Locally Linear Embedding for Exemplar-Based Spectral Conversion. | Yi-Chiao Wu, Hsin-Te Hwang, Chin-Cheng Hsu, Yu Tsao, Hsin-Min Wang |