| 2024 | ICASSP | Mapache: Masked Parallel Transformer for Advanced Speech Editing and Synthesis. | Guillermo Cmbara, Patrick Lumban Tobing, Mikolaj Babianski, Ravichander Vipperla, Duo Wang, Ron Shmelkin, Giuseppe Coccia, Orazio Angelini, Arnaud Joly, Mateusz Lajszczak, Vincent Pollet |
| 2023 | Interspeech | Expressive Machine Dubbing Through Phrase-level Cross-lingual Prosody Transfer. | Jakub Swiatkowski, Duo Wang, Mikolaj Babianski, Giuseppe Coccia, Patrick Lumban Tobing, Ravichander Vipperla, Viacheslav Klimkov, Vincent Pollet |
| 2023 | Interspeech | Cross-lingual Prosody Transfer for Expressive Machine Dubbing. | Jakub Swiatkowski, Duo Wang, Mikolaj Babianski, Patrick Lumban Tobing, Ravichander Vipperla, Vincent Pollet |
| 2022 | ICASSP | Direct Noisy Speech Modeling for Noisy-To-Noisy Voice Conversion. | Chao Xie, Yi-Chiao Wu, Patrick Lumban Tobing, Wen-Chin Huang, Tomoki Toda |
| 2021 | ICASSP | Crank: An Open-Source Software for Nonparallel Voice Conversion Based on Vector-Quantized Variational Autoencoder. | Kazuhiro Kobayashi, Wen-Chin Huang, Yi-Chiao Wu, Patrick Lumban Tobing, Tomoki Hayashi, Tomoki Toda |
| 2021 | Interspeech | High-Fidelity and Low-Latency Universal Neural Vocoder Based on Multiband WaveRNN with Data-Driven Linear Prediction for Discrete Waveform Modeling. | Patrick Lumban Tobing, Tomoki Toda |
| 2020 | ICASSP | Efficient Shallow Wavenet Vocoder Using Multiple Samples Output Based on Laplacian Distribution and Linear Prediction. | Patrick Lumban Tobing, Yi-Chiao Wu, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2020 | Interspeech | Cyclic Spectral Modeling for Unsupervised Unit Discovery into Voice Conversion with Excitation and Waveform Modeling. | Patrick Lumban Tobing, Tomoki Hayashi, Yi-Chiao Wu, Kazuhiro Kobayashi, Tomoki Toda |
| 2020 | Interspeech | A Cyclical Post-Filtering Approach to Mismatch Refinement of Neural Vocoder for Text-to-Speech Systems. | Yi-Chiao Wu, Patrick Lumban Tobing, Kazuki Yasuhara, Noriyuki Matsunaga, Yamato Ohtani, Tomoki Toda |
| 2019 | ASRU | Investigation of Shallow Wavenet Vocoder with Laplacian Distribution Output. | Patrick Lumban Tobing, Tomoki Hayashi, Tomoki Toda |
| 2019 | ICASSP | Voice Conversion with Cyclic Recurrent Neural Network and Fine-tuned Wavenet Vocoder. | Patrick Lumban Tobing, Yi-Chiao Wu, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2019 | Interspeech | Investigation of F0 Conditioning and Fully Convolutional Networks in Variational Autoencoder Based Voice Conversion. | Wen-Chin Huang, Yi-Chiao Wu, Chen-Chou Lo, Patrick Lumban Tobing, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda, Yu Tsao, Hsin-Min Wang |
| 2019 | Interspeech | Non-Parallel Voice Conversion with Cyclic Variational Autoencoder. | Patrick Lumban Tobing, Yi-Chiao Wu, Tomoki Hayashi, Kazuhiro Kobayashi, Tomoki Toda |
| 2019 | Interspeech | Quasi-Periodic WaveNet Vocoder: A Pitch Dependent Dilated Convolution Model for Parametric Speech Generation. | Yi-Chiao Wu, Tomoki Hayashi, Patrick Lumban Tobing, Kazuhiro Kobayashi, Tomoki Toda |
| 2018 | Interspeech | Collapsed Speech Segment Detection and Suppression for WaveNet Vocoder. | Yi-Chiao Wu, Kazuhiro Kobayashi, Tomoki Hayashi, Patrick Lumban Tobing, Tomoki Toda |
| 2016 | Interspeech | Acoustic-to-Articulatory Inversion Mapping Based on Latent Trajectory Gaussian Mixture Model. | Patrick Lumban Tobing, Tomoki Toda, Hirokazu Kameoka, Satoshi Nakamura |
| 2015 | Interspeech | Articulatory controllable speech modification based on Gaussian mixture models with direct waveform modification using spectrum differential. | Patrick Lumban Tobing, Kazuhiro Kobayashi, Tomoki Toda, Graham Neubig, Sakriani Sakti, Satoshi Nakamura |
| 2014 | Interspeech | Articulatory controllable speech modification based on statistical feature mapping with Gaussian mixture models. | Patrick Lumban Tobing, Tomoki Toda, Graham Neubig, Sakriani Sakti, Satoshi Nakamura, Ayu Purwarianti |