| 2025 | ASRU | The T12 System for AudioMOS Challenge 2025: Audio Aesthetics Score Prediction System Using KAN- and VERSA-based Models. | Katsuhiko Yamamoto, Koichi Miyazaki, Shogo Seki |
| 2025 | Interspeech | First Analyze Then Enhance: A Task-Aware System for Speech Separation, Denoising, and Dereverberation. | Shaoxiang Dang, Li Li, Shogo Seki, Hiroaki Kudo |
| 2024 | ICASSP | Remixed2remixed: Domain Adaptation for Speech Enhancement by Noise2noise Learning with Remixing. | Li Li, Shogo Seki |
| 2024 | Interspeech | Improved Remixing Process for Domain Adaptation-Based Speech Enhancement by Mitigating Data Imbalance in Signal-to-Noise Ratio. | Li Li, Shogo Seki |
| 2023 | ICASSP | Wave-U-Net Discriminator: Fast and Lightweight Discriminator for Generative Adversarial Network-Based Speech Synthesis. | Takuhiro Kaneko, Hirokazu Kameoka, Kou Tanaka, Shogo Seki |
| 2023 | ICASSP | JSV-VC: Jointly Trained Speaker Verification and Voice Conversion Models. | Shogo Seki, Hirokazu Kameoka, Kou Tanaka, Takuhiro Kaneko |
| 2023 | Interspeech | iSTFTNet2: Faster and More Lightweight iSTFT-Based Neural Vocoder Using 1D-2D CNN. | Takuhiro Kaneko, Hirokazu Kameoka, Kou Tanaka, Shogo Seki |
| 2023 | Interspeech | CFVC: Conditional Filtering for Controllable Voice Conversion. | Kou Tanaka, Takuhiro Kaneko, Hirokazu Kameoka, Shogo Seki |
| 2022 | ICASSP | Attentionpit: Soft Permutation Invariant Training for Audio Source Separation with Attention Mechanism. | Hirokazu Kameoka, Shogo Seki, Li Li, Chihiro Watanabe |
| 2022 | ICASSP | ISTFTNET: Fast and Lightweight Mel-Spectrogram Vocoder Incorporating Inverse Short-Time Fourier Transform. | Takuhiro Kaneko, Kou Tanaka, Hirokazu Kameoka, Shogo Seki |
| 2022 | ICASSP | HBP: An Efficient Block Permutation Solver Using Hungarian Algorithm and Spectrogram Inpainting for Multichannel Audio Source Separation. | Li Li, Hirokazu Kameoka, Shogo Seki |
| 2022 | ICASSP | Investigation And Comparison of Optimization Methods for Variational Autoencoder-Based Underdetermined Multichannel Source Separation. | Shogo Seki, Hirokazu Kameoka, Li Li |
| 2022 | Interspeech | CAUSE: Crossmodal Action Unit Sequence Estimation from Speech. | Hirokazu Kameoka, Takuhiro Kaneko, Shogo Seki, Kou Tanaka |
| 2022 | Interspeech | MISRNet: Lightweight Neural Vocoder Using Multi-Input Single Shared Residual Blocks. | Takuhiro Kaneko, Hirokazu Kameoka, Kou Tanaka, Shogo Seki |
| 2020 | Interspeech | Intelligibility Enhancement Based on Speech Waveform Modification Using Hearing Impairment. | Shu Hikosaka, Shogo Seki, Tomoki Hayashi, Kazuhiro Kobayashi, Kazuya Takeda, Hideki Banno, Tomoki Toda |
| 2020 | Interspeech | Semi-Supervised Self-Produced Speech Enhancement and Suppression Based on Joint Source Modeling of Air- and Body-Conducted Signals Using Variational Autoencoder. | Shogo Seki, Moe Takada, Tomoki Toda |
| 2019 | ICASSP | Joint Separation and Dereverberation of Reverberant Mixtures with Multichannel Variational Autoencoder. | Shota Inoue, Hirokazu Kameoka, Li Li, Shogo Seki, Shoji Makino |
| 2017 | SCA | Sketch-based 3D hair posing by contour drawings. | Shogo Seki, Takeo Igarashi |
| 2016 | Interspeech | Robust Example Search Using Bottleneck Features for Example-Based Speech Enhancement. | Atsunori Ogawa, Shogo Seki, Keisuke Kinoshita, Marc Delcroix, Takuya Yoshioka, Tomohiro Nakatani, Kazuya Takeda |