Gordon Wichern
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
51
Venues
7
Active years
2006–2025
Best venue rank
A*
Where they publish
Papers
51 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICASSP | 30+ Years of Source Separation Research: Achievements and Future Challenges. | Shoko Araki, Nobutaka Ito, Reinhold Haeb-Umbach, Gordon Wichern, Zhong-Qiu Wang, Yuki Mitsufuji |
| 2025 | ICASSP | No Class Left Behind: A Closer Look at Class Balancing for Audio Tagging. | Janek Ebbers, Franois G. Germain, Kevin Wilkinghoff, Gordon Wichern, Jonathan Le Roux |
| 2025 | ICASSP | Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization. | Yoshiki Masuyama, Gordon Wichern, Franois G. Germain, Christopher Ick, Jonathan Le Roux |
| 2025 | ICASSP | Leveraging Audio-Only Data for Text-Queried Target Sound Extraction. | Kohei Saijo, Janek Ebbers, Franois G. Germain, Sameer Khurana, Gordon Wichern, Jonathan Le Roux |
| 2025 | ICASSP | Task-Aware Unified Source Separation. | Kohei Saijo, Janek Ebbers, Franois G. Germain, Gordon Wichern, Jonathan Le Roux |
| 2025 | ICASSP | Keeping the Balance: Anomaly Score Calculation for Domain Generalization. | Kevin Wilkinghoff, Haici Yang, Janek Ebbers, Franois G. Germain, Gordon Wichern, Jonathan Le Roux |
| 2025 | Interspeech | HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement. | Amir Hussein, Sameer Khurana, Gordon Wichern, Franois G. Germain, Jonathan Le Roux |
| 2025 | Interspeech | Direction-Aware Neural Acoustic Fields for Few-Shot Interpolation of Ambisonic Impulse Responses. | Christopher Ick, Gordon Wichern, Yoshiki Masuyama, Franois G. Germain, Jonathan Le Roux |
| 2025 | Interspeech | Factorized RVQ-GAN For Disentangled Speech Tokenization. | Sameer Khurana, Dominik Klement, Antoine Laurent, Dominik Bobos, Juraj Novosad, Peter Gazdik, Ellen Zhang, Zili Huang, Amir Hussein, Ricard Marxer, Yoshiki Masuyama, Ryo Aihara, Chiori Hori, Franois G. Germain, Gordon Wichern, Jonathan Le Roux |
| 2025 | Interspeech | Investigating continuous autoregressive generative speech enhancement. | Haici Yang, Gordon Wichern, Ryo Aihara, Yoshiki Masuyama, Sameer Khurana, Franois G. Germain, Jonathan Le Roux |
| 2024 | ICASSP | Generation or Replication: Auscultating Audio Latent Diffusion Models. | Dimitrios Bralios, Gordon Wichern, Franois G. Germain, Zexu Pan, Sameer Khurana, Chiori Hori, Jonathan Le Roux |
| 2024 | ICASSP | Why Does Music Source Separation Benefit from Cacophony? | Chang-Bin Jeon, Gordon Wichern, Franois G. Germain, Jonathan Le Roux |
| 2024 | ICASSP | NIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization. | Yoshiki Masuyama, Gordon Wichern, Franois G. Germain, Zexu Pan, Sameer Khurana, Chiori Hori, Jonathan Le Roux |
| 2024 | ICASSP | NeuroHeed+: Improving Neuro-Steered Speaker Extraction with Joint Auditory Attention Detection. | Zexu Pan, Gordon Wichern, Franois G. Germain, Sameer Khurana, Jonathan Le Roux |
| 2024 | ICASSP | Late Audio-Visual Fusion for in-the-Wild Speaker Diarization. | Zexu Pan, Gordon Wichern, Franois G. Germain, Aswin Shanmugam Subramanian, Jonathan Le Roux |
| 2024 | ICASSP | Improving Audio Captioning Models with Fine-Grained Audio Features, Text Embedding Supervision, and LLM Mix-Up Augmentation. | Shih-Lun Wu, Xuankai Chang, Gordon Wichern, Jee-Weon Jung, Franois G. Germain, Jonathan Le Roux, Shinji Watanabe |
| 2024 | ICML | Deep Neural Room Acoustics Primitive. | Yuhang He, Anoop Cherian, Gordon Wichern, Andrew Markham |
| 2024 | Interspeech | Sound Event Bounding Boxes. | Janek Ebbers, Franois G. Germain, Gordon Wichern, Jonathan Le Roux |
| 2024 | Interspeech | ZeroST: Zero-Shot Speech Translation. | Sameer Khurana, Chiori Hori, Antoine Laurent, Gordon Wichern, Jonathan Le Roux |
| 2024 | Interspeech | PARIS: Pseudo-AutoRegressIve Siamese Training for Online Speech Separation. | Zexu Pan, Gordon Wichern, Franois G. Germain, Kohei Saijo, Jonathan Le Roux |
| 2024 | Interspeech | Enhanced Reverberation as Supervision for Unsupervised Speech Separation. | Kohei Saijo, Gordon Wichern, Franois G. Germain, Zexu Pan, Jonathan Le Roux |
| 2023 | ASRU | Scenario-Aware Audio-Visual TF-Gridnet for Target Speech Extraction. | Zexu Pan, Gordon Wichern, Yoshiki Masuyama, Franois G. Germain, Sameer Khurana, Chiori Hori, Jonathan Le Roux |
| 2023 | ICASSP | Reverberation as Supervision For Speech Separation. | Rohith Aralikatti, Christoph Bddeker, Gordon Wichern, Aswin Shanmugam Subramanian, Jonathan Le Roux |
| 2023 | ICASSP | Latent Iterative Refinement for Modular Source Separation. | Dimitrios Bralios, Efthymios Tzinis, Gordon Wichern, Paris Smaragdis, Jonathan Le Roux |
| 2023 | ICASSP | Paᗧ-HuBERT: Self-Supervised Music Source Separation Via Primitive Auditory Clustering And Hidden-Unit Bert. | Ke Chen, Gordon Wichern, Franois G. Germain, Jonathan Le Roux |
| 2023 | ICASSP | Hyperbolic Audio Source Separation. | Darius Petermann, Gordon Wichern, Aswin Shanmugam Subramanian, Jonathan Le Roux |
| 2023 | ICASSP | Optimal Condition Training for Target Source Separation. | Efthymios Tzinis, Gordon Wichern, Paris Smaragdis, Jonathan Le Roux |
| 2023 | ICASSP | Cold Diffusion for Speech Enhancement. | Hao Yen, Franois G. Germain, Gordon Wichern, Jonathan Le Roux |
| 2022 | ICASSP | The Cocktail Fork Problem: Three-Stem Audio Separation for Real-World Soundtracks. | Darius Petermann, Gordon Wichern, Zhong-Qiu Wang, Jonathan Le Roux |
| 2022 | ICASSP | Locate This, Not that: Class-Conditioned Sound Event DOA Estimation. | Olga Slizovskaia, Gordon Wichern, Zhong-Qiu Wang, Jonathan Le Roux |
| 2022 | Interspeech | Heterogeneous Target Speech Separation. | Efthymios Tzinis, Gordon Wichern, Aswin Shanmugam Subramanian, Paris Smaragdis, Jonathan Le Roux |
| 2021 | ICASSP | Transcription Is All You Need: Learning To Separate Musical Mixtures With Score As Supervision. | Yun-Ning Hung, Gordon Wichern, Jonathan Le Roux |
| 2020 | ICASSP | WHAMR!: Noisy and Reverberant Single-Channel Speech Separation. | Matthew Maciejewski, Gordon Wichern, Emmett McQuinn, Jonathan Le Roux |
| 2020 | ICASSP | Learning to Separate Sounds from Weakly Labeled Scenes. | Fatemeh Pishdadian, Gordon Wichern, Jonathan Le Roux |
| 2020 | Interspeech | All-in-One Transformer: Unifying Speech Recognition, Audio Tagging, and Event Detection. | Niko Moritz, Gordon Wichern, Takaaki Hori, Jonathan Le Roux |
| 2019 | ICASSP | Teacher-student Deep Clustering for Low-delay Single Channel Speech Separation. | Ryo Aihara, Toshiyuki Hanazawa, Yohei Okato, Gordon Wichern, Jonathan Le Roux |
| 2019 | ICASSP | End-to-end Audio Visual Scene-aware Dialog Using Multimodal Attention-based Video Features. | Chiori Hori, Huda AlAmri, Jue Wang, Gordon Wichern, Takaaki Hori, Anoop Cherian, Tim K. Marks, Vincent Cartillier, Raphael Gontijo Lopes, Abhishek Das, Irfan Essa, Dhruv Batra, Devi Parikh |
| 2019 | ICASSP | The Phasebook: Building Complex Masks via Discrete Representations for Source Separation. | Jonathan Le Roux, Gordon Wichern, Shinji Watanabe, Andy M. Sarroff, John R. Hershey |
| 2019 | ICASSP | Bootstrapping Single-channel Source Separation via Unsupervised Spatial Clustering on Stereo Mixtures. | Prem Seetharaman, Gordon Wichern, Jonathan Le Roux, Bryan Pardo |
| 2019 | ICASSP | Class-conditional Embeddings for Music Source Separation. | Prem Seetharaman, Gordon Wichern, Shrikant Venkataramani, Jonathan Le Roux |
| 2019 | Interspeech | WHAM!: Extending Speech Separation to Noisy Environments. | Gordon Wichern, Joe Antognini, Michael Flynn, Licheng Richard Zhu, Emmett McQuinn, Dwight Crow, Ethan Manilow, Jonathan Le Roux |
| 2018 | CVPR | Multimodal Attention for Fusion of Audio and Spatiotemporal Features for Video Description. | Chiori Hori, Takaaki Hori, Gordon Wichern, Jue Wang, Teng-Yok Lee, Anoop Cherian, Tim K. Marks |
| 2010 | ICASSP | Combining semantic, social, and acoustic similarity for retrieval of environmental sounds. | Brandon Mechtley, Gordon Wichern, Harvey D. Thornburg, Andreas Spanias |
| 2010 | ICASSP | Automatic audio tagging using covariate shift adaptation. | Gordon Wichern, Makoto Yamada, Harvey D. Thornburg, Masashi Sugiyama, Andreas Spanias |
| 2010 | ICASSP | Direct importance estimation with probabilistic principal component analyzers. | Makoto Yamada, Masashi Sugiyama, Gordon Wichern |
| 2010 | ICASSP | Acceleration of sequence kernel computation for real-time speaker identification. | Makoto Yamada, Masashi Sugiyama, Gordon Wichern, Tomoko Matsui |
| 2009 | ICASSP | Multi-channel audio segmentation for continuous observation and archival of large spaces. | Gordon Wichern, Harvey D. Thornburg, Andreas Spanias |
| 2008 | ICASSP | Fast query by example of environmental sounds via robust and efficient cluster-based indexing. | Jiachen Xue, Gordon Wichern, Harvey D. Thornburg, Andreas Spanias |
| 2007 | CBMI | Robust Multi-Features Segmentation and Indexing for Natural Sound Environments. | Gordon Wichern, Harvey D. Thornburg, Brandon Mechtley, Alex Fink, Kai Tu, Andreas Spanias |
| 2007 | IJCNN | An Operationally Adaptive System for Rapid Acoustic Transmission Loss Prediction. | Michael McCarron, Mahmood R. Azimi-Sadjadi, Gordon Wichern, Michael Mungiole |
| 2006 | IJCNN | An Environmentally Adaptive System for Rapid Acoustic Transmission Loss Prediction. | Gordon Wichern, Mahmood R. Azimi-Sadjadi, Michael Mungiole |