Skip to content

Gordon Wichern

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

51

Venues

7

Active years

2006–2025

Best venue rank

A*

Where they publish

Papers

51 indexed papers, newest first.

YearVenueTitleAuthors
2025ICASSP30+ Years of Source Separation Research: Achievements and Future Challenges.Shoko Araki, Nobutaka Ito, Reinhold Haeb-Umbach, Gordon Wichern, Zhong-Qiu Wang, Yuki Mitsufuji
2025ICASSPNo Class Left Behind: A Closer Look at Class Balancing for Audio Tagging.Janek Ebbers, Franois G. Germain, Kevin Wilkinghoff, Gordon Wichern, Jonathan Le Roux
2025ICASSPRetrieval-Augmented Neural Field for HRTF Upsampling and Personalization.Yoshiki Masuyama, Gordon Wichern, Franois G. Germain, Christopher Ick, Jonathan Le Roux
2025ICASSPLeveraging Audio-Only Data for Text-Queried Target Sound Extraction.Kohei Saijo, Janek Ebbers, Franois G. Germain, Sameer Khurana, Gordon Wichern, Jonathan Le Roux
2025ICASSPTask-Aware Unified Source Separation.Kohei Saijo, Janek Ebbers, Franois G. Germain, Gordon Wichern, Jonathan Le Roux
2025ICASSPKeeping the Balance: Anomaly Score Calculation for Domain Generalization.Kevin Wilkinghoff, Haici Yang, Janek Ebbers, Franois G. Germain, Gordon Wichern, Jonathan Le Roux
2025InterspeechHASRD: Hierarchical Acoustic and Semantic Representation Disentanglement.Amir Hussein, Sameer Khurana, Gordon Wichern, Franois G. Germain, Jonathan Le Roux
2025InterspeechDirection-Aware Neural Acoustic Fields for Few-Shot Interpolation of Ambisonic Impulse Responses.Christopher Ick, Gordon Wichern, Yoshiki Masuyama, Franois G. Germain, Jonathan Le Roux
2025InterspeechFactorized RVQ-GAN For Disentangled Speech Tokenization.Sameer Khurana, Dominik Klement, Antoine Laurent, Dominik Bobos, Juraj Novosad, Peter Gazdik, Ellen Zhang, Zili Huang, Amir Hussein, Ricard Marxer, Yoshiki Masuyama, Ryo Aihara, Chiori Hori, Franois G. Germain, Gordon Wichern, Jonathan Le Roux
2025InterspeechInvestigating continuous autoregressive generative speech enhancement.Haici Yang, Gordon Wichern, Ryo Aihara, Yoshiki Masuyama, Sameer Khurana, Franois G. Germain, Jonathan Le Roux
2024ICASSPGeneration or Replication: Auscultating Audio Latent Diffusion Models.Dimitrios Bralios, Gordon Wichern, Franois G. Germain, Zexu Pan, Sameer Khurana, Chiori Hori, Jonathan Le Roux
2024ICASSPWhy Does Music Source Separation Benefit from Cacophony?Chang-Bin Jeon, Gordon Wichern, Franois G. Germain, Jonathan Le Roux
2024ICASSPNIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization.Yoshiki Masuyama, Gordon Wichern, Franois G. Germain, Zexu Pan, Sameer Khurana, Chiori Hori, Jonathan Le Roux
2024ICASSPNeuroHeed+: Improving Neuro-Steered Speaker Extraction with Joint Auditory Attention Detection.Zexu Pan, Gordon Wichern, Franois G. Germain, Sameer Khurana, Jonathan Le Roux
2024ICASSPLate Audio-Visual Fusion for in-the-Wild Speaker Diarization.Zexu Pan, Gordon Wichern, Franois G. Germain, Aswin Shanmugam Subramanian, Jonathan Le Roux
2024ICASSPImproving Audio Captioning Models with Fine-Grained Audio Features, Text Embedding Supervision, and LLM Mix-Up Augmentation.Shih-Lun Wu, Xuankai Chang, Gordon Wichern, Jee-Weon Jung, Franois G. Germain, Jonathan Le Roux, Shinji Watanabe
2024ICMLDeep Neural Room Acoustics Primitive.Yuhang He, Anoop Cherian, Gordon Wichern, Andrew Markham
2024InterspeechSound Event Bounding Boxes.Janek Ebbers, Franois G. Germain, Gordon Wichern, Jonathan Le Roux
2024InterspeechZeroST: Zero-Shot Speech Translation.Sameer Khurana, Chiori Hori, Antoine Laurent, Gordon Wichern, Jonathan Le Roux
2024InterspeechPARIS: Pseudo-AutoRegressIve Siamese Training for Online Speech Separation.Zexu Pan, Gordon Wichern, Franois G. Germain, Kohei Saijo, Jonathan Le Roux
2024InterspeechEnhanced Reverberation as Supervision for Unsupervised Speech Separation.Kohei Saijo, Gordon Wichern, Franois G. Germain, Zexu Pan, Jonathan Le Roux
2023ASRUScenario-Aware Audio-Visual TF-Gridnet for Target Speech Extraction.Zexu Pan, Gordon Wichern, Yoshiki Masuyama, Franois G. Germain, Sameer Khurana, Chiori Hori, Jonathan Le Roux
2023ICASSPReverberation as Supervision For Speech Separation.Rohith Aralikatti, Christoph Bddeker, Gordon Wichern, Aswin Shanmugam Subramanian, Jonathan Le Roux
2023ICASSPLatent Iterative Refinement for Modular Source Separation.Dimitrios Bralios, Efthymios Tzinis, Gordon Wichern, Paris Smaragdis, Jonathan Le Roux
2023ICASSPPaᗧ-HuBERT: Self-Supervised Music Source Separation Via Primitive Auditory Clustering And Hidden-Unit Bert.Ke Chen, Gordon Wichern, Franois G. Germain, Jonathan Le Roux
2023ICASSPHyperbolic Audio Source Separation.Darius Petermann, Gordon Wichern, Aswin Shanmugam Subramanian, Jonathan Le Roux
2023ICASSPOptimal Condition Training for Target Source Separation.Efthymios Tzinis, Gordon Wichern, Paris Smaragdis, Jonathan Le Roux
2023ICASSPCold Diffusion for Speech Enhancement.Hao Yen, Franois G. Germain, Gordon Wichern, Jonathan Le Roux
2022ICASSPThe Cocktail Fork Problem: Three-Stem Audio Separation for Real-World Soundtracks.Darius Petermann, Gordon Wichern, Zhong-Qiu Wang, Jonathan Le Roux
2022ICASSPLocate This, Not that: Class-Conditioned Sound Event DOA Estimation.Olga Slizovskaia, Gordon Wichern, Zhong-Qiu Wang, Jonathan Le Roux
2022InterspeechHeterogeneous Target Speech Separation.Efthymios Tzinis, Gordon Wichern, Aswin Shanmugam Subramanian, Paris Smaragdis, Jonathan Le Roux
2021ICASSPTranscription Is All You Need: Learning To Separate Musical Mixtures With Score As Supervision.Yun-Ning Hung, Gordon Wichern, Jonathan Le Roux
2020ICASSPWHAMR!: Noisy and Reverberant Single-Channel Speech Separation.Matthew Maciejewski, Gordon Wichern, Emmett McQuinn, Jonathan Le Roux
2020ICASSPLearning to Separate Sounds from Weakly Labeled Scenes.Fatemeh Pishdadian, Gordon Wichern, Jonathan Le Roux
2020InterspeechAll-in-One Transformer: Unifying Speech Recognition, Audio Tagging, and Event Detection.Niko Moritz, Gordon Wichern, Takaaki Hori, Jonathan Le Roux
2019ICASSPTeacher-student Deep Clustering for Low-delay Single Channel Speech Separation.Ryo Aihara, Toshiyuki Hanazawa, Yohei Okato, Gordon Wichern, Jonathan Le Roux
2019ICASSPEnd-to-end Audio Visual Scene-aware Dialog Using Multimodal Attention-based Video Features.Chiori Hori, Huda AlAmri, Jue Wang, Gordon Wichern, Takaaki Hori, Anoop Cherian, Tim K. Marks, Vincent Cartillier, Raphael Gontijo Lopes, Abhishek Das, Irfan Essa, Dhruv Batra, Devi Parikh
2019ICASSPThe Phasebook: Building Complex Masks via Discrete Representations for Source Separation.Jonathan Le Roux, Gordon Wichern, Shinji Watanabe, Andy M. Sarroff, John R. Hershey
2019ICASSPBootstrapping Single-channel Source Separation via Unsupervised Spatial Clustering on Stereo Mixtures.Prem Seetharaman, Gordon Wichern, Jonathan Le Roux, Bryan Pardo
2019ICASSPClass-conditional Embeddings for Music Source Separation.Prem Seetharaman, Gordon Wichern, Shrikant Venkataramani, Jonathan Le Roux
2019InterspeechWHAM!: Extending Speech Separation to Noisy Environments.Gordon Wichern, Joe Antognini, Michael Flynn, Licheng Richard Zhu, Emmett McQuinn, Dwight Crow, Ethan Manilow, Jonathan Le Roux
2018CVPRMultimodal Attention for Fusion of Audio and Spatiotemporal Features for Video Description.Chiori Hori, Takaaki Hori, Gordon Wichern, Jue Wang, Teng-Yok Lee, Anoop Cherian, Tim K. Marks
2010ICASSPCombining semantic, social, and acoustic similarity for retrieval of environmental sounds.Brandon Mechtley, Gordon Wichern, Harvey D. Thornburg, Andreas Spanias
2010ICASSPAutomatic audio tagging using covariate shift adaptation.Gordon Wichern, Makoto Yamada, Harvey D. Thornburg, Masashi Sugiyama, Andreas Spanias
2010ICASSPDirect importance estimation with probabilistic principal component analyzers.Makoto Yamada, Masashi Sugiyama, Gordon Wichern
2010ICASSPAcceleration of sequence kernel computation for real-time speaker identification.Makoto Yamada, Masashi Sugiyama, Gordon Wichern, Tomoko Matsui
2009ICASSPMulti-channel audio segmentation for continuous observation and archival of large spaces.Gordon Wichern, Harvey D. Thornburg, Andreas Spanias
2008ICASSPFast query by example of environmental sounds via robust and efficient cluster-based indexing.Jiachen Xue, Gordon Wichern, Harvey D. Thornburg, Andreas Spanias
2007CBMIRobust Multi-Features Segmentation and Indexing for Natural Sound Environments.Gordon Wichern, Harvey D. Thornburg, Brandon Mechtley, Alex Fink, Kai Tu, Andreas Spanias
2007IJCNNAn Operationally Adaptive System for Rapid Acoustic Transmission Loss Prediction.Michael McCarron, Mahmood R. Azimi-Sadjadi, Gordon Wichern, Michael Mungiole
2006IJCNNAn Environmentally Adaptive System for Rapid Acoustic Transmission Loss Prediction.Gordon Wichern, Mahmood R. Azimi-Sadjadi, Michael Mungiole