Skip to content

Franois G. Germain

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

27

Venues

5

Active years

2013–2025

Best venue rank

A*

Where they publish

Papers

27 indexed papers, newest first.

YearVenueTitleAuthors
2025CVPRUWAV: Uncertainty-weighted Weakly-supervised Audio-Visual Video Parsing.Yung-Hsuan Lai, Janek Ebbers, Yu-Chiang Frank Wang, Franois G. Germain, Michael J. Jones, Moitreya Chatterjee
2025ICASSPNo Class Left Behind: A Closer Look at Class Balancing for Audio Tagging.Janek Ebbers, Franois G. Germain, Kevin Wilkinghoff, Gordon Wichern, Jonathan Le Roux
2025ICASSPRetrieval-Augmented Neural Field for HRTF Upsampling and Personalization.Yoshiki Masuyama, Gordon Wichern, Franois G. Germain, Christopher Ick, Jonathan Le Roux
2025ICASSPLeveraging Audio-Only Data for Text-Queried Target Sound Extraction.Kohei Saijo, Janek Ebbers, Franois G. Germain, Sameer Khurana, Gordon Wichern, Jonathan Le Roux
2025ICASSPTask-Aware Unified Source Separation.Kohei Saijo, Janek Ebbers, Franois G. Germain, Gordon Wichern, Jonathan Le Roux
2025ICASSPKeeping the Balance: Anomaly Score Calculation for Domain Generalization.Kevin Wilkinghoff, Haici Yang, Janek Ebbers, Franois G. Germain, Gordon Wichern, Jonathan Le Roux
2025InterspeechHASRD: Hierarchical Acoustic and Semantic Representation Disentanglement.Amir Hussein, Sameer Khurana, Gordon Wichern, Franois G. Germain, Jonathan Le Roux
2025InterspeechDirection-Aware Neural Acoustic Fields for Few-Shot Interpolation of Ambisonic Impulse Responses.Christopher Ick, Gordon Wichern, Yoshiki Masuyama, Franois G. Germain, Jonathan Le Roux
2025InterspeechFactorized RVQ-GAN For Disentangled Speech Tokenization.Sameer Khurana, Dominik Klement, Antoine Laurent, Dominik Bobos, Juraj Novosad, Peter Gazdik, Ellen Zhang, Zili Huang, Amir Hussein, Ricard Marxer, Yoshiki Masuyama, Ryo Aihara, Chiori Hori, Franois G. Germain, Gordon Wichern, Jonathan Le Roux
2025InterspeechInvestigating continuous autoregressive generative speech enhancement.Haici Yang, Gordon Wichern, Ryo Aihara, Yoshiki Masuyama, Sameer Khurana, Franois G. Germain, Jonathan Le Roux
2024ICASSPGeneration or Replication: Auscultating Audio Latent Diffusion Models.Dimitrios Bralios, Gordon Wichern, Franois G. Germain, Zexu Pan, Sameer Khurana, Chiori Hori, Jonathan Le Roux
2024ICASSPWhy Does Music Source Separation Benefit from Cacophony?Chang-Bin Jeon, Gordon Wichern, Franois G. Germain, Jonathan Le Roux
2024ICASSPNIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization.Yoshiki Masuyama, Gordon Wichern, Franois G. Germain, Zexu Pan, Sameer Khurana, Chiori Hori, Jonathan Le Roux
2024ICASSPNeuroHeed+: Improving Neuro-Steered Speaker Extraction with Joint Auditory Attention Detection.Zexu Pan, Gordon Wichern, Franois G. Germain, Sameer Khurana, Jonathan Le Roux
2024ICASSPLate Audio-Visual Fusion for in-the-Wild Speaker Diarization.Zexu Pan, Gordon Wichern, Franois G. Germain, Aswin Shanmugam Subramanian, Jonathan Le Roux
2024ICASSPImproving Audio Captioning Models with Fine-Grained Audio Features, Text Embedding Supervision, and LLM Mix-Up Augmentation.Shih-Lun Wu, Xuankai Chang, Gordon Wichern, Jee-Weon Jung, Franois G. Germain, Jonathan Le Roux, Shinji Watanabe
2024InterspeechSound Event Bounding Boxes.Janek Ebbers, Franois G. Germain, Gordon Wichern, Jonathan Le Roux
2024InterspeechPARIS: Pseudo-AutoRegressIve Siamese Training for Online Speech Separation.Zexu Pan, Gordon Wichern, Franois G. Germain, Kohei Saijo, Jonathan Le Roux
2024InterspeechEnhanced Reverberation as Supervision for Unsupervised Speech Separation.Kohei Saijo, Gordon Wichern, Franois G. Germain, Zexu Pan, Jonathan Le Roux
2023ASRUScenario-Aware Audio-Visual TF-Gridnet for Target Speech Extraction.Zexu Pan, Gordon Wichern, Yoshiki Masuyama, Franois G. Germain, Sameer Khurana, Chiori Hori, Jonathan Le Roux
2023ICASSPPaᗧ-HuBERT: Self-Supervised Music Source Separation Via Primitive Auditory Clustering And Hidden-Unit Bert.Ke Chen, Gordon Wichern, Franois G. Germain, Jonathan Le Roux
2023ICASSPCold Diffusion for Speech Enhancement.Hao Yen, Franois G. Germain, Gordon Wichern, Jonathan Le Roux
2021DAFXPractical Virtual Analog Modeling Using Mbius Transforms.Franois G. Germain
2019InterspeechSpeech Denoising with Deep Feature Losses.Franois G. Germain, Qifeng Chen, Vladlen Koltun
2016ICASSPEqualization matching of speech recordings in real-world environments.Franois G. Germain, Gautham J. Mysore, Takako Fujioka
2015ICASSPSpeaker and noise independent online single-channel speech enhancement.Franois G. Germain, Gautham J. Mysore
2013InterspeechSpeaker and noise independent voice activity detection.Franois G. Germain, Dennis L. Sun, Gautham J. Mysore