Skip to content

Timo Gerkmann

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

75

Venues

11

Active years

2006–2025

Best venue rank

A*

Where they publish

Papers

75 indexed papers, newest first.

YearVenueTitleAuthors
2025ASRUEnhancing In-the-Wild Speech Emotion Conversion with Resynthesis-based Duration Modeling.Navin Raj Prabhu, Danilo de Oliveira, Nale Lehmann-Willenbrock, Timo Gerkmann
2025ICASSPMask-Weighted Spatial Likelihood Coding for Speaker-Independent Joint Localization and Mask Estimation.Jakob Kienegger, Alina Mannanova, Timo Gerkmann
2025ICASSPInvestigating Training Objectives for Generative Speech Enhancement.Julius Richter, Danilo de Oliveira, Timo Gerkmann
2025ICASSPHRTF Estimation using a Score-based Prior.Etienne Thuillier, Jean-Marie Lemercier, Eloi Moliner, Timo Gerkmann, Vesa Vlimki
2025ICLRFlowDec: A flow-based full-band general audio codec with high perceptual quality.Simon Welker, Matthew Le, Ricky T. Q. Chen, Wei-Ning Hsu, Timo Gerkmann, Alexander Richard, Yi-Chiao Wu
2025InterspeechSteering Deep Non-Linear Spatially Selective Filters for Weakly Guided Extraction of Moving Speakers in Dynamic Scenarios.Jakob Kienegger, Timo Gerkmann
2025InterspeechDiffusion Buffer: Online Diffusion-based Speech Enhancement with Sub-Second Latency.Bunlong Lay, Rostilav Makarov, Timo Gerkmann
2025InterspeechReal-Time Diffusion Buffer for Speech Enhancement On A Laptop.Bunlong Lay, Rostilav Makarov, Timo Gerkmann
2025InterspeechNon-intrusive Speech Quality Assessment with Diffusion Models Trained on Clean Speech.Danilo de Oliveira, Julius Richter, Jean-Marie Lemercier, Simon Welker, Timo Gerkmann
2024ICASSPSingle and Few-Step Diffusion for Generative Speech Enhancement.Bunlong Lay, Jean-Marie Lemercier, Julius Richter, Timo Gerkmann
2024ICASSPDistilling Hubert with LSTMs via Decoupled Knowledge Distillation.Danilo de Oliveira, Timo Gerkmann
2024ICASSPA Flexible Online Framework for Projection-Based Stft Phase Retrieval.Tal Peer, Simon Welker, Johannes Kolhoff, Timo Gerkmann
2024ICASSPEMOCONV-Diff: Diffusion-Based Speech Emotion Conversion for Non-Parallel and in-the-Wild Data.Navin Raj Prabhu, Bunlong Lay, Simon Welker, Nale Lehmann-Willenbrock, Timo Gerkmann
2024ICASSPLive Iterative Ptychography with Projection-Based Algorithms.Simon Welker, Tal Peer, Henry N. Chapman, Timo Gerkmann
2024InterspeechAn Analysis of the Variance of Diffusion-based Speech Enhancement.Bunlong Lay, Timo Gerkmann
2024InterspeechThe PESQetarian: On the Relevance of Goodhart's Law for Speech Enhancement.Danilo de Oliveira, Simon Welker, Julius Richter, Timo Gerkmann
2024InterspeechEARS: An Anechoic Fullband Speech Dataset Benchmarked for Speech Enhancement and Dereverberation.Julius Richter, Yi-Chiao Wu, Steven Krenn, Simon Welker, Bunlong Lay, Shinji Watanabe, Alexander Richard, Timo Gerkmann
2023ICASSPUncertainty Estimation in Deep Speech Enhancement Using Complex Gaussian Mixture Models.Huajian Fang, Timo Gerkmann
2023ICASSPPartially Adaptive Multichannel Joint Reduction of Ego-Noise and Environmental Noise.Huajian Fang, Niklas Wittmer, Johannes Twiefel, Stefan Wermter, Timo Gerkmann
2023ICASSPAnalysing Diffusion-based Generative Approaches Versus Discriminative Approaches for Speech Restoration.Jean-Marie Lemercier, Julius Richter, Simon Welker, Timo Gerkmann
2023ICASSPDiffPhase: Generative Diffusion-Based STFT Phase Retrieval.Tal Peer, Simon Welker, Timo Gerkmann
2023ICASSPSpeech Signal Improvement Using Causal Generative Diffusion Models.Julius Richter, Simon Welker, Jean-Marie Lemercier, Bunlong Lay, Tal Peer, Timo Gerkmann
2023ICASSPSpatially Selective Deep Non-Linear Filters For Speaker Extraction.Kristina Tesch, Timo Gerkmann
2023ICMIAcoustic and Visual Knowledge Distillation for Contrastive Audio-Visual Localization.Ehsan Yaghoubi, Andr Peter Kelm, Timo Gerkmann, Simone Frintrop
2023InterspeechReducing the Prior Mismatch of Stochastic Differential Equations for Diffusion-based Speech Enhancement.Bunlong Lay, Simon Welker, Julius Richter, Timo Gerkmann
2023InterspeechExtending DNN-based Multiplicative Masking to Deep Subband Filtering for Improved Dereverberation.Jean-Marie Lemercier, Julian Tobergte, Timo Gerkmann
2023InterspeechAudio-Visual Speech Separation in Noisy Environments with a Lightweight Iterative Model.Hctor Martel, Julius Richter, Kai Li, Xiaolin Hu, Timo Gerkmann
2023InterspeechLeveraging Semantic Information for Efficient Self-Supervised Emotion Recognition with Audio-Textual Distilled Models.Danilo de Oliveira, Navin Raj Prabhu, Timo Gerkmann
2022ACIILabel Uncertainty Modeling and Prediction for Speech Emotion Recognition using t-Distributions.Navin Raj Prabhu, Nale Lehmann-Willenbrock, Timo Gerkmann
2022ICASSPIntegrating Statistical Uncertainty into Neural Network-Based Speech Enhancement.Huajian Fang, Tal Peer, Stefan Wermter, Timo Gerkmann
2022ICASSPCustomizable End-To-End Optimization Of Online Neural Network-Supported Dereverberation For Hearing Devices.Jean-Marie Lemercier, Joachim Thiemann, Raphael Koning, Timo Gerkmann
2022ICASSPDeep Iterative Phase Retrieval for Ptychography.Simon Welker, Tal Peer, Henry N. Chapman, Timo Gerkmann
2022IJCNNContinuous Phoneme Recognition based on Audio-Visual Modality Fusion.Julius Richter, Jeanine Liebold, Timo Gerkmann
2022InterspeechNeural Network-augmented Kalman Filtering for Robust Online Speech Dereverberation in Noisy Reverberant Environments.Jean-Marie Lemercier, Joachim Thiemann, Raphael Koning, Timo Gerkmann
2022InterspeechEfficient Transformer-based Speech Enhancement Using Long Frames and STFT Magnitudes.Danilo de Oliveira, Tal Peer, Timo Gerkmann
2022InterspeechEnd-To-End Label Uncertainty Modeling for Speech-based Arousal Recognition Using Bayesian Neural Networks.Navin Raj Prabhu, Guillaume Carbajal, Nale Lehmann-Willenbrock, Timo Gerkmann
2022InterspeechOn the Role of Spatial, Spectral, and Temporal Processing for DNN-based Non-linear Multi-channel Speech Enhancement.Kristina Tesch, Nils-Hendrik Mohrmann, Timo Gerkmann
2022InterspeechSpeech Enhancement with Score-Based Generative Models in the Complex STFT Domain.Simon Welker, Julius Richter, Timo Gerkmann
2022MMSPSpeech Enhancement Regularized by a Speaker Verification Model.Bunlong Lay, Timo Gerkmann
2021ICASSPGuided Variational Autoencoder for Speech Enhancement with a Supervised Classifier.Guillaume Carbajal, Julius Richter, Timo Gerkmann
2021ICASSPVariational Autoencoder for Speech Enhancement with a Noise-Aware Encoder.Huajian Fang, Guillaume Carbajal, Stefan Wermter, Timo Gerkmann
2021ICVSSee the Silence: Improving Visual-Only Voice Activity Detection by Optical Flow and RGB Fusion.Danu Caus, Guillaume Carbajal, Timo Gerkmann, Simone Frintrop
2020ICASSPA Multi-Phase Gammatone Filterbank for Speech Separation Via Tasnet.David Ditter, Timo Gerkmann
2020ICASSPNonlinear Spatial Filtering for Multichannel Speech Enhancement in Inhomogeneous Noise Fields.Kristina Tesch, Timo Gerkmann
2020ICPRImproving mix-and-separate training in audio-visual sound source separation with an object prior.Quan Nguyen, Julius Richter, Mikko Lauri, Timo Gerkmann, Simone Frintrop
2020InterspeechSpeech Enhancement with Stochastic Temporal Convolutional Networks.Julius Richter, Guillaume Carbajal, Timo Gerkmann
2020IROSRobust Robotic Pouring using Audition and Haptics.Hongzhuo Liang, Chuangchuang Zhou, Shuang Li, Xiaojian Ma, Norman Hendrich, Timo Gerkmann, Fuchun Sun, Marcus Stoffel, Jianwei Zhang
2019ICASSPAn Analysis of Noise-aware Features in Combination with the Size and Diversity of Training Data for DNN-based Speech Enhancement.Robert Rehr, Timo Gerkmann
2019InterspeechInfluence of Speaker-Specific Parameters on Speech Separation Systems.David Ditter, Timo Gerkmann
2019InterspeechOn Nonlinear Spatial Filtering in Multichannel Speech Enhancement.Kristina Tesch, Robert Rehr, Timo Gerkmann
2019IROSMaking Sense of Audio Vibration for Liquid Height Estimation in Robotic Pouring.Hongzhuo Liang, Shuang Li, Xiaojian Ma, Norman Hendrich, Timo Gerkmann, Fuchun Sun, Jianwei Zhang
2018ICASSPNonlinear Speech Enhancement Under Speech PSD Uncertainty.Martin Krawczyk-Becker, Timo Gerkmann
2018ICASSPWeighted and Multi-Task Loss for Rare Audio Event Detection.Huy Phan, Martin Krawczyk-Becker, Timo Gerkmann, Alfred Mertins
2017InterspeechMixMax Approximation as a Super-Gaussian Log-Spectral Amplitude Estimator for Speech Enhancement.Robert Rehr, Timo Gerkmann
2016ICASSPSparse reconstruction of quantized speech signals.Christoph Brauer, Timo Gerkmann, Dirk A. Lorenz
2016ICASSPPerceptual and instrumental evaluation of the perceived level of reverberation.Benjamin Cauchi, Hamza A. Javed, Timo Gerkmann, Simon Doclo, Stefan Goetze, Patrick A. Naylor
2016ICASSPSingle-microphone speech enhancement using MVDR filtering and Wiener post-filtering.Drte Fischer, Timo Gerkmann
2016ICASSPBIAS correction methods for adaptive recursive smoothing with applications in noise PSD estimation.Robert Rehr, Timo Gerkmann
2015ICASSPMulti-channel linear prediction-based speech dereverberation with low-rank power spectrogram approximation.Ante Jukic, Nasser Mohammadiha, Toon van Waterschoot, Timo Gerkmann, Simon Doclo
2015ICASSPUtilizing spectro-temporal correlations for an improved speech presence probability based noise power estimation.Martin Krawczyk-Becker, Drte Fischer, Timo Gerkmann
2015ICASSPMulti-channel PSD estimators for speech dereverberation - A theoretical and experimental comparison.Adam Kuklasinski, Simon Doclo, Timo Gerkmann, Sren Holdt Jensen, Jesper Jensen
2015ICASSPCepstral noise subtraction for robust automatic speech recognition.Robert Rehr, Timo Gerkmann
2015InterspeechLeast squares estimate of the initial phases in STFT based speech enhancement.Sidsel Marie Nrholm, Martin Krawczyk-Becker, Timo Gerkmann, Steven van de Par, Jesper Rindom Jensen, Mads Grsbll Christensen
2014ICASSPMMSE-optimal enhancement of complex speech coefficients with uncertain prior knowledge of the clean speech phase.Timo Gerkmann
2014ICASSPFrequency-domain single-channel inverse filtering for speech dereverberation: Theory and practice.Ina Kodrasi, Timo Gerkmann, Simon Doclo
2014ICASSPA posteriori voiced/unvoiced probability estimation based on a sinusoidal model.Robert Rehr, Martin Krawczyk, Timo Gerkmann
2013ICASSPOn the relation between speech corruption models in the spectral and the cepstral domain.Ramn Fernandez Astudillo, Timo Gerkmann
2013ICASSPPrivacy-preserving distributed speech enhancement forwireless sensor networks by processing in the encrypted domain.Richard C. Hendriks, Zekeriya Erkin, Timo Gerkmann
2012ICASSPImproved mmse-based noise PSD tracking using temporal cepstrum smoothing.Timo Gerkmann, Richard C. Hendriks
2011ICASSPEstimation of the noise correlation matrix.Richard C. Hendriks, Timo Gerkmann
2010ICASSPSpeech presence probability estimation based on temporal cepstrum smoothing.Timo Gerkmann, Martin Krawczyk, Rainer Martin
2009ICASSPMulti-microphone maximum a posteriori fundamental frequency estimation in the cepstral domain.Timo Gerkmann, Rainer Martin, Derya Dalga
2008ICASSPA novel a priori SNR estimation approach based on selective cepstro-temporal smoothing.Colin Breithaupt, Timo Gerkmann, Rainer Martin
2006ICASSPStatistical Inference of Missing Speech Data in the ICA Domain.Justinian Rosca, Timo Gerkmann, Doru-Cristian Balcan
2006InterspeechSoft decision combining for dual channel noise reduction.Timo Gerkmann, Rainer Martin