| 2025 | ASRU | Enhancing In-the-Wild Speech Emotion Conversion with Resynthesis-based Duration Modeling. | Navin Raj Prabhu, Danilo de Oliveira, Nale Lehmann-Willenbrock, Timo Gerkmann |
| 2025 | ICASSP | Mask-Weighted Spatial Likelihood Coding for Speaker-Independent Joint Localization and Mask Estimation. | Jakob Kienegger, Alina Mannanova, Timo Gerkmann |
| 2025 | ICASSP | Investigating Training Objectives for Generative Speech Enhancement. | Julius Richter, Danilo de Oliveira, Timo Gerkmann |
| 2025 | ICASSP | HRTF Estimation using a Score-based Prior. | Etienne Thuillier, Jean-Marie Lemercier, Eloi Moliner, Timo Gerkmann, Vesa Vlimki |
| 2025 | ICLR | FlowDec: A flow-based full-band general audio codec with high perceptual quality. | Simon Welker, Matthew Le, Ricky T. Q. Chen, Wei-Ning Hsu, Timo Gerkmann, Alexander Richard, Yi-Chiao Wu |
| 2025 | Interspeech | Steering Deep Non-Linear Spatially Selective Filters for Weakly Guided Extraction of Moving Speakers in Dynamic Scenarios. | Jakob Kienegger, Timo Gerkmann |
| 2025 | Interspeech | Diffusion Buffer: Online Diffusion-based Speech Enhancement with Sub-Second Latency. | Bunlong Lay, Rostilav Makarov, Timo Gerkmann |
| 2025 | Interspeech | Real-Time Diffusion Buffer for Speech Enhancement On A Laptop. | Bunlong Lay, Rostilav Makarov, Timo Gerkmann |
| 2025 | Interspeech | Non-intrusive Speech Quality Assessment with Diffusion Models Trained on Clean Speech. | Danilo de Oliveira, Julius Richter, Jean-Marie Lemercier, Simon Welker, Timo Gerkmann |
| 2024 | ICASSP | Single and Few-Step Diffusion for Generative Speech Enhancement. | Bunlong Lay, Jean-Marie Lemercier, Julius Richter, Timo Gerkmann |
| 2024 | ICASSP | Distilling Hubert with LSTMs via Decoupled Knowledge Distillation. | Danilo de Oliveira, Timo Gerkmann |
| 2024 | ICASSP | A Flexible Online Framework for Projection-Based Stft Phase Retrieval. | Tal Peer, Simon Welker, Johannes Kolhoff, Timo Gerkmann |
| 2024 | ICASSP | EMOCONV-Diff: Diffusion-Based Speech Emotion Conversion for Non-Parallel and in-the-Wild Data. | Navin Raj Prabhu, Bunlong Lay, Simon Welker, Nale Lehmann-Willenbrock, Timo Gerkmann |
| 2024 | ICASSP | Live Iterative Ptychography with Projection-Based Algorithms. | Simon Welker, Tal Peer, Henry N. Chapman, Timo Gerkmann |
| 2024 | Interspeech | An Analysis of the Variance of Diffusion-based Speech Enhancement. | Bunlong Lay, Timo Gerkmann |
| 2024 | Interspeech | The PESQetarian: On the Relevance of Goodhart's Law for Speech Enhancement. | Danilo de Oliveira, Simon Welker, Julius Richter, Timo Gerkmann |
| 2024 | Interspeech | EARS: An Anechoic Fullband Speech Dataset Benchmarked for Speech Enhancement and Dereverberation. | Julius Richter, Yi-Chiao Wu, Steven Krenn, Simon Welker, Bunlong Lay, Shinji Watanabe, Alexander Richard, Timo Gerkmann |
| 2023 | ICASSP | Uncertainty Estimation in Deep Speech Enhancement Using Complex Gaussian Mixture Models. | Huajian Fang, Timo Gerkmann |
| 2023 | ICASSP | Partially Adaptive Multichannel Joint Reduction of Ego-Noise and Environmental Noise. | Huajian Fang, Niklas Wittmer, Johannes Twiefel, Stefan Wermter, Timo Gerkmann |
| 2023 | ICASSP | Analysing Diffusion-based Generative Approaches Versus Discriminative Approaches for Speech Restoration. | Jean-Marie Lemercier, Julius Richter, Simon Welker, Timo Gerkmann |
| 2023 | ICASSP | DiffPhase: Generative Diffusion-Based STFT Phase Retrieval. | Tal Peer, Simon Welker, Timo Gerkmann |
| 2023 | ICASSP | Speech Signal Improvement Using Causal Generative Diffusion Models. | Julius Richter, Simon Welker, Jean-Marie Lemercier, Bunlong Lay, Tal Peer, Timo Gerkmann |
| 2023 | ICASSP | Spatially Selective Deep Non-Linear Filters For Speaker Extraction. | Kristina Tesch, Timo Gerkmann |
| 2023 | ICMI | Acoustic and Visual Knowledge Distillation for Contrastive Audio-Visual Localization. | Ehsan Yaghoubi, Andr Peter Kelm, Timo Gerkmann, Simone Frintrop |
| 2023 | Interspeech | Reducing the Prior Mismatch of Stochastic Differential Equations for Diffusion-based Speech Enhancement. | Bunlong Lay, Simon Welker, Julius Richter, Timo Gerkmann |
| 2023 | Interspeech | Extending DNN-based Multiplicative Masking to Deep Subband Filtering for Improved Dereverberation. | Jean-Marie Lemercier, Julian Tobergte, Timo Gerkmann |
| 2023 | Interspeech | Audio-Visual Speech Separation in Noisy Environments with a Lightweight Iterative Model. | Hctor Martel, Julius Richter, Kai Li, Xiaolin Hu, Timo Gerkmann |
| 2023 | Interspeech | Leveraging Semantic Information for Efficient Self-Supervised Emotion Recognition with Audio-Textual Distilled Models. | Danilo de Oliveira, Navin Raj Prabhu, Timo Gerkmann |
| 2022 | ACII | Label Uncertainty Modeling and Prediction for Speech Emotion Recognition using t-Distributions. | Navin Raj Prabhu, Nale Lehmann-Willenbrock, Timo Gerkmann |
| 2022 | ICASSP | Integrating Statistical Uncertainty into Neural Network-Based Speech Enhancement. | Huajian Fang, Tal Peer, Stefan Wermter, Timo Gerkmann |
| 2022 | ICASSP | Customizable End-To-End Optimization Of Online Neural Network-Supported Dereverberation For Hearing Devices. | Jean-Marie Lemercier, Joachim Thiemann, Raphael Koning, Timo Gerkmann |
| 2022 | ICASSP | Deep Iterative Phase Retrieval for Ptychography. | Simon Welker, Tal Peer, Henry N. Chapman, Timo Gerkmann |
| 2022 | IJCNN | Continuous Phoneme Recognition based on Audio-Visual Modality Fusion. | Julius Richter, Jeanine Liebold, Timo Gerkmann |
| 2022 | Interspeech | Neural Network-augmented Kalman Filtering for Robust Online Speech Dereverberation in Noisy Reverberant Environments. | Jean-Marie Lemercier, Joachim Thiemann, Raphael Koning, Timo Gerkmann |
| 2022 | Interspeech | Efficient Transformer-based Speech Enhancement Using Long Frames and STFT Magnitudes. | Danilo de Oliveira, Tal Peer, Timo Gerkmann |
| 2022 | Interspeech | End-To-End Label Uncertainty Modeling for Speech-based Arousal Recognition Using Bayesian Neural Networks. | Navin Raj Prabhu, Guillaume Carbajal, Nale Lehmann-Willenbrock, Timo Gerkmann |
| 2022 | Interspeech | On the Role of Spatial, Spectral, and Temporal Processing for DNN-based Non-linear Multi-channel Speech Enhancement. | Kristina Tesch, Nils-Hendrik Mohrmann, Timo Gerkmann |
| 2022 | Interspeech | Speech Enhancement with Score-Based Generative Models in the Complex STFT Domain. | Simon Welker, Julius Richter, Timo Gerkmann |
| 2022 | MMSP | Speech Enhancement Regularized by a Speaker Verification Model. | Bunlong Lay, Timo Gerkmann |
| 2021 | ICASSP | Guided Variational Autoencoder for Speech Enhancement with a Supervised Classifier. | Guillaume Carbajal, Julius Richter, Timo Gerkmann |
| 2021 | ICASSP | Variational Autoencoder for Speech Enhancement with a Noise-Aware Encoder. | Huajian Fang, Guillaume Carbajal, Stefan Wermter, Timo Gerkmann |
| 2021 | ICVS | See the Silence: Improving Visual-Only Voice Activity Detection by Optical Flow and RGB Fusion. | Danu Caus, Guillaume Carbajal, Timo Gerkmann, Simone Frintrop |
| 2020 | ICASSP | A Multi-Phase Gammatone Filterbank for Speech Separation Via Tasnet. | David Ditter, Timo Gerkmann |
| 2020 | ICASSP | Nonlinear Spatial Filtering for Multichannel Speech Enhancement in Inhomogeneous Noise Fields. | Kristina Tesch, Timo Gerkmann |
| 2020 | ICPR | Improving mix-and-separate training in audio-visual sound source separation with an object prior. | Quan Nguyen, Julius Richter, Mikko Lauri, Timo Gerkmann, Simone Frintrop |
| 2020 | Interspeech | Speech Enhancement with Stochastic Temporal Convolutional Networks. | Julius Richter, Guillaume Carbajal, Timo Gerkmann |
| 2020 | IROS | Robust Robotic Pouring using Audition and Haptics. | Hongzhuo Liang, Chuangchuang Zhou, Shuang Li, Xiaojian Ma, Norman Hendrich, Timo Gerkmann, Fuchun Sun, Marcus Stoffel, Jianwei Zhang |
| 2019 | ICASSP | An Analysis of Noise-aware Features in Combination with the Size and Diversity of Training Data for DNN-based Speech Enhancement. | Robert Rehr, Timo Gerkmann |
| 2019 | Interspeech | Influence of Speaker-Specific Parameters on Speech Separation Systems. | David Ditter, Timo Gerkmann |
| 2019 | Interspeech | On Nonlinear Spatial Filtering in Multichannel Speech Enhancement. | Kristina Tesch, Robert Rehr, Timo Gerkmann |
| 2019 | IROS | Making Sense of Audio Vibration for Liquid Height Estimation in Robotic Pouring. | Hongzhuo Liang, Shuang Li, Xiaojian Ma, Norman Hendrich, Timo Gerkmann, Fuchun Sun, Jianwei Zhang |
| 2018 | ICASSP | Nonlinear Speech Enhancement Under Speech PSD Uncertainty. | Martin Krawczyk-Becker, Timo Gerkmann |
| 2018 | ICASSP | Weighted and Multi-Task Loss for Rare Audio Event Detection. | Huy Phan, Martin Krawczyk-Becker, Timo Gerkmann, Alfred Mertins |
| 2017 | Interspeech | MixMax Approximation as a Super-Gaussian Log-Spectral Amplitude Estimator for Speech Enhancement. | Robert Rehr, Timo Gerkmann |
| 2016 | ICASSP | Sparse reconstruction of quantized speech signals. | Christoph Brauer, Timo Gerkmann, Dirk A. Lorenz |
| 2016 | ICASSP | Perceptual and instrumental evaluation of the perceived level of reverberation. | Benjamin Cauchi, Hamza A. Javed, Timo Gerkmann, Simon Doclo, Stefan Goetze, Patrick A. Naylor |
| 2016 | ICASSP | Single-microphone speech enhancement using MVDR filtering and Wiener post-filtering. | Drte Fischer, Timo Gerkmann |
| 2016 | ICASSP | BIAS correction methods for adaptive recursive smoothing with applications in noise PSD estimation. | Robert Rehr, Timo Gerkmann |
| 2015 | ICASSP | Multi-channel linear prediction-based speech dereverberation with low-rank power spectrogram approximation. | Ante Jukic, Nasser Mohammadiha, Toon van Waterschoot, Timo Gerkmann, Simon Doclo |
| 2015 | ICASSP | Utilizing spectro-temporal correlations for an improved speech presence probability based noise power estimation. | Martin Krawczyk-Becker, Drte Fischer, Timo Gerkmann |
| 2015 | ICASSP | Multi-channel PSD estimators for speech dereverberation - A theoretical and experimental comparison. | Adam Kuklasinski, Simon Doclo, Timo Gerkmann, Sren Holdt Jensen, Jesper Jensen |
| 2015 | ICASSP | Cepstral noise subtraction for robust automatic speech recognition. | Robert Rehr, Timo Gerkmann |
| 2015 | Interspeech | Least squares estimate of the initial phases in STFT based speech enhancement. | Sidsel Marie Nrholm, Martin Krawczyk-Becker, Timo Gerkmann, Steven van de Par, Jesper Rindom Jensen, Mads Grsbll Christensen |
| 2014 | ICASSP | MMSE-optimal enhancement of complex speech coefficients with uncertain prior knowledge of the clean speech phase. | Timo Gerkmann |
| 2014 | ICASSP | Frequency-domain single-channel inverse filtering for speech dereverberation: Theory and practice. | Ina Kodrasi, Timo Gerkmann, Simon Doclo |
| 2014 | ICASSP | A posteriori voiced/unvoiced probability estimation based on a sinusoidal model. | Robert Rehr, Martin Krawczyk, Timo Gerkmann |
| 2013 | ICASSP | On the relation between speech corruption models in the spectral and the cepstral domain. | Ramn Fernandez Astudillo, Timo Gerkmann |
| 2013 | ICASSP | Privacy-preserving distributed speech enhancement forwireless sensor networks by processing in the encrypted domain. | Richard C. Hendriks, Zekeriya Erkin, Timo Gerkmann |
| 2012 | ICASSP | Improved mmse-based noise PSD tracking using temporal cepstrum smoothing. | Timo Gerkmann, Richard C. Hendriks |
| 2011 | ICASSP | Estimation of the noise correlation matrix. | Richard C. Hendriks, Timo Gerkmann |
| 2010 | ICASSP | Speech presence probability estimation based on temporal cepstrum smoothing. | Timo Gerkmann, Martin Krawczyk, Rainer Martin |
| 2009 | ICASSP | Multi-microphone maximum a posteriori fundamental frequency estimation in the cepstral domain. | Timo Gerkmann, Rainer Martin, Derya Dalga |
| 2008 | ICASSP | A novel a priori SNR estimation approach based on selective cepstro-temporal smoothing. | Colin Breithaupt, Timo Gerkmann, Rainer Martin |
| 2006 | ICASSP | Statistical Inference of Missing Speech Data in the ICA Domain. | Justinian Rosca, Timo Gerkmann, Doru-Cristian Balcan |
| 2006 | Interspeech | Soft decision combining for dual channel noise reduction. | Timo Gerkmann, Rainer Martin |