| 2025 | Interspeech | SoundSculpt: Direction and Semantics Driven Ambisonic Target Sound Extraction. | Tuochao Chen, D. Shin, Hakan Erdogan, Sinan Hersek |
| 2024 | ICASSP | Binaural Angular Separation Network. | Yang Yang, George Sung, Shao-Fu Shih, Hakan Erdogan, Chehung Lee, Matthias Grundmann |
| 2024 | ICASSP | Quantifying The Effect Of Simulator-Based Data Augmentation For Speech Recognition On Augmented Reality Glasses. | Riku Arakawa, Mathieu Parvaix, Chiong Lai, Hakan Erdogan, Alex Olwal |
| 2023 | ICASSP | Guided Speech Enhancement Network. | Yang Yang, Shao-Fu Shih, Hakan Erdogan, Jamie Menjay Lin, Chehung Lee, Yunpeng Li, George Sung, Matthias Grundmann |
| 2023 | Interspeech | TokenSplit: Using Discrete Speech Representations for Direct, Refined, and Transcript-Conditioned Speech Separation and Recognition. | Hakan Erdogan, Scott Wisdom, Xuankai Chang, Zaln Borsos, Marco Tagliasacchi, Neil Zeghidour, John R. Hershey |
| 2022 | ICASSP | Adapting Speech Separation to Real-World Meetings using Mixture Invariant Training. | Aswin Sivaraman, Scott Wisdom, Hakan Erdogan, John R. Hershey |
| 2022 | Interspeech | CycleGAN-based Unpaired Speech Dereverberation. | Hannah Muckenhirn, Aleksandr Safin, Hakan Erdogan, Felix de Chaumont Quitry, Marco Tagliasacchi, Scott Wisdom, John R. Hershey |
| 2021 | ICASSP | End-To-End Diarization for Variable Number of Speakers with Local-Global Networks and Discriminative Speaker Embeddings. | Soumi Maiti, Hakan Erdogan, Kevin W. Wilson, Scott Wisdom, Shinji Watanabe, John R. Hershey |
| 2021 | ICASSP | Sound Event Detection and Separation: A Benchmark on Desed Synthetic Soundscapes. | Nicolas Turpault, Romain Serizel, Scott Wisdom, Hakan Erdogan, John R. Hershey, Eduardo Fonseca, Prem Seetharaman, Justin Salamon |
| 2021 | ICASSP | What's all the Fuss about Free Universal Sound Separation Data? | Scott Wisdom, Hakan Erdogan, Daniel P. W. Ellis, Romain Serizel, Nicolas Turpault, Eduardo Fonseca, Justin Salamon, Prem Seetharaman, John R. Hershey |
| 2021 | Interspeech | Continuous Speech Separation Using Speaker Inventory for Long Recording. | Cong Han, Yi Luo, Chenda Li, Tianyan Zhou, Keisuke Kinoshita, Shinji Watanabe, Marc Delcroix, Hakan Erdogan, John R. Hershey, Nima Mesgarani, Zhuo Chen |
| 2020 | ICASSP | Performance Study of a Convolutional Time-Domain Audio Separation Network for Real-Time Speech Denoising. | Samuel Sonning, Christian Schldt, Hakan Erdogan, Scott Wisdom |
| 2019 | ICASSP | SDR - Half-baked or Well Done? | Jonathan Le Roux, Scott Wisdom, Hakan Erdogan, John R. Hershey |
| 2019 | ICASSP | Single-channel Speech Extraction Using Speaker Inventory and Attention Network. | Xiong Xiao, Zhuo Chen, Takuya Yoshioka, Hakan Erdogan, Changliang Liu, Dimitrios Dimitriadis, Jasha Droppo, Yifan Gong |
| 2019 | ICASSP | Low-latency Speaker-independent Continuous Speech Separation. | Takuya Yoshioka, Zhuo Chen, Changliang Liu, Xiong Xiao, Hakan Erdogan, Dimitrios Dimitriadis |
| 2018 | ICASSP | Exploring Practical Aspects of Neural Mask-Based Beamforming for Far-Field Speech Recognition. | Christoph Bddeker, Hakan Erdogan, Takuya Yoshioka, Reinhold Haeb-Umbach |
| 2018 | ICASSP | Multi-Microphone Neural Speech Separation for Far-Field Multi-Talker Speech Recognition. | Takuya Yoshioka, Hakan Erdogan, Zhuo Chen, Fil Alleva |
| 2018 | Interspeech | Investigations on Data Augmentation and Loss Functions for Deep Learning Based Speech-Background Separation. | Hakan Erdogan, Takuya Yoshioka |
| 2018 | Interspeech | Recognizing Overlapped Speech in Meetings: A Multichannel Separation Approach Using Neural Networks. | Takuya Yoshioka, Hakan Erdogan, Zhuo Chen, Xiong Xiao, Fil Alleva |
| 2017 | ICASSP | Deep long short-term memory adaptive beamforming networks for multichannel robust speech recognition. | Zhong Meng, Shinji Watanabe, John R. Hershey, Hakan Erdogan |
| 2016 | CVPR | GMM-SVM Fingerprint Verification Based on Minutiae Only. | Berkay Topcu, Yusuf Ziya Isik, Hakan Erdogan |
| 2016 | ICASSP | Deep beamforming networks for multi-channel speech recognition. | Xiong Xiao, Shinji Watanabe, Hakan Erdogan, Liang Lu, John R. Hershey, Michael L. Seltzer, Guoguo Chen, Yu Zhang, Michael I. Mandel, Dong Yu |
| 2016 | Interspeech | Improved MVDR Beamforming Using Single-Channel Mask Prediction Networks. | Hakan Erdogan, John R. Hershey, Shinji Watanabe, Michael I. Mandel, Jonathan Le Roux |
| 2015 | ASRU | The MERL/SRI system for the 3RD CHiME challenge using beamforming, robust feature extraction, and advanced speech recognition. | Takaaki Hori, Zhuo Chen, Hakan Erdogan, John R. Hershey, Jonathan Le Roux, Vikramjit Mitra, Shinji Watanabe |
| 2015 | ICASSP | PLDA-based diarization of telephone conversations. | Ahmet Emin Bulut, Hakan Demir, Yusuf Ziya Isik, Hakan Erdogan |
| 2015 | ICASSP | Phase-sensitive and recognition-boosted speech separation using deep recurrent neural networks. | Hakan Erdogan, John R. Hershey, Shinji Watanabe, Jonathan Le Roux |
| 2015 | Interspeech | Speech enhancement and recognition using multi-task learning of long short-term memory recurrent neural networks. | Zhuo Chen, Shinji Watanabe, Hakan Erdogan, John R. Hershey |
| 2014 | AVSS | Counting people by clustering person detector outputs. | Ibrahim Saygin Topkaya, Hakan Erdogan, Fatih Murat Porikli |
| 2014 | ICASSP | Deep neural networks for single channel source separation. | Emad M. Grais, Mehmet Umut Sen, Hakan Erdogan |
| 2013 | ICASSP | A mixed integer linear programming formulation for the sparse recovery problem in compressed sensing. | Nazim Burak Karahanoglu, Hakan Erdogan, S. Ilker Birbil |
| 2013 | Interspeech | Discriminative nonnegative dictionary learning using cross-coherence penalties for single channel source separation. | Emad M. Grais, Hakan Erdogan |
| 2013 | Interspeech | Spectro-temporal post-enhancement using MMSE estimation in NMF based single-channel source separation. | Emad M. Grais, Hakan Erdogan |
| 2013 | ISVC | Detecting and Tracking Unknown Number of Objects with Dirichlet Process Mixture Models and Markov Random Fields. | Ibrahim Saygin Topkaya, Hakan Erdogan, Fatih Porikli |
| 2012 | Interspeech | Gaussian Mixture Gain Priors for Regularized Nonnegative Matrix Factorization in Single-Channel Source Separation. | Emad M. Grais, Hakan Erdogan |
| 2012 | Interspeech | Hidden Markov Models as Priors for Regularized Nonnegative Matrix Factorization in Single-Channel Source Separation. | Emad M. Grais, Hakan Erdogan |
| 2012 | LREC | SUTAV: A Turkish Audio-Visual Database. | Ibrahim Saygin Topkaya, Hakan Erdogan |
| 2011 | ICASSP | Compressed sensing signal recovery via A* Orthogonal Matching Pursuit. | Nazim Burak Karahanoglu, Hakan Erdogan |
| 2011 | ICASSP | Using multiple visual tandem streams in audio-visual speech recognition. | Ibrahim Saygin Topkaya, Hakan Erdogan |
| 2011 | Interspeech | Adaptation of Speaker-Specific Bases in Non-Negative Matrix Factorization for Single Channel Speech-Music Separation. | Emad M. Grais, Hakan Erdogan |
| 2011 | Interspeech | Single Channel Speech Music Separation Using Nonnegative Matrix Factorization with Sliding Windows and Spectral Masks. | Emad M. Grais, Hakan Erdogan |
| 2010 | ICPR | Semi-blind Speech-Music Separation Using Sparsity and Continuity Priors. | Hakan Erdogan, Emad M. Grais |
| 2010 | ICPR | A Unifying Framework for Learning the Linear Combiners for Classifier Ensembles. | Hakan Erdogan, Mehmet Umut Sen |
| 2010 | ICPR | Decision Fusion for Patch-Based Face Recognition. | Berkay Topcu, Hakan Erdogan |
| 2009 | ISVC | Probabilistic Facial Feature Extraction Using Joint Distribution of Location and Texture Information. | Mustafa Berkay Yilmaz, Hakan Erdogan, Mustafa Unel |
| 2008 | BMVC | Evolving Implicit Polynomial Interfaces. | Erol Ozgur, Mustafa Unel, Hakan Erdogan, Aytl Eril |
| 2008 | ICASSP | Using local temporal features of bounding boxes for walking/running classification. | Berkay Topcu, Hakan Erdogan |
| 2007 | ICASSP | Protein Fold Recognition using Residue-Based Alignments of Sequence and Secondary Structure. | Zafer Aydin, Hakan Erdogan, Yucel Altunbasak |
| 2005 | Interspeech | Regularizing linear discriminant analysis for speech recognition. | Hakan Erdogan |
| 2004 | ICASSP | Filler model based confidence measures for spoken dialogue systems: a case study for Turkish. | Aydin Akyol, Hakan Erdogan |
| 2002 | ICASSP | Turn-Based Language Modeling for spoken dialog systems. | Ruhi Sarikaya, Yuqing Gao, Hakan Erdogan, Michael Picheny |
| 2002 | Interspeech | Semantic structured language models. | Hakan Erdogan, Ruhi Sarikaya, Yuqing Gao, Michael Picheny |
| 2002 | Interspeech | Incremental on-line feature space MLLR adaptation for telephony speech recognition. | Yongxin Li, Hakan Erdogan, Yuqing Gao, Etienne Marcheret |
| 2001 | ICASSP | Rapid adaptation using penalized-likelihood methods. | Hakan Erdogan, Yuqing Gao, Michael Picheny |
| 2001 | ICASSP | Innovative approaches for large vocabulary name recognition. | Yuqing Gao, Bhuvana Ramabhadran, C. Julian Chen, Hakan Erdogan, Michael Picheny |
| 2001 | Interspeech | Recent advances in speech recognition system for IBM DARPA communicator. | Yuqing Gao, Hakan Erdogan, Yongxin Li, Vaibhava Goel, Michael Picheny |
| 2000 | ICASSP | Algorithms for joint estimation of attenuation and emission images in PET. | Hakan Erdogan, Jeffrey A. Fessler |
| 2000 | Interspeech | Weighted pairwise scatter to improve linear discriminant analysis. | Yongxin Li, Yuqing Gao, Hakan Erdogan |
| 1998 | ICIP | Accelerated Monotonic Algorithms for Transmission Tomography. | Hakan Erdogan, Jeffrey A. Fessler |