| 2026 | ACL | Beyond Transcripts: A Renewed Perspective on Audio Chaptering. | Fabian Retkowski, Maike Zfle, Thai-Binh Nguyen, Jan Niehues, Alexander Waibel |
| 2026 | EACL | BOOM: Beyond Only One Modality KIT's Multimodal Multilingual Lecture Companion. | Sai Koneru, Fabian Retkowski, Christian Huber, Lukas Hilgert, Seymanur Akti, Enes Yavuz Ugan, Alexander Waibel, Jan Niehues |
| 2026 | LREC | Paragraph Segmentation Revisited: Towards a Standard Task for Structuring Speech. | Fabian Retkowski, Alexander Waibel |
| 2026 | LREC | MUSCAT: MUltilingual, SCientific ConversATion Benchmark. | Supriti Sinhamahapatra, Thai-Binh Nguyen, Yigit Oguz, Enes Yavuz Ugan, Jan Niehues, Alexander Waibel |
| 2026 | WACV | CAPE: A CLIP-Aware Pointing Ensemble of Complementary Heatmap Cues for Embodied Reference Understanding. | Fevziye Irem Eyiokur, Dogucan Yaman, Hazim Kemal Ekenel, Alexander Waibel |
| 2025 | EMNLP | Few-Shot Learning Translation from New Languages. | Carlos Mullov, Alexander Waibel |
| 2025 | EMNLP | Summarizing Speech: A Comprehensive Survey. | Fabian Retkowski, Maike Zfle, Andreas Sudmann, Dinah Pfau, Shinji Watanabe, Jan Niehues, Alexander Waibel |
| 2025 | ICASSP | Continuously Learning New Words in Automatic Speech Recognition. | Christian Huber, Alexander Waibel |
| 2025 | ICASSP | Factorized-VITS: Decoupling Prosody and Text in End-to-End Speech Synthesis without External or Secondary Aligner. | Yining Liu, Alexander Waibel |
| 2025 | ICASSP | Improving Pronunciation and Accent Conversion through Knowledge Distillation And Synthetic Ground-Truth from Native TTS. | Tuan Nam Nguyen, Seymanur Akti, Ngoc-Quan Pham, Alexander Waibel |
| 2025 | ICASSP | MSA-ASR: Efficient Multilingual Speaker Attribution with frozen ASR Models. | Thai-Binh Nguyen, Alexander Waibel |
| 2025 | Interspeech | Towards Better Disentanglement in Non-Autoregressive Zero-Shot Expressive Voice Conversion. | Seymanur Akti, Tuan-Nam Nguyen, Alexander Waibel |
| 2025 | Interspeech | Streaming Non-Autoregressive Model for Accent Conversion and Pronunciation Improvement. | Tuan-Nam Nguyen, Ngoc-Quan Pham, Seymanur Akti, Alexander Waibel |
| 2025 | Interspeech | Cocktail-Party Audio-Visual Speech Recognition. | Thai-Binh Nguyen, Ngoc-Quan Pham, Alexander Waibel |
| 2025 | Interspeech | Weight Factorization and Centralization for Continual Learning in Speech Recognition. | Enes Yavuz Ugan, Ngoc-Quan Pham, Alexander Waibel |
| 2025 | Interspeech | From Speech Science to Language Transparence. | Alexander Waibel |
| 2025 | NAACL | Zero-Shot Strategies for Length-Controllable Summarization. | Fabian Retkowski, Alexander Waibel |
| 2024 | ACL | Decoupled Vocabulary Learning Enables Zero-Shot Translation from Unseen Languages. | Carlos Mullov, Ngoc-Quan Pham, Alexander Waibel |
| 2024 | COLING | DECM: Evaluating Bilingual ASR Performance on a Code-switching/mixing Benchmark. | Enes Yavuz Ugan, Ngoc-Quan Pham, Alexander Waibel |
| 2024 | CVPR | Audio-Visual Speech Representation Expert for Enhanced Talking Face Video Generation and Evaluation. | Dogucan Yaman, Fevziye Irem Eyiokur, Leonard Brmann, Seymanur Akti, Hazim Kemal Ekenel, Alexander Waibel |
| 2024 | EACL | From Text Segmentation to Smart Chaptering: A Novel Benchmark for Structuring Video Transcriptions. | Fabian Retkowski, Alexander Waibel |
| 2024 | ECCV | Audio-Driven Talking Face Generation with Stabilized Synchronization Loss. | Dogucan Yaman, Fevziye Irem Eyiokur, Leonard Brmann, Hazim Kemal Ekenel, Alexander Waibel |
| 2024 | EMNLP | SciEx: Benchmarking Large Language Models on Scientific Exams with Human Expert Grading and Automatic Grading. | Tu Anh Dinh, Carlos Mullov, Leonard Brmann, Zhaolin Li, Danni Liu, Simon Rei, Jueun Lee, Nathan Lerzer, Jianfeng Gao, Fabian Peller-Konrad, Tobias Rddiger, Alexander Waibel, Tamim Asfour, Michael Beigl, Rainer Stiefelhagen, Carsten Dachsbacher, Klemens Bhm, Jan Niehues |
| 2024 | ICASSP | Synthetic Conversations Improve Multi-Talker ASR. | Thai-Binh Nguyen, Alexander Waibel |
| 2024 | ICASSP | ConVoiFilter: A Case Study of Doing Cocktail Party Speech Recognition. | Thai-Binh Nguyen, Alexander Waibel |
| 2023 | EMNLP | End-to-End Evaluation for Low-Latency Simultaneous Speech Translation. | Christian Huber, Tu Anh Dinh, Carlos Mullov, Ngoc-Quan Pham, Thai Binh Nguyen, Fabian Retkowski, Stefan Constantin, Enes Yavuz Ugan, Danni Liu, Zhaolin Li, Sai Koneru, Jan Niehues, Alexander Waibel |
| 2023 | ICASSP | AdapITN: A Fast, Reliable, and Dynamic Adaptive Inverse Text Normalization. | Thai Binh Nguyen, Le Duc Minh Nhat, Quang Minh Nguyen, Quoc Truong Do, Chi Mai Luong, Alexander Waibel |
| 2023 | ICASSP | SYNTACC : Synthesizing Multi-Accent Speech By Weight Factorization. | Tuan-Nam Nguyen, Ngoc-Quan Pham, Alexander Waibel |
| 2023 | ICASSP | Face-Dubbing++: LIP-Synchronous, Voice Preserving Translation Of Videos. | Alexander Waibel, Moritz Behr, Dogucan Yaman, Fevziye Irem Eyiokur, Tuan-Nam Nguyen, Carlos Mullov, Mehmet Arif Demirtas, Alperen Kantarci, Stefan Constantin, Hazim Kemal Ekenel |
| 2022 | CVPR | Exposure Correction Model to Enhance Image Quality. | Fevziye Irem Eyiokur, Dogucan Yaman, Hazim Kemal Ekenel, Alexander Waibel |
| 2022 | CVPR | Alpha Matte Generation from Single Input for Portrait Matting. | Dogucan Yaman, Hazim Kemal Ekenel, Alexander Waibel |
| 2022 | Interspeech | Accent Conversion using Pre-trained Model and Synthesized Data from Voice Conversion. | Tuan-Nam Nguyen, Ngoc-Quan Pham, Alexander Waibel |
| 2022 | Interspeech | Adaptive multilingual speech recognition with pretrained models. | Ngoc-Quan Pham, Alexander Waibel, Jan Niehues |
| 2021 | ASRU | Instant One-Shot Word-Learning for Context-Specific Neural Sequence-to-Sequence Speech Recognition. | Christian Huber, Juan Hussain, Sebastian Stker, Alexander Waibel |
| 2020 | EAMT | Incorporating External Annotation to improve Named Entity Translation in NMT. | Maciej Modrzejewski, Miriam Exel, Bianka Buschbeck, Thanh-Le Ha, Alexander Waibel |
| 2019 | ACL | Paraphrases as Foreign Languages in Multilingual Neural Machine Translation. | Zhong Zhou, Matthias Sperber, Alexander Waibel |
| 2019 | ICMI | Connecting Humans with Humans: Multimodal, Multilingual, Multiparty Mediation. | Alexander Waibel |
| 2019 | NAACL | Fluent Translations from Disfluent Speech in End-to-End Speech Translation. | Elizabeth Salesky, Matthias Sperber, Alexander Waibel |
| 2018 | COLING | KIT Lecture Translator: Multilingual Speech Translation with One-Shot Learning. | Florian Dessloch, Thanh-Le Ha, Markus Mller, Jan Niehues, Thai Son Nguyen, Ngoc-Quan Pham, Elizabeth Salesky, Matthias Sperber, Sebastian Stker, Thomas Zenkel, Alexander Waibel |
| 2018 | LREC | KIT-Multi: A Translation-Oriented Multilingual Embedding Corpus. | Thanh-Le Ha, Jan Niehues, Matthias Sperber, Ngoc-Quan Pham, Alexander Waibel |
| 2017 | HAI | Keynote Talk. | Alexander Waibel |
| 2016 | ICASSP | An empirical exploration of CTC acoustic models. | Yajie Miao, Mohammad Gowayyed, Xingyu Na, Tom Ko, Florian Metze, Alexander Waibel |
| 2014 | ICMI | A World without Barriers: Connecting the World across Languages, Distances and Media. | Alexander Waibel |
| 1999 | Interspeech | Navigating German cities by spontaneous French queries. | Harouna Kabr, Alexander Waibel |
| 1990 | IJCNN | Phoneme-based word recognition by neural network - a step toward large vocabulary recognition. | Akihiro Hirai, Alexander Waibel |