| 2024 | ICASSP | Geodesic Interpolation of Frame-Wise Speaker Embeddings for the Diarization of Meeting Scenarios. | Tobias Cord-Landwehr, Christoph Bddeker, Catalin Zorila, Rama Doddipatla, Reinhold Haeb-Umbach |
| 2024 | Interspeech | Prompting Whisper for QA-driven Zero-shot End-to-end Spoken Language Understanding. | Mohan Li, Simon Keizer, Rama Doddipatla |
| 2023 | ASRU | Robust Recognition of Speaker Emotion With Difference Feature Extraction Using a Few Enrollment Utterances. | Daichi Hayakawa, Takehiko Kagoshima, Kenji Iwata, Norbert Braunschweiler, Rama Doddipatla |
| 2023 | ASRU | Towards a Unified End-to-End Language Understanding System for Speech and Text Inputs. | Mohan Li, Catalin Zorila, Cong-Thanh Do, Rama Doddipatla |
| 2023 | ICASSP | Frame-Wise and Overlap-Robust Speaker Embeddings for Meeting Diarization. | Tobias Cord-Landwehr, Christoph Bddeker, Catalin Zorila, Rama Doddipatla, Reinhold Haeb-Umbach |
| 2023 | ICASSP | Cumulative Attention Based Streaming Transformer ASR with Internal Language Model Joint Training and Rescoring. | Mohan Li, Cong-Thanh Do, Rama Doddipatla |
| 2023 | ICASSP | On the Effectiveness of Monoaural Target Source Extraction for Distant end-to-end Automatic Speech Recognition. | Catalin Zorila, Rama Doddipatla |
| 2023 | Interspeech | A Teacher-Student Approach for Extracting Informative Speaker Embeddings From Speech Mixtures. | Tobias Cord-Landwehr, Christoph Bddeker, Catalin Zorila, Rama Doddipatla, Reinhold Haeb-Umbach |
| 2023 | Interspeech | Domain Adaptive Self-supervised Training of Automatic Speech Recognition. | Cong-Thanh Do, Rama Doddipatla, Mohan Li, Thomas Hain |
| 2022 | ICASSP | Transformer-Based Streaming ASR with Cumulative Attention. | Mohan Li, Shucong Zhang, Catalin Zorila, Rama Doddipatla |
| 2022 | ICASSP | Speaker Reinforcement Using Target Source Extraction for Robust Automatic Speech Recognition. | Catalin Zorila, Rama Doddipatla |
| 2022 | Interspeech | Multiple-hypothesis RNN-T Loss for Unsupervised Fine-tuning and Self-training of Neural Transducer. | Cong-Thanh Do, Mohan Li, Rama Doddipatla |
| 2022 | Interspeech | On monoaural speech enhancement for automatic recognition of real noisy speech using mixture invariant training. | Jisi Zhang, Catalin Zorila, Rama Doddipatla, Jon Barker |
| 2021 | ASRU | A Study on Cross-Corpus Speech Emotion Recognition and Data Augmentation. | Norbert Braunschweiler, Rama Doddipatla, Simon Keizer, Svetlana Stoyanchev |
| 2021 | ASRU | Dialogue Strategy Adaptation to New Action Sets Using Multi-Dimensional Modelling. | Simon Keizer, Norbert Braunschweiler, Svetlana Stoyanchev, Rama Doddipatla |
| 2021 | ASRU | Improving HS-DACS Based Streaming Transformer ASR with Deep Reinforcement Learning. | Mohan Li, Rama Doddipatla |
| 2021 | ICASSP | Multiple-Hypothesis CTC-Based Semi-Supervised Adaptation of End-to-End Speech Recognition. | Cong-Thanh Do, Rama Doddipatla, Thomas Hain |
| 2021 | ICASSP | Head-Synchronous Decoding for Transformer-Based Streaming ASR. | Mohan Li, Catalin Zorila, Rama Doddipatla |
| 2021 | ICASSP | Action State Update Approach to Dialogue Management. | Svetlana Stoyanchev, Simon Keizer, Rama Doddipatla |
| 2021 | ICASSP | Train Your Classifier First: Cascade Neural Networks Training from Upper Layers to Lower Layers. | Shucong Zhang, Cong-Thanh Do, Rama Doddipatla, Erfan Loweimi, Peter Bell, Steve Renals |
| 2021 | ICASSP | Time-Domain Speech Extraction with Spatial Information and Multi Speaker Conditioning Mechanism. | Jisi Zhang, Catalin Zorila, Rama Doddipatla, Jon Barker |
| 2021 | Interspeech | Teacher-Student MixIT for Unsupervised and Semi-Supervised Speech Separation. | Jisi Zhang, Catalin Zorila, Rama Doddipatla, Jon Barker |
| 2020 | ICASSP | Learning Noise Invariant Features Through Transfer Learning For Robust End-to-End Speech Recognition. | Shucong Zhang, Cong-Thanh Do, Rama Doddipatla, Steve Renals |
| 2020 | ICASSP | On End-to-end Multi-channel Time Domain Speech Separation in Reverberant Environments. | Jisi Zhang, Catalin Zorila, Rama Doddipatla, Jon Barker |
| 2019 | ASRU | An Investigation into the Effectiveness of Enhancement in ASR Training and Test for Chime-5 Dinner Party Transcription. | Catalin Zorila, Christoph Bddeker, Rama Doddipatla, Reinhold Haeb-Umbach |
| 2019 | ICASSP | An Unsupervised Learning Approach to Neural-net-supported Wpe Dereverberation. | Petko Nikolov Petkov, Vasileios Tsiaras, Rama Doddipatla, Yannis Stylianou |
| 2019 | ICASSP | On Reducing the Effect of Speaker Overlap for Chime-5. | Catalin Zorila, Rama Doddipatla |
| 2017 | Interspeech | Speaker Adaptation in DNN-Based Speech Synthesis Using d-Vectors. | Rama Doddipatla, Norbert Braunschweiler, Ranniery Maia |
| 2016 | ICASSP | Speaker adaptive training in deep neural networks using speaker dependent bottleneck features. | Rama Doddipatla |
| 2015 | Interspeech | Noise-matched training of CRF based sentence end detection models. | Madina Hasan, Rama Doddipatla, Thomas Hain |
| 2014 | Interspeech | Speaker dependent bottleneck layer training for speaker adaptation in automatic speech recognition. | Rama Doddipatla, Madina Hasan, Thomas Hain |
| 2014 | Interspeech | Multi-pass sentence-end detection of lecture speech. | Madina Hasan, Rama Doddipatla, Thomas Hain |