Skip to content

Rama Doddipatla

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

32

Venues

3

Active years

2014–2024

Best venue rank

Multiconference

Where they publish

Papers

32 indexed papers, newest first.

YearVenueTitleAuthors
2024ICASSPGeodesic Interpolation of Frame-Wise Speaker Embeddings for the Diarization of Meeting Scenarios.Tobias Cord-Landwehr, Christoph Bddeker, Catalin Zorila, Rama Doddipatla, Reinhold Haeb-Umbach
2024InterspeechPrompting Whisper for QA-driven Zero-shot End-to-end Spoken Language Understanding.Mohan Li, Simon Keizer, Rama Doddipatla
2023ASRURobust Recognition of Speaker Emotion With Difference Feature Extraction Using a Few Enrollment Utterances.Daichi Hayakawa, Takehiko Kagoshima, Kenji Iwata, Norbert Braunschweiler, Rama Doddipatla
2023ASRUTowards a Unified End-to-End Language Understanding System for Speech and Text Inputs.Mohan Li, Catalin Zorila, Cong-Thanh Do, Rama Doddipatla
2023ICASSPFrame-Wise and Overlap-Robust Speaker Embeddings for Meeting Diarization.Tobias Cord-Landwehr, Christoph Bddeker, Catalin Zorila, Rama Doddipatla, Reinhold Haeb-Umbach
2023ICASSPCumulative Attention Based Streaming Transformer ASR with Internal Language Model Joint Training and Rescoring.Mohan Li, Cong-Thanh Do, Rama Doddipatla
2023ICASSPOn the Effectiveness of Monoaural Target Source Extraction for Distant end-to-end Automatic Speech Recognition.Catalin Zorila, Rama Doddipatla
2023InterspeechA Teacher-Student Approach for Extracting Informative Speaker Embeddings From Speech Mixtures.Tobias Cord-Landwehr, Christoph Bddeker, Catalin Zorila, Rama Doddipatla, Reinhold Haeb-Umbach
2023InterspeechDomain Adaptive Self-supervised Training of Automatic Speech Recognition.Cong-Thanh Do, Rama Doddipatla, Mohan Li, Thomas Hain
2022ICASSPTransformer-Based Streaming ASR with Cumulative Attention.Mohan Li, Shucong Zhang, Catalin Zorila, Rama Doddipatla
2022ICASSPSpeaker Reinforcement Using Target Source Extraction for Robust Automatic Speech Recognition.Catalin Zorila, Rama Doddipatla
2022InterspeechMultiple-hypothesis RNN-T Loss for Unsupervised Fine-tuning and Self-training of Neural Transducer.Cong-Thanh Do, Mohan Li, Rama Doddipatla
2022InterspeechOn monoaural speech enhancement for automatic recognition of real noisy speech using mixture invariant training.Jisi Zhang, Catalin Zorila, Rama Doddipatla, Jon Barker
2021ASRUA Study on Cross-Corpus Speech Emotion Recognition and Data Augmentation.Norbert Braunschweiler, Rama Doddipatla, Simon Keizer, Svetlana Stoyanchev
2021ASRUDialogue Strategy Adaptation to New Action Sets Using Multi-Dimensional Modelling.Simon Keizer, Norbert Braunschweiler, Svetlana Stoyanchev, Rama Doddipatla
2021ASRUImproving HS-DACS Based Streaming Transformer ASR with Deep Reinforcement Learning.Mohan Li, Rama Doddipatla
2021ICASSPMultiple-Hypothesis CTC-Based Semi-Supervised Adaptation of End-to-End Speech Recognition.Cong-Thanh Do, Rama Doddipatla, Thomas Hain
2021ICASSPHead-Synchronous Decoding for Transformer-Based Streaming ASR.Mohan Li, Catalin Zorila, Rama Doddipatla
2021ICASSPAction State Update Approach to Dialogue Management.Svetlana Stoyanchev, Simon Keizer, Rama Doddipatla
2021ICASSPTrain Your Classifier First: Cascade Neural Networks Training from Upper Layers to Lower Layers.Shucong Zhang, Cong-Thanh Do, Rama Doddipatla, Erfan Loweimi, Peter Bell, Steve Renals
2021ICASSPTime-Domain Speech Extraction with Spatial Information and Multi Speaker Conditioning Mechanism.Jisi Zhang, Catalin Zorila, Rama Doddipatla, Jon Barker
2021InterspeechTeacher-Student MixIT for Unsupervised and Semi-Supervised Speech Separation.Jisi Zhang, Catalin Zorila, Rama Doddipatla, Jon Barker
2020ICASSPLearning Noise Invariant Features Through Transfer Learning For Robust End-to-End Speech Recognition.Shucong Zhang, Cong-Thanh Do, Rama Doddipatla, Steve Renals
2020ICASSPOn End-to-end Multi-channel Time Domain Speech Separation in Reverberant Environments.Jisi Zhang, Catalin Zorila, Rama Doddipatla, Jon Barker
2019ASRUAn Investigation into the Effectiveness of Enhancement in ASR Training and Test for Chime-5 Dinner Party Transcription.Catalin Zorila, Christoph Bddeker, Rama Doddipatla, Reinhold Haeb-Umbach
2019ICASSPAn Unsupervised Learning Approach to Neural-net-supported Wpe Dereverberation.Petko Nikolov Petkov, Vasileios Tsiaras, Rama Doddipatla, Yannis Stylianou
2019ICASSPOn Reducing the Effect of Speaker Overlap for Chime-5.Catalin Zorila, Rama Doddipatla
2017InterspeechSpeaker Adaptation in DNN-Based Speech Synthesis Using d-Vectors.Rama Doddipatla, Norbert Braunschweiler, Ranniery Maia
2016ICASSPSpeaker adaptive training in deep neural networks using speaker dependent bottleneck features.Rama Doddipatla
2015InterspeechNoise-matched training of CRF based sentence end detection models.Madina Hasan, Rama Doddipatla, Thomas Hain
2014InterspeechSpeaker dependent bottleneck layer training for speaker adaptation in automatic speech recognition.Rama Doddipatla, Madina Hasan, Thomas Hain
2014InterspeechMulti-pass sentence-end detection of lecture speech.Madina Hasan, Rama Doddipatla, Thomas Hain