| 2024 | ACL | Speech vs. Transcript: Does It Matter for Human Annotators in Speech Summarization? | Roshan Sharma, Suwon Shon, Mark Lindsey, Hira Dhamyal, Bhiksha Raj |
| 2024 | ACL | On the Evaluation of Speech Foundation Models for Spoken Language Understanding. | Siddhant Arora, Ankita Pasad, Chung-Ming Chien, Jionghao Han, Roshan S. Sharma, Jee-weon Jung, Hira Dhamyal, William Chen, Suwon Shon, Hung-yi Lee, Karen Livescu, Shinji Watanabe |
| 2024 | ICASSP | Prompting Audios Using Acoustic Properties for Emotion Representation. | Hira Dhamyal, Benjamin Elizalde, Soham Deshmukh, Huaming Wang, Bhiksha Raj, Rita Singh |
| 2024 | Interspeech | SELM: Enhancing Speech Emotion Recognition for Out-of-Domain Scenarios. | Hazim T. Bukhari, Soham Deshmukh, Hira Dhamyal, Bhiksha Raj, Rita Singh |
| 2024 | NAACL | R-BASS : Relevance-aided Block-wise Adaptation for Speech Summarization. | Roshan Sharma, Ruchira Sharma, Hira Dhamyal, Rita Singh, Bhiksha Raj |
| 2022 | Interspeech | Positional Encoding for Capturing Modality Specific Cadence for Emotion Detection. | Hira Dhamyal, Bhiksha Raj, Rita Singh |
| 2021 | ASRU | Using Self Attention DNNs to Discover Phonemic Features for Audio Deep Fake Detection. | Hira Dhamyal, Ayesha Ali, Ihsan Ayyub Qazi, Agha Ali Raza |
| 2021 | Interspeech | Fake Audio Detection in Resource-Constrained Settings Using Microfeatures. | Hira Dhamyal, Ayesha Ali, Ihsan Ayyub Qazi, Agha Ali Raza |
| 2021 | Interspeech | Masked Proxy Loss for Text-Independent Speaker Verification. | Jiachen Lian, Aiswarya Vinod Kumar, Hira Dhamyal, Bhiksha Raj, Rita Singh |
| 2020 | Interspeech | The Phonetic Bases of Vocal Expressed Emotion: Natural versus Acted. | Hira Dhamyal, Shahan Ali Memon, Bhiksha Raj, Rita Singh |
| 2019 | ASRU | Optimizing Neural Network Embeddings Using a Pair-Wise Loss for Text-Independent Speaker Verification. | Hira Dhamyal, Tianyan Zhou, Bhiksha Raj, Rita Singh |