| 2024 | ICASSP | Leveraging Large Language Models for Exploiting ASR Uncertainty. | Pranay Dighe, Yi Su, Shangshang Zheng, Yunshu Liu, Vineet Garg, Xiaochuan Niu, Ahmed H. Tewfik |
| 2024 | ICASSP | Modality Drop-Out for Multimodal Device Directed Speech Detection Using Verbal and Non-Verbal Features. | Gautam Krishna, Sameer Dharur, Oggi Rudovic, Pranay Dighe, Saurabh Adya, Ahmed Hussen Abdelaziz, Ahmed H. Tewfik |
| 2023 | ICASSP | Audio-to-Intent Using Acoustic-Textual Subword Representations from End-to-End ASR. | Pranay Dighe, Prateeth Nayak, Oggi Rudovic, Erik Marchi, Xiaochuan Niu, Ahmed H. Tewfik |
| 2023 | ICASSP | Less Is More: A Unified Architecture for Device-Directed Speech Detection with Multiple Invocation Types. | Oggi Rudovic, Wonil Chang, Vineet Garg, Pranay Dighe, Pramod Simha, Jack Berkowitz, Ahmed Hussen Abdelaziz, Sachin Kajarekar, Erik Marchi, Saurabh Adya |
| 2022 | ICASSP | Streaming on-Device Detection of Device Directed Speech from Voice and Touch-Based Invocation. | Ognjen (Oggi) Rudovic, Akanksha Bindal, Vineet Garg, Pramod Simha, Pranay Dighe, Sachin Kajarekar |
| 2022 | Interspeech | Device-Directed Speech Detection: Regularization via Distillation for Weakly-Supervised Models. | Vineet Garg, Ognjen Rudovic, Pranay Dighe, Ahmed Hussen Abdelaziz, Erik Marchi, Saurabh Adya, Chandra Dhir, Ahmed H. Tewfik |
| 2021 | ICASSP | Knowledge Transfer for Efficient on-Device False Trigger Mitigation. | Pranay Dighe, Erik Marchi, Srikanth Vishnubhotla, Sachin Kajarekar, Devang Naik |
| 2021 | Interspeech | Streaming Transformer for Hardware Efficient Voice Trigger Detection and False Trigger Mitigation. | Vineet Garg, Wonil Chang, Siddharth Sigtia, Saurabh Adya, Pramod Simha, Pranay Dighe, Chandra Dhir |
| 2020 | ICASSP | Lattice-Based Improvements for Voice Triggering Using Graph Neural Networks. | Pranay Dighe, Saurabh Adya, Nuoyu Li, Srikanth Vishnubhotla, Devang Naik, Adithya Sagar, Ying Ma, Stephen Pulman, Jason D. Williams |
| 2020 | Interspeech | Complementary Language Model and Parallel Bi-LRNN for False Trigger Mitigation. | Rishika Agarwal, Xiaochuan Niu, Pranay Dighe, Srikanth Vishnubhotla, Sameer Badaskar, Devang Naik |
| 2019 | ICASSP | Analyzing Uncertainties in Speech Recognition Using Dropout. | Apoorv Vyas, Pranay Dighe, Sibo Tong, Herv Bourlard |
| 2017 | ICASSP | Low-rank and sparse soft targets to learn better DNN acoustic models. | Pranay Dighe, Afsaneh Asaei, Herv Bourlard |
| 2017 | Interspeech | Exploiting Eigenposteriors for Semi-Supervised Training of DNN Acoustic Models with Sequence Discrimination. | Pranay Dighe, Afsaneh Asaei, Herv Bourlard |
| 2016 | ICASSP | Exploiting low-dimensional structures to enhance DNN based acoustic modeling in speech recognition. | Pranay Dighe, Gil Luyet, Afsaneh Asaei, Herv Bourlard |
| 2016 | Interspeech | Low-Rank Representation of Nearest Neighbor Posterior Probabilities to Enhance DNN Based Acoustic Modeling. | Gil Luyet, Pranay Dighe, Afsaneh Asaei, Herv Bourlard |
| 2015 | Interspeech | Sparse modeling of posterior exemplars for keyword detection. | Dhananjay Ram, Afsaneh Asaei, Pranay Dighe, Herv Bourlard |
| 2014 | Interspeech | Detecting and labeling speakers on overlapping speech using vector taylor series. | Pranay Dighe, Marc Ferras, Herv Bourlard |
| 2012 | ICASSP | Audio event detection from acoustic unit occurrence patterns. | Anurag Kumar, Pranay Dighe, Rita Singh, Sourish Chaudhuri, Bhiksha Raj |
| 2012 | Interspeech | Language identification using spectro-temporal patch features. | Kamal Sahni, Pranay Dighe, Rita Singh, Bhiksha Raj |