| 2026 | EACL | SrcMix: Mixing of Related Source Languages Benefits Extremely Low-resource Machine Translation. | Sanjeev Kumar, Preethi Jyothi, Pushpak Bhattacharyya |
| 2026 | EACL | Post-ASR Correction in Hindi: Comparing Language Models and Large Language Models in Low-Resource Scenarios. | Rishabh Kumar, Amrith Krishna, Ganesh Ramakrishnan, Preethi Jyothi |
| 2026 | EACL | Improving Language Identification for Code-Switched Speech: The Pivotal Role of Accented English. | Adyasha Patra, Dhiraj Kumar Sah, Preethi Jyothi |
| 2025 | ACL | LexGen: Domain-aware Multilingual Lexicon Generation. | Ayush Maheshwari, Atul Kumar Singh, N. J. Karthika, Krishnakant Bhatt, Preethi Jyothi, Ganesh Ramakrishnan |
| 2025 | ACL | LoFTI: Localization and Factuality Transfer to Indian Locales. | Sona Elza Simon, Soumen Kumar Mondal, Abhishek Singhania, Sayambhu Sen, Preethi Jyothi |
| 2025 | COLING | CoSTA: Code-Switched Speech Translation using Aligned Speech-Text Interleaving. | Bhavani Shankar, Preethi Jyothi, Pushpak Bhattacharyya |
| 2025 | EMNLP | RECAST: Retrieval-Augmented Contextual ASR via Decoder-State Keyword Spotting. | Ashish R. Mittal, Sunita Sarawagi, Preethi Jyothi |
| 2025 | EMNLP | LASER: An LLM-based ASR Scoring and Evaluation Rubric. | Amruta Parulekar, Preethi Jyothi |
| 2025 | EMNLP | DeFT-X: Denoised Sparse Fine-Tuning for Zero-Shot Cross-Lingual Transfer. | Sona Elza Simon, Preethi Jyothi |
| 2025 | Interspeech | Skip-Salsa: Skip Synchronous Fusion of ASR LLM Decoders. | Ashish R. Mittal, Darshan Prabhu, Sunita Sarawagi, Preethi Jyothi |
| 2025 | NAACL | AMPS: ASR with Multimodal Paraphrase Supervision. | Abhishek Gupta, Amruta Parulekar, Sameep Chattopadhyay, Preethi Jyothi |
| 2024 | ACL | Boosting Zero-Shot Crosslingual Performance using LLM-Based Augmentations with Effective Data Selection. | Barah Fazili, Ashish Agrawal, Preethi Jyothi |
| 2024 | ACL | Part-of-speech Tagging for Extremely Low-resource Indian Languages. | Sanjeev Kumar, Preethi Jyothi, Pushpak Bhattacharyya |
| 2024 | ACL | DIMSIM: Distilled Multilingual Critics for Indic Text Simplification. | Sneha Mondal, Ritika, Ashish Agrawal, Preethi Jyothi, Aravindan Raghuveer |
| 2024 | ACL | In-context Mixing (ICM): Code-mixed Prompts for Multilingual LLMs. | Bhavani Shankar, Preethi Jyothi, Pushpak Bhattacharyya |
| 2024 | EACL | Translation Errors Significantly Impact Low-Resource Languages in Cross-Lingual Learning. | Ashish Sunil Agrawal, Barah Fazili, Preethi Jyothi |
| 2024 | EACL | STORiCo: Storytelling TTS for Hindi with Character Voice Modulation. | Pavan Tankala, Preethi Jyothi, Preeti Rao, Pushpak Bhattacharyya |
| 2024 | EMNLP | DictDis: Dictionary Constrained Disambiguation for Improved NMT. | Ayush Maheshwari, Preethi Jyothi, Ganesh Ramakrishnan |
| 2024 | Interspeech | Emotion Arithmetic: Emotional Speech Synthesis via Weight Space Interpolation. | Pavan Kalyan, Preeti Rao, Preethi Jyothi, Pushpak Bhattacharyya |
| 2024 | Interspeech | SALSA: Speedy ASR-LLM Synchronous Aggregation. | Ashish R. Mittal, Darshan Prabhu, Sunita Sarawagi, Preethi Jyothi |
| 2024 | Interspeech | Improving Self-supervised Pre-training using Accent-Specific Codebooks. | Darshan Prabhu, Abhishek Gupta, Omkar Nitsure, Preethi Jyothi, Sriram Ganapathy |
| 2024 | Interspeech | MULTI-CONVFORMER: Extending Conformer with Multiple Convolution Kernels. | Darshan Prabhu, Yifan Peng, Preethi Jyothi, Shinji Watanabe |
| 2023 | ACL | Adversarial Training for Low-Resource Disfluency Correction. | Vineet Bhat, Preethi Jyothi, Pushpak Bhattacharyya |
| 2023 | ACL | Improving Pretraining Techniques for Code-Switched NLP. | Richeek Das, Sahasra Ranjan, Shreya Pathak, Preethi Jyothi |
| 2023 | ACL | Zero-shot Cross-lingual Transfer With Learned Projections Using Unlabeled Target-Language Data. | Ujan Deb, Ridayesh Parab, Preethi Jyothi |
| 2023 | ACL | DITTO: Data-efficient and Fair Targeted Subset Selection for ASR Accent Adaptation. | Suraj Kothawade, Anmol Reddy Mekala, D. Chandra Sekhara Hetha Havya, Mayank Kothyari, Rishabh K. Iyer, Ganesh Ramakrishnan, Preethi Jyothi |
| 2023 | EMNLP | DISCO: A Large Scale Human Annotated Corpus for Disfluency Correction in Indo-European Languages. | Vineet Bhat, Preethi Jyothi, Pushpak Bhattacharyya |
| 2023 | EMNLP | Speech-enriched Memory for Inference-time Adaptation of ASR Models to Word Dictionaries. | Ashish R. Mittal, Sunita Sarawagi, Preethi Jyothi, George Saon, Gakuto Kurata |
| 2023 | EMNLP | Accented Speech Recognition With Accent-specific Codebooks. | Darshan Prabhu, Preethi Jyothi, Sriram Ganapathy, Vinit Unni |
| 2023 | ICASSP | Towards Zero-Shot Code-Switched Speech Recognition. | Brian Yan, Matthew Wiesner, Ondrej Klejch, Preethi Jyothi, Shinji Watanabe |
| 2023 | ICLR | In-Situ Text-Only Adaptation of Speech Models with Low-Overhead Speech Imputations. | Ashish R. Mittal, Sunita Sarawagi, Preethi Jyothi |
| 2023 | IJCAI | Temporally Aligning Long Audio Interviews with Questions: A Case Study in Multimodal Data Integration. | Piyush Singh Pasi, Karthikeya Battepati, Preethi Jyothi, Ganesh Ramakrishnan, Tanmay Mahapatra, Manoj Singh |
| 2023 | Interspeech | DisfluencyFixer: A tool to enhance Language Learning through Speech To Speech Disfluency Correction. | Vineet Bhat, Preethi Jyothi, Pushpak Bhattacharyya |
| 2023 | Interspeech | Unsupervised Code-switched Text Generation from Parallel Text. | Jie Chi, Brian Lu, Jason Eisner, Peter Bell, Preethi Jyothi, Ahmed M. Ali |
| 2023 | Interspeech | Narrator or Character: Voice Modulation in an Expressive Multi-speaker TTS. | Tankala Pavan Kalyan, Preeti Rao, Preethi Jyothi, Pushpak Bhattacharyya |
| 2023 | Interspeech | Improving RNN-Transducers with Acoustic LookAhead. | Vinit S. Unni, Ashish R. Mittal, Preethi Jyothi, Sunita Sarawagi |
| 2022 | ACL | Accurate Online Posterior Alignments for Principled Lexically-Constrained Decoding. | Soumya Chatterjee, Sunita Sarawagi, Preethi Jyothi |
| 2022 | COLING | Aligning Multilingual Embeddings for Improved Code-switched Natural Language Understanding. | Barah Fazili, Preethi Jyothi |
| 2022 | COLING | Zero-shot Disfluency Detection for Indian Languages. | Rohit Kundu, Preethi Jyothi, Pushpak Bhattacharyya |
| 2022 | EMNLP | Partitioned Gradient Matching-based Data Subset Selection for Compute-Efficient Robust ASR Training. | Ashish R. Mittal, Durga Sivasubramanian, Rishabh K. Iyer, Preethi Jyothi, Ganesh Ramakrishnan |
| 2022 | EMNLP | CoCoa: An Encoder-Decoder Model for Controllable Code-switched Generation. | Sneha Mondal, Ritika, Shreya Pathak, Preethi Jyothi, Aravindan Raghuveer |
| 2022 | ICASSP | Adaptive Discounting of Implicit Language Models in RNN-Transducers. | Vinit Unni, Shreya Khare, Ashish R. Mittal, Preethi Jyothi, Sunita Sarawagi, Samarth Bharadwaj |
| 2022 | Interspeech | SPLICEOUT: A Simple and Efficient Audio Augmentation Method. | Arjit Jain, Pranay Reddy Samala, Deepak Mittal, Preethi Jyothi, Maneesh Singh |
| 2022 | Interspeech | VAgyojaka: An Annotating and Post-Editing Tool for Automatic Speech Recognition. | Rishabh Kumar, Devaraja Adiga, Mayank Kothyari, Jatin Dalal, Ganesh Ramakrishnan, Preethi Jyothi |
| 2022 | Interspeech | Linguistically Informed Post-processing for ASR Error correction in Sanskrit. | Rishabh Kumar, Devaraja Adiga, Rishav Ranjan, Amrith Krishna, Ganesh Ramakrishnan, Pawan Goyal, Preethi Jyothi |
| 2021 | ACL | Automatic Speech Recognition in Sanskrit: A New Speech Corpus and Modelling Insights. | Devaraja Adiga, Rishabh Kumar, Amrith Krishna, Preethi Jyothi, Ganesh Ramakrishnan, Pawan Goyal |
| 2021 | ACL | From Machine Translation to Code-Switching: Generating High-Quality Code-Switched Text. | Ishan Tarunesh, Syamantak Kumar, Preethi Jyothi |
| 2021 | EACL | Disfluency Correction using Unsupervised and Semi-supervised Learning. | Nikhil Saini, Drumil Trivedi, Shreya Khare, Tejas I. Dhamecha, Preethi Jyothi, Samarth Bharadwaj, Pushpak Bhattacharyya |
| 2021 | EACL | Meta-Learning for Effective Multi-task and Multilingual Modelling. | Ishan Tarunesh, Sushil Khyalia, Vishwajeet Kumar, Ganesh Ramakrishnan, Preethi Jyothi |
| 2021 | ICASSP | Error-Driven Fixed-Budget ASR Personalization for Accented Speakers. | Abhijeet Awasthi, Aman Kansal, Sunita Sarawagi, Preethi Jyothi |
| 2021 | ICASSP | Collaborative Learning to Generate Audio-Video Jointly. | Vinod K. Kurmi, Vipul Bajaj, Badri N. Patro, K. S. Venkatesh, Vinay P. Namboodiri, Preethi Jyothi |
| 2021 | ICASSP | An Investigation of End-to-End Models for Robust Speech Recognition. | Archiki Prasad, Preethi Jyothi, Rajbabu Velmurugan |
| 2021 | ICMI | Cross Lingual Video and Text Retrieval: A New Benchmark Dataset and Algorithm. | Jayaprakash Akula, Abhishek Sharma, Rishabh Dabral, Preethi Jyothi, Ganesh Ramakrishnan |
| 2021 | IJCAI | Perturb, Predict & Paraphrase: Semi-Supervised Learning using Noisy Student for Image Captioning. | Arjit Jain, Pranay Reddy Samala, Preethi Jyothi, Deepak Mittal, Maneesh Kumar Singh |
| 2021 | Interspeech | Reduce and Reconstruct: ASR for Low-Resource Phonetic Languages. | Anuj Diwan, Preethi Jyothi |
| 2021 | Interspeech | MUCS 2021: Multilingual and Code-Switching ASR Challenges for Low Resource Indian Languages. | Anuj Diwan, Rakesh Vaideeswaran, Sanket Shah, Ankita Singh, Srinivasa Raghavan K. M., Shreya Khare, Vinit Unni, Saurabh Vyas, Akash Rajpuria, Chiranjeevi Yarra, Ashish R. Mittal, Prasanta Kumar Ghosh, Preethi Jyothi, Kalika Bali, Vivek Seshadri, Sunayana Sitaram, Samarth Bharadwaj, Jai Nanavati, Raoul Nanavati, Karthik Sankaranarayanan |
| 2021 | Interspeech | Low Resource ASR: The Surprising Effectiveness of High Resource Transliteration. | Shreya Khare, Ashish R. Mittal, Anuj Diwan, Sunita Sarawagi, Preethi Jyothi, Samarth Bharadwaj |
| 2021 | Interspeech | Cross-Modal Learning for Audio-Visual Video Parsing. | Jatin Lamba, Abhishek, Jayaprakash Akula, Rishabh Dabral, Preethi Jyothi, Ganesh Ramakrishnan |
| 2021 | SIGIR | Select, Substitute, Search: A New Benchmark for Knowledge-Augmented Visual Question Answering. | Aman Jain, Mayank Kothyari, Vishwajeet Kumar, Preethi Jyothi, Ganesh Ramakrishnan, Soumen Chakrabarti |
| 2020 | ACL | How Accents Confound: Probing for Accent Information in End-to-End Speech Recognition Systems. | Archiki Prasad, Preethi Jyothi |
| 2020 | ICASSP | Coupled Training of Sequence-to-Sequence Models for Accented Speech Recognition. | Vinit Unni, Nitish Joshi, Preethi Jyothi |
| 2020 | Interspeech | Black-Box Adaptation of ASR for Accented Speech. | Kartik Khandelwal, Preethi Jyothi, Abhijeet Awasthi, Sunita Sarawagi |
| 2020 | Interspeech | Caption Alignment for Low Resource Audio-Visual Data. | Vighnesh Reddy Konda, Mayur Warialani, Rakesh Prasanth Achari, Varad Bhatnagar, Jayaprakash Akula, Preethi Jyothi, Ganesh Ramakrishnan, Gholamreza Haffari, Pankaj Singh |
| 2020 | Interspeech | Improving Low Resource Code-Switched ASR Using Augmented Code-Switched TTS. | Yash Sharma, Basil Abraham, Karan Taneja, Preethi Jyothi |
| 2020 | LREC | Crowdsourcing Speech Data for Low-Resource Languages from Low-Income Workers. | Basil Abraham, Danish Goel, Divya Siddarth, Kalika Bali, Manu Chopra, Monojit Choudhury, Pratik Joshi, Preethi Jyothi, Sunayana Sitaram, Vivek Seshadri |
| 2019 | ACL | Cross-Lingual Training for Automatic Question Generation. | Vishwajeet Kumar, Nitish Joshi, Arijit Mukherjee, Ganesh Ramakrishnan, Preethi Jyothi |
| 2019 | Interspeech | Exploiting Monolingual Speech Corpora for Code-Mixed Speech Recognition. | Karan Taneja, Satarupa Guha, Preethi Jyothi, Basil Abraham |
| 2018 | EMNLP | Code-switched Language Models Using Dual RNNs and Same-Source Pretraining. | Saurabh Garg, Tanmay Parekh, Preethi Jyothi |
| 2018 | EMNLP | Revisiting the Importance of Encoding Logic Rules in Sentiment Classification. | Kalpesh Krishna, Preethi Jyothi, Mohit Iyyer |
| 2018 | ICLR | Generalizing Across Domains via Cross-Gradient Training. | Shiv Shankar, Vihari Piratla, Soumen Chakrabarti, Siddhartha Chaudhuri, Preethi Jyothi, Sunita Sarawagi |
| 2018 | Interspeech | Dual Language Models for Code Switched Speech Recognition. | Saurabh Garg, Tanmay Parekh, Preethi Jyothi |
| 2018 | Interspeech | Improved Accented Speech Recognition Using Accent Embeddings and Multi-task Learning. | Abhinav Jain, Minali Upreti, Preethi Jyothi |
| 2018 | Interspeech | Time Aggregation Operators for Multi-label Audio Event Detection. | Pankaj Joshi, Digvijaysingh Gautam, Ganesh Ramakrishnan, Preethi Jyothi |
| 2017 | ACSSC | Mismatched crowdsourcing: Mining latent skills to acquire speech transcriptions. | Mark Hasegawa-Johnson, Preethi Jyothi, Wenda Chen, Van Hai Do |
| 2017 | ASRU | Leveraging native language speech for accent identification using deep Siamese networks. | Aditya Siddhant, Preethi Jyothi, Sriram Ganapathy |
| 2017 | ICASSP | Low-resource grapheme-to-phoneme conversion using recurrent neural networks. | Preethi Jyothi, Mark Hasegawa-Johnson |
| 2016 | ICASSP | Adapting ASR for under-resourced languages using mismatched transcriptions. | Chunxi Liu, Preethi Jyothi, Hao Tang, Vimal Manohar, Rose Sloan, Tyler Kekona, Mark Hasegawa-Johnson, Sanjeev Khudanpur |
| 2016 | Interspeech | Automatic Speech Recognition Using Probabilistic Transcriptions in Swahili, Amharic, and Dinka. | Amit Das, Preethi Jyothi, Mark Hasegawa-Johnson |
| 2016 | ITA | Language coverage for mismatched crowdsourcing. | Lav R. Varshney, Preethi Jyothi, Mark Hasegawa-Johnson |
| 2015 | AAAI | Acquiring Speech Transcriptions Using Mismatched Crowdsourcing. | Preethi Jyothi, Mark Hasegawa-Johnson |
| 2015 | Interspeech | Transcribing continuous speech using mismatched crowdsourcing. | Preethi Jyothi, Mark Hasegawa-Johnson |
| 2015 | Interspeech | Improved hindi broadcast ASR by adapting the language model and pronunciation model using a priori syntactic and morphophonemic knowledge. | Preethi Jyothi, Mark Hasegawa-Johnson |
| 2013 | Interspeech | Discriminative training of WFST factors with application to pronunciation modeling. | Preethi Jyothi, Eric Fosler-Lussier, Karen Livescu |
| 2012 | ICASSP | Distributed discriminative language models for Google voice-search. | Preethi Jyothi, Leif Johnson, Ciprian Chelba, Brian Strope |
| 2012 | Interspeech | Discriminatively learning factorized finite state pronunciation models from dynamic Bayesian networks. | Preethi Jyothi, Eric Fosler-Lussier, Karen Livescu |
| 2012 | NAACL | Large-scale discriminative language model reranking for voice-search. | Preethi Jyothi, Leif Johnson, Ciprian Chelba, Brian Strope |
| 2011 | ICASSP | Lexical access experiments with context-dependent articulatory feature-based models. | Preethi Jyothi, Karen Livescu, Eric Fosler-Lussier |
| 2010 | Interspeech | Discriminative language modeling using simulated ASR errors. | Preethi Jyothi, Eric Fosler-Lussier |
| 2010 | NAACL | Investigations into the Crandem Approach to Word Recognition. | Rohit Prabhavalkar, Preethi Jyothi, William Hartmann, Jeremy Morris, Eric Fosler-Lussier |
| 2009 | Interspeech | A comparison of audio-free speech recognition error prediction methods. | Preethi Jyothi, Eric Fosler-Lussier |