Skip to content

Rita Singh

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

93

Venues

17

Active years

1998–2025

Best venue rank

A*

Where they publish

Papers

93 indexed papers, newest first.

YearVenueTitleAuthors
2025AAAIAudio Entailment: Assessing Deductive Reasoning for Audio Understanding.Soham Deshmukh, Shuo Han, Hazim T. Bukhari, Benjamin Elizalde, Hannes Gamper, Rita Singh, Bhiksha Raj
2025ACLLost in Transcription, Found in Distribution Shift: Demystifying Hallucination in Speech Foundation Models.Hanin Atwany, Abdul Waheed, Rita Singh, Monojit Choudhury, Bhiksha Raj
2025ACLOn the Robust Approximation of ASR Metrics.Abdul Waheed, Hanin Atwany, Rita Singh, Bhiksha Raj
2025ASRUCoLMbo: Speaker Language Model for Descriptive Profiling.Massa Baali, Shuo Han, Syed Abdul Hannan, Purusottam Samal, Karanveer Singh, Soham Deshmukh, Rita Singh, Bhiksha Raj
2025CIKMPlaceSim: An LLM-based Interactive Platform for Human Behavior Simulation in Physical Facilities.Suhyeon Lee, Youngjun Yu, Donghyuk Shin, Rita Singh
2025EMNLPSVeritas: Benchmark for Robust Speaker Verification under Diverse Conditions.Massa Baali, Sarthak Bisht, Francisco Teixeira, Kateryna Shapovalenko, Rita Singh, Bhiksha Raj
2025EMNLPCAARMA: Class Augmentation with Adversarial Mixup Regularization.Massa Baali, Xiang Li, Hao Chen, Syed Abdul Hannan, Rita Singh, Bhiksha Raj
2025EMNLPPhoniTale: Phonologically Grounded Mnemonic Generation for Typologically Distant Language Pairs.Sana Kang, Myeongseok Gwon, Su Young Kwon, Jaewook Lee, Andrew Lan, Bhiksha Raj, Rita Singh
2025ICASSPTessellated Linear Model for Age Prediction from Voice.Dareen Alharthi, Mahsa Zamani, Bhiksha Raj, Rita Singh
2025ICCVDoes Prior Data Matter? Exploring Joint Training in the Context of Few-Shot Class-Incremental Learning.Shiwon Kim, Dongjun Hwang, Sungwon Woo, Rita Singh
2025ICLRADIFF: Explaining audio difference using natural language.Soham Deshmukh, Shuo Han, Rita Singh, Bhiksha Raj
2024CVPRQDFormer: Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition.Xiang Li, Jinglu Wang, Xiaohao Xu, Xiulian Peng, Rita Singh, Yan Lu, Bhiksha Raj
2024ECCVRXiang Li, Kai Qiu, Jinglu Wang, Xiaohao Xu, Rita Singh, Kashu Yamazaki, Hao Chen, Xiaonan Huang, Bhiksha Raj
2024ICASSPImportance of Negative Sampling in Weak Label Learning.Ankit Shah, Fuyu Tang, Zelin Ye, Rita Singh, Bhiksha Raj
2024ICASSPTraining Audio Captioning Models without Audio.Soham Deshmukh, Benjamin Elizalde, Dimitra Emmanouilidou, Bhiksha Raj, Rita Singh, Huaming Wang
2024ICASSPPrompting Audios Using Acoustic Properties for Emotion Representation.Hira Dhamyal, Benjamin Elizalde, Soham Deshmukh, Huaming Wang, Bhiksha Raj, Rita Singh
2024ICASSPVocal Fold Dynamics for Automatic Detection of Amyotrophic Lateral Sclerosis from Voice.Jiayi Zhang, Rita Singh
2024ICMLA General Framework for Learning from Weak Supervision.Hao Chen, Jindong Wang, Lei Feng, Xiang Li, Yidong Wang, Xing Xie, Masashi Sugiyama, Rita Singh, Bhiksha Raj
2024ICMLCompleting Visual Objects via Bridging Generation and Segmentation.Xiang Li, Yinpeng Chen, Chung-Ching Lin, Hao Chen, Kai Hu, Rita Singh, Bhiksha Raj, Lijuan Wang, Zicheng Liu
2024InterspeechSELM: Enhancing Speech Emotion Recognition for Out-of-Domain Scenarios.Hazim T. Bukhari, Soham Deshmukh, Hira Dhamyal, Bhiksha Raj, Rita Singh
2024InterspeechPAM: Prompting Audio-Language Models for Audio Quality Assessment.Soham Deshmukh, Dareen Alharthi, Benjamin Elizalde, Hannes Gamper, Mahmoud Al Ismail, Rita Singh, Bhiksha Raj, Huaming Wang
2024InterspeechDomain Adaptation for Contrastive Audio-Language Models.Soham Deshmukh, Rita Singh, Bhiksha Raj
2024NAACLR-BASS : Relevance-aided Block-wise Adaptation for Speech Summarization.Roshan Sharma, Ruchira Sharma, Hira Dhamyal, Rita Singh, Bhiksha Raj
2023ASRUEspnet-Summ: Introducing a Novel Large Dataset, Toolkit, and a Cross-Corpora Evaluation of Speech Summarization Systems.Roshan S. Sharma, William Chen, Takatomo Kano, Ruchira Sharma, Siddhant Arora, Shinji Watanabe, Atsunori Ogawa, Marc Delcroix, Rita Singh, Bhiksha Raj
2023EMNLPToken Prediction as Implicit Classification to Identify LLM-Generated Text.Yutian Chen, Hao Kang, Vivian Zhai, Liangze Li, Rita Singh, Bhiksha Raj
2023EMNLPTowards Noise-Tolerant Speech-Referring Video Object Segmentation: Bridging Speech and Text.Xiang Li, Jinglu Wang, Xiaohao Xu, Muqiao Yang, Fan Yang, Yizhou Zhao, Rita Singh, Bhiksha Raj
2023ICCVPairwise Similarity Learning is SimPLE.Yandong Wen, Weiyang Liu, Yao Feng, Bhiksha Raj, Rita Singh, Adrian Weller, Michael J. Black, Bernhard Schlkopf
2023InterspeechBASS: Block-wise Adaptation for Speech Summarization.Roshan Sharma, Siddhant Arora, Kenneth Zheng, Shinji Watanabe, Rita Singh, Bhiksha Raj
2023InterspeechThe Hidden Dance of Phonemes and Visage: Unveiling the Enigmatic Link between Phonemes and Facial Features.Liao Qu, Xianwei Zou, Xiang Li, Yandong Wen, Rita Singh, Bhiksha Raj
2022ICLRSphereFace2: Binary Classification is All You Need for Deep Face Recognition.Yandong Wen, Weiyang Liu, Adrian Weller, Bhiksha Raj, Rita Singh
2022InterspeechPositional Encoding for Capturing Modality Specific Cadence for Emotion Detection.Hira Dhamyal, Bhiksha Raj, Rita Singh
2021ICASSPInterpreting Glottal Flow Dynamics for Detecting Covid-19 From Voice.Soham Deshmukh, Mahmoud Al Ismail, Rita Singh
2021ICASSPDetection of Covid-19 Through the Analysis of Vocal Fold Oscillations.Mahmoud Al Ismail, Soham Deshmukh, Rita Singh
2021ICCVSelf-Supervised 3D Face Reconstruction via Conditional Estimation.Yandong Wen, Weiyang Liu, Bhiksha Raj, Rita Singh
2021InterspeechImproving Weakly Supervised Sound Event Detection with Self-Supervised Auxiliary Tasks.Soham Deshmukh, Bhiksha Raj, Rita Singh
2021InterspeechGeneralized Spoofing Detection Inspired from Audio Generation Artifacts.Yang Gao, Tyler Vuong, Mahsa Elyasi, Gaurav Bharaj, Rita Singh
2021InterspeechMasked Proxy Loss for Text-Independent Speaker Verification.Jiachen Lian, Aiswarya Vinod Kumar, Hira Dhamyal, Bhiksha Raj, Rita Singh
2020ICASSPSpeech-Based Parameter Estimation of an Asymmetric Vocal Fold Oscillation Model and its Application in Discriminating Vocal Fold Pathologies.Wenbo Zhao, Rita Singh
2020ICPRHierarchical Routing Mixture of Experts.Wenbo Zhao, Yang Gao, Shahan Ali Memon, Bhiksha Raj, Rita Singh
2020InterspeechThe Phonetic Bases of Vocal Expressed Emotion: Natural versus Acted.Hira Dhamyal, Shahan Ali Memon, Bhiksha Raj, Rita Singh
2020InterspeechHide and Speak: Towards Deep Neural Networks for Speech Steganography.Felix Kreuk, Yossi Adi, Bhiksha Raj, Rita Singh, Joseph Keshet
2020ISVCControlled AutoEncoders to Generate Faces from Voices.Hao Liang, Lulan Yu, Guikang Xu, Bhiksha Raj, Rita Singh
2019ASRUOptimizing Neural Network Embeddings Using a Pair-Wise Loss for Text-Independent Speaker Verification.Hira Dhamyal, Tianyan Zhou, Bhiksha Raj, Rita Singh
2019ICASSPHuman Behaviour Recognition Using Wifi Channel State Information.Daanish Ali Khan, Saquib Razak, Bhiksha Raj, Rita Singh
2019ICLRDisjoint Mapping Network for Cross-modal Matching of Voices and Faces.Yandong Wen, Mahmoud Al Ismail, Weiyang Liu, Bhiksha Raj, Rita Singh
2019IJCNNNeural Regression Trees.Shahan Ali Memon, Wenbo Zhao, Bhiksha Raj, Rita Singh
2018ICASSPVoice Impersonation Using Generative Adversarial Networks.Yang Gao, Rita Singh, Bhiksha Raj
2018ICASSPA Corrective Learning Approach for Text-Independent Speaker Verification.Yandong Wen, Tianyan Zhou, Rita Singh, Bhiksha Raj
2017ICASSPSupervised monaural source separation based on autoencoders.Keiichi Osako, Yuki Mitsufuji, Rita Singh, Bhiksha Raj
2016ICASSPThe relationship of voice onset time and Voice Offset Time to physical age.Rita Singh, Joseph Keshet, Deniz Genaga, Bhiksha Raj
2016InterspeechEstimation of Children's Physical Characteristics from Their Voices.Jill Fain Lehman, Rita Singh
2015ICASSPFree energy for speech recognition.Rita Singh, Ken'ichi Kumatani
2015InterspeechKeyword spotting in multi-player voice driven games for children.Sundar Harshavardhan, Jill Fain Lehman, Rita Singh
2014IC2EAudio Classification with Thermodynamic Criteria.Rita Singh
2013ICASSPJoint constrained maximum likelihood regression for overlapping speech recognition.Ken'ichi Kumatani, Rita Singh, Friedrich Faubel, John W. McDonough, Youssef Oualil
2013InterspeechDiscriminatively trained dependency language modeling for conversational speech recognition.Benjamin Lambert, Bhiksha Raj, Rita Singh
2012ICASSPSpectrographic seam patterns for discriminative word spotting.Shubhranshu Barnwal, Kamal Sahni, Rita Singh, Bhiksha Raj
2012ICASSPAudio event detection from acoustic unit occurrence patterns.Anurag Kumar, Pranay Dighe, Rita Singh, Sourish Chaudhuri, Bhiksha Raj
2012ICASSPCompensating for denoising artifacts.Rita Singh
2012InterspeechExploiting Temporal Sequence Structure for Semantic Analysis of Multimedia.Sourish Chaudhuri, Rita Singh, Bhiksha Raj
2012InterspeechPlagiarism Detection in Polyphonic Music using Monaural Signal Separation.Soham De, Indradyumna Roy, Tarunima Prabhakar, Kriti Suneja, Sourish Chaudhuri, Rita Singh, Bhiksha Raj
2012InterspeechMicrophone Array Post-filter based on Spatially-Correlated Noise Measurements for Distant Speech Recognition.Ken'ichi Kumatani, Bhiksha Raj, Rita Singh, John W. McDonough
2012InterspeechLanguage identification using spectro-temporal patch features.Kamal Sahni, Pranay Dighe, Rita Singh, Bhiksha Raj
2012InterspeechA signal-separation-based array postfilter for distant speech recognition.Rita Singh, Ken'ichi Kumatani, John W. McDonough, Chen Liu
2011ICASSPAn iterative least-squares technique for dereverberation.Kshitiz Kumar, Bhiksha Raj, Rita Singh, Richard M. Stern
2011ICASSPGammatone sub-band magnitude-domain dereverberation for ASR.Kshitiz Kumar, Rita Singh, Bhiksha Raj, Richard M. Stern
2011ICASSPA paired test for recognizer selection with untranscribed data.Bhiksha Raj, Rita Singh, James Baker
2011InterspeechPhoneme-Dependent NMF for Speech Enhancement in Monaural Mixtures.Bhiksha Raj, Rita Singh, Tuomas Virtanen
2010ICASSPLatent-variable decomposition based dereverberation of monaural and multi-channel signals.Rita Singh, Bhiksha Raj, Paris Smaragdis
2010InterspeechCreating a linguistic plausibility dataset with non-expert annotators.Benjamin Lambert, Rita Singh, Bhiksha Raj
2010InterspeechNon-negative matrix factorization based compensation of music for automatic speech recognition.Bhiksha Raj, Tuomas Virtanen, Sourish Chaudhuri, Rita Singh
2010InterspeechThe use of sense in unsupervised training of acoustic models for ASR systems.Rita Singh, Benjamin Lambert, Bhiksha Raj
2009ICASSPA joint decoding algorithm for multiple-example-based addition of words to a pronunciation lexicon.Dhananjay Bansal, Nishanth Ulhas Nair, Rita Singh, Bhiksha Raj
2007ICASSPBandwidth Expansionwith a plya URN Model.Bhiksha Raj, Rita Singh, Madhusudana V. S. Shashanka, Paris Smaragdis
2007InterspeechProbabilistic deduction of symbol mappings for extension of lexicons.Rita Singh, Evandro B. Gouva, Bhiksha Raj
2005InterspeechRecognizing speech from simultaneous speakers.Bhiksha Raj, Rita Singh, Paris Smaragdis
2004ICASSPOn tracking noise with linear dynamical system models.Bhiksha Raj, Rita Singh, Richard M. Stern
2004InterspeechMaximum - likelihod adaptation of semi-continuous HMMs by latent variable decomposition of state distributions.Antoine Raux, Rita Singh
2003ICASSPTracking noise via dynamical systems with a continuum of states.Rita Singh, Bhiksha Raj
2003InterspeechDesign of the CMU sphinx-4 decoder.Paul Lamere, Philip Kwok, William Walker, Evandro B. Gouva, Rita Singh, Bhiksha Raj, Peter Wolf
2003InterspeechClassification with free energy at raised temperatures.Rita Singh, Manfred K. Warmuth, Bhiksha Raj, Paul Lamere
2002InterspeechRapid development of speech-to-speech translation systems.Alan W. Black, Ralf D. Brown, Robert E. Frederking, Kevin A. Lenzo, John Moody, Alexander I. Rudnicky, Rita Singh, Eric Steinbrecher
2002InterspeechCombining search spaces of heterogeneous recognizers for improved speech recogniton.Xiang Li, Rita Singh, Richard M. Stern
2001ICASSPTandem acoustic modeling in large-vocabulary recognition.Daniel P. W. Ellis, Rita Singh, Sunil Sivadas
2001ICASSPSpeech in Noisy Environments: robust automatic segmentation, feature extraction, and hypothesis combination.Rita Singh, Michael L. Seltzer, Bhiksha Raj, Richard M. Stern
2000ICASSPAutomatic generation of phone sets and lexical transcriptions.Rita Singh, Bhiksha Raj, Richard M. Stern
2000InterspeechPhone transition acoustic modeling: application to speaker independent and spontaneous speech systems.Jon P. Nedel, Rita Singh, Richard M. Stern
2000InterspeechAutomatic subword unit refinement for spontaneous speech recognition via phone splitting.Jon P. Nedel, Rita Singh, Richard M. Stern
2000InterspeechTask and domain specific modelling in the Carnegie Mellon communicator system.Alexander I. Rudnicky, Christina L. Bennett, Alan W. Black, Ananlada Chotimongkol, Kevin A. Lenzo, Alice Oh, Rita Singh
2000InterspeechStructured redefinition of sound units by merging and splitting for improved speech recognition.Rita Singh, Bhiksha Raj, Richard M. Stern
1999ICASSPAutomatic clustering and generation of contextual questions for tied states in hidden Markov models.Rita Singh, Bhiksha Raj, Richard M. Stern
1999InterspeechDomain adduced state tying for cross-domain acoustic modelling.Rita Singh, Bhiksha Raj, Richard M. Stern
1998InterspeechInference of missing spectrographic features for robust speech recognition.Bhiksha Raj, Rita Singh, Richard M. Stern