Skip to content

Ramani Duraiswami

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

79

Venues

24

Active years

2000–2026

Best venue rank

A*

Where they publish

Papers

79 indexed papers, newest first.

YearVenueTitleAuthors
2026AAAIMMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence.Sonal Kumar, Simon Sedlcek, Vaibhavi Lokegaonkar, Fernando Lpez, Wenyi Yu, Nishit Anand, Hyeonggon Ryu, Lichang Chen, Maxim Plicka, Miroslav Hlavcek, William Fineas Ellingwood, Sathvik Udupa, Siyuan Hou, Allison Ferner, Sara Barahona, Cecilia Bolaos, Satish Rahi, Laura Herrera-Alarcn, Satvik Dixit, Rupali S. Patil, Soham Deshmukh, Lasha Koroshinadze, Yao Liu, Leibny Paola Garca-Perera, Eleni Zanou, Themos Stafylakis, Joon Son Chung, David Harwath, Chao Zhang, Dinesh Manocha, Alicia Lozano-Diez, Santosh Kesiraju, Sreyan Ghosh, Ramani Duraiswami
2026ACLFIGMA: Towards FIne-Grained Music retrievAl.Nishit Anand, Ashish Seth, Sreyan Ghosh, Dinesh Manocha, Ramani Duraiswami
2026ACLPolyAudio: Advancing Multi-Audio Reasoning in Large Audio Language Models with Interleaved Multi-Audio Contexts.Sonal Kumar, Sreyan Ghosh, Yueqian Lin, S. Sakshi, Ashish Seth, Yiran Chen, Ramani Duraiswami, Dinesh Manocha
2025ACSSCA Scalable MVDR Beamforming Algorithm that is Linear in the number of Antennas.Sanjaya Herath, Armin Gerami, Kevin Wagner, Ramani Duraiswami, Christopher A. Metzler
2025EMNLPEGOILLUSION: Benchmarking Hallucinations in Egocentric Video Understanding.Ashish Seth, Utkarsh Tyagi, Ramaneswaran Selvakumar, Nishit Anand, Sonal Kumar, Sreyan Ghosh, Ramani Duraiswami, Chirag Agarwal, Dinesh Manocha
2025ICASSPTSPE: Task-Specific Prompt Ensemble for Improved Zero-Shot Audio Classification.Nishit Anand, Ashish Seth, Ramani Duraiswami, Dinesh Manocha
2025ICASSPEfficient Spatial Audio Rendering Via Differentiable FIR To IIR Estimation.Armin Gerami, Bowen Zhi, Dmitry N. Zotkin, Ramani Duraiswami
2025ICASSPReCLAP: Improving Zero Shot Audio Classification by Describing Sounds.Sreyan Ghosh, Sonal Kumar, Chandra Kiran Reddy Evuru, Oriol Nieto, Ramani Duraiswami, Dinesh Manocha
2025ICASSP3D Gaussian Splatting with Normal Information for Mesh Extraction and Improved Rendering.Meenakshi Krishnan, Liam Fowl, Ramani Duraiswami
2025ICLRMMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark.S. Sakshi, Utkarsh Tyagi, Sonal Kumar, Ashish Seth, Ramaneswaran Selvakumar, Oriol Nieto, Ramani Duraiswami, Sreyan Ghosh, Dinesh Manocha
2025NAACLProSE: Diffusion Priors for Speech Enhancement.Sonal Kumar, Sreyan Ghosh, Utkarsh Tyagi, Anton Jeran Ratnarajah, Chandra Kiran Reddy Evuru, Ramani Duraiswami, Dinesh Manocha
2024EMNLPGAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities.Sreyan Ghosh, Sonal Kumar, Ashish Seth, Chandra Kiran Reddy Evuru, Utkarsh Tyagi, S. Sakshi, Oriol Nieto, Ramani Duraiswami, Dinesh Manocha
2024ICASSPRecap: Retrieval-Augmented Audio Captioning.Sreyan Ghosh, Sonal Kumar, Chandra Kiran Reddy Evuru, Ramani Duraiswami, Dinesh Manocha
2024ICLRCompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models.Sreyan Ghosh, Ashish Seth, Sonal Kumar, Utkarsh Tyagi, Chandra Kiran Reddy Evuru, Ramaneswaran S., Sakshi Singh, Oriol Nieto, Ramani Duraiswami, Dinesh Manocha
2024ICMLA Closer Look at the Limitations of Instruction Tuning.Sreyan Ghosh, Chandra Kiran Reddy Evuru, Sonal Kumar, Ramaneswaran S., Deepali Aneja, Zeyu Jin, Ramani Duraiswami, Dinesh Manocha
2024InterspeechLipGER: Visually-Conditioned Generative Error Correction for Robust Automatic Speech Recognition.Sreyan Ghosh, Sonal Kumar, Ashish Seth, Purva Chiniya, Utkarsh Tyagi, Ramani Duraiswami, Dinesh Manocha
2023ICASSPRapid Audiometric Evaluation for Personalized Headphone Listening.Matthew J. Goupell, Marjan Davoodian, Sarah Weinstein, David Gadzinski, Dmitry N. Zotkin, Kaushik Sethunath, Ramani Duraiswami
2022ICASSPTowards Fast And Convenient End-To-End HRTF Personalization.Bowen Zhi, Dmitry N. Zotkin, Ramani Duraiswami
2018ICASSPSequential Direction Detection for Sound Scene Analysis.Nail A. Gumerov, Bowen Zhi, Ramani Duraiswami
2017ICASSPFast interpolation of bandlimited functions.Samuel F. Potter, Nail A. Gumerov, Ramani Duraiswami
2017ICASSPIncident field recovery for an arbitrary-shaped scatterer.Dmitry N. Zotkin, Nail A. Gumerov, Ramani Duraiswami
2015CHIHead-Mounted Display Visualizations to Support Sound Awareness for the Deaf and Hard of Hearing.Dhruv Jain, Leah Findlater, Jamie Gilkeson, Benjamin Holland, Ramani Duraiswami, Dmitry N. Zotkin, Christian Vogler, Jon E. Froehlich
2014ICASSPGaussian process models for HRTF based 3D sound localization.Yuancheng Luo, Dmitry N. Zotkin, Ramani Duraiswami
2013AISTATSFast Near-GRID Gaussian Process Regression.Yuancheng Luo, Ramani Duraiswami
2013ICASSPKernel regression for Head-Related Transfer Function interpolation and spectral extrema extraction.Yuancheng Luo, Dmitry N. Zotkin, Hal Daum III, Ramani Duraiswami
2012HPCCScalable Distributed Fast Multipole Methods.Qi Hu, Nail A. Gumerov, Ramani Duraiswami
2012ICASSPThe UMD-JHU 2011 speaker recognition system.Daniel Garcia-Romero, Xinhui Zhou, Dmitry N. Zotkin, Balaji Vasan Srinivasan, Yuancheng Luo, Sriram Ganapathy, Samuel Thomas, Sridhar Krishna Nemala, Garimella S. V. S. Sivaram, Majid Mirbagheri, Sri Harish Reddy Mallidi, Thomas Janu, Padmanabhan Rajan, Nima Mesgarani, Mounya Elhilali, Hynek Hermansky, Shihab A. Shamma, Ramani Duraiswami
2012SCAbstract: Scalable Fast Multipole Methods for Vortex Element Methods.Qi Hu, Nail A. Gumerov, Rio Yokota, Lorena A. Barba, Ramani Duraiswami
2012SCPoster: Scalable Fast Multipole Methods for Vortex Element Methods.Qi Hu, Nail A. Gumerov, Rio Yokota, Lorena A. Barba, Ramani Duraiswami
2011ASRULinear versus mel frequency cepstral coefficients for speaker recognition.Xinhui Zhou, Daniel Garcia-Romero, Ramani Duraiswami, Carol Y. Espy-Wilson, Shihab A. Shamma
2011ICASSPA partial least squares framework for speaker recognition.Balaji Vasan Srinivasan, Dmitry N. Zotkin, Ramani Duraiswami
2011InterspeechKernel Partial Least Squares for Speaker Recognition.Balaji Vasan Srinivasan, Daniel Garcia-Romero, Dmitry N. Zotkin, Ramani Duraiswami
2011SCScalable fast multipole methods on distributed heterogeneous architectures.Qi Hu, Nail A. Gumerov, Ramani Duraiswami
2010ICASSPAutomatic matched filter recovery via the audio camera.Adam O'Donovan, Ramani Duraiswami, Dmitry N. Zotkin
2010ICASSPKernelized Rnyi distance for speaker recognition.Balaji Vasan Srinivasan, Ramani Duraiswami, Dmitry N. Zotkin
2009ICASSPModal expansion of HRTFs: Continuous representation in frequency-range-angle.Wen Zhang, Thushara D. Abhayapala, Rodney A. Kennedy, Ramani Duraiswami
2009ICASSPPlane-wave decomposition of a sound scene using a cylindrical microphone array.Dmitry N. Zotkin, Ramani Duraiswami
2009ICCVEfficient subset selection via the kernelized Rnyi distance.Balaji Vasan Srinivasan, Ramani Duraiswami
2008CVPRCanny edge detection on NVIDIA CUDA.Yuancheng Luo, Ramani Duraiswami
2008ICASSPImaging concert hall acoustics using visual and audio cameras.Adam O'Donovan, Ramani Duraiswami, Dmitry N. Zotkin
2008ICASSPSound field decomposition using spherical microphone arrays.Dmitry N. Zotkin, Ramani Duraiswami, Nail A. Gumerov
2008WACVTracking Down Under: Following the Satin Bowerbird.Aniruddha Kembhavi, Ryan Farrell, Yuancheng Luo, David W. Jacobs, Ramani Duraiswami, Larry S. Davis
2007CVPRMicrophone Arrays as Generalized Cameras for Integrated Audio Visual Processing.Adam O'Donovan, Ramani Duraiswami, Jan Neumann
2007CVPRMultimodal Tracking for Smart Videoconferencing and Video Surveillance.Dmitry N. Zotkin, Vikas C. Raykar, Ramani Duraiswami, Larry S. Davis
2007ICASSPFast Multipole Accelerated Boundary Elements for Numerical Computation of the Head Related Transfer Function.Nail A. Gumerov, Ramani Duraiswami, Dmitry N. Zotkin
2007ICASSPEfficient Conversion of X.Y Surround Sound Content to Binaural Head-Tracked Form for HRTF-Enabled Playback.Dmitry N. Zotkin, Ramani Duraiswami, Nail A. Gumerov
2006ICASSPHeadphone-Based Reproduction of 3D Auditory Scenes Captured by Spherical/Hemispherical Microphone Arrays.Zhiyun Li, Ramani Duraiswami
2006ICASSPFrequency Independent Flexible Spherical Beamforming Via Rbf Fitting.Arkady Yerukhimovich, Ramani Duraiswami, Nail A. Gumerov, Dmitry N. Zotkin
2006SDMFast optimal bandwidth selection for kernel density estimation.Vikas C. Raykar, Ramani Duraiswami
2005CVPREfficient Mean-Shift Tracking via a New Similarity Measure.Changjiang Yang, Ramani Duraiswami, Larry S. Davis
2005ICASSPThe manifolds of spatial hearing.Ramani Duraiswami, Vikas C. Raykar
2005ICASSPA robust and self-reconfigurable design of spherical microphone array for multi-resolution beamforming.Zhiyun Li, Ramani Duraiswami
2005ICASSPApproximate expressions for the mean and the covariance of the maximum likelihood estimator for acoustic source localization.Vikas C. Raykar, Ramani Duraiswami
2005ICCVFast Multiple Object Tracking via a Hierarchical Particle Filter.Changjiang Yang, Ramani Duraiswami, Larry S. Davis
2004ECCVStructure of Applicable Surfaces from Single Views.Nail A. Gumerov, Ali Zandifar, Ramani Duraiswami, Larry S. Davis
2004ICASSPInterpolation and range extrapolation of HRTFs [head related transfer functions].Ramani Duraiswami, Dmitry N. Zotkin, Nail A. Gumerov
2004ICASSPFlexible layout and optimal cancellation of the orthonormality error for spherical microphone arrays.Zhiyun Li, Ramani Duraiswami, Elena Grassi, Larry S. Davis
2004ICASSPAutomatic position calibration of multiple microphones.Vikas C. Raykar, Ramani Duraiswami
2004ICIPMulti-level fast multipole method for thin plate spline evaluation.Ali Zandifar, Ser-Nam Lim, Ramani Duraiswami, Nail A. Gumerov, Larry S. Davis
2004ISMARRecording and Reproducing High Order Surround Auditory Scenes for Mixed and Augmented Reality.Zhiyun Li, Ramani Duraiswami, Larry S. Davis
2003CVPRSimultaneous Pose and Correspondence Determination using Line Feature.Philip David, Daniel DeMenthon, Ramani Duraiswami, Hanan Samet
2003CVPRProbabilistic Tracking in Joint Feature-Spatial Spaces.Ahmed M. Elgammal, Ramani Duraiswami, Larry S. Davis
2003ICASSPPitch and timbre manipulations using cortical representation of sound.Dmitry N. Zotkin, Shihab A. Shamma, Powen Ru, Ramani Duraiswami, Larry S. Davis
2003ICCVImproved Fast Gauss Transform and Efficient Kernel Density Estimation.Changjiang Yang, Ramani Duraiswami, Nail A. Gumerov, Larry S. Davis
2003ICIPMean-shift analysis using quasiNewton methods.Changjiang Yang, Ramani Duraiswami, Daniel DeMenthon, Larry S. Davis
2003InterspeechTracking a moving speaker using excitation source information.Vikas C. Raykar, Ramani Duraiswami, B. Yegnanarayana, S. R. Mahadeva Prasanna
2002ECCVSoftPOSIT: Simultaneous Pose and Correspondence Determination.Philip David, Daniel DeMenthon, Ramani Duraiswami, Hanan Samet
2002ICASSPNumerical study of the influence of the torso on the HRTF.Nail A. Gumerov, Ramani Duraiswami, Zhihui Tang
2002ICASSPCreation of virtual auditory spaces.Dmitry N. Zotkin, Ramani Duraiswami, Larry S. Davis
2002ICMIA Video Based Interface to Textual Information for the Visually Impaired.Ali Zandifar, Ramani Duraiswami, Antoine Chahine, Larry S. Davis
2002ICPRNear-Optimal Regularization Parameters for Applications in Computer Vision.Changjiang Yang, Ramani Duraiswami, Larry S. Davis
2002ICPRVirtual Audio System Customization Using Visual Matching of Ear Parameters.Dmitry N. Zotkin, Ramani Duraiswami, Larry S. Davis, Ankur Mohan, Vikas C. Raykar
2001CVPREfficient Non-Parametric Adaptive Color Modeling Using Fast Gauss Transform.Ahmed M. Elgammal, Ramani Duraiswami, Larry S. Davis
2001ICASSPActive speech source localization by a dual coarse-to-fine search.Ramani Duraiswami, Dmitry N. Zotkin, Larry S. Davis
2001ICASSPMultimodal localization of a flying bat.Kaushik Ghose, Dmitry N. Zotkin, Ramani Duraiswami, Cynthia F. Moss
2001ICASSPModeling the effect of a nearby boundary on the HRTF.Nail A. Gumerov, Ramani Duraiswami
2000ECCVQuasi-Random Sampling for Condensation.Vasanth Philomin, Ramani Duraiswami, Larry S. Davis
2000ICPRTracking Humans from a Moving Platform.Larry S. Davis, Vasanth Philomin, Ramani Duraiswami
2000SMCAn audio-video front-end for multimedia applications.Dmitry N. Zotkin, Ramani Duraiswami, Larry S. Davis, Ismail Haritaoglu