Ramani Duraiswami
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
79
Venues
24
Active years
2000–2026
Best venue rank
A*
Where they publish
- MulticonferenceICASSP36 papers
- A*CVPR7 papers
- AInterspeech3 papers
- ASC3 papers
- A*ICCV3 papers
- A*ECCV3 papers
- BICPR3 papers
- A*ACL2 papers
- A*EMNLP2 papers
- A*ICLR2 papers
- BICIP2 papers
- A*AAAI1 paper
- NationalACSSC1 paper
- ANAACL1 paper
- A*ICML1 paper
- A*CHI1 paper
- AAISTATS1 paper
- CHPCC1 paper
- CASRU1 paper
- AWACV1 paper
- ASDM1 paper
- A*ISMAR1 paper
- BICMI1 paper
- BSMC1 paper
Papers
79 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | AAAI | MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence. | Sonal Kumar, Simon Sedlcek, Vaibhavi Lokegaonkar, Fernando Lpez, Wenyi Yu, Nishit Anand, Hyeonggon Ryu, Lichang Chen, Maxim Plicka, Miroslav Hlavcek, William Fineas Ellingwood, Sathvik Udupa, Siyuan Hou, Allison Ferner, Sara Barahona, Cecilia Bolaos, Satish Rahi, Laura Herrera-Alarcn, Satvik Dixit, Rupali S. Patil, Soham Deshmukh, Lasha Koroshinadze, Yao Liu, Leibny Paola Garca-Perera, Eleni Zanou, Themos Stafylakis, Joon Son Chung, David Harwath, Chao Zhang, Dinesh Manocha, Alicia Lozano-Diez, Santosh Kesiraju, Sreyan Ghosh, Ramani Duraiswami |
| 2026 | ACL | FIGMA: Towards FIne-Grained Music retrievAl. | Nishit Anand, Ashish Seth, Sreyan Ghosh, Dinesh Manocha, Ramani Duraiswami |
| 2026 | ACL | PolyAudio: Advancing Multi-Audio Reasoning in Large Audio Language Models with Interleaved Multi-Audio Contexts. | Sonal Kumar, Sreyan Ghosh, Yueqian Lin, S. Sakshi, Ashish Seth, Yiran Chen, Ramani Duraiswami, Dinesh Manocha |
| 2025 | ACSSC | A Scalable MVDR Beamforming Algorithm that is Linear in the number of Antennas. | Sanjaya Herath, Armin Gerami, Kevin Wagner, Ramani Duraiswami, Christopher A. Metzler |
| 2025 | EMNLP | EGOILLUSION: Benchmarking Hallucinations in Egocentric Video Understanding. | Ashish Seth, Utkarsh Tyagi, Ramaneswaran Selvakumar, Nishit Anand, Sonal Kumar, Sreyan Ghosh, Ramani Duraiswami, Chirag Agarwal, Dinesh Manocha |
| 2025 | ICASSP | TSPE: Task-Specific Prompt Ensemble for Improved Zero-Shot Audio Classification. | Nishit Anand, Ashish Seth, Ramani Duraiswami, Dinesh Manocha |
| 2025 | ICASSP | Efficient Spatial Audio Rendering Via Differentiable FIR To IIR Estimation. | Armin Gerami, Bowen Zhi, Dmitry N. Zotkin, Ramani Duraiswami |
| 2025 | ICASSP | ReCLAP: Improving Zero Shot Audio Classification by Describing Sounds. | Sreyan Ghosh, Sonal Kumar, Chandra Kiran Reddy Evuru, Oriol Nieto, Ramani Duraiswami, Dinesh Manocha |
| 2025 | ICASSP | 3D Gaussian Splatting with Normal Information for Mesh Extraction and Improved Rendering. | Meenakshi Krishnan, Liam Fowl, Ramani Duraiswami |
| 2025 | ICLR | MMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark. | S. Sakshi, Utkarsh Tyagi, Sonal Kumar, Ashish Seth, Ramaneswaran Selvakumar, Oriol Nieto, Ramani Duraiswami, Sreyan Ghosh, Dinesh Manocha |
| 2025 | NAACL | ProSE: Diffusion Priors for Speech Enhancement. | Sonal Kumar, Sreyan Ghosh, Utkarsh Tyagi, Anton Jeran Ratnarajah, Chandra Kiran Reddy Evuru, Ramani Duraiswami, Dinesh Manocha |
| 2024 | EMNLP | GAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities. | Sreyan Ghosh, Sonal Kumar, Ashish Seth, Chandra Kiran Reddy Evuru, Utkarsh Tyagi, S. Sakshi, Oriol Nieto, Ramani Duraiswami, Dinesh Manocha |
| 2024 | ICASSP | Recap: Retrieval-Augmented Audio Captioning. | Sreyan Ghosh, Sonal Kumar, Chandra Kiran Reddy Evuru, Ramani Duraiswami, Dinesh Manocha |
| 2024 | ICLR | CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models. | Sreyan Ghosh, Ashish Seth, Sonal Kumar, Utkarsh Tyagi, Chandra Kiran Reddy Evuru, Ramaneswaran S., Sakshi Singh, Oriol Nieto, Ramani Duraiswami, Dinesh Manocha |
| 2024 | ICML | A Closer Look at the Limitations of Instruction Tuning. | Sreyan Ghosh, Chandra Kiran Reddy Evuru, Sonal Kumar, Ramaneswaran S., Deepali Aneja, Zeyu Jin, Ramani Duraiswami, Dinesh Manocha |
| 2024 | Interspeech | LipGER: Visually-Conditioned Generative Error Correction for Robust Automatic Speech Recognition. | Sreyan Ghosh, Sonal Kumar, Ashish Seth, Purva Chiniya, Utkarsh Tyagi, Ramani Duraiswami, Dinesh Manocha |
| 2023 | ICASSP | Rapid Audiometric Evaluation for Personalized Headphone Listening. | Matthew J. Goupell, Marjan Davoodian, Sarah Weinstein, David Gadzinski, Dmitry N. Zotkin, Kaushik Sethunath, Ramani Duraiswami |
| 2022 | ICASSP | Towards Fast And Convenient End-To-End HRTF Personalization. | Bowen Zhi, Dmitry N. Zotkin, Ramani Duraiswami |
| 2018 | ICASSP | Sequential Direction Detection for Sound Scene Analysis. | Nail A. Gumerov, Bowen Zhi, Ramani Duraiswami |
| 2017 | ICASSP | Fast interpolation of bandlimited functions. | Samuel F. Potter, Nail A. Gumerov, Ramani Duraiswami |
| 2017 | ICASSP | Incident field recovery for an arbitrary-shaped scatterer. | Dmitry N. Zotkin, Nail A. Gumerov, Ramani Duraiswami |
| 2015 | CHI | Head-Mounted Display Visualizations to Support Sound Awareness for the Deaf and Hard of Hearing. | Dhruv Jain, Leah Findlater, Jamie Gilkeson, Benjamin Holland, Ramani Duraiswami, Dmitry N. Zotkin, Christian Vogler, Jon E. Froehlich |
| 2014 | ICASSP | Gaussian process models for HRTF based 3D sound localization. | Yuancheng Luo, Dmitry N. Zotkin, Ramani Duraiswami |
| 2013 | AISTATS | Fast Near-GRID Gaussian Process Regression. | Yuancheng Luo, Ramani Duraiswami |
| 2013 | ICASSP | Kernel regression for Head-Related Transfer Function interpolation and spectral extrema extraction. | Yuancheng Luo, Dmitry N. Zotkin, Hal Daum III, Ramani Duraiswami |
| 2012 | HPCC | Scalable Distributed Fast Multipole Methods. | Qi Hu, Nail A. Gumerov, Ramani Duraiswami |
| 2012 | ICASSP | The UMD-JHU 2011 speaker recognition system. | Daniel Garcia-Romero, Xinhui Zhou, Dmitry N. Zotkin, Balaji Vasan Srinivasan, Yuancheng Luo, Sriram Ganapathy, Samuel Thomas, Sridhar Krishna Nemala, Garimella S. V. S. Sivaram, Majid Mirbagheri, Sri Harish Reddy Mallidi, Thomas Janu, Padmanabhan Rajan, Nima Mesgarani, Mounya Elhilali, Hynek Hermansky, Shihab A. Shamma, Ramani Duraiswami |
| 2012 | SC | Abstract: Scalable Fast Multipole Methods for Vortex Element Methods. | Qi Hu, Nail A. Gumerov, Rio Yokota, Lorena A. Barba, Ramani Duraiswami |
| 2012 | SC | Poster: Scalable Fast Multipole Methods for Vortex Element Methods. | Qi Hu, Nail A. Gumerov, Rio Yokota, Lorena A. Barba, Ramani Duraiswami |
| 2011 | ASRU | Linear versus mel frequency cepstral coefficients for speaker recognition. | Xinhui Zhou, Daniel Garcia-Romero, Ramani Duraiswami, Carol Y. Espy-Wilson, Shihab A. Shamma |
| 2011 | ICASSP | A partial least squares framework for speaker recognition. | Balaji Vasan Srinivasan, Dmitry N. Zotkin, Ramani Duraiswami |
| 2011 | Interspeech | Kernel Partial Least Squares for Speaker Recognition. | Balaji Vasan Srinivasan, Daniel Garcia-Romero, Dmitry N. Zotkin, Ramani Duraiswami |
| 2011 | SC | Scalable fast multipole methods on distributed heterogeneous architectures. | Qi Hu, Nail A. Gumerov, Ramani Duraiswami |
| 2010 | ICASSP | Automatic matched filter recovery via the audio camera. | Adam O'Donovan, Ramani Duraiswami, Dmitry N. Zotkin |
| 2010 | ICASSP | Kernelized Rnyi distance for speaker recognition. | Balaji Vasan Srinivasan, Ramani Duraiswami, Dmitry N. Zotkin |
| 2009 | ICASSP | Modal expansion of HRTFs: Continuous representation in frequency-range-angle. | Wen Zhang, Thushara D. Abhayapala, Rodney A. Kennedy, Ramani Duraiswami |
| 2009 | ICASSP | Plane-wave decomposition of a sound scene using a cylindrical microphone array. | Dmitry N. Zotkin, Ramani Duraiswami |
| 2009 | ICCV | Efficient subset selection via the kernelized Rnyi distance. | Balaji Vasan Srinivasan, Ramani Duraiswami |
| 2008 | CVPR | Canny edge detection on NVIDIA CUDA. | Yuancheng Luo, Ramani Duraiswami |
| 2008 | ICASSP | Imaging concert hall acoustics using visual and audio cameras. | Adam O'Donovan, Ramani Duraiswami, Dmitry N. Zotkin |
| 2008 | ICASSP | Sound field decomposition using spherical microphone arrays. | Dmitry N. Zotkin, Ramani Duraiswami, Nail A. Gumerov |
| 2008 | WACV | Tracking Down Under: Following the Satin Bowerbird. | Aniruddha Kembhavi, Ryan Farrell, Yuancheng Luo, David W. Jacobs, Ramani Duraiswami, Larry S. Davis |
| 2007 | CVPR | Microphone Arrays as Generalized Cameras for Integrated Audio Visual Processing. | Adam O'Donovan, Ramani Duraiswami, Jan Neumann |
| 2007 | CVPR | Multimodal Tracking for Smart Videoconferencing and Video Surveillance. | Dmitry N. Zotkin, Vikas C. Raykar, Ramani Duraiswami, Larry S. Davis |
| 2007 | ICASSP | Fast Multipole Accelerated Boundary Elements for Numerical Computation of the Head Related Transfer Function. | Nail A. Gumerov, Ramani Duraiswami, Dmitry N. Zotkin |
| 2007 | ICASSP | Efficient Conversion of X.Y Surround Sound Content to Binaural Head-Tracked Form for HRTF-Enabled Playback. | Dmitry N. Zotkin, Ramani Duraiswami, Nail A. Gumerov |
| 2006 | ICASSP | Headphone-Based Reproduction of 3D Auditory Scenes Captured by Spherical/Hemispherical Microphone Arrays. | Zhiyun Li, Ramani Duraiswami |
| 2006 | ICASSP | Frequency Independent Flexible Spherical Beamforming Via Rbf Fitting. | Arkady Yerukhimovich, Ramani Duraiswami, Nail A. Gumerov, Dmitry N. Zotkin |
| 2006 | SDM | Fast optimal bandwidth selection for kernel density estimation. | Vikas C. Raykar, Ramani Duraiswami |
| 2005 | CVPR | Efficient Mean-Shift Tracking via a New Similarity Measure. | Changjiang Yang, Ramani Duraiswami, Larry S. Davis |
| 2005 | ICASSP | The manifolds of spatial hearing. | Ramani Duraiswami, Vikas C. Raykar |
| 2005 | ICASSP | A robust and self-reconfigurable design of spherical microphone array for multi-resolution beamforming. | Zhiyun Li, Ramani Duraiswami |
| 2005 | ICASSP | Approximate expressions for the mean and the covariance of the maximum likelihood estimator for acoustic source localization. | Vikas C. Raykar, Ramani Duraiswami |
| 2005 | ICCV | Fast Multiple Object Tracking via a Hierarchical Particle Filter. | Changjiang Yang, Ramani Duraiswami, Larry S. Davis |
| 2004 | ECCV | Structure of Applicable Surfaces from Single Views. | Nail A. Gumerov, Ali Zandifar, Ramani Duraiswami, Larry S. Davis |
| 2004 | ICASSP | Interpolation and range extrapolation of HRTFs [head related transfer functions]. | Ramani Duraiswami, Dmitry N. Zotkin, Nail A. Gumerov |
| 2004 | ICASSP | Flexible layout and optimal cancellation of the orthonormality error for spherical microphone arrays. | Zhiyun Li, Ramani Duraiswami, Elena Grassi, Larry S. Davis |
| 2004 | ICASSP | Automatic position calibration of multiple microphones. | Vikas C. Raykar, Ramani Duraiswami |
| 2004 | ICIP | Multi-level fast multipole method for thin plate spline evaluation. | Ali Zandifar, Ser-Nam Lim, Ramani Duraiswami, Nail A. Gumerov, Larry S. Davis |
| 2004 | ISMAR | Recording and Reproducing High Order Surround Auditory Scenes for Mixed and Augmented Reality. | Zhiyun Li, Ramani Duraiswami, Larry S. Davis |
| 2003 | CVPR | Simultaneous Pose and Correspondence Determination using Line Feature. | Philip David, Daniel DeMenthon, Ramani Duraiswami, Hanan Samet |
| 2003 | CVPR | Probabilistic Tracking in Joint Feature-Spatial Spaces. | Ahmed M. Elgammal, Ramani Duraiswami, Larry S. Davis |
| 2003 | ICASSP | Pitch and timbre manipulations using cortical representation of sound. | Dmitry N. Zotkin, Shihab A. Shamma, Powen Ru, Ramani Duraiswami, Larry S. Davis |
| 2003 | ICCV | Improved Fast Gauss Transform and Efficient Kernel Density Estimation. | Changjiang Yang, Ramani Duraiswami, Nail A. Gumerov, Larry S. Davis |
| 2003 | ICIP | Mean-shift analysis using quasiNewton methods. | Changjiang Yang, Ramani Duraiswami, Daniel DeMenthon, Larry S. Davis |
| 2003 | Interspeech | Tracking a moving speaker using excitation source information. | Vikas C. Raykar, Ramani Duraiswami, B. Yegnanarayana, S. R. Mahadeva Prasanna |
| 2002 | ECCV | SoftPOSIT: Simultaneous Pose and Correspondence Determination. | Philip David, Daniel DeMenthon, Ramani Duraiswami, Hanan Samet |
| 2002 | ICASSP | Numerical study of the influence of the torso on the HRTF. | Nail A. Gumerov, Ramani Duraiswami, Zhihui Tang |
| 2002 | ICASSP | Creation of virtual auditory spaces. | Dmitry N. Zotkin, Ramani Duraiswami, Larry S. Davis |
| 2002 | ICMI | A Video Based Interface to Textual Information for the Visually Impaired. | Ali Zandifar, Ramani Duraiswami, Antoine Chahine, Larry S. Davis |
| 2002 | ICPR | Near-Optimal Regularization Parameters for Applications in Computer Vision. | Changjiang Yang, Ramani Duraiswami, Larry S. Davis |
| 2002 | ICPR | Virtual Audio System Customization Using Visual Matching of Ear Parameters. | Dmitry N. Zotkin, Ramani Duraiswami, Larry S. Davis, Ankur Mohan, Vikas C. Raykar |
| 2001 | CVPR | Efficient Non-Parametric Adaptive Color Modeling Using Fast Gauss Transform. | Ahmed M. Elgammal, Ramani Duraiswami, Larry S. Davis |
| 2001 | ICASSP | Active speech source localization by a dual coarse-to-fine search. | Ramani Duraiswami, Dmitry N. Zotkin, Larry S. Davis |
| 2001 | ICASSP | Multimodal localization of a flying bat. | Kaushik Ghose, Dmitry N. Zotkin, Ramani Duraiswami, Cynthia F. Moss |
| 2001 | ICASSP | Modeling the effect of a nearby boundary on the HRTF. | Nail A. Gumerov, Ramani Duraiswami |
| 2000 | ECCV | Quasi-Random Sampling for Condensation. | Vasanth Philomin, Ramani Duraiswami, Larry S. Davis |
| 2000 | ICPR | Tracking Humans from a Moving Platform. | Larry S. Davis, Vasanth Philomin, Ramani Duraiswami |
| 2000 | SMC | An audio-video front-end for multimedia applications. | Dmitry N. Zotkin, Ramani Duraiswami, Larry S. Davis, Ismail Haritaoglu |