| 2025 | An Efficient and Streaming Audio Visual Active Speaker Detection System. | Arnav Kundu, Yanzi Jin, Mohammad Hossein Sekhavat, Maxwell Horton, Danny Tormoen, Devang Naik |
| 2025 | A Proximal Variable Smoothing for Nonsmooth Minimization Involving Weakly Convex Composite with MIMO Application. | Keita Kume, Isao Yamada |
| 2025 | Addressing Emotion Ambiguity and Annotator Subjectivity for Enhanced Speech Emotion Labeling. | Pooja Kumawat, Aurobinda Routray |
| 2025 | Performance Evaluation of SLAM-ASR: The Good, the Bad, the Ugly, and the Way Forward. | Shashi Kumar, Iuliia Thorbecke, Sergio Burdisso, Esa Villatoro-Tello, Manjunath K. E, Kadri Hacioglu, Pradeep Rangappa, Petr Motlcek, Aravind Ganapathiraju, Andreas Stolcke |
| 2025 | Confidence-Enhanced Models for Indian Art Music Analysis. | Sumit Kumar, Parampreet Singh, Vipul Arora |
| 2025 | Using RLHF to align speech enhancement approaches to mean-opinion quality scores. | Anurag Kumar, Andrew Perrault, Donald S. Williamson |
| 2025 | SEAL: Speaker Error Correction using Acoustic-conditioned Large Language Models. | Anurag Kumar, Rohit Paturi, Amber Afshan, Sundararajan Srinivasan |
| 2025 | XLSR-Transducer: Streaming ASR for Self-Supervised Pretrained Models. | Shashi Kumar, Srikanth R. Madikeri, Juan Zuluaga-Gomez, Esa Villatoro-Tello, Iuliia Thorbecke, Petr Motlcek, Manjunath K. E, Aravind Ganapathiraju |
| 2025 | DRSFANet: Dual-Path CNN with Residual and Frequency Attention for Image Denoising. | Manish Kumar, Suman Kumar Maji, Anirban Saha |
| 2025 | Multiscale Adaptive Channel Estimation for OTFS. | Prabhat Kumar, Chandra R. Murthy |
| 2025 | Beyond Uniformity: Deblurring Images With Complex Noise Patterns Using Half Quadratic Splitting. | Avinash Kumar, Koyyada Dinesh Kumar, Sujit Kumar Sahoo |
| 2025 | Outage Analysis of IRS-Aided Wireless Energy Transfer Under Correlation and Imperfect CSI. | Chandan Kumar, Salil Kashyap |
| 2025 | Practical Radar Sensing Using Two Stage Neural Network for Denoising OTFS Signals. | Ashok S. Kumar, Sheetal Kalyani |
| 2025 | Spherical Sector Harmonics Based Acoustic Source Localization Using Circular Array. | Deepika Kumari, Lalan Kumar |
| 2025 | Generative Speech Foundation Model Pretraining for High-Quality Speech Extraction and Restoration. | Pin-Jui Ku, Alexander H. Liu, Roman Korostik, Sung-Feng Huang, Szu-Wei Fu, Ante Jukic |
| 2025 | Interpolation for Weight-Constrained Nested Arrays Having Non-Central ULA Segments in the Coarray. | Pranav Kulkarni, P. P. Vaidyanathan |
| 2025 | Generalized Constructions of Weight-Constrained Sparse Arrays. | Pranav Kulkarni, P. P. Vaidyanathan |
| 2025 | An Explicit Consistency-Preserving Loss Function for Phase Reconstruction and Speech Enhancement. | Pin-Jui Ku, Chun-Wei Ho, Hao Yen, Sabato Marco Siniscalchi, Chin-Hui Lee |
| 2025 | Detecting and Defending Against Adversarial Attacks on Automatic Speech Recognition via Diffusion Models. | Nikolai Lund Khne, Astrid H. F. Kitchena, Marie S. Jensen, Mikkel S. L. Brndt, Martin Gonzalez, Christophe A. N. Biscio, Zheng-Hua Tan |
| 2025 | Gradient-Oriented Clustered Federated Learning With Efficient Knowledge Sharing in Non-IID Settings. | Kenta Kubota, Ren Togo, Keisuke Maeda, Takahiro Ogawa, Miki Haseyama |
| 2025 | Performance Bounds for Single-Bit AOA-Based RF Source Localization. | Shaunak Kubal, Anastasia Lavrenko, Andr Kokkeler |
| 2025 | Can Large Audio-Language Models Truly Hear? Tackling Hallucinations with Multi-Task Assessment and Stepwise Audio Reasoning. | Chun-Yi Kuan, Hung-Yi Lee |
| 2025 | Unveiling the Pruning Risks on Privacy Vulnerabilities of Deep Neural Networks. | Wenxin Kuang, Qizhuang Liang, Peng Sun, Wei Fu, Qiao Hu, Yupeng Hu |
| 2025 | Two-stream Semantic Alignment Networks for Multi-label Image Classification. | Wenlan Kuang, Zhixin Li |
| 2025 | On Investigating a Better Audio Representation for Mood Classification in Indian Popular Music. | Arathi K, Hotha Durga Swetha, V. G. Vaishali, Kalyan Munukutla, Abhijith V, Gurram Shalini, Joe Cheri Ross |