Skip to content

IEEE International Conference on Acoustics, Speech and Signal Processing

ICASSP

Multiconference

CORE rank

CORE rank (raw)

Multiconference

Fields of research

Computer Vision and Multimedia Computation

Papers indexed

56,707

1976–2025

Papers per year

19763,371 peak2025

ICASSP papers

56,707 records sourced from DBLP. Search titles, filter by year, sort by recency.

YearTitleAuthors
2025Hypothesis Clustering and Merging: Novel MultiTalker Speech Recognition with Speaker Tokens.Yosuke Kashiwagi, Hayato Futami, Emiru Tsunoo, Siddhant Arora, Shinji Watanabe
2025A Modified Gain Normalized Step Size Adaptive Algorithm for Improved Online Secondary Path Modelling in Active Noise Control.Asutosh Kar, Gagandeep Singh, Pradeep K. Shill, Somanath Pradhan, Vasundhara, Mads Grsbll Christensen
2025Influence of Oropharyngeal Esophageal Cavity Geometry and Beak Angle on Vocal Tract Resonance of Birds using Computational Modeling.Noumida Abdul Kareem, Rajeev Rajan
2025Simultaneous Music Separation and Generation Using Multi-Track Latent Diffusion Models.Tornike Karchkhadze, Mohammad Rasool Izadi, Shlomo Dubnov
2025Identifying Adversarial Attacks in Crowdsourcing via Dense Subgraph Detection.Abdullah Karaaslanli, Panagiotis A. Traganitis, Aritra Konar
2025Graph Refinement in Latent Space: A Hypergraph Convolution for Underwater Object Detection.Meghna Kapoor, Badri Narayan Subudhi, Ankur Bansal
2025K-HashFed: Communication Efficient Federated Learning through Gradient Clustering and Hashing.Ayshika Kapoor, Dheeraj Kumar
2025Input Uncertainty Attribution by Uncertainty Propagation.Benedikt Kantz, Sophie Steger, Clemens Staudinger, Christoph Feilmayr, Johannes Wachlmayr, Alexander Haberl, Stefan Schuster, Franz Pernkopf
2025Bridging Speech and Text Foundation Models with ReShape Attention.Takatomo Kano, Atsunori Ogawa, Marc Delcroix, William Chen, Ryo Fukuda, Kohei Matsuura, Takanori Ashihara, Shinji Watanabe
2025Modeling of HR Filters for Audio Object Rendering.Sumeyra Demir Kanik, Erlendur Karlsson, Tomas Toftgard, Erik Norvell
2025Enhancing Image Editing with Chain-of-Thought Reasoning and Multimodal Large Language Models.Mengxue Kang, Xinyu Zhang, Fei Wei, Shuang Xu, Yuhe Liu
2025FADEL: Uncertainty-aware Fake Audio Detection with Evidential Deep Learning.Ju Yeon Kang, Ji Won Yoon, Semin Kim, Min Hyun Han, Nam Soo Kim
2025StrucFormer: Structural Prior Guided Transformer for Mobile Crowdsensing Data Inference.Xu Kang, Shouceng Tian, Feifei Kou, Lei Shi, Jiadong Ren
2025Face-StyleSpeech: Enhancing Zero-shot Speech Synthesis from Face Images with Improved Face-to-Speech Mapping.Minki Kang, Wooseok Han, Eunho Yang
2025Mask augmented Object-Centric Contrastive Learning for Amodal Instance Segmentation.Tomokazu Kaneko, Ryosuke Sakai, Takashi Shibata, Soma Shiraishi
2025SBL Algorithms for the Multiple Measurement Vector Problem: New Modeling and Inference Methods.Vinay Kanakeri, Florian Meyer, Bhaskar D. Rao
2025Neuromorphic Unlimited Sampling for High-Dynamic-Range Video Acquisition.Abijith Jagannath Kamath, Chandra Sekhar Seelamantula
2025MorphFader: Enabling Fine-grained Controllable Morphing with Text-to-Audio Models.Purnima Kamath, Chitralekha Gupta, Suranga Nanayakkara
2025On the Design of Weakly-Convex Regularizers for Solving Linear Inverse Problems.Abijith Jagannath Kamath, Abhishek Shreekant Bhandiwad, Chandra Sekhar Seelamantula
2025Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection.Yacouba Kaloga, Shakeel A. Sheikh, Ina Kodrasi
2025Towards Context-aware EEG-based Emotion Recognition Models: Personality and Emotional Intelligence as Context.Kannadasan Kalidasan, Nikita Rajesh Verma, Jainendra Shukla
2025FARE: A Deep Learning-Based Framework for Radar-Based Face Recognition and Out-of-Distribution Detection.Sabri Mustafa Kahya, Boran Hamdi Sivrikaya, Muhammet Sami Yavuz, Eckehard G. Steinbach
2025Addressing Speed-Induced Dispersion in Stepped-Frequency PMCW Radar Systems.Moritz Kahlert, Tai Fei, Claas Tebruegge, Shunqiao Sun, Markus Gardill
2025Audio Codec Augmentation for Robust Collaborative Watermarking of Speech Synthesis.Lauri Juvela, Xin Wang
2025Text-Aware Adapter for Few-Shot Keyword Spotting.Youngmoon Jung, Jinyoung Lee, Seungjin Lee, Myunghun Jung, Yong-Hyeok Lee, Hoon-Young Cho
2,2262,250 of 56,707← PreviousNext →

Comparable venues

Other A*/A conferences filed under the same field of research.