| 2025 | ICASSP | A decade of DCASE: Achievements, practices, evaluations and future challenges. | Annamaria Mesaros, Romain Serizel, Toni Heittola, Tuomas Virtanen, Mark D. Plumbley |
| 2025 | ICASSP | FlowSep: Language-Queried Sound Separation with Rectified Flow Matching. | Yi Yuan, Xubo Liu, Haohe Liu, Mark D. Plumbley, Wenwu Wang |
| 2025 | ICASSP | Sound-VECaps: Improving Audio Generation with Visually Enhanced Captions. | Yi Yuan, Dongya Jia, Xiaobin Zhuang, Yuanzhe Chen, Zhuo Chen, Yuping Wang, Yuxuan Wang, Xubo Liu, Xiyuan Kang, Mark D. Plumbley, Wenwu Wang |
| 2025 | Interspeech | EnvSDD: Benchmarking Environmental Sound Deepfake Detection. | Han Yin, Yang Xiao, Rohan Kumar Das, Jisheng Bai, Haohe Liu, Wenwu Wang, Mark D. Plumbley |
| 2025 | MMSP | Music Source Restoration. | Yongyi Zang, Zheqi Dai, Mark D. Plumbley, Qiuqiang Kong |
| 2024 | AAAI | Learning Temporal Resolution in Spectrogram for Audio Classification. | Haohe Liu, Xubo Liu, Qiuqiang Kong, Wenwu Wang, Mark D. Plumbley |
| 2024 | ICASSP | Audiosr: Versatile Audio Super-Resolution at Scale. | Haohe Liu, Ke Chen, Qiao Tian, Wenwu Wang, Mark D. Plumbley |
| 2024 | ICASSP | Retrieval-Augmented Text-to-Audio Generation. | Yi Yuan, Haohe Liu, Xubo Liu, Qiushi Huang, Mark D. Plumbley, Wenwu Wang |
| 2024 | Interspeech | Efficient CNNs with Quaternion Transformations and Pruning for Audio Tagging. | Aryan Chaudhary, Arshdeep Singh, Vinayak Abrol, Mark D. Plumbley |
| 2024 | Interspeech | PFCA-Net: Pyramid Feature Fusion and Cross Content Attention Network for Automated Audio Captioning. | Jianyuan Sun, Wenwu Wang, Mark D. Plumbley |
| 2024 | Interspeech | Neural Compression Augmentation for Contrastive Audio Representation Learning. | Zhaoyu Wang, Haohe Liu, Harry Coppock, Bjrn W. Schuller, Mark D. Plumbley |
| 2024 | Interspeech | Efficient Audio Captioning with Encoder-Level Knowledge Distillation. | Xuenan Xu, Haohe Liu, Mengyue Wu, Wenwu Wang, Mark D. Plumbley |
| 2023 | ICASSP | Simple Pooling Front-Ends for Efficient Audio Classification. | Xubo Liu, Haohe Liu, Qiuqiang Kong, Xinhao Mei, Mark D. Plumbley, Wenwu Wang |
| 2023 | ICASSP | Efficient Similarity-Based Passive Filter Pruning for Compressing CNNS. | Arshdeep Singh, Mark D. Plumbley |
| 2023 | ICML | AudioLDM: Text-to-Audio Generation with Latent Diffusion Models. | Haohe Liu, Zehua Chen, Yi Yuan, Xinhao Mei, Xubo Liu, Danilo P. Mandic, Wenwu Wang, Mark D. Plumbley |
| 2023 | Interspeech | Adapting Language-Audio Models as Few-Shot Audio Learners. | Jinhua Liang, Xubo Liu, Haohe Liu, Huy Phan, Emmanouil Benetos, Mark D. Plumbley, Wenwu Wang |
| 2023 | Interspeech | Visually-Aware Audio Captioning With Adaptive Audio-Visual Attention. | Xubo Liu, Qiushi Huang, Xinhao Mei, Haohe Liu, Qiuqiang Kong, Jianyuan Sun, Shengchen Li, Tom Ko, Yu Zhang, H. Lilian Tang, Mark D. Plumbley, Volkan Kili, Wenwu Wang |
| 2023 | Interspeech | Ontology-aware Learning and Evaluation for Audio Tagging. | Haohe Liu, Qiuqiang Kong, Xubo Liu, Xinhao Mei, Wenwu Wang, Mark D. Plumbley |
| 2023 | Interspeech | Dual Transformer Decoder based Features Fusion Network for Automated Audio Captioning. | Jianyuan Sun, Xubo Liu, Xinhao Mei, Volkan Kili, Mark D. Plumbley, Wenwu Wang |
| 2022 | ICASSP | A Track-Wise Ensemble Event Independent Network for Polyphonic Sound Event Localization and Detection. | Jinbo Hu, Yin Cao, Ming Wu, Qiuqiang Kong, Feiran Yang, Mark D. Plumbley, Jun Yang |
| 2022 | ICASSP | Diverse Audio Captioning Via Adversarial Training. | Xinhao Mei, Xubo Liu, Jianyuan Sun, Mark D. Plumbley, Wenwu Wang |
| 2022 | Interspeech | Separate What You Describe: Language-Queried Audio Source Separation. | Xubo Liu, Haohe Liu, Qiuqiang Kong, Xinhao Mei, Jinzheng Zhao, Qiushi Huang, Mark D. Plumbley, Wenwu Wang |
| 2022 | Interspeech | On Metric Learning for Audio-Text Cross-Modal Retrieval. | Xinhao Mei, Xubo Liu, Jianyuan Sun, Mark D. Plumbley, Wenwu Wang |
| 2022 | Interspeech | A Passive Similarity based CNN Filter Pruning for Efficient Acoustic Scene Classification. | Arshdeep Singh, Mark D. Plumbley |
| 2021 | ICASSP | An Improved Event-Independent Network for Polyphonic Sound Event Localization and Detection. | Yin Cao, Turab Iqbal, Qiuqiang Kong, Fengyan An, Wenwu Wang, Mark D. Plumbley |
| 2021 | ICASSP | Weighted Magnitude-Phase Loss for Speech Dereverberation. | Jingshu Zhang, Mark D. Plumbley, Wenwu Wang |
| 2020 | ICASSP | Learning With Out-of-Distribution Data for Audio Classification. | Turab Iqbal, Yin Cao, Qiuqiang Kong, Mark D. Plumbley, Wenwu Wang |
| 2020 | ICASSP | Source Separation with Weakly Labelled Data: an Approach to Computational Auditory Scene Analysis. | Qiuqiang Kong, Yuxuan Wang, Xuchen Song, Yin Cao, Wenwu Wang, Mark D. Plumbley |
| 2019 | ICASSP | Sound Event Detection with Sequentially Labelled Data Based on Connectionist Temporal Classification and Unsupervised Clustering. | Yuanbo Hou, Qiuqiang Kong, Shengchen Li, Mark D. Plumbley |
| 2019 | ICASSP | Acoustic Scene Generation with Conditional Samplernn. | Qiuqiang Kong, Yong Xu, Turab Iqbal, Yin Cao, Wenwu Wang, Mark D. Plumbley |
| 2019 | ICASSP | Generalisation in Environmental Sound Classification: The 'Making Sense of Sounds' Data Set and Challenge. | Christian Kroos, Oliver Bones, Yin Cao, Lara Harris, Philip J. B. Jackson, William J. Davies, Wenwu Wang, Trevor J. Cox, Mark D. Plumbley |
| 2019 | ICASSP | Acoustic Event Detection from Weakly Labeled Data Using Auditory Salience. | Zuzanna Podwinska, Iwona Sobieraj, Bruno M. Fazenda, William J. Davies, Mark D. Plumbley |
| 2019 | ICASSP | Attention-based Atrous Convolutional Neural Networks: Visualisation and Understanding Perspectives of Acoustic Scenes. | Zhao Ren, Qiuqiang Kong, Jing Han, Mark D. Plumbley, Bjrn W. Schuller |
| 2019 | IJCAI | Single-Channel Signal Separation and Deconvolution with Generative Adversarial Networks. | Qiuqiang Kong, Yong Xu, Philip J. B. Jackson, Wenwu Wang, Mark D. Plumbley |
| 2018 | ACSSC | Predicting the perceived level of reverberation using machine learning. | Saeid Safavi, Andy Pearce, Wenwu Wang, Mark D. Plumbley |
| 2018 | HCI | Supporting Audiography: Design of a System for Sentimental Sound Recording, Classification and Playback. | Tijs Duel, David M. Frohlich, Christian Kroos, Yong Xu, Philip J. B. Jackson, Mark D. Plumbley |
| 2018 | ICASSP | Large-Scale Weakly Supervised Audio Classification Using Gated Convolutional Neural Network. | Yong Xu, Qiuqiang Kong, Wenwu Wang, Mark D. Plumbley |
| 2018 | ICASSP | Synthesis of Images by Two-Stage Generative Adversarial Networks. | Qiang Huang, Philip J. B. Jackson, Mark D. Plumbley, Wenwu Wang |
| 2018 | ICASSP | Audio Set Classification with Attention Model: A Probabilistic Perspective. | Qiuqiang Kong, Yong Xu, Wenwu Wang, Mark D. Plumbley |
| 2018 | ICASSP | A Joint Separation-Classification Model for Sound Event Detection of Weakly Labelled Data. | Qiuqiang Kong, Yong Xu, Wenwu Wang, Mark D. Plumbley |
| 2018 | ICASSP | Inexact Proximal Operators for 𝓁 | Cian O'Brien, Mark D. Plumbley |
| 2018 | ICASSP | Orthogonality-Regularized Masked NMF for Learning on Weakly Labeled Audio Data. | Iwona Sobieraj, Lucas Rencker, Mark D. Plumbley |
| 2018 | ICASSP | BSS Eval or Peass? Predicting the Perception of Singing-Voice Separation. | Dominic Ward, Hagen Wierstorf, Russell D. Mason, Emad M. Grais, Mark D. Plumbley |
| 2017 | ICASSP | Assessment of musical noise using localization of isolated peaks in time-frequency domain. | Ronan Hamon, Valentin Emiya, Lucas Rencker, Wenwu Wang, Mark D. Plumbley |
| 2017 | ICASSP | Fast tagging of natural sounds using marginal co-regularization. | Qiang Huang, Yong Xu, Philip J. B. Jackson, Wenwu Wang, Mark D. Plumbley |
| 2017 | ICASSP | A joint detection-classification model for audio tagging of weakly labelled data. | Qiuqiang Kong, Yong Xu, Wenwu Wang, Mark D. Plumbley |
| 2017 | ICASSP | A greedy algorithm with learned statistics for sparse signal reconstruction. | Lucas Rencker, Wenwu Wang, Mark D. Plumbley |
| 2017 | IJCNN | Convolutional gated recurrent neural network incorporating spatial features for audio tagging. | Yong Xu, Qiuqiang Kong, Qiang Huang, Wenwu Wang, Mark D. Plumbley |
| 2017 | Interspeech | Learning the Mapping Function from Voltage Amplitudes to Sensor Positions in 3D-EMA Using Deep Neural Networks. | Christian Kroos, Mark D. Plumbley |
| 2017 | Interspeech | Attention and Localization Based on a Deep Convolutional Recurrent Model for Weakly Supervised Audio Tagging. | Yong Xu, Qiuqiang Kong, Qiang Huang, Wenwu Wang, Mark D. Plumbley |
| 2017 | MMSP | Binaural and log-power spectra features with deep neural networks for speech-noise separation. | Alfredo Zermini, Qingju Liu, Yong Xu, Mark D. Plumbley, Dave Betts, Wenwu Wang |
| 2016 | ICASSP | Detection of overlapping acoustic events using a temporally-constrained probabilistic model. | Emmanouil Benetos, Grgoire Lafay, Mathieu Lagrange, Mark D. Plumbley |
| 2016 | Interspeech | Combining Mask Estimates for Single Channel Audio Source Separation Using Deep Neural Networks. | Emad M. Grais, Gerard Roma, Andrew J. R. Simpson, Mark D. Plumbley |
| 2015 | ICASSP | A dynamic programming variant of non-negative matrix deconvolution for the transcription of struck string instruments. | Sebastian Ewert, Mark D. Plumbley, Mark B. Sandler |
| 2015 | ICASSP | Non-negative matrix factorisation incorporating greedy hellinger sparse coding applied to polyphonic music transcription. | Ken O'Hanlon, Mark B. Sandler, Mark D. Plumbley |
| 2014 | ICASSP | Accounting for phase cancellations in non-negative matrix factorization using weighted distances. | Sebastian Ewert, Mark D. Plumbley, Mark B. Sandler |
| 2014 | ICASSP | Improving instrument recognition in polyphonic music through system integration. | Dimitrios Giannoulis, Emmanouil Benetos, Anssi Klapuri, Mark D. Plumbley |
| 2014 | ICASSP | Polyphonic piano transcription using non-negative Matrix Factorisation with group sparsity. | Ken O'Hanlon, Mark D. Plumbley |
| 2014 | Interspeech | Phase-based harmonic/percussive separation. | Estefana Cano, Mark D. Plumbley, Christian Dittmar |
| 2013 | ICASSP | Score informed audio source separation using constrained nonnegative matrix factorization and score synthesis. | Joachim Fritsch, Mark D. Plumbley |
| 2013 | ICASSP | Recognition of harmonic sounds in polyphonic audio using a missing feature approach. | Dimitrios Giannoulis, Anssi Klapuri, Mark D. Plumbley |
| 2013 | ICASSP | Behavior of greedy sparse representation algorithms on nested supports. | Boris Mailh, Bob L. T. Sturm, Mark D. Plumbley |
| 2013 | ICASSP | Automatic Music Transcription using row weighted decompositions. | Ken O'Hanlon, Mark D. Plumbley |
| 2013 | ICASSP | Improved multiple birdsong tracking with distribution derivative method and Markov renewal process clustering. | Dan Stowell, Saso Musevic, Jordi Bonada, Mark D. Plumbley |
| 2012 | ICASSP | Sound Software: Towards software reuse in audio and music research. | Chris Cannam, Lus Figueira, Mark D. Plumbley |
| 2012 | ICASSP | Analysis-based sparse reconstruction with synthesis-based solvers. | Nicolae Cleju, Maria G. Jafari, Mark D. Plumbley |
| 2012 | ICASSP | Instrumentation-based music similarity using sparse representations. | Hiromasa Fujihara, Anssi Klapuri, Mark D. Plumbley |
| 2012 | ICASSP | INK-SVD: Learning incoherent dictionaries for sparse representations. | Boris Mailh, Daniele Barchiesi, Mark D. Plumbley |
| 2012 | ICASSP | Structured sparsity for automatic music transcription. | Ken O'Hanlon, Hidehisa Nagano, Mark D. Plumbley |
| 2011 | ICASSP | A constrained matching pursuit approach to audio declipping. | Amir Adler, Valentin Emiya, Maria G. Jafari, Michael Elad, Rmi Gribonval, Mark D. Plumbley |
| 2011 | ICASSP | Dictionary learning of convolved signals. | Daniele Barchiesi, Mark D. Plumbley |
| 2011 | ICASSP | Separating sources from sequentially acquired mixtures of heart signals. | Fbio de Lima Hedayioglu, Maria G. Jafari, Sandra da Silva Mattos, Mark D. Plumbley, Miguel T. Coimbra |
| 2010 | ICASSP | Note onset detection using rhythmic structure. | Norberto Degara, Antonio S. Pena, Matthew E. P. Davies, Mark D. Plumbley |
| 2010 | ICASSP | Gradient Polytope Faces Pursuit for large scale sparse recovery problems. | Aris Gretsistas, Ivan Damnjanovic, Mark D. Plumbley |
| 2010 | ICASSP | An L1 criterion for dictionary learning by subspace identification. | Florent Jaillet, Rmi Gribonval, Mark D. Plumbley, Hadi Zayyani |
| 2010 | ICASSP | Performance following: Tracking a performance without a score. | Adam M. Stark, Mark D. Plumbley |
| 2009 | ICASSP | Benchmarking flexible adaptive time-frequency transforms for underdetermined audio source separation. | Andrew Nesbit, Emmanuel Vincent, Mark D. Plumbley |
| 2009 | ICASSP | Using phase linearity in frequency-domain ICA to tackle the permutation problem. | Keisuke Toyama, Mark D. Plumbley |
| 2009 | IDA | Extension of Sparse, Adaptive Signal Decompositions to Semi-blind Audio Source Separation. | Andrew Nesbit, Emmanuel Vincent, Mark D. Plumbley |
| 2009 | IDA | Estimating Phase Linearity in the Frequency-Domain ICA Demixing Matrix. | Keisuke Toyama, Mark D. Plumbley |
| 2008 | ICANN | Natural Conjugate Gradient on Complex Flag Manifolds for Complex Independent Subspace Analysis. | Yasunori Nishimori, Shotaro Akaho, Mark D. Plumbley |
| 2008 | ICASSP | Oracle estimation of adaptive cosine packet transforms for underdetermined audio source separation. | Andrew Nesbit, Mark D. Plumbley |
| 2007 | ICASSP | On the Use of Entropy for Beat Tracking Evaluation. | Matthew E. P. Davies, Mark D. Plumbley |
| 2007 | ICASSP | Flag Manifolds for Subspace ICA Problems. | Yasunori Nishimori, Shotaro Akaho, Samer A. Abdallah, Mark D. Plumbley |
| 2007 | ICASSP | Geometry and Manifolds for Independent Component Analysis. | Mark D. Plumbley |
| 2005 | ICASSP | Beat tracking with a two state model [music applications]. | Matthew E. P. Davies, Mark D. Plumbley |
| 2000 | IJCNN | On-Line Connectionist Q-Learning Produces Unreliable Performance with A Synonym Finding Task. | Ian Johnson, Mark D. Plumbley |
| 1997 | ICASSP | Communications and neural networks: theory and practice. | Mark D. Plumbley |
| 1991 | ICASSP | The effect of receptor signal-to-noise levels on optimal filtering in a sensory system. | Mark D. Plumbley, Frank Fallside |