| 2025 | EMNLP | iKnow-audio: Integrating Knowledge Graphs with Audio-Language Models. | Michel Olvera, Changhong Wang, Paraskevas Stamatiadis, Gal Richard, Slim Essid |
| 2025 | ICASSP | Perceptual Noise-Masking with Music through Deep Spectral Envelope Shaping. | Clmentine Berger, Roland Badeau, Slim Essid |
| 2025 | ICASSP | O-EENC-SD: Efficient Online End-to-End Neural Clustering for Speaker Diarization. | Elio Gruttadauria, Mathieu Fontaine, Jonathan Le Roux, Slim Essid |
| 2025 | ICASSP | Multiple Choice Learning for Efficient Speech Separation with Many Speakers. | David Perera, Franois Derrida, Tho Mariotte, Gal Richard, Slim Essid |
| 2025 | ICASSP | Masked Latent Prediction and Classification for Self-Supervised Audio Representation Learning. | Aurian Quelennec, Pierre Chouteau, Geoffroy Peeters, Slim Essid |
| 2025 | ICASSP | Contrastive Knowledge Distillation for Embedding Refinement in Personalized Speech Enhancement. | Thomas Serre, Mathieu Fontaine, ric Benhaim, Slim Essid |
| 2025 | Interspeech | MTSE: Multi-Target Speaker Extraction for Conversation Scenarios. | Thomas Serre, Mathieu Fontaine, Eric Benhaim, Slim Essid |
| 2024 | CVPR | Collaborating Foundation Models for Domain Generalized Semantic Segmentation. | Yasser Benigmim, Subhankar Roy, Slim Essid, Vicky Kalogeiton, Stphane Lathuilire |
| 2024 | ICASSP | Adapting Pitch-Based Self Supervised Learning Models for Tempo Estimation. | Antonin Gagner, Slim Essid, Geoffroy Peeters |
| 2024 | ICASSP | Online Speaker Diarization of Meetings Guided by Speech Separation. | Elio Gruttadauria, Mathieu Fontaine, Slim Essid |
| 2024 | ICASSP | On The Choice of the Optimal Temporal Support for Audio Classification with Pre-Trained Embeddings. | Aurian Quelennec, Michel Olvera, Geoffroy Peeters, Slim Essid |
| 2024 | ICASSP | A Lightweight Dual-Stage Framework for Personalized Speech Enhancement Based on Deepfilternet2. | Thomas Serre, Mathieu Fontaine, ric Benhaim, Geoffroy Dutour, Slim Essid |
| 2024 | ICML | Winner-takes-all learners are geometry-aware conditional density estimators. | Victor Letzelter, David Perera, Cdric Rommel, Mathieu Fontaine, Slim Essid, Gal Richard, Patrick Prez |
| 2023 | CVPR | One-shot Unsupervised Domain Adaptation with Personalized Diffusion Models. | Yasser Benigmim, Subhankar Roy, Slim Essid, Vicky Kalogeiton, Stphane Lathuilire |
| 2023 | ICASSP | Cosmopolite Sound Monitoring (CoSMo): A Study of Urban Sound Event Detection Systems Generalizing to Multiple Cities. | Florian Angulo, Slim Essid, Geoffroy Peeters, Christophe Mietlicki |
| 2023 | ICASSP | Fine-Tuning Strategies for Faster Inference Using Speech Self-Supervised Models: A Comparative Study. | Salah Zaiem, Robin Algayres, Titouan Parcollet, Slim Essid, Mirco Ravanelli |
| 2023 | Interspeech | Speech Self-Supervised Representation Benchmarking: Are We Doing it Right? | Salah Zaiem, Youcef Kemiche, Titouan Parcollet, Slim Essid, Mirco Ravanelli |
| 2023 | Interspeech | Automatic Data Augmentation for Domain Adapted Fine-Tuning of Self-Supervised Speech Representations. | Salah Zaiem, Titouan Parcollet, Slim Essid |
| 2022 | Interspeech | Automatic Data Augmentation Selection and Parametrization in Contrastive Self-Supervised Speech Representation Learning. | Salah Zaiem, Titouan Parcollet, Slim Essid |
| 2022 | LREC | Opinions in Interactions : New Annotations of the SEMAINE Database. | Valentin Barrire, Slim Essid, Chlo Clavel |
| 2021 | ICASSP | Neuro-Steered Music Source Separation With EEG-Based Auditory Attention Decoding And Contrastive-NMF. | Giorgia Cantisani, Slim Essid, Gal Richard |
| 2021 | ICASSP | Distributed Speech Separation in Spatially Unconstrained Microphone Arrays. | Nicolas Furnon, Romain Serizel, Irina Illina, Slim Essid |
| 2021 | Interspeech | Conditional Independence for Pretext Task Selection in Self-Supervised Speech Representation Learning. | Salah Zaiem, Titouan Parcollet, Slim Essid |
| 2020 | ICASSP | DNN-based Distributed Multichannel Mask Estimation for Speech Enhancement in Microphone Arrays. | Nicolas Furnon, Romain Serizel, Irina Illina, Slim Essid |
| 2019 | EMNLP | From the Token to the Review: A Hierarchical Multimodal approach to Opinion Mining. | Alexandre Garcia, Pierre Colombo, Florence d'Alch-Buc, Slim Essid, Chlo Clavel |
| 2019 | ICASSP | A Music Structure Informed Downbeat Tracking System Using Skip-chain Conditional Random Fields and Deep Learning. | Magdalena Fuentes, Brian McFee, Hlne C. Crayencour, Slim Essid, Juan Pablo Bello |
| 2018 | CVPR | Weakly Supervised Representation Learning for Unsynchronized Audio-Visual Events. | Sanjeel Parekh, Slim Essid, Alexey Ozerov, Ngoc Q. K. Duong, Patrick Prez, Gal Richard |
| 2018 | ICASSP | Attitude Classification in Adjacency Pairs of a Human-Agent Interaction with Hidden Conditional Random Fields. | Valentin Barrire, Chlo Clavel, Slim Essid |
| 2018 | ICASSP | An Ensemble Learning Approach to Detect Epileptic Seizures from Long Intracranial EEG Recordings. | Jean-Baptiste Schiratti, Jean-Eudes Le Douget, Michel Le Van Quyen, Slim Essid, Alexandre Gramfort |
| 2018 | ICML | Structured Output Learning with Abstention: Application to Accurate Opinion Prediction. | Alexandre Garcia, Chlo Clavel, Slim Essid, Florence d'Alch-Buc |
| 2017 | ICASSP | Overlapping sound event detection with supervised Nonnegative Matrix Factorization. | Victor Bisot, Slim Essid, Gal Richard |
| 2017 | ICASSP | Motion informed audio source separation. | Sanjeel Parekh, Slim Essid, Alexey Ozerov, Ngoc Q. K. Duong, Patrick Prez, Gal Richard |
| 2017 | ICASSP | Supervised group nonnegative matrix factorisation with similarity constraints and applications to speaker identification. | Romain Serizel, Victor Bisot, Slim Essid, Gal Richard |
| 2017 | ICMI | UE-HRI: a new dataset for the study of user engagement in spontaneous human-robot interactions. | Atef Ben Youssef, Chlo Clavel, Slim Essid, Miriam Bilac, Marine Chamoux, Angelica Lim |
| 2017 | Interspeech | Opinion Dynamics Modeling for Movie Review Transcripts Classification with Hidden Conditional Random Fields. | Valentin Barrire, Chlo Clavel, Slim Essid |
| 2016 | ICASSP | Acoustic scene classification with matrix factorization for unsupervised feature learning. | Victor Bisot, Romain Serizel, Slim Essid, Gal Richard |
| 2016 | ICASSP | Group nonnegative matrix factorisation with speaker and session variability compensation for speaker identification. | Romain Serizel, Slim Essid, Gal Richard |
| 2016 | ICIP | Machine listening techniques as a complement to video image analysis in forensics. | Romain Serizel, Victor Bisot, Slim Essid, Gal Richard |
| 2015 | ICASSP | A Conditional Random Field system for beat tracking. | Thomas Fillon, Cyril Joder, Simon Durand, Slim Essid |
| 2014 | ICASSP | Assessment of new spectral features for eeg-based emotion recognition. | Anne-Claire Conneau, Slim Essid |
| 2014 | ICASSP | Gesture recognition using a NMF-based representation of motion-traces extracted from depth silhouettes. | Aymeric Masurelle, Slim Essid, Gal Richard |
| 2014 | ICASSP | Piecewise constant nonnegative matrix factorization. | Nicolas Seichepine, Slim Essid, Cdric Fvotte, Olivier Capp |
| 2013 | ICASSP | Non-negative matrix factorization for single-channel EEG artifact rejection. | Ccilia Damon, Antoine Liutkus, Alexandre Gramfort, Slim Essid |
| 2013 | ICASSP | Probabilistic dance performance alignment by fusion of multimodal features. | Anglique Dremeau, Slim Essid |
| 2013 | ICASSP | Soft nonnegative matrix co-factorizationwith application to multimodal speaker diarization. | Nicolas Seichepine, Slim Essid, Cdric Fvotte, Olivier Capp |
| 2012 | ICASSP | A single-class SVM based algorithm for computing an identifiable NMF. | Slim Essid |
| 2012 | ICASSP | An advanced virtual dance performance evaluator. | Slim Essid, Dimitrios S. Alexiadis, Robin Tournemenne, Marc Gowing, Philip Kelly, David S. Monaghan, Petros Daras, Anglique Dremeau, Noel E. O'Connor |
| 2012 | ICASSP | A regressive boosting approach to automatic audio tagging based on soft annotator fusion. | Rmi Foucard, Slim Essid, Mathieu Lagrange, Gal Richard |
| 2012 | ICIP | Decomposing the video editing structure of a talk-show using nonnegative matrix factorization. | Slim Essid, Cdric Fvotte |
| 2011 | ICASSP | Hidden Discrete Tempo Model: A tempo-aware timing model for audio-to-score alignment. | Cyril Joder, Slim Essid, Gal Richard |
| 2010 | ICASSP | A comparative study of tonal acoustic features for a symbolic level music-to-score alignment. | Cyril Joder, Slim Essid, Gal Richard |
| 2010 | ICIP | Robust visual features for the multimodal identification of unregistered speakers in TV talk-shows. | Flicien Vallet, Slim Essid, Jean Carrive, Gal Richard |
| 2009 | ICASSP | Incorporating prior knowledge on the digital media creation process into audio classifiers. | Maxime Lardeur, Slim Essid, Gal Richard, Martin Haller, Thomas Sikora |
| 2008 | ICIP | A collaborative approach to automatic rushes video summarization. | Werner Bailer, Emilie Dumont, Slim Essid, Bernard Mrialdo |
| 2007 | ICASSP | Combined Supervised and Unsupervised Approaches for Automatic Segmentation of Radiophonic Audio Streams. | Gal Richard, Mathieu Ramona, Slim Essid |
| 2006 | ICASSP | Hierarchical Classification of Musical Instruments on Solo Recordings. | Slim Essid, Gal Richard, Bertrand David |
| 2005 | ICASSP | Instrument recognition in polyphonic music. | Slim Essid, Gal Richard, Bertrand David |