Themos Stafylakis
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
53
Venues
10
Active years
2007–2026
Best venue rank
A*
Where they publish
Papers
53 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | AAAI | MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence. | Sonal Kumar, Simon Sedlcek, Vaibhavi Lokegaonkar, Fernando Lpez, Wenyi Yu, Nishit Anand, Hyeonggon Ryu, Lichang Chen, Maxim Plicka, Miroslav Hlavcek, William Fineas Ellingwood, Sathvik Udupa, Siyuan Hou, Allison Ferner, Sara Barahona, Cecilia Bolaos, Satish Rahi, Laura Herrera-Alarcn, Satvik Dixit, Rupali S. Patil, Soham Deshmukh, Lasha Koroshinadze, Yao Liu, Leibny Paola Garca-Perera, Eleni Zanou, Themos Stafylakis, Joon Son Chung, David Harwath, Chao Zhang, Dinesh Manocha, Alicia Lozano-Diez, Santosh Kesiraju, Sreyan Ghosh, Ramani Duraiswami |
| 2025 | ASRU | State-of-the-art Embeddings with Video-free Segmentation of the Source VoxCeleb Data. | Sara Barahona, Ladislav Mosner, Themos Stafylakis, Oldrich Plchot, Junyi Peng, Luks Burget, Jan Cernock |
| 2025 | ICASSP | CA-MHFA: A Context-Aware Multi-Head Factorized Attentive Pooling for SSL-Based Speaker Verification. | Junyi Peng, Ladislav Mosner, Lin Zhang, Oldrich Plchot, Themos Stafylakis, Luks Burget, Jan Cernock |
| 2025 | Interspeech | Analysis of ABC Frontend Audio Systems for the NIST-SRE24. | Sara Barahona, Anna Silnova, Ladislav Mosner, Junyi Peng, Oldrich Plchot, Johan Rohdin, Lin Zhang, Jiangyu Han, Petr Plka, Federico Landini, Luks Burget, Themos Stafylakis, Sandro Cumani, Dominik Bobos, Miroslav Hlavcek, Martin Kodovsky, Toms Pavlcek |
| 2025 | Interspeech | Synthetic Speech Source Tracing using Metric Learning. | Dimitrios Koutsianos, Stavros Zacharopoulos, Yannis Panagakis, Themos Stafylakis |
| 2025 | SIGdial | Building Open-Retrieval Conversational Question Answering Systems by Generating Synthetic Data and Decontextualizing User Questions. | Christos Vlachos, Nikolaos Stylianou, Alexandra Fiotaki, Spiros Methenitis, Elisavet Palogiannidi, Themos Stafylakis, Ion Androutsopoulos |
| 2024 | ACL | Comparing Data Augmentation Methods for End-to-End Task-Oriented Dialog Systems. | Christos Vlachos, Themos Stafylakis, Ion Androutsopoulos |
| 2024 | Interspeech | Challenging margin-based speaker embedding extractors by using the variational information bottleneck. | Themos Stafylakis, Anna Silnova, Johan Rohdin, Oldrich Plchot, Luks Burget |
| 2023 | EMNLP | A Simple Baseline for Knowledge-Based Visual Question Answering. | Alexandros Xenos, Themos Stafylakis, Ioannis Patras, Georgios Tzimiropoulos |
| 2023 | ICASSP | Speech-Based Emotion Recognition with Self-Supervised Models Using Attentive Channel-Wise Correlations and Label Smoothing. | Sofoklis Kakouros, Themos Stafylakis, Ladislav Mosner, Luks Burget |
| 2023 | ICASSP | Parameter-Efficient Transfer Learning of Pre-Trained Transformer Models for Speaker Verification Using Adapters. | Junyi Peng, Themos Stafylakis, Rongzhi Gu, Oldrich Plchot, Ladislav Mosner, Luks Burget, Jan Cernock |
| 2023 | Interspeech | Description and Analysis of ABC Submission to NIST LRE 2022. | Pavel Matejka, Anna Silnova, Josef Slavcek, Ladislav Mosner, Oldrich Plchot, Michal Klco, Junyi Peng, Themos Stafylakis, Luks Burget |
| 2023 | Interspeech | Improving Speaker Verification with Self-Pretrained Transformer Models. | Junyi Peng, Oldrich Plchot, Themos Stafylakis, Ladislav Mosner, Luks Burget, Jan Cernock |
| 2022 | Interspeech | Probabilistic Spherical Discriminant Analysis: An Alternative to PLDA for length-normalized embeddings. | Niko Brummer, Albert Swart, Ladislav Mosner, Anna Silnova, Oldrich Plchot, Themos Stafylakis, Luks Burget |
| 2022 | Interspeech | Training speaker embedding extractors using multi-speaker audio with unknown speaker boundaries. | Themos Stafylakis, Ladislav Mosner, Oldrich Plchot, Johan Rohdin, Anna Silnova, Luks Burget, Jan Cernock |
| 2021 | Interspeech | Speaker Embeddings by Modeling Channel-Wise Correlations. | Themos Stafylakis, Johan Rohdin, Luks Burget |
| 2020 | BMVC | Seeing wake words: Audio-visual Keyword Spotting. | Liliane Momeni, Triantafyllos Afouras, Themos Stafylakis, Samuel Albanie, Andrew Zisserman |
| 2020 | ICASSP | End-to-End Architectures for ASR-Free Spoken Language Understanding. | Elisavet Palogiannidi, Ioannis Gkinis, George Mastrapas, Petr Mizera, Themos Stafylakis |
| 2019 | ICASSP | Speaker Verification Using End-to-end Adversarial Language Adaptation. | Johan Rohdin, Themos Stafylakis, Anna Silnova, Hossein Zeinali, Luks Burget, Oldrich Plchot |
| 2019 | ICASSP | How to Improve Your Speaker Embeddings Extractor in Generic Toolkits. | Hossein Zeinali, Luks Burget, Johan Rohdin, Themos Stafylakis, Jan Honza Cernock |
| 2019 | Interspeech | Privacy-Preserving Speaker Recognition with Cohort Score Normalisation. | Andreas Nautsch, Jose Patino, Amos Treiber, Themos Stafylakis, Petr Mizera, Massimiliano Todisco, Thomas Schneider, Nicholas W. D. Evans |
| 2019 | Interspeech | Self-Supervised Speaker Embeddings. | Themos Stafylakis, Johan Rohdin, Oldrich Plchot, Petr Mizera, Luks Burget |
| 2019 | Interspeech | Detecting Spoofing Attacks Using VGG and SincNet: BUT-Omilia Submission to ASVspoof 2019 Challenge. | Hossein Zeinali, Themos Stafylakis, Georgia Athanasopoulou, Johan Rohdin, Ioannis Gkinis, Luks Burget, Jan Cernock |
| 2018 | ECCV | Zero-Shot Keyword Spotting for Visual Speech Recognition In-the-wild. | Themos Stafylakis, Georgios Tzimiropoulos |
| 2018 | ICASSP | End-to-End Audiovisual Speech Recognition. | Stavros Petridis, Themos Stafylakis, Pingchuan Ma, Feipeng Cai, Georgios Tzimiropoulos, Maja Pantic |
| 2018 | ICASSP | Deep Word Embeddings for Visual Speech Recognition. | Themos Stafylakis, Georgios Tzimiropoulos |
| 2017 | Interspeech | The I4U Mega Fusion and Collaboration for NIST Speaker Recognition Evaluation 2016. | Kong-Aik Lee, Ville Hautamki, Tomi Kinnunen, Anthony Larcher, Chunlei Zhang, Andreas Nautsch, Themos Stafylakis, Gang Liu, Mickal Rouvier, Wei Rao, Federico Alegre, J. Ma, Man-Wai Mak, Achintya Kumar Sarkar, Hctor Delgado, Rahim Saeidi, Hagai Aronowitz, Aleksandr Sizov, Hanwu Sun, Trung Hieu Nguyen, Guangsen Wang, Bin Ma, Ville Vestman, Md. Sahidullah, M. Halonen, Anssi Kanervisto, Gal Le Lan, Fahimeh Bahmaninezhad, Sergey Isadskiy, Christian Rathgeb, Christoph Busch, Georgios Tzimiropoulos, Q. Qian, Z. Wang, Q. Zhao, T. Wang, H. Li, J. Xue, S. Zhu, R. Jin, T. Zhao, Pierre-Michel Bousquet, Moez Ajili, Waad Ben Kheder, Driss Matrouf, Zhi Hao Lim, Chenglin Xu, Haihua Xu, Xiong Xiao, Eng Siong Chng, Benoit G. B. Fauve, Kaavya Sriskandaraja, Vidhyasaharan Sethu, W. W. Lin, Dennis Alexander Lehmann Thomsen, Zheng-Hua Tan, Massimiliano Todisco, Nicholas W. D. Evans, Haizhou Li, John H. L. Hansen, Jean-Franois Bonastre, Eliathamby Ambikairajah |
| 2017 | Interspeech | Combining Residual Networks with LSTMs for Lipreading. | Themos Stafylakis, Georgios Tzimiropoulos |
| 2016 | ICASSP | Towards PLDA-RBM based speaker recognition in mobile environment: Designing stacked/deep PLDA-RBM systems. | Andreas Nautsch, Hong Hao, Themos Stafylakis, Christian Rathgeb, Christoph Busch |
| 2015 | ICASSP | JFA modeling with left-to-right structure and a new backend for text-dependent speaker recognition. | Patrick Kenny, Themos Stafylakis, Jahangir Alam, Marcel Kockmann |
| 2015 | Interspeech | Development of CRIM system for the automatic speaker verification spoofing and countermeasures challenge 2015. | Md. Jahangir Alam, Patrick Kenny, Gautam Bhattacharya, Themos Stafylakis |
| 2015 | Interspeech | Combining amplitude and phase-based features for speaker verification with short duration utterances. | Md. Jahangir Alam, Patrick Kenny, Themos Stafylakis |
| 2015 | Interspeech | An i-vector backend for speaker verification. | Patrick Kenny, Themos Stafylakis, Md. Jahangir Alam, Marcel Kockmann |
| 2015 | Interspeech | The reddots data collection for speaker recognition. | Kong-Aik Lee, Anthony Larcher, Guangsen Wang, Patrick Kenny, Niko Brmmer, David A. van Leeuwen, Hagai Aronowitz, Marcel Kockmann, Carlos Vaquero, Bin Ma, Haizhou Li, Themos Stafylakis, Md. Jahangir Alam, Albert Swart, Javier Perez |
| 2015 | Interspeech | JFA for speaker recognition with random digit strings. | Themos Stafylakis, Patrick Kenny, Md. Jahangir Alam, Marcel Kockmann |
| 2014 | ICASSP | Unscented transform for ivector-based noisy speaker recognition. | David Martnez Gonzlez, Luks Burget, Themos Stafylakis, Yun Lei, Patrick Kenny, Eduardo Lleida |
| 2014 | ICASSP | I-vector-based speaker adaptation of deep neural networks for French broadcast audio transcription. | Vishwa Gupta, Patrick Kenny, Pierre Ouellet, Themos Stafylakis |
| 2014 | ICASSP | JFA-based front ends for speaker recognition. | Patrick Kenny, Themos Stafylakis, Pierre Ouellet, Md. Jahangir Alam |
| 2014 | Interspeech | In-domain versus out-of-domain training for text-dependent JFA. | Patrick Kenny, Themos Stafylakis, Md. Jahangir Alam, Pierre Ouellet, Marcel Kockmann |
| 2013 | ICASSP | PLDA for speaker verification with utterances of arbitrary duration. | Patrick Kenny, Themos Stafylakis, Pierre Ouellet, Md. Jahangir Alam, Pierre Dumouchel |
| 2013 | ICASSP | Efficient iterative mean shift based cosine dissimilarity for multi-recording speaker clustering. | Mohammed Senoussaoui, Patrick Kenny, Pierre Dumouchel, Themos Stafylakis |
| 2013 | ICASSP | Compensation for inter-frame correlations in speaker diarization and recognition. | Themos Stafylakis, Patrick Kenny, Vishwa Gupta, Pierre Dumouchel |
| 2013 | Interspeech | Text-dependent speaker recognition using PLDA with uncertainty propagation. | Themos Stafylakis, Patrick Kenny, Pierre Ouellet, Javier Perez, Marcel Kockmann, Pierre Dumouchel |
| 2012 | ICASSP | Music tempo estimation and beat tracking by applying source separation and metrical relations. | Aggelos Gkiokas, Vassilis Katsouros, George Carayannis, Themos Stafylakis |
| 2012 | Interspeech | PLDA using Gaussian Restricted Boltzmann Machines with application to Speaker Verification. | Themos Stafylakis, Patrick Kenny, Mohammed Senoussaoui, Pierre Dumouchel |
| 2011 | ICASSP | Closed-form expressions vs. BIC: A comparison for speaker clustering. | Themos Stafylakis, Xavier Anguera Mir, Vassilis Katsouros, George Carayannis |
| 2011 | ICDAR | Enhancing Handwritten Word Segmentation by Employing Local Spatial Features. | Fotini Simistira, Vassilis Papavassiliou, Themos Stafylakis, Vassilis Katsouros |
| 2010 | ICASSP | A new penalty term for the BIC with respect to speaker diarization. | Themos Stafylakis, Georgios Tzimiropoulos, Vassilis Katsouros, George Carayannis |
| 2010 | Interspeech | Improvements to the equal-parameter BIC for speaker diarization. | Themos Stafylakis, Xavier Anguera |
| 2009 | Interspeech | Redefining the Bayesian information criterion for speaker diarisation. | Themos Stafylakis, Vassilis Katsouros, George Carayannis |
| 2008 | ICASSP | Robust text-line and word segmentation for handwritten documents images. | Themos Stafylakis, Vassilis Papavassiliou, Vassilis Katsouros, George Carayannis |
| 2007 | ASRU | Efficient combination of parametric spaces, models and metrics for speaker diarization | Themos Stafylakis, Vassilis Katsouros, George Carayannis |
| 2007 | ICDAR | A Parametric Spectral-Based Method for Verification of Text in Videos. | Vassilis Papavassiliou, Themos Stafylakis, Vassilis Katsouros, George Carayannis |