| 2026 | EACL | RoSE: Round-robin Synthetic Data Evaluation for Selecting LLM Generators without Human Test Sets. | Jn Cegin, Branislav Pecher, Ivan Srba, Jakub Simko |
| 2026 | EACL | MultiCW: A Large-Scale Balanced Benchmark Dataset for Training Robust Check-Worthiness Detection Models. | Martin Hyben, Sebastian Kula, Jn Cegin, Jakub Simko, Ivan Srba, Rbert Mro |
| 2026 | EACL | Better as Generators Than Classifiers: Leveraging LLMs and Synthetic Data for Low-Resource Multilingual Classification. | Branislav Pecher, Jn Cegin, Rbert Belanec, Ivan Srba, Jakub Simko, Mria Bielikov |
| 2025 | EMNLP | A Rigorous Evaluation of LLM Data Generation Strategies for Low-Resource Languages. | Tatiana Anikina, Jn Cegin, Jakub Simko, Simon Ostermann |
| 2025 | EMNLP | Use Random Selection for Now: Investigation of Few-Shot Selection Strategies in LLM-based Text Augmentation. | Jn Cegin, Branislav Pecher, Jakub Simko, Ivan Srba, Mria Bielikov, Peter Brusilovsky |
| 2025 | NAACL | LLMs vs Established Text Augmentation Techniques for Classification: When do the Benefits Outweight the Costs? | Jn Cegin, Jakub Simko, Peter Brusilovsky |
| 2024 | ACL | Effects of diversity incentives on sample diversity and downstream model performance in LLM-based text augmentation. | Jn Cegin, Branislav Pecher, Jakub Simko, Ivan Srba, Mria Bielikov, Peter Brusilovsky |
| 2024 | EAMT | Multilinguality in the VIGILANT project. | Brendan Spillane, Carolina Scarton, Rbert Mro, Petar Ivanov, Andrey Tagarev, Jakub Simko, Ibrahim Abu Farha, Gary Munnelly, Filip Uhlrik, Freddy Heppell |
| 2024 | EMNLP | Authorship Obfuscation in Multilingual Machine-Generated Text Detection. | Dominik Macko, Rbert Mro, Adaku Uchendu, Ivan Srba, Jason Samuel Lucas, Michiharu Yamashita, Nafis Irtiza Tripto, Dongwon Lee, Jakub Simko, Mria Bielikov |
| 2024 | EMNLP | Fighting Randomness with Randomness: Mitigating Optimisation Instability of Fine-Tuning using Delayed Ensemble and Noisy Interpolation. | Branislav Pecher, Jn Cegin, Rbert Belanec, Jakub Simko, Ivan Srba, Mria Bielikov |
| 2023 | EMNLP | ChatGPT to Replace Crowdsourcing of Paraphrases for Intent Classification: Higher Diversity and Comparable Model Robustness. | Jn Cegin, Jakub Simko, Peter Brusilovsky |
| 2023 | EMNLP | MULTITuDE: Large-Scale Multilingual Machine-Generated Text Detection Benchmark. | Dominik Macko, Rbert Mro, Adaku Uchendu, Jason Samuel Lucas, Michiharu Yamashita, Mats Pikuliak, Ivan Srba, Thai Le, Dongwon Lee, Jakub Simko, Mria Bielikov |
| 2023 | EMNLP | Multilingual Previously Fact-Checked Claim Retrieval. | Mats Pikuliak, Ivan Srba, Rbert Mro, Timo Hromadka, Timotej Smolen, Martin Melisek, Ivan Vykopal, Jakub Simko, Juraj Podrouzek, Mria Bielikov |
| 2022 | IJCAI | Black-box Audit of YouTube's Video Recommendation: Investigation of Misinformation Filter Bubble Dynamics (Extended Abstract). | Mats Tomlein, Branislav Pecher, Jakub Simko, Ivan Srba, Rbert Mro, Elena Stefancova, Michal Kompan, Andrea Hrckova, Juraj Podrouzek, Mria Bielikov |
| 2022 | SIGIR | Monant Medical Misinformation Dataset: Mapping Articles to Fact-Checked Claims. | Ivan Srba, Branislav Pecher, Mats Tomlein, Rbert Mro, Elena Stefancova, Jakub Simko, Mria Bielikov |
| 2021 | RecSys | An Audit of Misinformation Filter Bubbles on YouTube: Bubble Bursting and Recent Behavior Changes. | Mats Tomlein, Branislav Pecher, Jakub Simko, Ivan Srba, Rbert Mro, Elena Stefancova, Michal Kompan, Andrea Hrckova, Juraj Podrouzek, Mria Bielikov |
| 2019 | ADBIS | Web-Navigation Skill Assessment Through Eye-Tracking Data. | Patrik Hlavac, Jakub Simko, Mria Bielikov |
| 2018 | RecSys | Analysis of User Behavior in Interfaces with Recommended Items: An Eye-tracking Study. | Pter Gspr, Michal Kompan, Jakub Simko, Mria Bielikov |
| 2013 | ICCCI | Classsourcing: Crowd-Based Validation of Question-Answer Learning Objects. | Jakub Simko, Marin Simko, Mria Bielikov, Jakub Sevcech, Roman Burger |