| 2026 | ACL | Language Models Learn Universal Representations of Numbers and Here's Why You Should Care. | Michal Stefnik, Timothee Mickus, Marek Kadlck, Bertram Hjer, Michal Spiegel, Ral Vzquez, Aman Sinha, Josef Kuchar, Philipp Mondorf, Pontus Stenetorp |
| 2026 | LREC | VectorEdits: A Dataset and Benchmark for Instruction-Based Editing of Vector Graphics. | Josef Kuchar, Marek Kadlck, Michal Spiegel, Michal Stefnik |
| 2025 | EMNLP | Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers. | Marek Kadlck, Michal Stefnik, Timothee Mickus, Josef Kuchar, Michal Spiegel |
| 2025 | EMNLP | Can Out-of-Distribution Evaluations Uncover Reliance on Prediction Shortcuts? A Case Study in Question Answering. | Michal Stefnik, Timothee Mickus, Michal Spiegel, Marek Kadlck, Josef Kuchar |
| 2025 | EMNLP | Towards the Roots of the Negation Problem: A Multilingual NLI Dataset and Model Scaling Analysis. | Tereza Vrabcov, Marek Kadlck, Petr Sojka, Michal Stefnik, Michal Spiegel |
| 2024 | ACL | Concept-aware Data Construction Improves In-context Learning of Language Models. | Michal Stefnik, Marek Kadlck, Petr Sojka |
| 2024 | EACL | Think Twice: Measuring the Efficiency of Eliminating Prediction Shortcuts of Question Answering Models. | Luks Mikula, Michal Stefnik, Marek Petrovic, Petr Sojka |
| 2024 | EMNLP | Self-training Language Models for Arithmetic Reasoning. | Marek Kadlck, Michal Stefnik |
| 2023 | ACL | People and Places of Historical Europe: Bootstrapping Annotation Pipeline and a New Corpus of Named Entities in Late Medieval Texts. | Vit Novotny, Kristina Luger, Michal Stefnik, Tereza Vrabcov, Ales Hork |
| 2023 | ACL | Soft Alignment Objectives for Robust Adaptation of Language Generation. | Michal Stefnik, Marek Kadlck, Petr Sojka |
| 2023 | EMNLP | Calc-X and Calcformers: Empowering Arithmetical Chain-of-Thought through Interaction with Symbolic Systems. | Marek Kadlck, Michal Stefnik, Ondrej Sotolr, Vlastimil Martinek |
| 2022 | ACL | Adaptor: Objective-Centric Adaptation Framework for Language Models. | Michal Stefnik, Vt Novotn, Nikola Groverov, Petr Sojka |
| 2022 | NAACL | Methods for Estimating and Improving Robustness of Language Models. | Michal Stefnik |
| 2021 | RANLP | One Size Does Not Fit All: Finding the Optimal Subword Sizes for FastText Models across Languages. | Vt Novotn, Eniafe Festus Ayetiran, Dalibor Bacovsk, Dvid Luptk, Michal Stefnik, Petr Sojka |