| 2026 | LREC | Towards Expectation Detection in Language: A Case Study on Treatment Expectations in Reddit. | Aswathy Velutharambath, Amelie Whrl |
| 2025 | ACL | Which Demographics do LLMs Default to During Annotation? | Johannes Schfer, Aidan Combs, Christopher Bagdon, Jiahui Li, Nadine Probol, Lynn Greschner, Sean Papay, Yarik Menchaca Resendiz, Aswathy Velutharambath, Amelie Whrl, Sabine Weber, Roman Klinger |
| 2024 | ACL | Understanding Fine-grained Distortions in Reports of Scientific Findings. | Amelie Whrl, Dustin Wright, Roman Klinger, Isabelle Augenstein |
| 2024 | COLING | Can Factual Statements Be Deceptive? The DeFaBel Corpus of Belief-based Deception. | Aswathy Velutharambath, Roman Klinger, Amelie Whrl |
| 2024 | EACL | What Makes Medical Claims (Un)Verifiable? Analyzing Entity and Relation Properties for Fact Verification. | Amelie Whrl, Yarik Menchaca Resendiz, Lara Grimminger, Roman Klinger |
| 2024 | EMNLP | How Entangled is Factuality and Deception in German? | Aswathy Velutharambath, Amelie Whrl, Roman Klinger |
| 2022 | LREC | CoVERT: A Corpus of Fact-checked Biomedical COVID-19 Tweets. | Isabelle Mohr, Amelie Whrl, Roman Klinger |
| 2022 | LREC | Recovering Patient Journeys: A Corpus of Biomedical Entities and Relations on Twitter (BEAR). | Amelie Whrl, Roman Klinger |