| 2026 | EACL | PictureStories: Predicting the Task Adherence of Language Learner Answers to a Picture Story-Based Writing Task. | Marie Bexte, Andrew Caines, Diane Nicholls, Paula Buttery, Torsten Zesch |
| 2025 | AIED | Second Workshop on the Automatic Evaluation of Learning and Assessment Content. | Luca Benedetto, Diana Galvn-Sosa, George Dueas, Shiva Taslimipoor, Gabrielle Gaudeau, Andrew Caines, Anastassia Loukina, Torsten Zesch |
| 2024 | AIED | Workshop on Automatic Evaluation of Learning and Assessment Content. | Luca Benedetto, Shiva Taslimipoor, Andrew Caines, Diana Galvn-Sosa, George Dueas, Anastassia Loukina, Torsten Zesch |
| 2024 | COLING | EVil-Probe - a Composite Benchmark for Extensive Visio-Linguistic Probing. | Marie Bexte, Andrea Horbach, Torsten Zesch |
| 2024 | COLING | Comprehensive Study on German Language Models for Clinical and Biomedical Text Understanding. | Ahmad Idrissi-Yaghir, Amin Dada, Henning Schfer, Kamyar Arzideh, Giulia Baldini, Jan Trienes, Max Hasin, Jeanette Bewersdorff, Cynthia S. Schmidt, Marie Bauer, Kaleb E. Smith, Jiang Bian, Yonghui Wu, Jrg Schltterer, Torsten Zesch, Peter A. Horn, Christin Seifert, Felix Nensa, Jens Kleesiek, Christoph M. Friedrich |
| 2024 | COLING | Every Verb in Its Right Place? A Roadmap for Operationalizing Developmental Stages in the Acquisition of L2 German. | Josef Ruppenhofer, Matthias Schwendemann, Annette Portmann, Katrin Wisniewski, Torsten Zesch |
| 2024 | EACL | Text or Image? What is More Important in Cross-Domain Generalization Capabilities of Hate Meme Detection Models? | Piush Aggarwal, Jawar Mehrabanian, Weigang Huang, zge Alaam, Torsten Zesch |
| 2024 | EACL | Rainbow - A Benchmark for Systematic Testing of How Sensitive Visio-Linguistic Models are to Color Naming. | Marie Bexte, Andrea Horbach, Torsten Zesch |
| 2024 | EACL | Unraveling the Dynamics of Semi-Supervised Hate Speech Detection: The Impact of Unlabeled Data Characteristics and Pseudo-Labeling Strategies. | Florian Ludwig, Klara Dolos, Ana Alves-Pinto, Torsten Zesch |
| 2023 | ACL | Similarity-Based Content Scoring - A more Classroom-Suitable Alternative to Instance-Based Scoring? | Marie Bexte, Andrea Horbach, Torsten Zesch |
| 2023 | WWW | HateProof: Are Hateful Meme Detection Systems really Robust? | Piush Aggarwal, Pranit Chawla, Mithun Das, Punyajoy Saha, Binny Mathew, Torsten Zesch, Animesh Mukherjee |
| 2022 | COLING | Analyzing the Real Vulnerability of Hate Speech Detection Systems against Targeted Intentional Noise. | Piush Aggarwal, Torsten Zesch |
| 2022 | ICFHR | CNN-Based Ruled Line Removal in Handwritten Documents. | Christian Gold, Torsten Zesch |
| 2022 | LREC | LeSpell - A Multi-Lingual Benchmark Corpus of Spelling Errors to Develop Spellchecking Methods for Learner Language. | Marie Bexte, Ronja Laarmann-Quante, Andrea Horbach, Torsten Zesch |
| 2021 | ICDAR | Personalizing Handwriting Recognition Systems with Limited User-Specific Samples. | Christian Gold, Dario van den Boom, Torsten Zesch |
| 2020 | COLING | Don't take "nswvtnvakgxpm" for an answer -The surprising vulnerability of automatic content scoring systems to adversarial input. | Yuning Ding, Brian Riordan, Andrea Horbach, Aoife Cahill, Torsten Zesch |
| 2020 | ICFHR | Exploring the Impact of Handwriting Recognition on the Automated Scoring of Handwritten Student Answers. | Christian Gold, Torsten Zesch |
| 2020 | ICPR | Fully vs. Weakly Supervised Caries Localization in Smartphone Images with CNNs. | Duc Duy Pham, Jonas Mller, Piush Aggarwal, Amit Khatri, Mayank Sharma, Torsten Zesch, Josef Pauli |
| 2020 | IJCNLP | Chinese Content Scoring: Open-Access Datasets and Features on Different Segmentation Levels. | Yuning Ding, Andrea Horbach, Torsten Zesch |
| 2020 | LREC | Decomposing and Comparing Meaning Relations: Paraphrasing, Textual Entailment, Contradiction, and Specificity. | Venelin Kovatchev, Darina Gold, Maria Antnia Mart, Maria Salam, Torsten Zesch |
| 2019 | NAACL | From legal to technical concept: Towards an automated classification of German political Twitter postings as criminal offenses. | Frederike Zufall, Tobias Horsmann, Torsten Zesch |
| 2019 | RANLP | Divide and Extract - Disentangling Clause Splitting and Proposition Extraction. | Darina Gold, Torsten Zesch |
| 2018 | LREC | DeepTC - An Extension of DKPro Text Classification for Fostering Reproducibility of Deep Learning Experiments. | Tobias Horsmann, Torsten Zesch |
| 2018 | LREC | Quantifying Qualitative Data for Understanding Controversial Issues. | Michael Wojatzki, Saif M. Mohammad, Torsten Zesch, Svetlana Kiritchenko |
| 2018 | LREC | ESCRITO - An NLP-Enhanced Educational Scoring Toolkit. | Torsten Zesch, Andrea Horbach |
| 2017 | EMNLP | Do LSTMs really work so well for PoS tagging? - A replication study. | Tobias Horsmann, Torsten Zesch |
| 2017 | RANLP | Same same, but different: Compositionality of paraphrase granularity levels. | Darina Benikova, Torsten Zesch |
| 2016 | COLING | Assigning Fine-grained PoS Tags based on High-precision Coarse-grained Tagging. | Tobias Horsmann, Torsten Zesch |
| 2016 | COLING | Predicting proficiency levels in learner writings by transferring a linguistic complexity model from expert-written coursebooks. | Ildik Piln, Elena Volodina, Torsten Zesch |
| 2016 | LREC | FlexTag: A Highly Flexible PoS Tagging Framework. | Torsten Zesch, Tobias Horsmann |
| 2014 | ACL | DKPro TC: A Java-based Framework for Supervised Learning Experiments on Textual Data. | Johannes Daxenberger, Oliver Ferschke, Iryna Gurevych, Torsten Zesch |
| 2014 | ACL | DKPro Keyphrases: Flexible and Reusable Keyphrase Extraction Experiments. | Nicolai Erbs, Pedro Bispo Santos, Iryna Gurevych, Torsten Zesch |
| 2013 | ACL | DKPro Similarity: An Open Source Framework for Text Similarity. | Daniel Br, Torsten Zesch, Iryna Gurevych |
| 2013 | ACL | Recognizing Partial Textual Entailment. | Omer Levy, Torsten Zesch, Ido Dagan, Iryna Gurevych |
| 2013 | ACL | DKPro WSD: A Generalized UIMA-based Framework for Word Sense Disambiguation. | Tristan Miller, Nicolai Erbs, Hans-Peter Zorn, Torsten Zesch, Iryna Gurevych |
| 2013 | IJCNLP | Cognate Production using Character-based Machine Translation. | Lisa Beinborn, Torsten Zesch, Iryna Gurevych |
| 2013 | RANLP | Hierarchy Identification for Automatically Generating Table-of-Contents. | Nicolai Erbs, Iryna Gurevych, Torsten Zesch |
| 2012 | COLING | Text Reuse Detection using a Composition of Text Similarity Measures. | Daniel Br, Torsten Zesch, Iryna Gurevych |
| 2012 | COLING | Using Distributional Similarity for Lexical Expansion in Knowledge-based Word Sense Disambiguation. | Tristan Miller, Chris Biemann, Torsten Zesch, Iryna Gurevych |
| 2012 | EACL | Measuring Contextual Fitness Using Error Contexts Extracted from the Wikipedia Revision History. | Torsten Zesch |
| 2011 | ACL | Wikulu: An Extensible Architecture for Integrating Natural Language Processing Techniques with Wikis. | Daniel Br, Nicolai Erbs, Torsten Zesch, Iryna Gurevych |
| 2011 | ACL | Wikipedia Revision Toolkit: Efficiently Accessing Wikipedia's Edit History. | Oliver Ferschke, Torsten Zesch, Iryna Gurevych |
| 2011 | CICLING | Combining Heterogeneous Knowledge Resources for Improved Distributional Semantic Models. | Gyrgy Szarvas, Torsten Zesch, Iryna Gurevych |
| 2011 | RANLP | A Reflective View on Text Similarity. | Daniel Br, Torsten Zesch, Iryna Gurevych |
| 2010 | LREC | The More the Better? Assessing the Influence of Wikipedia's Growth on Semantic Relatedness Measures. | Torsten Zesch, Iryna Gurevych |
| 2009 | RANLP | Approximate Matching for Evaluating Keyphrase Extraction. | Torsten Zesch, Iryna Gurevych |
| 2008 | AAAI | Using Wiktionary for Computing Semantic Relatedness. | Torsten Zesch, Christof Mller, Iryna Gurevych |
| 2008 | LREC | Extracting Lexical Semantic Knowledge from Wikipedia and Wiktionary. | Torsten Zesch, Christof Mller, Iryna Gurevych |
| 2007 | ACL | What to be? - Electronic Career Guidance Based on Semantic Relatedness. | Iryna Gurevych, Christof Mller, Torsten Zesch |
| 2007 | EMNLP | Cross-Lingual Distributional Profiles of Concepts for Measuring Semantic Distance. | Saif M. Mohammad, Iryna Gurevych, Graeme Hirst, Torsten Zesch |
| 2007 | NAACL | Comparing Wikipedia and German Wordnet by Evaluating Semantic Relatedness on Multiple Datasets. | Torsten Zesch, Iryna Gurevych, Max Mhlhuser |