| 2020 | LREC | ProGene - A Large-scale, High-Quality Protein-Gene Annotated Benchmark Corpus. | Erik Faessler, Luise Modersohn, Christina Lohr, Udo Hahn |
| 2020 | SIGIR | What Makes a Top-Performing Precision Medicine Search Engine?: Tracing Main System Features in a Systematic Way. | Erik Faessler, Michel Oleynik, Udo Hahn |
| 2018 | DocEng | Annotation Data Management with JeDIS. | Erik Faessler, Udo Hahn |
| 2017 | ACL | Semedico: A Comprehensive Semantic Search Engine for the Life Sciences. | Erik Faessler, Udo Hahn |
| 2016 | EKAW | Selecting and Tailoring Ontologies with JOYCE. | Erik Faessler, Friederike Klan, Alsayed Algergawy, Birgitta Knig-Ries, Udo Hahn |
| 2016 | LREC | UIMA-Based JCoRe 2.0 Goes GitHub and Maven Central ― State-of-the-Art Software Resource Engineering and Distribution of NLP Pipelines. | Udo Hahn, Franz Matthies, Erik Faessler, Johannes Hellrich |
| 2014 | LREC | Disclose Models, Hide the Data - How to Make Use of Confidential Corpora without Seeing Sensitive Raw Data. | Erik Faessler, Johannes Hellrich, Udo Hahn |
| 2012 | AMIA | Active Learning-Based Corpus Annotation - The PathoJen Experience. | Udo Hahn, Elena Beisswanger, Ekaterina Buyko, Erik Faessler |
| 2012 | LREC | Iterative Refinement and Quality Checking of Annotation Guidelines - How to Deal Effectively with Semantically Sloppy Named Entity Types, such as Pathological Phenomena. | Udo Hahn, Elena Beisswanger, Ekaterina Buyko, Erik Faessler, Jenny Traumller, Susann Schrder, Kerstin Hornbostel |