| 2024 | NAACL | A Measure for Transparent Comparison of Linguistic Diversity in Multilingual NLP Data Sets. | Tanja Samardzic, Ximena Gutierrez-Vasques, Christian Bentz, Steven Moran, Olga Pelloni |
| 2023 | CoNLL | The Zipfian Challenge: Learning the statistical fingerprint of natural languages. | Christian Bentz |
| 2022 | LREC | TeDDi Sample: Text Data Diversity Sample for Language Comparison and Multilingual NLP. | Steven Moran, Christian Bentz, Ximena Gutierrez-Vasques, Olga Pelloni, Tanja Samardzic |
| 2021 | EACL | From characters to words: the turning point of BPE merges. | Ximena Gutierrez-Vasques, Christian Bentz, Olga Sozinova, Tanja Samardzic |
| 2020 | COLING | Grammatical error detection in transcriptions of spoken English. | Andrew Caines, Christian Bentz, Kate M. Knill, Marek Rei, Paula Buttery |
| 2016 | LREC | Crowdsourcing a Multi-lingual Speech Corpus: Recording, Transcription and Annotation of the CrowdIS Corpora. | Andrew Caines, Christian Bentz, Calbert Graham, Tim Polzehl, Paula Buttery |
| 2013 | CogSci | Beyond Rule versus Rote? Processing of Distinctive Dative and Genitive Case Markers in German. | Christian Bentz |
| 2013 | CogSci | Large-Scale Empricial Analyses of the Abstract/Concrete Distinction. | Felix Hill, Anna Korhonen, Christian Bentz |