| 2024 | INLG | Exploring the impact of data representation on neural data-to-text generation. | David M. Howcroft, Lewis N. Watson, Olesia Nedopas, Dimitra Gkatzia |
| 2024 | INLG | Automatic Metrics in Natural Language Generation: A survey of Current Evaluation Practices. | Patrcia Schmidtov, Saad Mahamood, Simone Balloccu, Ondrej Dusek, Albert Gatt, Dimitra Gkatzia, David M. Howcroft, Ondrej Pltek, Adarsa Sivaprasad |
| 2023 | INLG | LOWRECORP: the Low-Resource NLG Corpus Building Challenge. | Khyathi Raghavi Chandu, David M. Howcroft, Dimitra Gkatzia, Yi-Ling Chung, Yufang Hou, Chris Chinenye Emezue, Pawan Rajpoot, Tosin P. Adewumi |
| 2023 | INLG | Building a dual dataset of text- and image-grounded conversations and summarisation in Gidhlig (Scottish Gaelic). | David M. Howcroft, William Lamb, Anna Groundwater, Dimitra Gkatzia |
| 2021 | ACL | OTTers: One-turn Topic Transitions for Open-Domain Dialogue. | Karin Sevegnani, David M. Howcroft, Ioannis Konstas, Verena Rieser |
| 2021 | EMNLP | What happens if you treat ordinal ratings as interval data? Human evaluations in NLP are even more under-powered than you think. | David M. Howcroft, Verena Rieser |
| 2020 | INLG | Disentangling the Properties of Human Evaluation Methods: A Classification System to Support Comparability, Meta-Evaluation and Reproducibility Testing. | Anya Belz, Simon Mille, David M. Howcroft |
| 2020 | INLG | Twenty Years of Confusion in Human Evaluation: NLG Needs Evaluation Sheets and Standardised Definitions. | David M. Howcroft, Anya Belz, Miruna-Adriana Clinciu, Dimitra Gkatzia, Sadid A. Hasan, Saad Mahamood, Simon Mille, Emiel van Miltenburg, Sashank Santhanam, Verena Rieser |
| 2019 | INLG | Semantic Noise Matters for Neural Natural Language Generation. | Ondrej Dusek, David M. Howcroft, Verena Rieser |
| 2018 | INLG | Toward Bayesian Synchronous Tree Substitution Grammars for Sentence Planning. | David M. Howcroft, Dietrich Klakow, Vera Demberg |
| 2017 | EACL | Psycholinguistic Models of Sentence Processing Improve Sentence Readability Ranking. | David M. Howcroft, Vera Demberg |
| 2017 | INLG | G-TUNA: a corpus of referring expressions in German, including duration information. | David M. Howcroft, Jorrig Vogels, Vera Demberg |
| 2017 | Interspeech | The Extended SPaRKy Restaurant Corpus: Designing a Corpus with Variable Information Density. | David M. Howcroft, Dietrich Klakow, Vera Demberg |
| 2016 | COLING | From OpenCCG to AI Planning: Detecting Infeasible Edges in Sentence Generation. | Maximilian Schwenger, lvaro Torralba, Jrg Hoffmann, David M. Howcroft, Vera Demberg |