Skip to content

Craig Thomson

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

18

Venues

6

Active years

2016–2026

Best venue rank

A*

Where they publish

Papers

18 indexed papers, newest first.

YearVenueTitleAuthors
2026ACLAutomatic Paper Analysis and Categorisation for Systematic Reviews with Combined Reasoning-Augmented SFT and DAPO RL.Michela Lorandi, Anya Belz, Simon Mille, Craig Thomson
2025ACLStandard Quality Criteria Derived from Current NLP Evaluations for Guiding Evaluation Design and Grounding Comparability and AI Compliance Assessments.Anya Belz, Simon Mille, Craig Thomson
2025EMNLPEvolving Stances on Reproducibility: A Longitudinal Study of NLP and ML Researchers' Views and Experience of Reproducibility.Craig Thomson, Ehud Reiter, Joo Sedoc, Anya Belz
2025INLGAssessing Semantic Consistency in Data-to-Text Generation: A Meta-Evaluation of Textual, Semantic and Model-Based Metrics.Rudali Huidrom, Michela Lorandi, Simon Mille, Craig Thomson, Anya Belz
2024INLGQCET: An Interactive Taxonomy of Quality Criteria for Comparable and Repeatable Evaluation of NLP Systems.Anya Belz, Simon Mille, Craig Thomson, Rudali Huidrom
2024INLGFilling Gaps in Wikipedia: Leveraging Data-to-Text Generation to Improve Encyclopedic Coverage of Underrepresented Groups.Simon Mille, Massimiliano Pronesti, Craig Thomson, Michela Lorandi, Sophie Fitzpatrick, Rudali Huidrom, Mohammed Sabry, Amy O'Riordan, Anya Belz
2024INLG(Mostly) Automatic Experiment Execution for Human Evaluations of NLP Systems.Craig Thomson, Anya Belz
2023ACLNon-Repeatable Experiments and Non-Reproducible Results: The Reproducibility Crisis in Human Evaluation in NLP.Anya Belz, Craig Thomson, Ehud Reiter, Simon Mille
2023INLGEnhancing factualness and controllability of Data-to-Text Generation via data Views and constraints.Craig Thomson, Clment Rebuffel, Ehud Reiter, Laure Soulier, Somayajulu Sripada, Patrick Gallinari
2021INLGUnderreporting of errors in NLG output, and what to do about it.Emiel van Miltenburg, Miruna-Adriana Clinciu, Ondrej Dusek, Dimitra Gkatzia, Stephanie Inglis, Leo Leppnen, Saad Mahamood, Emma Manning, Stephanie Schoch, Craig Thomson, Luou Wen
2021INLGGeneration Challenges: Results of the Accuracy Evaluation Shared Task.Craig Thomson, Ehud Reiter
2021SINMin-max Training: Adversarially Robust Learning Models for Network Intrusion Detection Systems.Sam Grierson, Craig Thomson, Pavlos Papadopoulos, Bill Buchanan
2020INLGShared Task on Evaluating Accuracy.Ehud Reiter, Craig Thomson
2020INLGA Gold Standard Methodology for Evaluating Accuracy in Data-To-Text Systems.Craig Thomson, Ehud Reiter
2020INLGStudying the Impact of Filling Information Gaps on the Output Quality of Neural Data-to-Text.Craig Thomson, Zhijie Zhao, Somayajulu Sripada
2018HPCCPerformance Investigation of RPL Routing in Pipeline Monitoring WSNs.Ahmed Yassin Al-Dubai, Isam Wadhaj, Wajeb Gharibi, Craig Thomson
2018INLGComprehension Driven Document Planning in Natural Language Generation Systems.Craig Thomson, Ehud Reiter, Somayajulu Sripada
2016AICCSAPerformance evaluation of RPL metrics in environments with strained transmission ranges.Craig Thomson, Isam Wadhaj, Imed Romdhani, Ahmed Yassin Al-Dubai