| 2026 | ACL | LoVeC: Reinforcement Learning for Better Verbalized Confidence in Long-Form Generation. | Caiqi Zhang, Xiaochen Zhu, Chengzu Li, Nigel Collier, Andreas Vlachos |
| 2026 | ACL | Demystifying Multi-Agent Debate: The Role of Confidence and Diversity. | Xiaochen Zhu, Caiqi Zhang, Yizhou Chi, Tom Stafford, Nigel Collier, Andreas Vlachos |
| 2026 | EACL | Uncertainty Quantification for Evaluating Gender Bias in Machine Translation. | Ieva Staliunaite, Julius Cheng, Andreas Vlachos |
| 2025 | ACL | Mitigating Shortcut Learning with InterpoLated Learning. | Michalis Korakakis, Andreas Vlachos, Adrian Weller |
| 2025 | ACL | Causal Estimation of Tokenisation Bias. | Pietro Lesci, Clara Meister, Thomas Hofmann, Andreas Vlachos, Tiago Pimentel |
| 2025 | ACL | Segment-Level Diffusion: A Framework for Controllable Long-Form Generation with Diffusion Language Models. | Xiaochen Zhu, Georgi Karadzhov, Chenxi Whitehouse, Andreas Vlachos |
| 2025 | ACL | Conformity in Large Language Models. | Xiaochen Zhu, Caiqi Zhang, Tom Stafford, Nigel Collier, Andreas Vlachos |
| 2025 | EMNLP | PledgeTracker: A System for Monitoring the Fulfilment of Pledges. | Yulong Chen, Michael Sejr Schlichtkrull, Zhenyun Deng, David P. A. Corney, Nasim Asl, Joshua Salisbury, Andrew Dudfield, Andreas Vlachos |
| 2025 | EMNLP | Social Good or Scientific Curiosity? Uncovering the Research Framing Behind NLP Artefacts. | Eric Chamoun, Nedjma Ousidhoum, Michael Sejr Schlichtkrull, Andreas Vlachos |
| 2025 | EMNLP | Improving Zero-shot Sentence Decontextualisation with Content Selection and Planning. | Zhenyun Deng, Yulong Chen, Andreas Vlachos |
| 2025 | EMNLP | TCP: a Benchmark for Temporal Constraint-Based Planning. | Zifeng Ding, Sikuan Yan, Moy Yuan, Xianglong Hu, Fangru Lin, Andreas Vlachos |
| 2025 | EMNLP | TSVer: A Benchmark for Fact Verification Against Time-Series Evidence. | Marek Strong, Andreas Vlachos |
| 2025 | EMNLP | MPTA: MultiTask Personalization Assessment. | Matthieu Tehenan, Eric Chamoun, Andreas Vlachos |
| 2025 | NAACL | A Bayesian Optimization Approach to Machine Translation Reranking. | Julius Cheng, Maike Zfle, Vilm Zouhar, Andreas Vlachos |
| 2025 | NAACL | Dis2Dis: Explaining Ambiguity in Fact-Checking. | Ieva Staliunaite, Andreas Vlachos |
| 2024 | ACL | Automated Focused Feedback Generation for Scientific Writing Assistance. | Eric Chamoun, Michael Sejr Schlichtkrull, Andreas Vlachos |
| 2024 | ACL | Document-level Claim Extraction and Decontextualisation for Fact-Checking. | Zhenyun Deng, Michael Sejr Schlichtkrull, Andreas Vlachos |
| 2024 | ACL | Causal Estimation of Memorisation Profiles. | Pietro Lesci, Clara Meister, Thomas Hofmann, Andreas Vlachos, Tiago Pimentel |
| 2024 | CogSci | The effect of diversity on group decision-making. | Georgi Karadzhov, Andreas Vlachos, Tom Stafford |
| 2024 | EACL | Measuring Uncertainty in Neural Machine Translation with Similarity-Sensitive Entropy. | Julius Cheng, Andreas Vlachos |
| 2024 | EMNLP | ALVIN: Active Learning Via INterpolation. | Michalis Korakakis, Andreas Vlachos, Adrian Weller |
| 2024 | EMNLP | Zero-Shot Fact Verification via Natural Logic and Large Language Models. | Marek Strong, Rami Aly, Andreas Vlachos |
| 2024 | EMNLP | Do We Need Language-Specific Fact-Checking Models? The Case of Chinese. | Caiqi Zhang, Zhijiang Guo, Andreas Vlachos |
| 2024 | EMNLP | An LLM Feature-based Framework for Dialogue Constructiveness Assessment. | Lexin Zhou, Youmna Farag, Andreas Vlachos |
| 2024 | NAACL | AnchorAL: Computationally Efficient Active Learning for Large and Imbalanced Datasets. | Pietro Lesci, Andreas Vlachos |
| 2023 | ACL | Improving the robustness of NLI models with minimax training. | Michalis Korakakis, Andreas Vlachos |
| 2023 | EMNLP | Multimodal Automated Fact-Checking: A Survey. | Mubashara Akhtar, Michael Sejr Schlichtkrull, Zhijiang Guo, Oana Cocarascu, Elena Simperl, Andreas Vlachos |
| 2023 | EMNLP | QA-NatVer: Question Answering for Natural Logic-based Fact Verification. | Rami Aly, Marek Strong, Andreas Vlachos |
| 2023 | EMNLP | Automated Fact-Checking in Dialogue: Are Specialized Models Needed? | Eric Chamoun, Marzieh Saeidi, Andreas Vlachos |
| 2023 | EMNLP | Faster Minimum Bayes Risk Decoding with Confidence-based Pruning. | Julius Cheng, Andreas Vlachos |
| 2023 | EMNLP | The Intended Uses of Automated Fact-Checking Artefacts: Why, How and Who. | Michael Sejr Schlichtkrull, Nedjma Ousidhoum, Andreas Vlachos |
| 2022 | ACL | Leveraging Wikipedia article evolution for promotional tone detection. | Christine de Kock, Andreas Vlachos |
| 2022 | EMNLP | Natural Logic-guided Autoregressive Multi-hop Document Retrieval for Fact Verification. | Rami Aly, Andreas Vlachos |
| 2022 | EMNLP | Opening up Minds with Argumentative Dialogues. | Youmna Farag, Charlotte O. Brand, Jacopo Amidei, Paul Piwek, Tom Stafford, Svetlana Stoyanchev, Andreas Vlachos |
| 2022 | EMNLP | How to disagree well: Investigating the dispute tactics used on Wikipedia. | Christine de Kock, Andreas Vlachos |
| 2022 | EMNLP | Improving Scheduled Sampling with Elastic Weight Consolidation for Neural Machine Translation. | Michalis Korakakis, Andreas Vlachos |
| 2022 | EMNLP | Varifocal Question Generation for Fact-checking. | Nedjma Ousidhoum, Zhangdie Yuan, Andreas Vlachos |
| 2022 | SIGdial | What makes you change your mind? An empirical investigation in online group decision-making conversations. | Georgi Karadzhov, Tom Stafford, Andreas Vlachos |
| 2021 | ACL | Leveraging Type Descriptions for Zero-shot Named Entity Recognition and Classification. | Rami Aly, Andreas Vlachos, Ryan McDonald |
| 2021 | ACL | Survival text regression for time-to-event prediction in conversations. | Christine de Kock, Andreas Vlachos |
| 2021 | ACL | Evidence-based Factual Error Correction. | James Thorne, Andreas Vlachos |
| 2021 | EACL | Incremental Beam Manipulation for Natural Language Generation. | James Hargreaves, Andreas Vlachos, Guy Emerson |
| 2021 | EACL | I Beg to Differ: A study of constructive disagreement in online conversations. | Christine de Kock, Andreas Vlachos |
| 2021 | EACL | Elastic weight consolidation for better bias inoculation. | James Thorne, Andreas Vlachos |
| 2021 | EMNLP | Cross-Policy Compliance Detection via Question Answering. | Marzieh Saeidi, Majid Yazdani, Andreas Vlachos |
| 2020 | EMNLP | Generating Fact Checking Briefs. | Angela Fan, Aleksandra Piktus, Fabio Petroni, Guillaume Wenzek, Marzieh Saeidi, Andreas Vlachos, Antoine Bordes, Sebastian Riedel |
| 2019 | ACL | Merge and Label: A Novel Neural Network Architecture for Nested NER. | Joseph Fisher, Andreas Vlachos |
| 2019 | ACL | HighRES: Highlight-based Reference-less Evaluation of Summarization. | Hardy, Shashi Narayan, Andreas Vlachos |
| 2019 | ACL | Model-Agnostic Meta-Learning for Relation Classification with Limited Supervision. | Abiola Obamuyide, Andreas Vlachos |
| 2019 | EMNLP | Incorporating Label Dependencies in Multilabel Stance Detection. | William Ferreira, Andreas Vlachos |
| 2019 | EMNLP | Neural Generative Rhetorical Structure Parsing. | Amandla Mabona, Laura Rimell, Stephen Clark, Andreas Vlachos |
| 2019 | EMNLP | Evaluating adversarial attacks against multiple fact verification systems. | James Thorne, Andreas Vlachos, Christos Christodoulopoulos, Arpit Mittal |
| 2019 | NAACL | Strong Baselines for Complex Word Identification across Multiple Languages. | Pierre Finnimore, Elisabeth Fritzsch, Daniel King, Alison Sneyd, Aneeq Ur Rehman, Fernando Alva-Manchego, Andreas Vlachos |
| 2019 | NAACL | Generating Token-Level Explanations for Natural Language Inference. | James Thorne, Andreas Vlachos, Christos Christodoulopoulos, Arpit Mittal |
| 2019 | WWW | Automated Fact Checking in the News Room. | Sebastio Miranda, David Nogueira, Afonso Mendes, Andreas Vlachos, Andrew Secker, Rebecca Garrett, Jeff Mitchell, Zita Marinho |
| 2018 | CIKM | Automated Fact Checking. | Andreas Vlachos |
| 2018 | COLING | Topic or Style? Exploring the Most Useful Features for Authorship Attribution. | Yunita Sari, Mark Stevenson, Andreas Vlachos |
| 2018 | COLING | Automated Fact Checking: Task Formulations, Methods and Future Directions. | James Thorne, Andreas Vlachos |
| 2018 | EMNLP | Guided Neural Language Generation for Abstractive Summarization using Abstract Meaning Representation. | Hardy, Andreas Vlachos |
| 2018 | EMNLP | Zero-shot Relation Classification as Textual Entailment. | Abiola Obamuyide, Andreas Vlachos |
| 2018 | EMNLP | The Fact Extraction and VERification (FEVER) Shared Task. | James Thorne, Andreas Vlachos, Oana Cocarascu, Christos Christodoulopoulos, Arpit Mittal |
| 2018 | NAACL | FEVER: a Large-scale Dataset for Fact Extraction and VERification. | James Thorne, Andreas Vlachos, Christos Christodoulopoulos, Arpit Mittal |
| 2017 | EACL | The SUMMA Platform Prototype. | Renars Liepins, Ulrich Germann, Guntis Barzdins, Alexandra Birch, Steve Renals, Susanne Weber, Peggy van der Kreeft, Herv Bourlard, Joo Prieto, Ondrej Klejch, Peter Bell, Alexandros Lazaridis, Afonso Mendes, Sebastian Riedel, Mariana S. C. Almeida, Pedro Balage, Shay B. Cohen, Tomasz Dwojak, Philip N. Garner, Andreas Giefer, Marcin Junczys-Dowmunt, Hina Imran, David Nogueira, Ahmed M. Ali, Sebastio Miranda, Andrei Popescu-Belis, Lesly Miculicich Werlen, Nikos Papasarantopoulos, Abiola Obamuyide, Clive Jones, Fahim Dalvi, Andreas Vlachos, Yang Wang, Sibo Tong, Rico Sennrich, Nikolaos Pappas, Shashi Narayan, Marco Damonte, Nadir Durrani, Sameer Khurana, Ahmed Abdelali, Hassan Sajjad, Stephan Vogel, David Sheppey, Chris Hernon, Jeff Mitchell |
| 2017 | EACL | Continuous N-gram Representations for Authorship Attribution. | Yunita Sari, Andreas Vlachos, Mark Stevenson |
| 2017 | EACL | An Extensible Framework for Verification of Numerical Claims. | James Thorne, Andreas Vlachos |
| 2017 | EMNLP | Fake news stance detection using stacked ensemble of classifiers. | James Thorne, Mingjie Chen, Giorgos Myrianthous, Jiashu Pu, Xiaoxuan Wang, Andreas Vlachos |
| 2016 | ACL | Noise reduction and targeted exploration in imitation learning for Abstract Meaning Representation parsing. | James Goodman, Andreas Vlachos, Jason Naradowsky |
| 2016 | COLING | Imitation learning for language generation from unaligned data. | Gerasimos Lampouras, Andreas Vlachos |
| 2016 | EMNLP | Stance Detection with Bidirectional Conditional Encoding. | Isabelle Augenstein, Tim Rocktschel, Andreas Vlachos, Kalina Bontcheva |
| 2016 | EMNLP | Timeline extraction using distant supervision and joint inference. | Savelie Cornegruta, Andreas Vlachos |
| 2016 | NAACL | Emergent: a novel data-set for stance classification. | William Ferreira, Andreas Vlachos |
| 2015 | ACL | Matrix and Tensor Factorization Methods for Natural Language Processing. | Guillaume Bouchard, Jason Naradowsky, Sebastian Riedel, Tim Rocktschel, Andreas Vlachos |
| 2015 | ACL | Dependency Recurrent Neural Language Models for Sentence Completion. | Piotr Mirowski, Andreas Vlachos |
| 2015 | EMNLP | Extracting Relations between Non-Standard Entities using Distant Supervision and Imitation Learning. | Isabelle Augenstein, Andreas Vlachos, Diana Maynard |
| 2015 | EMNLP | A Strong Lexical Matching Method for the Machine Comprehension Test. | Ellery Smith, Nicola Greco, Matko Bosnjak, Andreas Vlachos |
| 2015 | EMNLP | Identification and Verification of Simple Claims about Statistical Properties. | Andreas Vlachos, Sebastian Riedel |
| 2014 | ACL | Fact Checking: Task definition and dataset construction. | Andreas Vlachos, Sebastian Riedel |
| 2013 | ACL | Semantic Parsing as Machine Translation. | Jacob Andreas, Andreas Vlachos, Stephen Clark |
| 2013 | EMNLP | Dependency Language Models for Sentence Completion. | Joseph Gubbins, Andreas Vlachos |
| 2011 | CoNLL | Search-based Structured Prediction applied to Biomedical Event Extraction. | Andreas Vlachos, Mark Craven |
| 2011 | EMNLP | Evaluating unsupervised learning for natural language processing tasks. | Andreas Vlachos |
| 2010 | CoNLL | Detecting Speculative Language Using Syntactic Dependencies and Logistic Regression. | Andreas Vlachos, Mark Craven |
| 2009 | EMNLP | The infinite HMM for unsupervised PoS tagging. | Jurgen Van Gael, Andreas Vlachos, Zoubin Ghahramani |
| 2006 | PSB | Bootstrapping the Recognition and Anaphoric Linking of Named Entities in Drosophila Articles. | Andreas Vlachos, Caroline Gasperin, Ian Lewin, Ted Briscoe |