| 2026 | ACL | How Value Induction Reshapes LLM Behavior. | Arnav Arora, Natalie Schluter, Katherine Metcalf, Maartje ter Hoeve |
| 2025 | ACL | Steering into New Embedding Spaces: Analyzing Cross-Lingual Alignment Induced by Model Interventions in Multilingual Language Models. | Anirudh Sundar, Sinead Williamson, Katherine Metcalf, Barry-John Theobald, Skyler Seto, Masha Fedzechkina |
| 2025 | ICML | Aligning LLMs by Predicting Preferences from User Writing Samples. | Stphane Aroca-Ouellette, Natalie Mackraz, Barry-John Theobald, Katherine Metcalf |
| 2025 | ICML | Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs. | Yinong Oliver Wang, Nivedha Sivakumar, Falaah Arif Khan, Katherine Metcalf, Adam Golinski, Natalie Mackraz, Barry-John Theobald, Luca Zappella, Nicholas Apostoloff |
| 2024 | AAAI | Can You Rely on Synthetic Labellers in Preference-Based Reinforcement Learning? It's Complicated. | Katherine Metcalf, Miguel Sarabia, Masha Fedzechkina, Barry-John Theobald |
| 2024 | EMNLP | On the Limited Generalization Capability of the Implicit Reward Model Induced by Direct Preference Optimization. | Yong Lin, Skyler Seto, Maartje ter Hoeve, Katherine Metcalf, Barry-John Theobald, Xuan Wang, Yizhe Zhang, Chen Huang, Tong Zhang |
| 2024 | ICLR | Hindsight PRIORs for Reward Learning from Human Preferences. | Mudit Verma, Katherine Metcalf |
| 2024 | ICML | Whispering Experts: Neural Interventions for Toxicity Mitigation in Language Models. | Xavier Suau, Pieter Delobelle, Katherine Metcalf, Armand Joulin, Nicholas Apostoloff, Luca Zappella, Pau Rodrguez |
| 2023 | CoRL | Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards. | Katherine Metcalf, Miguel Sarabia, Natalie Mackraz, Barry-John Theobald |
| 2023 | ICASSP | On the Role of LIP Articulation in Visual Speech Perception. | Zakaria Aldeneh, Masha Fedzechkina, Skyler Seto, Katherine Metcalf, Miguel Sarabia, Nicholas Apostoloff, Barry-John Theobald |
| 2019 | IJCAI | Unsupervised Hierarchical Temporal Abstraction by Simultaneously Learning Expectations and Representations. | Katherine Metcalf, David Leake |
| 2019 | Interspeech | Mirroring to Build Trust in Digital Assistants. | Katherine Metcalf, Barry-John Theobald, Garrett Weinberg, Robert Lee, Ing-Marie Jonsson, Russ Webb, Nicholas Apostoloff |
| 2018 | ICCBR | Embedded Word Representations for Rich Indexing: A Case Study for Medical Records. | Katherine Metcalf, David Leake |
| 2017 | CogSci | Modelling Unsupervised Event Segmentation: Learning Event Boundaries from Prediction Errors. | Katherine Metcalf, David Leake |
| 2016 | ECAI | A Computational Method for Extracting, Representing, and Predicting Social Closeness. | Katherine Metcalf, David B. Leake |
| 2013 | JURIX | Automated Methods for Extracting and Expanding Lists in Regulatory Text. | Alan Buabuchachart, Nina Charness, Katherine Metcalf, Leora Morgenstern |
| 2013 | JURIX | Classification of Regulatory Paragraphs by Discourse Structure, Reference Structure, and Regulation Type. | Alan Buabuchachart, Katherine Metcalf, Nina Charness, Leora Morgenstern |