| 2026 | ACL | CUB: Benchmarking Context Utilisation Techniques for Language Models. | Lovisa Hagstrm, Youna Kim, Haeun Yu, Sang-goo Lee, Richard Johansson, Hyunsoo Cho, Isabelle Augenstein |
| 2025 | ACL | A Reality Check on Context Utilisation for Retrieval-Augmented Generation. | Lovisa Hagstrm, Sara Vera Marjanovic, Haeun Yu, Arnav Arora, Christina Lioma, Maria Maistro, Pepa Atanasova, Isabelle Augenstein |
| 2025 | ACL | Fact Recall, Heuristics or Pure Guesswork? Precise Interpretations of Language Models for Fact Completion. | Denitsa Saynova, Lovisa Hagstrm, Moa Johansson, Richard Johansson, Marco Kuhlmann |
| 2023 | EMNLP | The Effect of Scaling, Retrieval Augmentation and Form on the Factual Consistency of Language Models. | Lovisa Hagstrm, Denitsa Saynova, Tobias Norlund, Moa Johansson, Richard Johansson |
| 2022 | ACL | What do Models Learn From Training on More Than Text? Measuring Visual Commonsense Knowledge. | Lovisa Hagstrm, Richard Johansson |
| 2022 | COLING | How to Adapt Pre-trained Vision-and-Language Models to a Text-only Input? | Lovisa Hagstrm, Richard Johansson |