| 2025 | ACL | Compute Optimal Scaling of Skills: Knowledge vs Reasoning. | Nicholas Roberts, Niladri S. Chatterji, Sharan Narang, Mike Lewis, Dieuwke Hupkes |
| 2025 | NAACL | Lost in Inference: Rediscovering the Role of Natural Language Inference for Large Language Models. | Lovish Madaan, David Esiobu, Pontus Stenetorp, Barbara Plank, Dieuwke Hupkes |
| 2024 | ACL | Interpretability of Language Models via Task Spaces. | Lucas Weber, Jaap Jumelet, Elia Bruni, Dieuwke Hupkes |
| 2023 | CoNLL | The Validity of Evaluation Results: Assessing Concurrence Across Compositionality Benchmarks. | Kaiser Sun, Adina Williams, Dieuwke Hupkes |
| 2023 | CoNLL | Mind the instructions: a holistic evaluation of consistency and interactions in prompt-based learning. | Lucas Weber, Elia Bruni, Dieuwke Hupkes |
| 2023 | EMNLP | Memorisation Cartography: Mapping out the Memorisation-Generalisation Continuum in Neural Machine Translation. | Verna Dankers, Ivan Titov, Dieuwke Hupkes |
| 2023 | ICLR | Neural Agents Struggle to Take Turns in Bidirectional Emergent Communication. | Valentin Taillandier, Dieuwke Hupkes, Benot Sagot, Emmanuel Dupoux, Paul Michel |
| 2022 | ACL | The Paradox of the Compositionality of Natural Language: A Neural Machine Translation Case Study. | Verna Dankers, Elia Bruni, Dieuwke Hupkes |
| 2022 | CogSci | Evaluating locality in NMT models. | Itay Itzhak, Koustuv Sinha, Brenden M. Lake, Adina Williams, Dieuwke Hupkes |
| 2022 | COLING | Can Transformers Process Recursive Nested Constructions, Like Humans? | Yair Lakretz, Tho Desbordes, Dieuwke Hupkes, Stanislas Dehaene |
| 2022 | EMNLP | The Curious Case of Absolute Position Embeddings. | Koustuv Sinha, Amirhossein Kazemnejad, Siva Reddy, Joelle Pineau, Dieuwke Hupkes, Adina Williams |
| 2022 | IJCNLP | Text Characterization Toolkit (TCT). | Daniel Simig, Tianlu Wang, Verna Dankers, Peter Henderson, Khuyagbaatar Batsuren, Dieuwke Hupkes, Mona T. Diab |
| 2021 | ACL | Language Models Use Monotonicity to Assess NPI Licensing. | Jaap Jumelet, Milica Denic, Jakub Szymanik, Dieuwke Hupkes, Shane Steinert-Threlkeld |
| 2021 | CoNLL | Generalising to German Plural Noun Classes, from the Perspective of a Recurrent Neural Network. | Verna Dankers, Anna Langedijk, Kate McCurdy, Adina Williams, Dieuwke Hupkes |
| 2021 | EACL | Co-evolution of language and agents in referential games. | Gautier Dagan, Dieuwke Hupkes, Elia Bruni |
| 2021 | EACL | Language Modelling as a Multi-Task Problem. | Lucas Weber, Jaap Jumelet, Elia Bruni, Dieuwke Hupkes |
| 2021 | EMNLP | Masked Language Modeling and the Distributional Hypothesis: Order Word Matters Pre-training for Little. | Koustuv Sinha, Robin Jia, Dieuwke Hupkes, Joelle Pineau, Adina Williams, Douwe Kiela |
| 2020 | ACL | Location Attention for Extrapolation to Longer Sequences. | Yann Dubois, Gautier Dagan, Dieuwke Hupkes, Elia Bruni |
| 2020 | EMNLP | Internal and External Pressures on Language Emergence: Least Effort, Object Constancy and Frequency. | Diana Rodrguez Luna, Edoardo Maria Ponti, Dieuwke Hupkes, Elia Bruni |
| 2020 | EMNLP | The Grammar of Emergent Languages. | Oskar van der Wal, Silvan de Boer, Elia Bruni, Dieuwke Hupkes |
| 2020 | IJCAI | Compositionality Decomposed: How do Neural Networks Generalise? (Extended Abstract). | Dieuwke Hupkes, Verna Dankers, Mathijs Mul, Elia Bruni |
| 2019 | CoNLL | Analysing Neural Language Models: Contextual Decomposition Reveals Default Reasoning in Number and Gender Assignment. | Jaap Jumelet, Willem H. Zuidema, Dieuwke Hupkes |
| 2019 | NAACL | The emergence of number and syntax units in LSTM language models. | Yair Lakretz, Germn Kruszewski, Theo Desbordes, Dieuwke Hupkes, Stanislas Dehaene, Marco Baroni |
| 2018 | EMNLP | Under the Hood: Using Diagnostic Classifiers to Investigate and Improve how Language Models Track Agreement Information. | Mario Giulianelli, Jack Harding, Florian Mohnert, Dieuwke Hupkes, Willem H. Zuidema |
| 2018 | EMNLP | Analysing the potential of seq-to-seq models for incremental interpretation in task-oriented dialogue. | Dieuwke Hupkes, Sanne Bouwmeester, Raquel Fernndez |
| 2018 | EMNLP | Do Language Models Understand Anything? On the Ability of LSTMs to Understand Negative Polarity Items. | Jaap Jumelet, Dieuwke Hupkes |
| 2018 | IJCAI | Visualisation and 'Diagnostic Classifiers' Reveal how Recurrent and Recursive Neural Networks Process Hierarchical Structure (Extended Abstract). | Dieuwke Hupkes, Willem H. Zuidema |
| 2016 | LREC | POS-tagging of Historical Dutch. | Dieuwke Hupkes, Rens Bod |