| 2026 | ACL | From If-Statements to ML Pipelines: Revisiting Bias in Code-Generation. | Minh Duc Bui, Xenia Heilmann, Mattia Cerrato, Manuel Mager, Katharina von der Wense |
| 2026 | ACL | Large Language Models Are Overconfident in Their Own Responses. | Mario Sanz-Guerrero, Manuel Mager, Katharina von der Wense |
| 2026 | LREC | Meenz bleibt Meenz, but Large Language Models Do Not Speak Its Dialect. | Minh Duc Bui, Manuel Mager, Peter Herbert Kann, Katharina von der Wense |
| 2025 | ACL | On Generalization across Measurement Systems: LLMs Entail More Test-Time Compute for Underrepresented Cultures. | Minh Duc Bui, Kyung Eun Park, Goran Glavas, Fabian David Schmidt, Katharina von der Wense |
| 2025 | ACL | Understanding the Gap: an Analysis of Research Collaborations in NLP and Language Documentation. | Luke Gessler, Alexis Palmer, Katharina von der Wense |
| 2025 | ACL | CLIX: Cross-Lingual Explanations of Idiomatic Expressions. | Aaron Gluck, Katharina von der Wense, Maria Leonor Pacheco |
| 2025 | ACL | The Effectiveness of Uncased Tokeniziaion for Clinical Notes. | Cory Paik, Katharina von der Wense |
| 2025 | ACL | MALAMUTE: A Multilingual, Highly-granular, Template-free, Education-based Probing Dataset. | Sagi Shaier, George Arthur Baker, Chiranthan Sridhar, Lawrence Hunter, Katharina von der Wense |
| 2025 | ACL | Improving Low-Resource Morphological Inflection via Self-Supervised Objectives. | Adam Wiemerslage, Katharina von der Wense |
| 2025 | COLING | From Priest to Doctor: Domain Adaptation for Low-Resource Neural Machine Translation. | Ali Marashian, Enora Rice, Luke Gessler, Alexis Palmer, Katharina von der Wense |
| 2025 | COLING | Measuring Contextual Informativeness in Child-Directed Text. | Maria R. Valentini, Ta Y. Wright, Ali Marashian, Jennifer Weber, Eliana Colunga, Katharina von der Wense |
| 2025 | EMNLP | ReSeeding Latent States for Sequential Language Understanding. | Stphane Aroca-Ouellette, Katharina von der Wense, Alessandro Roncone |
| 2025 | EMNLP | Molecular String Representation Preferences in Pretrained LLMs: A Comparative Study in Zero- & Few-Shot Molecular Property Prediction. | George Arthur Baker, Mario Sanz-Guerrero, Katharina von der Wense |
| 2025 | EMNLP | Large Language Models Discriminate Against Speakers of German Dialects. | Minh Duc Bui, Carolin Holtermann, Valentin Hofmann, Anne Lauscher, Katharina von der Wense |
| 2025 | EMNLP | Model-Based Ranking of Source Languages for Zero-Shot Cross-Lingual Transfer. | Abteen Ebrahimi, Adam Wiemerslage, Katharina von der Wense |
| 2025 | EMNLP | Linguistic Alignment Predicts Learning in Small Group Tutoring Sessions. | Dorothea French, Robert G. Moulder, Kelechi Ezema, Katharina von der Wense, Sidney K. D'Mello |
| 2025 | EMNLP | Interdisciplinary Research in Conversation: A Case Study in Computational Morphology for Language Documentation. | Enora Rice, Katharina von der Wense, Alexis Palmer |
| 2025 | EMNLP | Mind the Gap: A Closer Look at Tokenization for Multiple-Choice Question Answering with LLMs. | Mario Sanz-Guerrero, Minh Duc Bui, Katharina von der Wense |
| 2025 | ICLR | More Experts Than Galaxies: Conditionally-Overlapping Experts with Biologically-Inspired Fixed Routing. | Sagi Shaier, Francisco Pereira, Katharina von der Wense, Lawrence Hunter, Matt Jones |
| 2025 | IJCAI | Implicitly Aligning Humans and Autonomous Agents through Shared Task Abstractions. | Stphane Aroca-Ouellette, Miguel Aroca-Ouellette, Katharina von der Wense, Alessandro Roncone |
| 2025 | IJCNLP | Mitigating Label Length Bias in Large Language Models. | Mario Sanz-Guerrero, Katharina von der Wense |
| 2025 | NAACL | Multi³Hate: Multimodal, Multilingual, and Multicultural Hate Speech Detection with Vision-Language Models. | Minh Duc Bui, Katharina von der Wense, Anne Lauscher |
| 2024 | ACL | TAMS: Translation-Assisted Morphological Segmentation. | Enora Rice, Ali Marashian, Luke Gessler, Alexis Palmer, Katharina von der Wense |
| 2024 | ACL | It Is Not About What You Say, It Is About How You Say It: A Surprisingly Simple Approach for Improving Reading Comprehension. | Sagi Shaier, Lawrence Hunter, Katharina von der Wense |
| 2024 | CogSci | Evaluating LLMs as Tools to Support Early Vocabulary Learning. | Jennifer Weber, Maria R. Valentini, Ta Wright, Katharina von der Wense, Eliana Colunga |
| 2024 | EACL | Comparing Template-based and Template-free Language Model Probing. | Sagi Shaier, Kevin Bennett, Lawrence Hunter, Katharina von der Wense |
| 2024 | EACL | Desiderata For The Context Use Of Question Answering Systems. | Sagi Shaier, Lawrence Hunter, Katharina von der Wense |
| 2024 | EACL | Quantifying the Hyperparameter Sensitivity of Neural Networks for Character-level Sequence-to-Sequence Tasks. | Adam Wiemerslage, Kyle Gorman, Katharina von der Wense |
| 2024 | EMNLP | Getting The Most Out of Your Training Data: Exploring Unsupervised Tasks for Morphological Inflection. | Abhishek Purushothama, Adam Wiemerslage, Katharina von der Wense |
| 2024 | NAACL | Zero-Shot vs. Translation-Based Cross-Lingual Transfer: The Case of Lexical Gaps. | Abteen Ebrahimi, Katharina von der Wense |
| 2024 | RO-MAN | Eyes on the Game: Deciphering Implicit Human Signals to Infer Human Proficiency, Trust, and Intent. | Nikhil Hulle, Stphane Aroca-Ouellette, Anthony J. Ries, Jake Brawer, Katharina von der Wense, Alessandro Roncone |
| 2023 | EMNLP | On the Automatic Generation and Simplification of Children's Stories. | Maria R. Valentini, Jennifer Weber, Jesus Salcido, Ta Wright, Eliana Colunga, Katharina von der Wense |