| 2026 | ACL | On Emergent Social World Models - Evidence for Functional Integration of Theory of Mind and Pragmatic Reasoning in Language Models. | Polina Tsvilodub, Jan-Felix Klumpp, Amir Pour, Jennifer Hu, Michael Franke |
| 2025 | CogSci | Making Sense of Nonsense. | Jennifer Hu, Felix A. Sosa, Tomer D. Ullman |
| 2025 | CogSci | Language production is harder than comprehension for children and language models. | Jennifer Hu, Alvin Wei Ming Tan, Steven Y. Feng, Michael C. Frank |
| 2025 | CogSci | The Uncanny Valley meets the Humorous Hill: Things are funny when they match a pattern but fall short on quality. | Antara Raaghavi Bhattacharya, Jennifer Hu, Tomer D. Ullman |
| 2025 | NAACL | One fish, two fish, but not the whole sea: Alignment reduces language models' conceptual diversity. | Sonia K. Murthy, Tomer D. Ullman, Jennifer Hu |
| 2024 | CogSci | Shades of Zero: Distinguishing impossibility from inconceivability. | Jennifer Hu, Felix A. Sosa, Tomer D. Ullman |
| 2024 | CogSci | The Task Task: Creative problem generation in humans and language models. | Junyi Chu, Jennifer Hu, Tomer D. Ullman |
| 2023 | ACL | A fine-grained comparison of pragmatic language understanding in humans and language models. | Jennifer Hu, Sammy Floyd, Olessia Jouravlev, Evelina Fedorenko, Edward Gibson |
| 2023 | ACL | I Cast Detect Thoughts: Learning to Converse and Guide with Intents and Theory-of-Mind in Dungeons and Dragons. | Pei Zhou, Andrew Zhu, Jennifer Hu, Jay Pujara, Xiang Ren, Chris Callison-Burch, Yejin Choi, Prithviraj Ammanabrolu |
| 2023 | EMNLP | Pragmatics in Language Grounding: Phenomena, Tasks, and Modeling Approaches. | Daniel Fried, Nicholas Tomlin, Jennifer Hu, Roma Patel, Aida Nematzadeh |
| 2023 | EMNLP | Prompting is not a substitute for probability measurements in large language models. | Jennifer Hu, Roger Levy |
| 2022 | CogSci | Teasing apart models of pragmatics using optimal reference game design. | Irene Zhou, Jennifer Hu, Roger Levy, Noga Zaslavsky |
| 2021 | CogSci | Competition from novel features drives scalar inferences in reference games. | Jennifer Hu, Noga Zaslavsky, Roger Levy |
| 2021 | CogSci | Empirical Support for a Rate-Distortion Account of Pragmatic Reasoning. | Irene Zhou, Jennifer Hu, Roger Levy, Noga Zaslavsky |
| 2021 | EMNLP | Controlled Evaluation of Grammatical Knowledge in Mandarin Chinese Language Models. | Yiwen Wang, Jennifer Hu, Roger Levy, Peng Qian |
| 2020 | ACL | SyntaxGym: An Online Platform for Targeted Evaluation of Language Models. | Jon Gauthier, Jennifer Hu, Ethan Wilcox, Peng Qian, Roger Levy |
| 2020 | ACL | A Systematic Assessment of Syntactic Generalization in Neural Language Models. | Jennifer Hu, Jon Gauthier, Peng Qian, Ethan Wilcox, Roger Levy |
| 2020 | CogSci | On the Predictive Power of Neural Language Models for Human Real-Time Comprehension Behavior. | Ethan Wilcox, Jon Gauthier, Jennifer Hu, Peng Qian, Roger Levy |
| 2019 | CogSci | Separating object resonance and room reverberation in impact sounds. | Jennifer Hu, James Traer, Josh H. McDermott |
| 2018 | NAACL | Generating Bilingual Pragmatic Color References. | Will Monroe, Jennifer Hu, Andrew Jong, Christopher Potts |