| 2025 | AAAI | To Err Is AI: A Case Study Informing LLM Flaw Reporting Practices. | Sean McGregor, Allyson Ettinger, Nick Judd, Paul Albee, Liwei Jiang, Kavel Rao, William H. Smith, Shayne Longpre, Avijit Ghosh, Christopher Fiorelli, Michelle Hoang, Sven Cattell, Nouha Dziri |
| 2025 | ICLR | AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web Text. | Ximing Lu, Melanie Sclar, Skyler Hallinan, Niloofar Mireshghallah, Jiacheng Liu, Seungju Han, Allyson Ettinger, Liwei Jiang, Khyathi Raghavi Chandu, Nouha Dziri, Yejin Choi |
| 2024 | EMNLP | Experimental Contexts Can Facilitate Robust Semantic Property Inference in Language Models, but Inconsistently. | Kanishka Misra, Allyson Ettinger, Kyle Mahowald |
| 2024 | ICLR | The Generative AI Paradox: "What It Can Create, It May Not Understand". | Peter West, Ximing Lu, Nouha Dziri, Faeze Brahman, Linjie Li, Jena D. Hwang, Liwei Jiang, Jillian Fisher, Abhilasha Ravichander, Khyathi Raghavi Chandu, Benjamin Newman, Pang Wei Koh, Allyson Ettinger, Yejin Choi |
| 2024 | NAACL | When Hindsight is Not 20/20: Testing Limits on Reflective Thinking in Large Language Models. | Yanhong Li, Chenghao Yang, Allyson Ettinger |
| 2023 | ACL | Counterfactual reasoning: Testing language models' understanding of hypothetical scenarios. | Jiaxuan Li, Lang Yu, Allyson Ettinger |
| 2023 | EACL | COMPS: Conceptual Minimal Pair Sentences for testing Robust Property Knowledge and its Inheritance in Pre-trained Language Models. | Kanishka Misra, Julia Rayz, Allyson Ettinger |
| 2023 | EMNLP | "You Are An Expert Linguistic Annotator": Limits of LLMs as Analyzers of Abstract Meaning Representation. | Allyson Ettinger, Jena D. Hwang, Valentina Pyatkin, Chandra Bhagavatula, Yejin Choi |
| 2023 | EMNLP | Can You Follow Me? Testing Situational Understanding for ChatGPT. | Chenghao Yang, Allyson Ettinger |
| 2023 | IJCNLP | Linguistic Productivity: the Case of Determiners in English. | Raquel G. Alhama, Ruthe Foushee, Daniel Byrne, Allyson Ettinger, Susan Goldin-Meadow, Afra Alishahi |
| 2022 | CogSci | A Property Induction Framework for Neural Language Models. | Kanishka Misra, Julia Rayz, Allyson Ettinger |
| 2022 | COLING | "No, They Did Not": Dialogue Response Dynamics in Pre-trained Language Models. | Sanghee J. Kim, Lang Yu, Allyson Ettinger |
| 2021 | ACL | On the Interplay Between Fine-tuning and Composition in Transformers. | Lang Yu, Allyson Ettinger |
| 2021 | CogSci | Do language models learn typicality judgments from text? | Kanishka Misra, Allyson Ettinger, Julia Rayz |
| 2021 | CoNLL | Pragmatic competence of pre-trained language models through the lens of discourse connectives. | Lalchand Pandia, Yan Cong, Allyson Ettinger |
| 2021 | EMNLP | Sorting through the noise: Testing robustness of information processing in pre-trained language models. | Lalchand Pandia, Allyson Ettinger |
| 2020 | ACL | Spying on Your Neighbors: Fine-grained Probing of Contextual Embeddings for Information about Surrounding Words. | Josef Klafka, Allyson Ettinger |
| 2020 | ACL | PeTra: A Sparsely Supervised Memory Model for People Tracking. | Shubham Toshniwal, Allyson Ettinger, Kevin Gimpel, Karen Livescu |
| 2020 | CogSci | Exploring Lexical Relations in BERT using Semantic Priming. | Kanishka Misra, Allyson Ettinger, Julia Rayz |
| 2020 | EMNLP | Exploring BERT's sensitivity to lexical cues using tests from semantic priming. | Kanishka Misra, Allyson Ettinger, Julia Rayz |
| 2020 | EMNLP | Learning to Ignore: Long Document Coreference with Bounded Memory Neural Networks. | Shubham Toshniwal, Sam Wiseman, Allyson Ettinger, Karen Livescu, Kevin Gimpel |
| 2020 | EMNLP | Assessing Phrasal Representation and Composition in Transformers. | Lang Yu, Allyson Ettinger |
| 2018 | COLING | Assessing Composition in Sentence Vector Representations. | Allyson Ettinger, Ahmed Elgohary, Colin Phillips, Philip Resnik |
| 2016 | CogSci | Modeling N400 amplitude using vector space models of word representation. | Allyson Ettinger, Naomi Feldman, Philip Resnik, Colin Phillips |
| 2016 | NAACL | Retrofitting Sense-Specific Word Vectors Using Parallel Text. | Allyson Ettinger, Philip Resnik, Marine Carpuat |
| 2015 | NAACL | Dialogue focus tracking for zero pronoun resolution. | Sudha Rao, Allyson Ettinger, Hal Daum III, Philip Resnik |