| 2026 | ACL | Translation or Recitation? Calibrating Evaluation Scores for Machine Translation of Extremely Low-Resource Languages. | Danlu Chen, Ka Sing He, Jiahe Tian, Chenghao Xiao, Zhaofeng Wu, Taylor Berg-Kirkpatrick, Freda Shi |
| 2026 | ACL | How Tokenization Limits Phonological Knowledge Representation in Language Models and How to Improve Them. | Disen Liao, Freda Shi |
| 2026 | ACL | On the Effect of Hyperparameters in Language Modeling for Computational Linguistics. | Ruoxi Ning, Yongpeng Zhu, Qingcheng Zeng, Tatsuki Kuribayashi, Freda Shi |
| 2025 | AAAI | Learning Language Structures Through Grounding. | Freda Shi |
| 2025 | ACL | SpaRE: Enhancing Spatial Reasoning in Vision-Language Models with Synthetic Data. | Michael Ogezi, Freda Shi |
| 2025 | ACL | FORG3D: Flexible Object Rendering for Generating Vision-Language Spatial Reasoning Data from 3D Scenes. | Oscar Pang, Freda Shi |
| 2025 | ACL | Blessing of Multilinguality: A Systematic Analysis of Multilingual In-Context Learning. | Yilei Tu, Andrew Xue, Freda Shi |
| 2025 | ACL | Logical forms complement probability in understanding language model (and human) performance. | Yixuan Wang, Freda Shi |
| 2025 | EMNLP | From Behavioral Performance to Internal Competence: Interpreting Vision-Language Models with VLM-Lens. | Hala Sheta, Eric Huang, Shuyu Wu, Ilia Alenabi, Jiajun Hong, Ryker Lin, Ruoxi Ning, Daniel Wei, Jialin Yang, Jiawei Zhou, Ziqiao Ma, Freda Shi |
| 2025 | EMNLP | Distribution Prompting: Understanding the Expressivity of Language Models Through the Next-Token Distributions They Can Produce. | Haojin Wang, Zining Zhu, Freda Shi |
| 2025 | EMNLP | LingGym: How Far Are LLMs from Thinking Like Field Linguists? | Changbing Yang, Franklin Ma, Freda Shi, Jian Zhu |
| 2025 | ICLR | Do Vision-Language Models Represent Space and How? Evaluating Spatial Frame of Reference under Ambiguities. | Zheyuan Zhang, Fengyuan Hu, Jayjun Lee, Freda Shi, Parisa Kordjamshidi, Joyce Chai, Ziqiao Ma |
| 2024 | ACL | LogogramNLP: Comparing Visual and Textual Representations of Ancient Logographic Writing Systems for NLP. | Danlu Chen, Freda Shi, Aditi Agarwal, Jacobo Myerston, Taylor Berg-Kirkpatrick |
| 2024 | ACL | Structured Tree Alignment for Evaluation of (Speech) Constituency Parsing. | Freda Shi, Kevin Gimpel, Karen Livescu |
| 2023 | ASRU | Audio-Visual Neural Syntax Acquisition. | Cheng-I Jeff Lai, Freda Shi, Puyuan Peng, Yoon Kim, Kevin Gimpel, Shiyu Chang, Yung-Sung Chuang, Saurabhchand Bhati, David D. Cox, David Harwath, Yang Zhang, Karen Livescu, James R. Glass |
| 2023 | ICLR | InCoder: A Generative Model for Code Infilling and Synthesis. | Daniel Fried, Armen Aghajanyan, Jessy Lin, Sida Wang, Eric Wallace, Freda Shi, Ruiqi Zhong, Scott Yih, Luke Zettlemoyer, Mike Lewis |
| 2023 | ICLR | Language models are multilingual chain-of-thought reasoners. | Freda Shi, Mirac Suzgun, Markus Freitag, Xuezhi Wang, Suraj Srivats, Soroush Vosoughi, Hyung Won Chung, Yi Tay, Sebastian Ruder, Denny Zhou, Dipanjan Das, Jason Wei |
| 2023 | ICML | Large Language Models Can Be Easily Distracted by Irrelevant Context. | Freda Shi, Xinyun Chen, Kanishka Misra, Nathan Scales, David Dohan, Ed H. Chi, Nathanael Schrli, Denny Zhou |
| 2022 | ACL | Substructure Distribution Projection for Zero-Shot Cross-Lingual Dependency Parsing. | Freda Shi, Kevin Gimpel, Karen Livescu |
| 2022 | EMNLP | Natural Language to Code Translation with Execution. | Freda Shi, Daniel Fried, Marjan Ghazvininejad, Luke Zettlemoyer, Sida I. Wang |