| 2025 | ACL | Memorization vs. Reasoning: Updating LLMs with New Knowledge. | Aochong Oliver Li, Tanya Goyal |
| 2025 | ACL | RefreshKV: Updating Small KV Cache During Long-form Generation. | Fangyuan Xu, Tanya Goyal, Eunsol Choi |
| 2025 | EMNLP | DCRM: A Heuristic to Measure Response Pair Quality in Preference Optimization. | Chengyu Huang, Tanya Goyal |
| 2025 | EMNLP | The Progress Illusion: Revisiting meta-evaluation standards of LLM evaluators. | Tianruo Rose Xu, Vedant Gaur, Liu Leqi, Tanya Goyal |
| 2025 | NAACL | Challenges in Trustworthy Human Evaluation of Chatbots. | Wenting Zhao, Alexander M. Rush, Tanya Goyal |
| 2024 | EMNLP | LitSearch: A Retrieval Benchmark for Scientific Literature Search. | Anirudh Ajith, Mengzhou Xia, Alexis Chevalier, Tanya Goyal, Danqi Chen, Tianyu Gao |
| 2024 | EMNLP | One Thousand and One Pairs: A "novel" challenge for long-context language models. | Marzena Karpinska, Katherine Thai, Kyle Lo, Tanya Goyal, Mohit Iyyer |
| 2024 | ICLR | BooookScore: A systematic exploration of book-length summarization in the era of LLMs. | Yapei Chang, Kyle Lo, Tanya Goyal, Mohit Iyyer |
| 2024 | ICLR | Evaluating Large Language Models at Evaluating Instruction Following. | Zhiyuan Zeng, Jiatong Yu, Tianyu Gao, Yu Meng, Tanya Goyal, Danqi Chen |
| 2023 | ACL | Understanding Factual Errors in Summarization: Errors, Summarizers, Datasets, Error Detectors. | Liyan Tang, Tanya Goyal, Alexander R. Fabbri, Philippe Laban, Jiacheng Xu, Semih Yavuz, Wojciech Kryscinski, Justin F. Rousseau, Greg Durrett |
| 2023 | EACL | Shortcomings of Question Answering Based Factuality Frameworks for Error Localization. | Ryo Kamoi, Tanya Goyal, Greg Durrett |
| 2023 | EMNLP | WiCE: Real-World Entailment for Claims in Wikipedia. | Ryo Kamoi, Tanya Goyal, Juan Diego Rodriguez, Greg Durrett |
| 2022 | ACL | Training Dynamics for Text Summarization Models. | Tanya Goyal, Jiacheng Xu, Junyi Jessy Li, Greg Durrett |
| 2022 | EMNLP | SNaC: Coherence Error Detection for Narrative Summarization. | Tanya Goyal, Junyi Jessy Li, Greg Durrett |
| 2022 | EMNLP | FALTE: A Toolkit for Fine-grained Annotation for Long Text Evaluation. | Tanya Goyal, Junyi Jessy Li, Greg Durrett |
| 2022 | EMNLP | HydraSum: Disentangling Style Features in Text Summarization with Multi-Decoder Models. | Tanya Goyal, Nazneen Rajani, Wenhao Liu, Wojciech Kryscinski |
| 2021 | NAACL | Annotating and Modeling Fine-grained Factuality in Summarization. | Tanya Goyal, Greg Durrett |
| 2020 | ACL | Neural Syntactic Preordering for Controlled Paraphrase Generation. | Tanya Goyal, Greg Durrett |
| 2020 | EMNLP | Evaluating Factuality in Generation with Dependency-level Entailment. | Tanya Goyal, Greg Durrett |
| 2019 | ACL | Embedding Time Expressions for Deep Temporal Ordering Models. | Tanya Goyal, Greg Durrett |
| 2018 | HCOMP | Your Behavior Signals Your Reliability: Modeling Crowd Behavioral Traces to Ensure Quality Relevance Annotations. | Tanya Goyal, Tyler McDonnell, Mcahid Kutlu, Tamer Elsayed, Matthew Lease |
| 2018 | PAKDD | Harvesting Knowledge from Cultural Heritage Artifacts in Museums of India. | Abhilasha Sancheti, Paridhi Maheshwari, Rajat Chaturvedi, Anish V. Monsy, Tanya Goyal, Balaji Vasan Srinivasan |
| 2017 | EMNLP | An Empirical Analysis of Edit Importance between Document Versions. | Tanya Goyal, Sachin Kelkar, Manas Agarwal, Jeenu Grover |
| 2017 | PAKDD | Preventing Inadvertent Information Disclosures via Automatic Security Policies. | Tanya Goyal, Sanket Mehta, Balaji Vasan Srinivasan |
| 2016 | IUI | Environment Specific Content Rendering & Transformation. | Balaji Vasan Srinivasan, Tanya Goyal, Varun Syal, Shubhankar Suman Singh, Vineet Sharma |