| 2026 | ACL | AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering? | Maharshi Gor, Yoo Yeon Sung, Yu Hou, Eve Fleisig, Zhu Irene Ying, Tianyi Zhou, Jordan Lee Boyd-Graber |
| 2025 | ACL | GRACE: A Granular Benchmark for Evaluating Model Calibration against Human Calibration. | Yoo Yeon Sung, Eve Fleisig, Yu Hou, Ishan Upadhyay, Jordan Lee Boyd-Graber |
| 2025 | NAACL | Is your benchmark truly adversarial? AdvScore: Evaluating Human-Grounded Adversarialness. | Yoo Yeon Sung, Maharshi Gor, Eve Fleisig, Ishani Mondal, Jordan Lee Boyd-Graber |
| 2024 | EMNLP | Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination. | Eve Fleisig, Genevieve Smith, Madeline Bossi, Ishita Rustagi, Xavier Yin, Dan Klein |
| 2024 | EMNLP | Accurate and Data-Efficient Toxicity Prediction when Annotators Disagree. | Harbani Jaggi, Kashyap Coimbatore Murali, Eve Fleisig, Erdem Biyik |
| 2024 | NAACL | The Perspectivist Paradigm Shift: Assumptions and Challenges of Capturing Human Labels. | Eve Fleisig, Su Lin Blodgett, Dan Klein, Zeerak Talat |
| 2024 | NAACL | First Tragedy, then Parse: History Repeats Itself in the New Era of Large Language Models. | Naomi Saphra, Eve Fleisig, Kyunghyun Cho, Adam Lopez |
| 2024 | NAACL | Ghostbuster: Detecting Text Ghostwritten by Large Language Models. | Vivek Verma, Eve Fleisig, Nicholas Tomlin, Dan Klein |
| 2023 | ACL | FairPrism: Evaluating Fairness-Related Harms in Text Generation. | Eve Fleisig, Aubrie Amstutz, Chad Atalla, Su Lin Blodgett, Hal Daum III, Alexandra Olteanu, Emily Sheng, Dan Vann, Hanna M. Wallach |
| 2023 | EMNLP | When the Majority is Wrong: Modeling Annotator Disagreement for Subjective Tasks. | Eve Fleisig, Rediet Abebe, Dan Klein |
| 2023 | EMNLP | Incorporating Worker Perspectives into MTurk Annotation Practices for NLP. | Olivia Huang, Eve Fleisig, Dan Klein |
| 2023 | EMNLP | Centering the Margins: Outlier-Based Identification of Harmed Populations in Toxicity Detection. | Vyoma Raman, Eve Fleisig, Dan Klein |