| 2021 | Persistent Anti-Muslim Bias in Large Language Models. | Abubakar Abid, Maheen Farooqi, James Zou |
| 2021 | The Grey Hoodie Project: Big Tobacco, Big Tech, and the Threat on Academic Integrity. | Mohamed Abdalla, Moustafa Abdalla |
| 2021 | FaiR-N: Fair and Robust Neural Networks for Structured Data. | Shubham Sharma, Alan H. Gee, David Paydarfar, Joydeep Ghosh |
| 2021 | Becoming Good at AI for Good. | Meghana Kshirsagar, Caleb Robinson, Siyu Yang, Shahrzad Gholami, Ivan S. Klyuzhin, Sumit Mukherjee, Md Nasir, Anthony Ortiz, Felipe Oviedo, Darren Tanner, Anusua Trivedi, Yixi Xu, Ming Zhong, Bistra Dilkina, Rahul Dodhia, Juan M. Lavista Ferres |
| 2021 | Accounting for Model Uncertainty in Algorithmic Discrimination. | Junaid Ali, Preethi Lahoti, Krishna P. Gummadi |
| 2020 | Arbiter: A Domain-Specific Language for Ethical Machine Learning. | Julian Zucker, Myraeka d'Leeuwen |
| 2020 | Deepfakes for Medical Video De-Identification: Privacy Protection and Diagnostic Information Preservation. | Bingquan Zhu, Hao Fang, Yanan Sui, Luming Li |
| 2020 | Assessing Post-hoc Explainability of the BKT Algorithm. | Tongyu Zhou, Haoyu Sheng, Iris Howley |
| 2020 | Different "Intelligibility" for Different Folks. | Yishan Zhou, David Danks |
| 2020 | U.S. Public Opinion on the Governance of Artificial Intelligence. | Baobao Zhang, Allan Dafoe |
| 2020 | Joint Optimization of AI Fairness and Utility: A Human-Centered Approach. | Yunfeng Zhang, Rachel K. E. Bellamy, Kush R. Varshney |
| 2020 | A Deontic Logic for Programming Rightful Machines. | Ava Thomas Wright |
| 2020 | Conservative Agency via Attainable Utility Preservation. | Alexander Matt Turner, Dylan Hadfield-Menell, Prasad Tadepalli |
| 2020 | Social and Governance Implications of Improved Data Efficiency. | Aaron D. Tucker, Markus Anderljung, Allan Dafoe |
| 2020 | Why Reliabilism Is not Enough: Epistemic and Moral Justification in Machine Learning. | Andrew Smart, Larry James, Ben Hutchinson, Simone Wu, Shannon Vallor |
| 2020 | Fooling LIME and SHAP: Adversarial Attacks on Post hoc Explanation Methods. | Dylan Slack, Sophie Hilgard, Emily Jia, Sameer Singh, Himabindu Lakkaraju |
| 2020 | Meta Decision Trees for Explainable Recommendation Systems. | Eyal Shulman, Lior Wolf |
| 2020 | The Offense-Defense Balance of Scientific Knowledge: Does Publishing AI Research Reduce Misuse? | Toby Shevlane, Allan Dafoe |
| 2020 | Data Augmentation for Discrimination Prevention and Bias Disambiguation. | Shubham Sharma, Yunfeng Zhang, Jess M. Ros Aliaga, Djallel Bouneffouf, Vinod Muthusamy, Kush R. Varshney |
| 2020 | CERTIFAI: A Common Framework to Provide Explanations and Analyse the Fairness and Robustness of Black-box Models. | Shubham Sharma, Jette Henderson, Joydeep Ghosh |
| 2020 | Trade-offs in Fair Redistricting. | Zachary Schutzman |
| 2020 | What's Next for AI Ethics, Policy, and Governance? A Global Overview. | Daniel S. Schiff, Justin Biddle, Jason Borenstein, Kelly Laas |
| 2020 | Balancing the Tradeoff Between Clustering Value and Interpretability. | Sandhya Saisubramanian, Sainyam Galhotra, Shlomo Zilberstein |
| 2020 | Human Comprehension of Fairness in Machine Learning. | Debjani Saha, Candice Schumann, Duncan C. McElfresh, John P. Dickerson, Michelle L. Mazurek, Michael Carl Tschantz |
| 2020 | Saving Face: Investigating the Ethical Concerns of Facial Recognition Auditing. | Inioluwa Deborah Raji, Timnit Gebru, Margaret Mitchell, Joy Buolamwini, Joonseok Lee, Emily Denton |