Peter Henderson
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
43
Venues
20
Active years
1971–2025
Best venue rank
A*
Where they publish
Papers
43 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | AIES | The Model Hears You: Audio Language Model Deployments Should Consider the Principle of Least Privilege. | Luxi He, Xiangyu Qi, Michel Liao, Inyoung Cheong, Prateek Mittal, Danqi Chen, Peter Henderson |
| 2025 | ICLR | Fantastic Copyrighted Beasts and How (Not) to Generate Them. | Luxi He, Yangsibo Huang, Weijia Shi, Tinghao Xie, Haotian Liu, Yue Wang, Luke Zettlemoyer, Chiyuan Zhang, Danqi Chen, Peter Henderson |
| 2025 | ICLR | Safety Alignment Should be Made More Than Just a Few Tokens Deep. | Xiangyu Qi, Ashwinee Panda, Kaifeng Lyu, Xiao Ma, Subhrajit Roy, Ahmad Beirami, Prateek Mittal, Peter Henderson |
| 2025 | ICLR | On Evaluating the Durability of Safeguards for Open-Weight LLMs. | Xiangyu Qi, Boyi Wei, Nicholas Carlini, Yangsibo Huang, Tinghao Xie, Luxi He, Matthew Jagielski, Milad Nasr, Prateek Mittal, Peter Henderson |
| 2025 | ICLR | SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal. | Tinghao Xie, Xiangyu Qi, Yi Zeng, Yangsibo Huang, Udari Madhushani Sehwag, Kaixuan Huang, Luxi He, Boyi Wei, Dacheng Li, Ying Sheng, Ruoxi Jia, Bo Li, Kai Li, Danqi Chen, Peter Henderson, Prateek Mittal |
| 2025 | ICML | Position: In-House Evaluation Is Not Enough. Towards Robust Third-Party Evaluation and Flaw Disclosure for General-Purpose AI. | Shayne Longpre, Kevin Klyman, Ruth Elisabeth Appel, Sayash Kapoor, Rishi Bommasani, Michelle Sahar, Sean McGregor, Avijit Ghosh, Borhane Blili-Hamelin, Nathan Butters, Alondra Nelson, Amit Elazari, Andrew Sellars, Casey John Ellis, Dane Sherrets, Dawn Song, Harley Geiger, Ilona Cohen, Lauren McIlvenny, Madhulika Srikumar, Mark M. Jaycox, Markus Anderljung, Nadine Farid Johnson, Nicholas Carlini, Nicolas Miailhe, Nik Marda, Peter Henderson, Rebecca S. Portnoff, Rebecca Weiss, Victoria Westerhoff, Yacine Jernite, Rumman Chowdhury, Percy Liang, Arvind Narayanan |
| 2025 | NAACL | LawInstruct: A Resource for Studying Language Model Adaptation to the Legal Domain. | Joel Niklaus, Lucia Zheng, Arya D. McCarthy, Christopher Hahn, Brian M. Rosen, Peter Henderson, Daniel E. Ho, Garrett Honke, Percy Liang, Christopher D. Manning |
| 2024 | AAAI | Visual Adversarial Examples Jailbreak Aligned Large Language Models. | Xiangyu Qi, Kaixuan Huang, Ashwinee Panda, Peter Henderson, Mengdi Wang, Prateek Mittal |
| 2024 | ICLR | Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To! | Xiangyu Qi, Yi Zeng, Tinghao Xie, Pin-Yu Chen, Ruoxi Jia, Prateek Mittal, Peter Henderson |
| 2024 | ICML | Position: On the Societal Impact of Open Foundation Models. | Sayash Kapoor, Rishi Bommasani, Kevin Klyman, Shayne Longpre, Ashwin Ramaswami, Peter Cihon, Aspen K. Hopkins, Kevin Bankston, Stella Biderman, Miranda Bogen, Rumman Chowdhury, Alex Engler, Peter Henderson, Yacine Jernite, Seth Lazar, Stefano Maffulli, Alondra Nelson, Joelle Pineau, Aviya Skowron, Dawn Song, Victor Storchan, Daniel Zhang, Daniel E. Ho, Percy Liang, Arvind Narayanan |
| 2024 | ICML | Position: A Safe Harbor for AI Evaluation and Red Teaming. | Shayne Longpre, Sayash Kapoor, Kevin Klyman, Ashwin Ramaswami, Rishi Bommasani, Borhane Blili-Hamelin, Yangsibo Huang, Aviya Skowron, Zheng Xin Yong, Suhas Kotha, Yi Zeng, Weiyan Shi, Xianjun Yang, Reid Southen, Alexander Robey, Patrick Chao, Diyi Yang, Ruoxi Jia, Daniel Kang, Sandy Pentland, Arvind Narayanan, Percy Liang, Peter Henderson |
| 2024 | ICML | Assessing the Brittleness of Safety Alignment via Pruning and Low-Rank Modifications. | Boyi Wei, Kaixuan Huang, Yangsibo Huang, Tinghao Xie, Xiangyu Qi, Mengzhou Xia, Prateek Mittal, Mengdi Wang, Peter Henderson |
| 2023 | AAAI | Integrating Reward Maximization and Population Estimation: Sequential Decision-Making for Internal Revenue Service Audit Selection. | Peter Henderson, Ben Chugg, Brandon R. Anderson, Kristen M. Altenburger, Alex Turk, John Guyton, Jacob S. Goldin, Daniel E. Ho |
| 2023 | AAAI | Entropy Regularization for Population Estimation. | Ben Chugg, Peter Henderson, Jacob S. Goldin, Daniel E. Ho |
| 2023 | AIES | Self-Destructing Models: Increasing the Costs of Harmful Dual Uses of Foundation Models. | Peter Henderson, Eric Mitchell, Christopher D. Manning, Dan Jurafsky, Chelsea Finn |
| 2022 | IJCNLP | Text Characterization Toolkit (TCT). | Daniel Simig, Tianlu Wang, Verna Dankers, Peter Henderson, Khuyagbaatar Batsuren, Dieuwke Hupkes, Mona T. Diab |
| 2021 | ICAIL | When does pretraining help?: assessing self-supervised learning for law and the CaseHOLD dataset of 53, 000+ legal holdings. | Lucia Zheng, Neel Guha, Brandon R. Anderson, Peter Henderson, Daniel E. Ho |
| 2020 | EMNLP | With Little Power Comes Great Responsibility. | Dallas Card, Peter Henderson, Urvashi Khandelwal, Robin Jia, Kyle Mahowald, Dan Jurafsky |
| 2019 | ICML | Separable value functions across time-scales. | Joshua Romoff, Peter Henderson, Ahmed Touati, Yann Ollivier, Joelle Pineau, Emma Brunskill |
| 2018 | AAAI | OptionGAN: Learning Joint Reward-Policy Options Using Generative Adversarial Inverse Reinforcement Learning. | Peter Henderson, Wei-Di Chang, Pierre-Luc Bacon, David Meger, Joelle Pineau, Doina Precup |
| 2018 | AAAI | Deep Reinforcement Learning That Matters. | Peter Henderson, Riashat Islam, Philip Bachman, Joelle Pineau, Doina Precup, David Meger |
| 2018 | AIES | Ethical Challenges in Data-Driven Dialogue Systems. | Peter Henderson, Koustuv Sinha, Nicolas Angelard-Gontier, Nan Rosemary Ke, Genevieve Fried, Ryan Lowe, Joelle Pineau |
| 2018 | CoRL | Reward Estimation for Variance Reduction in Deep Reinforcement Learning. | Joshua Romoff, Peter Henderson, Alexandre Pich, Vincent Franois-Lavet, Joelle Pineau |
| 2018 | ICLR | Reward Estimation for Variance Reduction in Deep Reinforcement Learning. | Joshua Romoff, Alexandre Pich, Peter Henderson, Vincent Franois-Lavet, Joelle Pineau |
| 2018 | IROS | Cost Adaptation for Robust Decentralized Swarm Behaviour. | Peter Henderson, Matthew Vertescher, David Meger, Mark Coates |
| 2017 | IROS | Underwater multi-robot convoying using visual tracking by detection. | Florian Shkurti, Wei-Di Chang, Peter Henderson, Md Jahidul Islam, Juan Camilo Gamboa Higuera, Jimmy Li, Travis Manderson, Anqi Xu, Gregory Dudek, Junaed Sattar |
| 2009 | ICSR | Consistency Checking for Component Reuse in Open Systems. | Peter Henderson, Matthew J. Henderson |
| 2009 | SEKE | Collaborative Development of System Architecture - a Tool for Coping with Inconsistency. | Peter Henderson, Matthew J. Henderson |
| 2008 | COMPSAC | Utilising Located Functions to Model and Optimise Distributed Computations. | Stephen Crouch, Peter Henderson, Robert John Walters |
| 2008 | SEKE | System Architecture Induces Document Architecture. | Peter Henderson, Nishadi De Silva |
| 2007 | COMPSAC | DataWarp: Empowering Applications to Make Progress in the Face of Contradictory or Inconsistent Data. | Stephen Crouch, Peter Henderson, Robert John Walters |
| 2007 | SAC | Selecting a distributed agreement algorithm. | Robert John Walters, Peter Henderson, Stephen Crouch |
| 2005 | AINA | A Practical Modelling Notation for Secure Distributed Computation. | Yih-Jiun Lee, Peter Henderson |
| 2004 | COMPSAC | Implementing Hierarchical Features in a Graphically Based Formal Modelling Language. | Peter Henderson, Robert John Walters, Stephen Crouch |
| 2004 | ICSR | Reusable Web Services. | Peter Henderson, Jingtao Yang |
| 2003 | COMPSAC | Effects of Introducing Survival Behaviours into Automated Negotiators. | Peter Henderson, Stephen Crouch, Robert John Walters, Qinglai Ni |
| 2003 | DAIS | DataWarp: Building Applications Which Make Progress in an Inconsistent World. | Peter Henderson, Robert John Walters, Stephen Crouch, Qinglai Ni |
| 2002 | ICECCS | Reasoning about Asynchronous Behaviour in Distributed Systems. | Peter Henderson |
| 1999 | RSP | System Design Validation Using Formal Models. | Peter Henderson, Robert John Walters |
| 1998 | ICSR | Laws for dynamic systems. | Peter Henderson |
| 1995 | ICECCS | POSD-a notation for presenting complex systems of processes. | Peter Henderson, Graham D. Pratten |
| 1976 | POPL | A Lazy Evaluator. | Peter Henderson, James H. Morris Jr. |
| 1971 | IJCAI | Derived Semantics for Some Programming Language Constructs. | Peter Henderson |