Skip to content

Supporting Human Raters with the Detection of Harmful Content Using Large Language Models.

Kurt Thomas, Patrick Gage Kelley, David Tao, Sarah Meiklejohn, Owen Vallis, Shunwen Tan, Blaz Bratanic, Felipe Tiengo Ferreira, Vijay Kumar Eranti, Elie Bursztein

VenueA*SP
Year2025
ProceedingsSP

Browse the full SP paper archive.