Skip to content

Watching the AI Watchdogs: A Fairness and Robustness Analysis of AI Safety Moderation Classifiers.

Akshit Achara, Anshuman Chhabra

VenueANAACL
Year2025
ProceedingsNAACL (Short Papers)

Browse the full NAACL paper archive.