Skip to content

Interpretability of LLM Classifiers via the Rational Inattention Theory with Application to Hate Speech Detection.

Yuan Zhao, Ali Abdi

VenueA*ACL
Year2026
ProceedingsACL (4)

Browse the full ACL paper archive.