Skip to content

Correcting Negative Bias in Large Language Models through Negative Attention Score Alignment.

Sangwon Yu, Jongyoon Song, Bongkyu Hwang, Hoyoung Kang, Sooah Cho, Junhwa Choi, Seongho Joe, Taehee Lee, Youngjune Gwon, Sungroh Yoon

VenueANAACL
Year2025
ProceedingsNAACL (Long Papers)

Browse the full NAACL paper archive.