Skip to content

HarmRLVR: Weaponizing Verifiable Rewards for Harmful LLM Alignment.

Yuexiao Liu, Lijun Li, Xingjun Wang, Jing Shao

VenueA*ACL
Year2026
ProceedingsACL (1)

Browse the full ACL paper archive.