Skip to content

CRDA: Content Risk Drift Assessment of Large Language Models through Adversarial Multi-Agent Interaction.

Zongzhen Liu, Guoyi Li, Bingkang Shi, Xiaodan Zhang, Jingguo Ge, Yulei Wu, Honglei Lyu

VenueBIJCNN
Year2024
ProceedingsIJCNN

Browse the full IJCNN paper archive.