CRDA: Content Risk Drift Assessment of Large Language Models through Adversarial Multi-Agent Interaction.
Zongzhen Liu, Guoyi Li, Bingkang Shi, Xiaodan Zhang, Jingguo Ge, Yulei Wu, Honglei Lyu
Browse the full IJCNN paper archive.
Zongzhen Liu, Guoyi Li, Bingkang Shi, Xiaodan Zhang, Jingguo Ge, Yulei Wu, Honglei Lyu
Browse the full IJCNN paper archive.