Obedience or Vigilance? How Large Language Models React to Malicious Multiple-Choice Options (Student Abstract).
Yow-Fu Liou, Yu-Chien Tang, An-Zi Yen
Browse the full AAAI paper archive.
Yow-Fu Liou, Yu-Chien Tang, An-Zi Yen
Browse the full AAAI paper archive.