Whose Values Prevail? Bias in Large Language Model Value Alignment.
Ruoxi Qi, Gleb Papyshev, Kellee Tsai, Antoni B. Chan, Janet H. Hsiao
Browse the full CogSci paper archive.
Ruoxi Qi, Gleb Papyshev, Kellee Tsai, Antoni B. Chan, Janet H. Hsiao
Browse the full CogSci paper archive.