Structured, flexible, and robust: benchmarking and improving large language models towards more human-like behavior in out-of-distribution reasoning tasks.
Katherine M. Collins, Catherine Wong, Jiahai Feng, Megan Wei, Josh Tenenbaum
Browse the full CogSci paper archive.