TY - RPRT TI - Structured, flexible, and robust: benchmarking and improving large language models towards more human-like behavior in out-of-distribution reasoning tasks AU - Katherine M. Collins AU - Catherine Wong AU - Jiahai Feng AU - Megan Wei AU - Joshua B. Tenenbaum PY - 2022 UR - https://arxiv.org/abs/2205.05718 ID - 2205.05718 ER -