TY - RPRT TI - Learning Human Rewards by Inferring Their Latent Intelligence Levels in Multi-Agent Games: A Theory-of-Mind Approach with Application to Driving Data AU - Ran Tian AU - Masayoshi Tomizuka AU - Liting Sun PY - 2021 UR - https://arxiv.org/abs/2103.04289 ID - 2103.04289 ER -