TY - RPRT TI - Learning Reward Functions by Integrating Human Demonstrations and Preferences AU - Malayandi Palan AU - Nicholas C. Landolfi AU - Gleb Shevchuk AU - Dorsa Sadigh PY - 2019 UR - https://arxiv.org/abs/1906.08928 ID - 1906.08928 ER -