TY - RPRT TI - Multi-Objective Reinforcement Learning for Large-Scale Tote Allocation in Human-Robot Collaborative Fulfillment Centers AU - Sikata Sengupta AU - Guangyi Liu AU - Omer Gottesman AU - Joseph W Durham AU - Michael Kearns AU - Aaron Roth AU - Michael Caldara PY - 2026 UR - https://arxiv.org/abs/2602.24182 ID - 2602.24182 ER -