TY - RPRT TI - Offline Reinforcement Learning for Warehouse SLAM Throughput Control AU - Tina Dongxu Li AU - Mouhacine Benosman AU - Rajat Kumar AU - Kevin Tan AU - Ken Meszaros AU - Trevor Dardik PY - 2026 UR - https://arxiv.org/abs/2606.23978 ID - 2606.23978 ER -