TY - RPRT TI - Bridging Reinforcement Learning and Optimal Control via Feasible Action Mapping AU - Stefan Richter AU - Alberto Giammarino AU - Guillem Torrente AU - Sam Blakeman AU - Peter Dürr PY - 2026 UR - https://arxiv.org/abs/2607.23930 ID - 2607.23930 ER -