TY - RPRT TI - Decision-Focused On-Policy Learning for Contextual Linear Optimization with Partial Feedback AU - Wyame Benslimane AU - Tinghan Ye AU - Pascal Van Hentenryck AU - Paul Grigas PY - 2026 UR - https://arxiv.org/abs/2606.01081 ID - 2606.01081 ER -