TY - RPRT TI - Adaptive Multi-Fidelity Reinforcement Learning for Variance Reduction in Engineering Design Optimization AU - Akash Agrawal AU - Christopher McComb PY - 2025 UR - https://arxiv.org/abs/2503.18229 ID - 2503.18229 ER -