TY - RPRT TI - Memory-Based Advantage Shaping for LLM-Guided Reinforcement Learning AU - Narjes Nourzad AU - Carlee Joe-Wong PY - 2026 UR - https://arxiv.org/abs/2602.17931 ID - 2602.17931 ER -