TY - RPRT TI - Cache-Efficient Posterior Sampling for Reinforcement Learning with LLM-Derived Priors Across Discrete and Continuous Domains AU - Ibne Farabi Shihab AU - Sanjeda Akter AU - Anuj Sharma PY - 2025 UR - https://arxiv.org/abs/2505.07274 ID - 2505.07274 ER -