TY - RPRT TI - Reinforcement Learning-based Recommender Systems with Large Language Models for State Reward and Action Modeling AU - Jie Wang AU - Alexandros Karatzoglou AU - Ioannis Arapakis AU - Joemon M. Jose PY - 2024 UR - https://arxiv.org/abs/2403.16948 ID - 2403.16948 ER -