TY - RPRT TI - Relative Entropy Regularized Reinforcement Learning for Efficient Encrypted Policy Synthesis AU - Jihoon Suh AU - Yeongjun Jang AU - Kaoru Teranishi AU - Takashi Tanaka PY - 2025 DO - 10.1109/lcsys.2025.3578573 UR - https://arxiv.org/abs/2506.12358 ID - 2506.12358 ER -