TY - RPRT TI - Heuristic Algorithm-based Action Masking Reinforcement Learning (HAAM-RL) with Ensemble Inference Method AU - Kyuwon Choi AU - Cheolkyun Rho AU - Taeyoun Kim AU - Daewoo Choi PY - 2024 UR - https://arxiv.org/abs/2403.14110 ID - 2403.14110 ER -