TY - RPRT TI - Learning to Factor Policies and Action-Value Functions: Factored Action Space Representations for Deep Reinforcement learning AU - Sahil Sharma AU - Aravind Suresh AU - Rahul Ramesh AU - Balaraman Ravindran PY - 2017 UR - https://arxiv.org/abs/1705.07269 ID - 1705.07269 ER -