TY - RPRT TI - A Method to Improve the Performance of Reinforcement Learning Based on the Y Operator for a Class of Stochastic Differential Equation-Based Child-Mother Systems AU - Cheng Yin AU - Yi Chen PY - 2025 UR - https://arxiv.org/abs/2311.04014 ID - 2311.04014 ER -