TY - RPRT TI - Reinforcement Learning-Guided Evolutionary Policy Optimization for Preference-Adjustable Heterogeneous Agile Earth Observation Satellite Scheduling AU - He Wang AU - Junyu Wu AU - Hui Li AU - Yanjie Song AU - Witold Pedrycz AU - Liang Li PY - 2026 UR - https://arxiv.org/abs/2608.24470 ID - 2608.24470 ER -