TY - RPRT TI - Freshness-Aware Prioritized Experience Replay for LLM/VLM Reinforcement Learning AU - Weiyu Ma AU - Yongcheng Zeng AU - Yan Song AU - Xinyu Cui AU - Jian Zhao AU - Xuhui Liu AU - Mohamed Elhoseiny PY - 2026 UR - https://arxiv.org/abs/2604.16918 ID - 2604.16918 ER -