TY - RPRT TI - From Verdict to Process: Agentic Reinforcement Learning for Multi-Stage Fact Verification AU - Rongxin Yang AU - Shenghong He AU - Siyuan Zhu AU - Chao Yu PY - 2026 UR - https://arxiv.org/abs/2606.13262 ID - 2606.13262 ER -