TY - RPRT TI - MetaResearcher: Scaling Deep Research via Self-Reflective Reinforcement Learning in Adversarial Virtual Environments AU - Wei Yu AU - Suxing Liu AU - Minjie Yu AU - Jiahao Wang AU - Zhijian Zheng AU - Haocheng Deng AU - Bing Li PY - 2026 UR - https://arxiv.org/abs/2606.19893 ID - 2606.19893 ER -