TY - RPRT TI - PEARL: Plan Exploration and Adaptive Reinforcement Learning for Multihop Tool Use AU - Qihao Wang AU - Mingzhe Lu AU - Jiayue Wu AU - Yue Hu AU - Yanbing Liu PY - 2026 UR - https://arxiv.org/abs/2601.20439 ID - 2601.20439 ER -