TY - RPRT TI - Learning to Construct Knowledge through Sparse Reference Selection with Reinforcement Learning AU - Shao-An Yin PY - 2025 UR - https://arxiv.org/abs/2509.05874 ID - 2509.05874 ER -