TY - RPRT TI - Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading AU - Minrui Xu AU - Dusit Niyato AU - Christopher G. Brinton PY - 2025 UR - https://arxiv.org/abs/2501.14205 ID - 2501.14205 ER -