TY - RPRT TI - Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference AU - Ruokai Yin AU - Sattwik Deb Mishra AU - Xuan Zuo AU - Hokchhay Tann AU - Preyas Shah AU - Apala Guha PY - 2025 UR - https://arxiv.org/abs/2509.00217 ID - 2509.00217 ER -