TY - RPRT TI - On-Device Deep Reinforcement Learning for Decentralized Task Offloading Performance trade-offs in the training process AU - Gorka Nieto AU - Idoia de la Iglesia AU - Cristina Perfecto AU - Unai Lopez-Novoa PY - 2026 UR - https://arxiv.org/abs/2601.03976 ID - 2601.03976 ER -