TY - RPRT TI - Distilling Knowledge from Large Language Models into Lightweight Reinforcement Learning Agents for Autonomous Cyber Operations AU - Konur Tholl AU - François Rivest AU - Mariam El Mezouar AU - Adrian Taylor AU - Ranwa Al Mallah PY - 2026 UR - https://arxiv.org/abs/2607.28826 ID - 2607.28826 ER -