TY - RPRT TI - Multi-Agent LLM Governance for Safe Two-Timescale Reinforcement Learning in SDN-IoT Defense AU - Saeid Jamshidi AU - Negar Shahabi AU - Foutse Khomh AU - Carol Fung AU - Mohammad Hamdaqa PY - 2026 UR - https://arxiv.org/abs/2604.01127 ID - 2604.01127 ER -