arXiv · 2609.32061
Contract monitoring: governing AI via separation of powers
Abstract
We propose an AI safety framework that binds worker agents to contracts specifying their permitted actions. We show how these contracts can be enforced and specified by assigning distinct responsibilities to monitor agents and judges, and asymmetric computational resources to monitors and workers. Our framework allows us to empirically measure statistical safety guarantees. The framework applies to a wide range of settings, including code security and escape-the-box scenarios.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Enric Boix-Adsera. 2026-09-25. Contract monitoring: governing AI via separation of powers. https://arxiv.org/abs/2609.32061
Cite the original work for its findings. Save a collection to share your selection of sources.