
EB
Enric Boix-Adsera
· 1 min read
ResearcharXiv cs.AI
Contract monitoring: governing AI via separation of powers
arXiv:2609.32061v1 Announce Type: new
Abstract: We propose an AI safety framework that binds worker agents to contracts specifying their permitted actions. We show how these contracts can be enforced and specified by assigning distinct responsibilities to monitor agents and judges, and asymmetric computational resources to monitors and workers. Our framework allows us to empirically measure statistical safety guarantees. The framework applies to a wide range of settings, including code security and escape-the-box scenarios.
Original source
This story was published by arXiv cs.AI and written by Enric Boix-Adsera. SyncAI.news shows a preview; the complete article is on the publisher's site.
Read the full story on arxiv.org


