Paper: 2607.07695 Authors: Yujiao Chen Categories: cs.AI, cs.GT, cs.MA
The Gap
Multi-agent AI safety research has been laser-focused on two layers: model capabilities (can the model refuse harmful requests? does RLHF actually align it?) and architecture design (how do agents coordinate? what communication protocols work?). But there’s a third layer that’s been almost entirely untested: the deployment rules — the institutional constraints that govern what happens when agents