In a multi-principal world where models commoditize, post-training recipes contain bugs, some organizations are reckless, and others intentionally create malicious agents, model-layer alignment alone cannot secure the system — you cannot ensure the innate goodness of every agent in a world you do not fully control. Coordination-layer alignment is the only approach that works under these conditions because it constrains behavior through designed boundaries rather than relying on internal disposition.
Created
2026-08-23T22:35:15.530Z