The core thesis
Chaos vs. order.
The central conflict of this paper is between the chaos of AI, the actor, and the order of a predictable guard, the governor.
- Before: chance
- A fuzzy, shifting cloud. It changes shape when you touch it. A black box where inputs go in and unpredictable outputs come out.
- After: rules
- A rigid, crystal-clear structure. A glass box where every piece of data can be traced.
The physics of failure
Safety that depends on server load is not safety. It is luck.
Today’s AI safety tools are flawed at the root. These are not bugs to fix. They are physics to replace.
- The drift problem
- On a GPU, the math changes with server load. 1 + 1 does not always equal 2. Safety checks drift by up to 21.4% under real production load.
- Compound failure
- A 99% safe model is fine for chat. It is fatal for an agent doing 50 tasks. Each step leaks safety. By step 50, you are at coin-flip odds.
- The rare-fact problem
- AI must make things up about rare facts. It is a statistical law. Your company’s data is rare, so the AI will make things up about it.
Field manuals
Four teams. One standard.
Each appendix turns the physics into the language of one role, so every team can see its risk and its way out.
For general counsel
- The liability shift: from publisher to operator
- Why built-in safety fails the foreseeability test
- The glass box defense: chain of custody
- From hidden liability to insurable assets
- The end of the black box defense
For actuaries and risk officers
- The black box pricing crisis
- Why standard deviation fails for AI
- The Net Insurable Token framework
- Risk decay curves and reserve release
- AI’s steam boiler moment
For engineering leads and CISOs
- Floating-point math and safety drift
- Load-proof guard architecture
- Hot-swappable protection
- Training small guards from big ones
- Replaying any decision exactly
For auditors
- Why chance-based logs are a weakness
- Checking every decision, not a sample
- Replaying the past
- Mapping to SOC 2, ISO, SOX, and GDPR
- The control function exemption
The verdict
The standard has shifted.
- The risks are known.
- Hiding how AI decides is negligence.
- A predictable guard is the new baseline.
The ways AI fails are now known, so not knowing is no longer a defense. Running a black box when a glass box is available invites real damages. Chance-based guardrails are not enough.