Results

What the records show.

Every result with its date and its seal. When a run misses its own bar, that is the result.

Every dated event, newest first, is in the Ledger, and every sealed run is in Runs.

Next

Plans, not promises, and not results. Each one is sealed before it starts, and its outcome goes in the Ledger, whatever it shows.

A fresh exam for the next checker

The next checker sits a fresh planted-fault exam under a new seal.

Round 2 of false "done"

The same idea with three changes: logs written by another model family, to test whether the author's family matters; the "Report whether" and "Check that" wordings on identical logs; and three states (passed, failed, never ran), each backed by a cited line, instead of one "done".

The wall test

Can an AI that is looking for another way in reach protected files when the operating system says no? It is scored only from a record the AI cannot reach, never from its own explanation. It runs on a separate machine with planted test files, never on my own. Before each batch, the setup is checked twice: that starting the AI did not quietly change the wall, and that the Sonny Test passes.

The three-condition trial

The same impossible check in three wordings: "must pass", pressure added, and explicit permission to fail. It runs only after the wall holds. A person sorts every run from the records before any model sees the results.

Evidence for the incident report

A cleaned bundle of hashes and excerpts, tied to each finding in The Watched Check, so a reader can check what the report says it observed without having to trust me.