Record class: Previous-protocol construction record. Presentation revision 3: terminology and claim-boundary corrections applied per review; revisions 1 and 2 preserved byte-exact in this archive; underlying record bytes unchanged. Record class detail: previous-protocol construction record. Built under the pre-v1.3 machinery; the evaluator it constructed cannot verify the run that built it, exactly as this page already states. Historical evidence label: server_observed_proxy; current interpretation: engine capture over a worker-authored proxy signal, effective class proxy when the required negative probe is bound and valid.
Read this as
Atlas North Institute · field record No. 3 · July 14, 2026

The referee, born on the record.

In 63 minutes, a governed run built gad-evaluate: the separately executable evaluator that verifies Governed AI Development records on a stranger's machine. Fifteen nodes, zero retries, zero halts, 61 chained entries. The run that ended the era of "trust our tooling" was itself governed by the tooling it retires. Every verification, class label, and warning is preserved and inspectable on this page.

What happened

On the night of July 13, the reference machinery of Governed AI Development executed a frozen fifteen-node plan and built the discipline's missing piece: a small, separately executable evaluator that answers three questions about any governed record, on any computer, with no access to the producer's software or secrets. Is the record intact and internally lawful? How did the run actually end? Was the contract satisfied? It signs its answers. The same night, on a second machine running a different operating system, a cold rebuild passed all seventeen exit checks, and the evaluator rejected nine deliberately broken records, each for the correct stated reason, caught a single flipped byte, and refused to emit an unsigned attestation, because, in its own words, an unsigned attestation is not an attestation.

15 / 15
nodes verified done
0
retries · 0 halts · 0 waivers
61
chained entries, seq 0 through 60
63 min
initialize to final verification
23
engine-executed checks, every one negative-probed
17 / 17
phase-gate checks, passed twice: second time cold, on a second machine and OS
In plain terms: this run's history is hash-chained end to end and verified at export, which is tamper evidence within the engagement trust model, stated exactly as that: no independent time anchor is claimed for this record. Every claim of finished work was checked by machinery outside the worker's session, and every checker first demonstrated a failure under its planted-defect probe. This record does not prove the code is semantically correct against the specification (adversarial conformance vectors and a second evaluator are chartered for that), does not prove the efficiency estimate below (that is labeled commentary), and cannot be verified by the evaluator it built: the referee does not bless its own birth certificate; Conformance Trial 0 judges these books. A record that claimed more would be worth less.

Why a clean record is believable here

A vendor showing you a flawless run has shown you nothing: flawless is what fabrication looks like too. This record's clean report borrows its credibility from its siblings. The honesty box above never disappears from these records, and in record No. 1 it reported a mid-run executor crash; in No. 2, the failures being remediated. Same machinery, same box, same refusal to hide. When that machinery reports a clean run, the report means something, because you have seen what it says when things go wrong.

And this plan was not trusted on sight. Before launch, the engine's own strict validator rejected the first draft: the genesis node did not prove its baseline commit, and every checker carried a stronger evidence label than an executor-authored tool can honestly hold. The plan that ran was the corrected one, every checker demoted to its true class and required to demonstrate, against a committed planted defect, that it can fail. Fifteen green nodes, each one green only after its checker demonstrated a failure under its planted-defect probe. see the run's first warning, honestly recorded →

On efficiency, labeled as commentary

What follows is analysis, not chained evidence: the record proves the run's duration and interaction pattern; it cannot prove a counterfactual that never executed. The governed run cost one launch and fifteen brief supervised checkpoints, because the frozen plan carried the entire specification and the executor never had to be told anything twice. The same scope executed ungoverned, estimated from this operator's own prior build rhythm, runs seventy to one hundred ten authored prompts across several days, most of them specification retyping and correction cycles. The largest item in that estimate is exactly what this record shows never happened: zero retries means the correction loop, which is most of what ungoverned prompting is, never ran once. Roughly a four-to-six-times specification efficiency, and the ungoverned version produces this record at no price: never.

Identity: what ran, exactly

Frozen plan (authored identity)0683ba78ad295e08fa6077d585951044100606ebf8360ab8193ef208258920b0
Master spec hash71a0c23460ea2da1faef5c9dcc5ee89c1dec6477ecc8f2061425ef1f809eed4a
Genesis hash (chain root)dc9e7bddb9e3bfa191dd16a526ee0e8f807bfbd95bb1c8d2e8633eaede6eb852
Engine · cockpitAtlas Orchestrator 1.5.0 · Waypoint 1.1.0
Nodes · retry cap · mode15 · 3 · supervised, scope enforced, server verification default
Repository at startnone: genesis initialized the repository; the one pre-seeded entry (docs/) was warned about on the chain, entry 1

The fifteen nodes, timed from the chain

Durations are claim-to-verification, computed from the entries below. Every check was run by the engine against executor-authored tooling, so every row carries the honest class for that arrangement, server_observed_proxy, and every checker holds a negative probe: a committed planted defect it must detect, or its passes do not count.

#nodetitledurationchecksevidence class
01n01-scaffoldGenesis: repo scaffold, script surface, standing rules8m 20s2server_observed_proxy
02n02-register-mirrorRegister verification: seeded REG-5 closure and REG-6 through REG-91m 21s1server_observed_proxy
03n03-schema-signed-witnessSchemas: signed-object and witness-entry10m 28s2server_observed_proxy
04n04-schema-planSchema: plan (P)1m 05s2server_observed_proxy
05n05-schema-bundle-envelopeSchemas: core-bundle and envelope (REG-7, REG-8 form)1m 24s2server_observed_proxy
06n06-schema-operator-policy-attestationSchemas: operator, policy, attestation1m 49s2server_observed_proxy
07n07-canonicalizationCanonicalization: RFC 8785 base, domain tags, vectors2m 57s1server_observed_proxy
08n08-transition-tableTransition table for L(C) and invalid traces (REG-1, REG-3, RUL-2/3/4)5m 31s2server_observed_proxy
09n09-registriesPredicate and evidence-strength registries1m 53s1server_observed_proxy
10n10-default-policyDefault GAD-4 trust policy1m 04s1server_observed_proxy
11n11-fixtures-validValid fixture envelopes: COMPLETED, HALTED, INCOMPLETE3m 14s2server_observed_proxy
12n12-fixtures-invalidInvalid fixtures: one per R-condition2m 20s1server_observed_proxy
13n13-evaluator-coregad-evaluate core: R1 through R9, OUTCOME, CONTRACT_SATISFIED5m 38s2server_observed_proxy
14n14-evaluator-cli-qCLI and attestation Q2m 24s1server_observed_proxy
15n15-phase-gatePhase 1 exit gate: the full acceptance sweep1m 27s1server_observed_proxy

The boundary this record refuses to blur

The referee cannot bless its own birth certificate. This run's witness is a record in the previous protocol: a keyed, HMAC-linked chain, verified end to end at export, tamper-evident within the engagement trust model, and carrying no independent time anchor. It is not the public, signature-authenticated witness the new specification defines, and gad-evaluate cannot verify the record of the run that built it. The old machinery kept honest books on the construction of its successor; judging those books against the new specification, gaps published either way, is the chartered job of Conformance Trial 0. No artifact in this discipline is grandfathered into the standard it precedes, including this one.

What this record proves, and what it does not

It proves the plan was frozen and hash-addressed before work began; that every completion claim was checked by machinery outside the worker's session with its evidence class stated per node; that every checker demonstrated the capacity to fail before its passes counted; that the history is chained end to end under the engagement trust model; and that the resulting code passed its full exit gate twice, the second time cold, on a machine its builder never touched. It does not prove semantic correctness against the specification, the efficiency estimate above, or this record's own conformance to the specification its contents implement. Those verdicts belong to the adversarial vectors, the second evaluator, and Trial 0, all chartered, none skipped.

The complete chain, entry by entry

All 61 entries, seq 0 through 60, exactly as written at the moment. Click any entry for its raw form and its link to the next.

Verification panel (presentation rev 3)
Presentation: field-record-3-referee.html · rev 3
Underlying run project: project-2-1783990995972 (digest-matched in the supplied runs corpus)
Matched digest: dc9e7bddb9e3bfa191dd16a526ee0e8f807bfbd95bb1c8d2e8633eaede6eb852 located in the project logs
Protocol: previous-protocol construction; no v1.3 conformance claimed for this run
Anchor: none claimed (stated on the page)