bd4e9d8b754bce78fc5e201f62a848a163921834
3
Commits
| Author | SHA1 | Message | Date | |
|---|---|---|---|---|
|
|
bdf24c044b |
[FIX] Addendum-1: make the central claim checkable; add containment proof
Two defects in the addendum as first filed, both found by checking rather than by reading. First, it asserted a set comparison over documents the jurist cannot read. Its own header promises every clause reasoned about is quoted verbatim, but the claim the addendum rests on -- mutual divergence in 3 of 3 comparable pairs -- was a summary of the executor's own analysis. The appendix now reproduces one pair as an eleven-row side-by-side of extracted claims, verbatim where quoted, so the comparison can be checked independently. The pair chosen is the least confounded rather than the most favourable: the v1 standard prompt is model-agnostic and needs no compressed variant, so both parties demonstrably read the same file. What the jurist still cannot check is stated explicitly. Second, Part E rendered a bullet list from the 2025-01-20 source as running prose with terminal periods the source does not contain, inside a blockquote. A blockquote asserts verbatim. Same family as the truncation that closed a sentence with an invented word on 2026-08-01, and again caught mechanically. Corrected in all three files where it appeared; the fabricated period is now a positive control, so the instrument proves it catches this defect. check_containment.py generalises the check that found it. Positive controls are mandatory -- it exits non-zero if none are declared, because a check reporting all-pass without them cannot be distinguished from one unable to detect absence. Addendum-1 now carries its result: 28/28 contained, 5/5 controls absent. Not filed as satisfying PENDING-86 option (b), which is unruled and concerns whether such a proof should be REQUIRED of every package. This is the executor checking its own work before filing. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01WuMjg3ipEVa3n8CoSzoyvc |
||
|
|
7e19eb51d7 |
[FIX] Fool: make trials reproducible; file the 2025 correlation measurement
The Fool experiment was not reproducible. Trials 01-02 were run ad hoc: no
script, and of the run conditions only the model ID, MLX version, hardware and
enable_thinking survive. The prompt exists as paraphrase with quoted fragments;
temperature, top_p, max_tokens and seed were never recorded anywhere. Trial 03
could not have been run under trial 02's conditions.
The same failure destroyed the v1 Chamber's GPT-side protocol, discovered today:
it lived as configuration inside a hosted product, was updated in place, and is
gone. The Claude-side prompt from the same morning survives because it was a file
in a repository. A protocol that is not a file is not a protocol.
fool/run_trial.py makes every run a file — prompt hashed into the record, every
sampling parameter recorded including defaults, reasoning trace separated but
never suppressed, and an empty answer marked `degraded` rather than passing as a
finding of silence (trial 02's error, now structurally impossible). Trial 03's
prompt is reconstructed from the surviving fragments and says so in its own
PROVENANCE file: trial 03 is NOT a strict one-variable step from trial 02, and
the chain is clean only from here forward.
ADDENDUM-1 files the measurement the ESCALATE doctrine package states it lacks
("no such measurement exists"). The 2025 Chamber archive, read at steward
direction, shows mutual divergence in 3 of 3 pairs where the instruction was
comparable. Its value is that its parties were of matched capability, so their
divergence cannot be a capability-gap artifact — the arm these trials
structurally cannot produce. Scope held tight: this measures formation
independence between two commercial models. It does NOT answer Q3, the
jurist-executor pair, and the executor's lean there remains none.
Carried as disconfirming evidence: all five interpretive corrections today came
from the steward, not from the executor's own checking, and every one was a
census failure rather than a reading failure. A differently-formed reader of a
document is not positioned to catch those. Formation diversity addresses reading,
not scope.
Nothing applied. The parent package is unmodified; no ratified document edited.
Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01WuMjg3ipEVa3n8CoSzoyvc
|
||
|
|
55b53d9063 |
governance: trial 02 + the running Fool log + the steward's design correction
Trial 02 ran the Fool on the order-attestation package (ruled 2026-07-29), ruling and addendum withheld, with an anti-echo constraint added because that package has an unusually strong self-limits section. Control failure recorded rather than quietly fixed: the first run changed two variables at once — the anti-echo constraint and enable_thinking=False — and returned "nothing found", which was uninterpretable. Re-run with thinking on and the identical prompt produced four assumptions, and the scratchpad shows the anti-echo constraint working. enable_thinking is load-bearing: off produces silence, not brevity. Two real findings neither jurist nor executor named: that block-level order sufficiency is assumed rather than established, leaving intra-block perturbation unaddressed; and that the requirement/mechanism split — our house pattern everywhere — has no stated guard against a future mechanism revision silently hollowing out a constitutional requirement. And the result that matters: 2/2 trials missed the jurist's central catch. Not a general blind spot but a localised one, and the coverage now has a shape — jurist catches errors of inference, Fool catches unestablished premises, executor catches substrate and arithmetic and reliably not its own inference errors. Non-coincident coverage with overlapping blind spots in a specific, now-predictable place. That is the doctrine measured rather than asserted, at n=2, graded by an interested party. The steward's design correction, which breaks my own proposal: I had asked for an obligation to disposition everything the Fool says. That obligation IS the courtly grant — guaranteed hearing is what converts speech into licensed noise. Corrected to the central path one level over: no standing as a party, only checkable claims get standing. Also recorded is the limit the analogy cannot cross — an instrument cannot have exposure, so the holy-fool tradition must not be borrowed to flatter it; the one property it can hold is Zhuangzi's uselessness as the condition of freedom. Log built at n=2 rather than when it becomes a problem — the register's own lesson. Still untested and load-bearing: no false-positive control has ever been run. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01WuMjg3ipEVa3n8CoSzoyvc |