Files
dotfiles/claude/memory/project-studium-engine.md
T

19 KiB
Raw Blame History

name, description, metadata
name description metadata
project-studium-engine Canonical Studium Engine workstream tracker — build state, governing instruments, live blocker, and the chronological log of substantive moves. Established 2026-08-07 at the MEMORY.md trim, filling the gap MEMORY.md had flagged as 'no tracker file yet'; state seeded verbatim from the MEMORY.md line it replaces.
node_type type originSessionId modified
memory project 033cfe63-c9d0-4fad-accf-c45de561f09a 2026-08-07T15:08:08.484Z

Studium Engine — canonical workstream tracker

Established 2026-08-07, at the steward-directed MEMORY.md trim. Until now the engine had no tracker file, so its build state lived inline in MEMORY.md (one 950-character line) and in per-session memories. That is two update surfaces and no canonical one — the drift shape recorded as skill-harvest proposal #183. This file is now the canonical surface; MEMORY.md carries only a pointer.

This file holds state. It does not hold the why or the law:


Current state (as of 2026-08-07 wake)

Build: Steps 0–7 built. Corpus CLEAN and gate-validated 13/13.

Governing instruments: V1 verify-quote + fidelity_equivalence@3 — ratified 2026-08-05, GOVERNING (REVIEWED-87 placed 2026-08-06).

⚠ @3 is under challenge. PENDING-111 + a full jurist package filed 2026-08-06 (studium-engine/docs/fidelity-3-literal-asterisk-JURIST-PACKAGE-2026-08-06.md, containment 13/13): _MARKUP_EMPHASIS = re.compile(r"[_*]") strips the escaped literal \*, erasing Alexander's confidence rating (two asterisks = a true invariant, one = progress, none = far from invariant; "Using this book", pp. 14–15). Measured 81 / 114 / 54 across the manifested corpus. The ruling sets the V-track course. Verified 2026-08-07: no ruling yet — zero occurrences of PENDING-111 in ~/dotfiles/REVIEWED.md. It does not gate N1.

Read side — the parse fix LANDED, 27b79ca. retrieve.py now accepts a sentence; the query was previously passed to MATCH ? where FTS5 parses it as a query expression, so ? and : were syntax errors. Result, answer-keyed against tests/chavruta_harness.py:

crashes 26 → 0
HIT 0/22 ← the number to beat
MISLOCATED 0
FALSE-POSITIVE 0 (across all 5 items where silence is the correct answer)

Both deciding buckets empty ⇒ the revert condition was not met. The engine now grounds nothing honestly — every question needs 13–19 terms to co-occur. Recorded explicitly as the number to beat so a later pass cannot mistake silence for progress.

Test floor (the read side's first): tests/test_retrieve.py (21 checks) + tests/chavruta_harness.py. test_conjunction_is_monotonic is the tripwire against every future answer-more change — it must stay green.

Live blocker: PENDING-97 — retrieval AND-s bare tokens and has no semantic layer. Now reachable and measurable for the first time (the 13–19 term conjunctions above).

⚡ The embedding arm already scores 22/22 recall@20, voice-scoped, on the identical 22 items where FTS scores 0/22 (corpus/measure-rerank-voicescoped.json, verified same id-set). Capability measured in June, never landed. recall@20 means the right passage is in the top 20 alongside nineteen others — the answer-more direction — so V2 is the gate that makes surfacing it safe. Landing embeddings first would be the make-the-demo-nicer move.

NEXT: N1 → V2.

  • N1 — the navigation-tree builder. Contract written: docs/spec/n0-navigation-tree-contract.md §1 (the tree) and §2 (the four primitives: list-children, open-node, expand-to-parent, load-whole-work), including that open-node refuses citable text for a non-citable node. Do not re-derive it. ⚠ N1 will NOT move 0/22 — that is N2, where the reasoner navigates; N1 lays the ground it walks on. ⚠ Build the tree from the READING INDEX (chamber-library/reading-indices/*.yaml — Alexander: all 253, re-found by name, ascending order verified, sha-bound), not by parsing headings: heading text hits three documented OCR defects (179 misnumbered 178 → a duplicate node; 187 lost its ##; 195 has no heading). The harness's heading regex is a working reference, not the input.
  • V2 — gold set + pre-registered thresholds. Follows N1 rather than PENDING-97 for the reason above.

OPEN THREADS — the stack as of 2026-08-07 (captured mid-session, before the TEI detour)

Written because the session went N1 → @3 → R0 → TEI-decision and each step opened threads. That accumulation is the thing the steward abhors; this block is the guard.

Owed to the steward, not startable by the executor

  1. REVIEWED-87 amendment — DRAFTED, NOT PLACED. studium-engine/docs/REVIEWED-87-amendment-DRAFT-2026-08-07.md §A is the block to place. ~/REVIEWED.md is [ESCALATE], steward's hand.
  2. Three measured findings for relay to the jurist (same draft, §B): the package's ≡ claim is false and its own table refutes it; condition (a)'s "verdicts that may have overclaimed" has an empty referent (the risk ran the other way — false refusals); Q1's grounds hold on the corpus side only. Q3 census: escaped emphasis in 3 of 13 sources, not "Alexander only".
  3. The TEI ruling itself — agreed in shape (MD+sidecar canonical, proxy trigger retired, deferral becomes a design window), not yet recorded. Blocked on the mechanism below.

Executor-startable, in rough priority 4. N2 — the agentic navigation loop + embedding entry-finder fallback; discovery emits hypothesis-labelled output. This is the next N-track station and the one that can move 0/22. 5. V2 — gold set + pre-registered thresholds; the gate that makes surfacing the embedding arm safe (22/22 recall@20 measured in June, never landed). 6. Alexander front_matter re-anchor — all five anchors stale (+20/+20/+22/+26/+32). Now mechanical: python3 -m engine.reading_index recover proposes; nothing is applied. Unblocks → 7. The "Using this book" FIX — a composite-span misclassification, not a D-4 policy change. D-4's own text: "citability is a function of convocation, not an intrinsic byte property", with citable:false reserved for matter that is nobody's quotable voice. Alexander's framing essays are his own words (the sidecar's own note says so). Partition the span; don't flip a flag. 8. R0 emit / migration — reading_index emit <id> renders native R0; nothing has been written to chamber-library (D-3). Steward review before any write. 9. The collision census — count characters ambiguous between markdown syntax and authorial content, per source. The number the TEI question will eventually turn on; deferred behind the mechanism by steward call 2026-08-07. 10. 56 regions unverified (Mauss 23 + after-the-reply 33) — editorial/synthetic titles, so name-landing cannot test them. A content probe at the declared boundary is owed.

Not this thread: the Seb package · the L2 design note · PENDING-109 census + PENDING-104 brief, both still needing dates, not "later."

Chronological log

Append substantive moves here at /wrap-up — not only to "Current state" above. A tracker with two update surfaces drifts between them (skill-harvest #183).

2026-08-08 — disposition (vi) RULED (REVIEWED-97), and voice: is the convocation key (824139d)

Ruled and placed. REVIEWED-97 (PENDING-113) disposes kind (vi): four slots, each one job — identity voice: · relation quoted_by: · category (traditional / non-individual-origin, held out of the key) · per-source prose note. Reasoning of record: docs/voice-non-individual-origin-2026-08-08.md. Substrate findings kept in docs/vi-disposition-DRAFT-2026-08-08.md (superseded in part).

The correction that mattered. The jurist's first structure put the category pair in voice:. voice: is what retrieve.py:216 filters on, so that would have made the Havámál and the Mahābhārata one convocable speaker. Measured before asserting: glidden spans 5 sources, weil 2 — correct, one person each. Aggregation principle, jurist's phrasing: individual-author voices aggregate at the person because a person is real and singular; traditional matter has no such person, so identity lives at the work.

Refuted by measurement: option C (omit voice, let the relation carry it) — omission resolves to the host via sec.get("voice", catalog.get("voice")). Control 4/4.

Corrections found in filed records. Surah LXIV (at-Taghābun), not CXIV — the sidecar title was wrong and had reached REVIEWED-96, PENDING-113 and memory; it voided the jurist's worked provenance note, which was built on the "Say" formula absent from the quoted passage. quotation-poet-jurist reclassified from "unnamed individual" to traditional matter by one footnote. Naming evidence for six of nine blocks sits inside the fenced apparatus, engine-unreachable.

Fourth defect in 118f411. L850 ("M. Cahen nous signale aussi la strophe 145 :") is Mauss's own prose, fenced inside the Havámál block — found only because the steward corrected a framing about language. And the body → body-01..13 split left test_navigate.py red for a full day (stale hardcoded node id; the containment invariant itself verified intact). Re-run the fleet after any sidecar/corpus change — proposed as a pre-commit hook extension.

Filed: PENDING-114 (scripture quoted unmarked in Harrison — Mark 16:7–8 served as voice: harrison) → steward AUTHORIZED (b)+(c), REVIEWED-98 drafted. PENDING-115 (two step-3 blockers: ROLE_CLASS has no quotation key, so such sections are searchable while classified outside the declared scope; and the warrant scope is computed per source, so a sub-source voice overclaims — 191 chunks of Mauss would warrant a Havámál silence).

NEXT (steward-agreed order): the quotation-in × translation-of jurist package — all twelve blocks are translated matter and role is single-valued, so (vi) is decided but inapplicable until it is ruled. Then PENDING-114 (b), validation phase first.

2026-08-07 night — V2 preconditions worked; the corpus answered with a bigger question

P4 censused 14 sources (parsed, not grepped — grep -c '^ - id:' gives 22 and reproduces the design addendum's "21" error; 8 numeric ids live under corpus_findings). P1 verified clean. P5 → span-binding, 73dfef3: P5's content_located: 6 is byte-locatability; Tier-2 gold needs a bound span — 15 of 17 bind (11 distinct spans), each corroborated twice (verify_quote locates + line sits at stated − 1). corpus/mauss-phase2-spans.yaml. ⚠ The first pass bound 0 of 11 on a criterion inherited from Tier-1; the control (6 known answers, 6/6) is what corrected it. P7, 2a45c26: fr tagged 1 A : 9 B — inverting P7's own prediction — and §6.3's French method produces B by construction, so ~8 stratum-A pairs must be authored and nothing schedules them. en is NOT taggable (no spans, only division anchors; EN divisions run ~29× the FR spans). P6 still 0 bytes.

Then the corpus-wide finding. Mauss's body had no role: quotation region, so §7.4(i)'s provenance join did not exist. Fixed G&G (57090ab) and Mauss (118f411) — and the jurist ruled the fix the wrong instrument. REVIEWED-96: D-4's convocation governs; citable: false is for matter that is nobody's voice; §4.1 case 2 keeps a non-host voice quotable under a relation. Q3's rule: a quoted span grounds the host's reproduction, never the quoted author's authorship. Q2 deferred — the chunk invariant is derived; carry provenance at the span layer instead. Q4 binds borrowed authority only; (vi) anonymous/traditional matter GATES the Mauss remediation. Q5 upgraded to BLOCKING. PENDING-113 lodged with the remediation order.

⚠ 118f411 was mislabelled [FIX] — corpus-wide policy under a scoped label, overriding a ratified default on the authority of a document whose front matter forbids acting on it pre-review, and destroying the only human-verified §7.4(i) negative (the Havámál, identified as a gold-negative candidate six hours earlier). Commits STAND pending the ruled order: (vi) → re-tag → only then citable: true.

2026-08-07 — PENDING-111 RULED, @3 corrected in place (4be9378)

Jurist: Q1 AUTHORIZE (narrow to unescaped delimiters), Q2 correction-in-place not an @4 bump (mechanism defect against standing doctrine), Q3 census follows non-gating, Q4 steward's. FIDELITY_VERSION stays @3; @4 reserved. Implemented as one left-to-right scan, not lookbehind-plus-unescape (that form mis-reads \\*). Falsifier incl. the jurist's nested case; suite 33, fleet 153/153; gold 6/17 before and after — no verdict moved either way.

⚠ The REVIEWED-87 amendment is DRAFTED, NOT PLACED — docs/REVIEWED-87-amendment-DRAFT-2026-08-07.md. ~/REVIEWED.md is [ESCALATE], steward's hand; a jurist sign-off does not authorize a REVIEWED write.

⚡ Three measured findings contradict the package's own premises (draft §B, for relay): (1) the COMPOST\* ≡ COMPOST\*\* ≡ COMPOST claim is FALSE — old @3 gave three distinct strings and the package's own Part I table printed the refutation; (2) so condition (a)'s "verdicts that may have overclaimed" has an empty referent — the real failure was false refusals, the opposite risk direction; (3) Q1's grounds hold on the corpus side only — a human's bare COMPOST** still normalizes to COMPOST. Census: escaped emphasis in 3 of 13 sources, not "Alexander only" (Alexander 293 · Musil 16 · Arendt 1); ratings 83/114/56 over all 253.

Also opened by this ruling (Q4): D-4's primary text says "citability is a function of convocation, not an intrinsic byte property" and reserves citable: false for matter that is nobody's quotable voice. Alexander's framing essays are his own words — the sidecar's own note says so — so fencing them is a misclassification against D-4, not D-4 working. The frontmatter section is a composite span (YAML+TOC furniture + four Alexander essays) never partitioned. ⚠ Blocked: its partition points live in the reading index's front_matter block, and all five of those anchors are stale (offsets +20/+20/+22/+26/+32; three land on blank lines) while the patterns block in the same file is exact 253/253 — the 2026-06-12 re-anchor was partial and the manifest reports one status, RE-ANCHORED-BOUND, for a file bound in one region and stale in another.

2026-08-07 — N1 BUILT

engine/navigate.py + tests/test_navigate.py (32 checks) + docs/spec/n1-navigation-tree-note.md. Tree: 9 works · 13 expressions · 359 divisions · 5,685 spans; four N0 primitives + a browsable CLI. Fleet 142/142, retrieval untouched. 0/22 unchanged, as expected.

The derivation that mattered: every sidecar declares exactly ONE served section (Alexander's body = 10,832 lines, "Patterns 1-253"), so the sidecar is the envelope and the reading index is the articulation. Adapters declared per index filename; unknown shape → UndeclaredIndexShape.

Three defects, all caught by measurement, none by reading the code: 455 spans orphaned in gaps between declared divisions → 314 more in the no-sidecar source → citability reimplemented and diverged from chunker.section_is_served (latent). In the first two load_whole_work would have silently under-returned. test_every_drawer_is_reachable is the invariant; red-witnessed at 1,970.

Owed / open: ⚠ 6 of 8 indexed sources cannot be name-tested (editorial or synthetic division titles) — reported as an OPEN GAP, a content probe at the declared boundary is owed. ⚠ Mauss's reading index is not sha-bound to the manifested file yet the manifest declares VERIFIED-BOUND. Two constraint-candidates for the steward: (1) Alexander's confidence rating is in the source but not declared by the reading index, so the tree cannot carry it without a chamber-side change — converging with PENDING-111, where @3 erases the same semantic; (2) a manifest role: reading-source with no sidecar is chunked as role: text, citable: true — citable by absence.

2026-08-07 — tracker established

Created at the steward-directed MEMORY.md trim. State above seeded from the MEMORY.md line it replaces plus session-2026-08-06-evening-the-asterisk-that-carried-meaning.md; nothing dropped. Substrate-verified at creation: PENDING-111 has no ruling; engine/ contains no navigation module and the four N0 primitives appear only in docs (N1 genuinely unbuilt); the N0 contract and the Alexander reading index both exist at the paths named.

2026-08-06 evening — the parse fix landed; the asterisk that carried meaning

27b79ca — 26 crashes → 0, HIT 0/22, MISLOCATED 0, FALSE-POSITIVE 0. Semantics measured unchanged (old path vs new over all 27 items, not one disagreement); whole-query phrasing rejected because it answers less (1 where the conjunction returns 4). Silence path reached for the first time by long questions, so a silence now names its term count and states it cannot distinguish "the voice is silent" from "the terms did not co-occur"; an unsearchable query is marked ✗ NOT SEARCHED, never coverage-warranted. Then the steward's printed A Pattern Language exposed the @3 asterisk defect → PENDING-111 + jurist package. Manifest fixes 41527be, cbd6a9b (a disarmed absent-sidecar tripwire; a usage fact in a bibliographic field; a defect record cited at corpus_findings[1], a key existing in zero files fleet-wide). Full account: session-2026-08-06-evening-the-asterisk-that-carried-meaning.

Before 2026-08-06

Not reconstructed here. Per-session memories carry it (session-*.md, 2026-07-05 onward for the Stage-1 rebuild), together with studium-engine/docs/stage-1-rebuild-plan-2026-07-05.md and docs/tool-evolution-log.md. Backfill on demand rather than speculatively.