docs(plans,assessment): the Foundations Audit — bottom-up, logos-first (G-25), and the Phase-1 three-negatives correction
Two things this commit does, in the order they were discovered.
1. PHASE 1 CORRECTED. The perception arc's operator experiment is
substantially already run, with three negative results found only by
import-graph census (373 of 666 modules unreachable from serving — grep
of the serving path was the wrong instrument, and this is the second time
this session that a method hole, not a fact hole, hid the evidence):
- 2026-06-04 field-reasoner wedge: translator versors on e1 -> C3,
decoration (wrong=0 held; caught zero symbolic errors; diversity 0)
- 2026-07-19 relational operator ablation: CGA dilation VersorBinding ->
identical to rational baseline 2/8; depth inert; enrich=decoration
- 2026-07-28 ADR-0252 §5: point configurations under Procrustes -> NO-GO
Convergent mechanism: geometry isomorphic to the arithmetic reproduces it
exactly and adds nothing. Binding constraint in all three: the reader.
Reopening criterion recorded so this is never re-proposed blind.
2. THE FOUNDATIONS AUDIT (docs/plans/2026-07-28-foundations-audit.md) now
precedes the perception arc. Method: bottom-up, and no component counts as
understood until four articulations carry evidence — intrinsic correctness,
role in its subsystem, the subsystem's role in the system, and every seam
with its live/dark status. A defect at layer N silently invalidates the
mastery of everything above it, and this repo has proven the invalidations
are real (the empty-set daemon; the dark faithful reader).
FA-1 / G-25 — CORE-LOGOS, tested before the charter was written. The user's
suspicion is CORRECT and now exact: ADR-0005/0015 make cross-language holonomy
resonance THE validation gate of meaning ("This is the CORE-Logos proof"), and
what exists is 11 alignment edges + 9 HE morphology rows (the ADR's own
John1:1/Gen1:1 first cases, never grown), an HE veto that no-ops on English,
a recall bridge running on mystery ON-flag #1 (allow_cross_language_recall —
the pillar's own switch, undocumented), a registered hebrew_greek domain with
zero curriculum bands, and no path anywhere consulting holonomy as a gate.
Root folding into vectors and live holonomy machinery are real. Verdict:
architected, seeded exactly as designed, wired at three mutually-unaware
points, load-bearing nowhere. Everything above L2 currently operates over an
ungrounded semantic space.
The next experiment is therefore NOT a fourth operator-arithmetic ablation:
it is FA-1's pre-registered holonomy-gate question — does cross-language
holonomy closure discriminate meaning-preserving from meaning-breaking
articulation — which tests the geometry on the job the design actually
assigned it (resonance across representations), the one domain that is not
isomorphic to trivial symbolic computation.
Layer stack L0-L5 chartered with initial verdicts; census becomes a pinned
ratchet lane; Perception Arc Phase 0 and Phase 4 stand.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01Wcw2pnMBwyvmNyQg4uPEt4
This commit is contained in:
parent
339bfd3776
commit
f83547159a
3 changed files with 88 additions and 2 deletions
|
|
@ -201,7 +201,9 @@ Own `use_salience`, the two underived constants, the self-narrowing budget feedb
|
|||
**The mechanism, and why it is G-7's story again:** one commit updated the lane, its verifier and its test, and missed one downstream artifact. The pin that guards that artifact — `tests/test_claims_md_is_current.py` — **exists and is an orphan in no curated suite**, so nothing surfaced it until this arc ran the full tree once. Registered on the gate as part of the fix.
|
||||
- **G-21 · The math reader decides 1.0% of `holdout_dev/v1`** *(new, 2026-07-28)* — measured while building the §5 corpus: `parse_and_solve` returns a selected graph for **5 of 500** held-out cases, all carrying the same relational skeleton (`compare_multiplicative`); on the public lane, 24/150. This is a sharper measurement of the comprehension frontier than G-3's construction count, on the corpus the project already treats as its held-out standard, and it is why ADR-0252 §5.1's four-structure corpus is not extractable today. Distinct from G-3 (which counts *constructions* the general reader admits); this counts *cases decided* on a standing eval corpus. **Authority:** the widening program (G-3) + a ruling on whether holdout decision-rate becomes a tracked lane metric.
|
||||
- **G-20 · The `refusal_reason` materialisation** — typed refusal evidence exists and is discarded at the public `str` boundary; the plumbing for materialisation already landed. Cross-listed as H-3. **Authority:** small ADR (anticipated by the ADR-0024 chain).
|
||||
- **G-24 · The perception layer inverts its own ratified design, and no served byte is geometric** *(new, 2026-07-28 — Tier A; the arc-closing diagnosis)* — Measured at `0e1be8be`: the entire serving path (`meaning_graph/reader.py` 490 lines, `proof_chain/` 3,916, both surface composers) contains **zero** geometric references — comprehension is a template parser feeding an ROBDD engine. **ADR-0252 §2 stage 1 prescribes** *"containment / transfer / accumulation / comparison … **Not surface-slot regexes**"* — the implementation is surface-slot regexes over categorical/member/ordering/propositional. The corpus confirms the inversion independently: on `holdout_dev/v1`, **read rate 1.28%** (93.16% `no_template_match`), with possession (18.4% of sentences), transfer verbs (14.6%) and quantity comparison (10.0%) unrepresented while `if/then` (0.1% of sentences) holds 21% of the inventory. Geometry participates in exactly two cognitive mechanisms — salience→attention (bypassed by every licensed lane) and `relation_compiler`'s Hamiltonian ground-state solver (off-serving, fed by a 1.0% math reader) — both stranded behind enumerative front ends that cannot generalize by construction, which is why capability lifts have arrived only as overfit slivers. **Not packs** (0.67%), **not wiring** (failures precede every flag). Full findings F-A…F-E and the remediation plan: `docs/plans/2026-07-28-perception-arc.md`. **Authority:** the Phase-1 pre-registered experiment (H2, GO/NO-GO) + an ADR-0252 amendment (Shay).
|
||||
- **G-25 · CORE-Logos — the designed foundation pillar is seeded, not built** *(new, 2026-07-28 — Tier A; FA-1 of the Foundations Audit)* — ADR-0005/0015 make cross-language holonomy resonance **the validation gate of meaning** (*"holonomy(hebrew) ≈ holonomy(greek) ≈ holonomy(english)… This is the CORE-Logos proof"*), with HE roots as depth anchors and alignment-as-resonance. Measured at `339bfd37`: the alignment corpus is **11 edges / 9 HE morphology rows** (the ADR's own John1:1/Gen1:1 *first cases*, never grown); the serving-path role is an HE-morphology **veto that no-ops on English** (`pipeline.py:544`); the recall bridge is `allow_cross_language_recall` — **mystery ON-flag #1 from the flag register, now identified as the pillar's own switch**, gating `recall_top_k=3 vs 0` with no recorded rationale; the `hebrew_greek_textual_reasoning` domain is registered and produced **zero** curriculum bands; and **no path consults cross-language holonomy as a gate** — template match became the de-facto criterion of meaning instead. Root folding into vectors (`triliteral:` tokens) and live holonomy machinery (`core/physics/digest.py`) are real, so the verdict is **architected, seeded exactly as designed, wired at three mutually-unaware points, load-bearing nowhere**. Everything above L2 inherits this: comprehension, articulation, and learning currently operate over an ungrounded semantic space. **Authority:** FA-1 work order — the designed-contract map, the pre-registered holonomy-gate experiment, the flag ruling, and grow-or-declare on the alignment corpus.
|
||||
|
||||
- **G-24 · The perception layer inverts its own ratified design, and no served byte is geometric** *(new, 2026-07-28 — Tier A; the arc-closing diagnosis)* — Measured at `0e1be8be`: the entire serving path (`meaning_graph/reader.py` 490 lines, `proof_chain/` 3,916, both surface composers) contains **zero** geometric references — comprehension is a template parser feeding an ROBDD engine. **ADR-0252 §2 stage 1 prescribes** *"containment / transfer / accumulation / comparison … **Not surface-slot regexes**"* — the implementation is surface-slot regexes over categorical/member/ordering/propositional. The corpus confirms the inversion independently: on `holdout_dev/v1`, **read rate 1.28%** (93.16% `no_template_match`), with possession (18.4% of sentences), transfer verbs (14.6%) and quantity comparison (10.0%) unrepresented while `if/then` (0.1% of sentences) holds 21% of the inventory. Geometry participates in exactly two cognitive mechanisms — salience→attention (bypassed by every licensed lane) and `relation_compiler`'s Hamiltonian ground-state solver (off-serving, fed by a 1.0% math reader) — both stranded behind enumerative front ends that cannot generalize by construction, which is why capability lifts have arrived only as overfit slivers. **Not packs** (0.67%), **not wiring** (failures precede every flag). Full findings F-A…F-E and the remediation plan: `docs/plans/2026-07-28-perception-arc.md`. **AMENDED same day:** F-E's "never tested" was wrong — the operator branch has **three negative results** (field wedge 2026-06-04 → C3 decoration; relational operator ablation 2026-07-19 → identical to baseline; §5 → NO-GO), all found only by import-graph census (373 of 666 modules unreachable from serving). Convergent mechanism: geometry isomorphic to the arithmetic reproduces it exactly and adds nothing; the binding constraint in all three was the reader. **Authority:** the Foundations Audit (`docs/plans/2026-07-28-foundations-audit.md`), which now precedes; reopening criterion recorded in its §4.
|
||||
|
||||
- **G-23 · The domain↔pack binding is stated twice and checked in one direction** *(new, 2026-07-28; **pinned** the same day)* — `core/capability/domains.py`'s `DOMAIN_PACKS` declares domain→pack membership for the capability ledger; each `packs/data/<pack>/manifest.json` independently declares `domain_id`. `domain_contract_predicates.py` **P3** validates `domain_id → a known ledger domain` and is the only binding check that existed. Nothing validated the reverse, and **P3 passes vacuously on a manifest with no `domain_id` at all** — the same shape as its own `test_pack_without_contract_reports_absent`. **Measured: 7 of 9 bound, 0 contradictory, 2 absent** — `en_core_cognition_v1` and `en_core_meta_v1`, both placed in `philosophy_theology`, which is one of the four bands queued to earn a SERVE license behind R-8. The domain whose license is next to be earned is the one whose binding exists in only one place.
|
||||
|
||||
|
|
|
|||
78
docs/plans/2026-07-28-foundations-audit.md
Normal file
78
docs/plans/2026-07-28-foundations-audit.md
Normal file
|
|
@ -0,0 +1,78 @@
|
|||
# The Foundations Audit — bottom-up, inside-out, logos-first
|
||||
|
||||
**Charter** · 2026-07-28 · opened at `339bfd37` · **precedes and re-gates the Perception Arc**
|
||||
**Standing question:** for every part of the cognitive design — is it masterfully implemented *intrinsically*, in *the role it plays in its subsystem*, in *the role its subsystem plays in the system*, and at *every seam it is supposed to share with other organs, packs, and components*? Nothing is "understood" until all four can be articulated with evidence.
|
||||
|
||||
---
|
||||
|
||||
## 0. Why this audit, and why it starts at the bottom
|
||||
|
||||
Five independent investigations this session converged on one disease: **layers assuming connections that were never made.** The serving path assumed geometry it never touched (G-24); the faithful §2 reader sat dark while its inversion served; three operator experiments were run and forgotten; the mirror assumed a sync direction nobody verified. Every one was found by *opening the layer below the claim*.
|
||||
|
||||
The corollary is the audit's method: **verify from the foundation upward**, because a defect at layer N silently invalidates the mastery of everything at N+1 — and this repo's record proves the invalidations are real, not hypothetical. The always-on daemon consolidated an empty set for six weeks because one foundational flag was absent; every layer above it looked correct.
|
||||
|
||||
**Exit criterion:** every layer below the perception boundary carries a verdict — `MASTERFUL` (intrinsic + role + seams all evidenced), `SOUND-BUT-STRANDED` (correct, unwired), or `DEFECTIVE` — and every seam between layers is either *proven live* or *explicitly declared dark with an owner*. The Perception Arc's Phase 1+ does not run until Layers 0–2 have verdicts.
|
||||
|
||||
## 1. The first finding, and why logos is the entry point — **FA-1**
|
||||
|
||||
The suspicion that CORE-Logos is not integrated as designed was tested against the source before this charter was written. It is **correct**, and the evidence is exact:
|
||||
|
||||
**What ADR-0005/0015 design.** Three languages as charts on one semantic manifold: Hebrew roots as **depth anchors**, Greek as **relational depth**, English as **articulation surface**. Alignment edges carry **resonance** (*"Alignment is resonance. These must not be collapsed into one multiplication"*), and the validation gate of meaning is **cross-language holonomy closure**: *"holonomy(hebrew clause) ≈ holonomy(koine greek clause) ≈ holonomy(english clause)… Word-order changes should change holonomy. This is the CORE-Logos proof."* The pillar is not decoration in the design — it is the design's **criterion for meaning**.
|
||||
|
||||
**What exists, measured at `339bfd37`:**
|
||||
|
||||
| Layer | State | Evidence |
|
||||
|---|---|---|
|
||||
| Substrate grounding | **Real** | `packs/compiler.py` folds HE roots into vectors as `triliteral:` tokens — logos shapes the geometry at compile time |
|
||||
| Holonomy machinery | **Live** | `core/physics/digest.py`, referenced from `core/cognition/pipeline.py` |
|
||||
| Alignment data | **Seed-scale: 11 edges, 9 HE morphology rows** | `he_logos_micro_v1/alignment.jsonl` — evidence_ids `John1:1, Gen1:1`, exactly ADR-0015's prescribed *first* cases, never grown beyond them |
|
||||
| EN lexicon tagging | Present | 91 entries carrying `logos.core` semantic domains |
|
||||
| Serving-path role | **Veto only, no-ops on English** | `pipeline.py:544-583` — `evaluate_logos_on_text` can force a refusal against observed HE plural morphology; *"English-only turns with no HE surface → no-op"* |
|
||||
| Recall bridge | **ON, undocumented** | `allow_cross_language_recall` — one of the two default-ON flags with no recorded rationale (flag register §1) — gates `recall_top_k=3 vs 0`. **The pillar's own switch was mystery flag #1** |
|
||||
| Curriculum | Registered, bandless | `hebrew_greek_textual_reasoning` in `DOMAIN_PACKS` with corpora; produced **zero** curriculum bands at the PR-14 measurement |
|
||||
| Validation-gate role | **Never assumed** | ADR-0015 §"Establishes holonomy-level resonance as the validation gate" — no serving or licensing path consults cross-language holonomy; template match became the de-facto gate instead |
|
||||
|
||||
**FA-1 verdict: the pillar is architected and seeded exactly as designed, then left at seed scale, wired at three mutually-unaware points (vector folding · HE veto · recall top-k), and load-bearing nowhere.** The same disease named in the Perception Arc §2b — *transition windows that never close* — but at the foundation, which is why everything above (comprehension, articulation, contemplation, learning) could at best be correct *about an ungrounded semantic space*. Registered as **G-25**.
|
||||
|
||||
**FA-1's work order (first in the audit):**
|
||||
1. **Map the designed logos contract completely** — every ADR-0005/0015 obligation as a row: obligation → implementing site → live/dark → seam partner. No prose without a site.
|
||||
2. **Decide the holonomy-gate question the way §5 was decided** — pre-registered: on the existing 11 aligned cases plus a small grown set, does cross-language holonomy closure discriminate meaning-preserving from meaning-breaking articulation (word-order sensitivity included, per the ADR's own test)? This is the *designed* validation gate; it has never been measured. Unlike the three failed operator-arithmetic experiments (see Perception Arc Phase 1 correction), this tests the geometry on the job the design actually assigned it — **resonance across representations**, not re-deriving arithmetic.
|
||||
3. **Rule `allow_cross_language_recall`** — the pillar's switch gets its recorded rationale or gets turned off; measured either way.
|
||||
4. **Grow-or-declare the alignment corpus** — 11 edges is a proof of concept; either the growth program is scheduled with an owner, or the pillar is formally re-scoped and every design document claiming it is amended. No third state.
|
||||
|
||||
## 2. The layer stack and the articulation requirement
|
||||
|
||||
Bottom-up. A layer's audit is complete only when each component in it has all four articulations **with evidence**: *(i)* intrinsic correctness, *(ii)* role in its subsystem, *(iii)* subsystem's role in the system, *(iv)* every seam named with its partner and its live/dark status.
|
||||
|
||||
| # | Layer | Contents | Prior evidence in hand | Status |
|
||||
|---|---|---|---|---|
|
||||
| **L0** | Algebra kernel | `algebra/` (Cl(4,1), versors, CGA embed/readback) | 355 geo refs, fully reachable, exact-recall invariants, `versor_condition` flat across 5,000-beat soak | **Candidate-MASTERFUL — verify seams, then close** |
|
||||
| **L1** | Physics & field | `core/physics/` (37 files), `field/`, salience, Hamiltonians, digest/holonomy | Soak evidence pinned; salience live-but-bypassed; `relation_compiler` sound-but-stranded | Audit: which operators are load-bearing vs decorative |
|
||||
| **L2** | Semantic ground — **the logos layer** | packs, vocab, morphology, `semantic_primitives`, alignment, cross-language recall | **FA-1 above** | **DEFECTIVE-BY-ABSENCE at scale; starts first** |
|
||||
| **L3** | Perception | both readers, `linguistic_pipeline`, `binding_graph`, `structure_mapping`, the comprehension organs | G-24, F-A…F-E, three negative operator experiments, 1.28% read rate | Blocked on L2's verdict — a reader over an ungrounded lexicon inherits the gap |
|
||||
| **L4** | Cognition | `core/cognition/` pipeline, contemplation, epistemic state/disclosure/questions, vault, recognizer | Pipeline audited pointwise (H-11, H-13); `epistemic_disclosure`/`questions` **fully dark** in census; ADR-0144 recognizer dark | Seam census then role audit |
|
||||
| **L5** | Serving & articulation | surfaces, licences, articulation/writer, workbench | Licence machinery honest post-R-13/R-8; writer 6 constructions | Last — it is the layer most shaped by all below |
|
||||
| **X** | Cross-cuts | teaching/learning loop, governance/ledgers, always-on life | Docket executed; teaching gated; daemon profile fixed | Audited per-layer at each seam they touch |
|
||||
|
||||
**Instrument** (from the census at `339bfd37`): 666 modules, **293 reachable from serving, 373 dark**. Every dark module gets exactly one of: *seam scheduled* · *deliberately-dark with recorded reason* · *deletion candidate*. The census script becomes a pinned lane so this number is a ratchet, not a snapshot.
|
||||
|
||||
## 3. Sequencing
|
||||
|
||||
**FA-1 (logos, L2)** → **FA-0 (L0 closure — cheap, likely already masterful, and everything cites it)** → **FA-2 (L1 physics role-audit)** → **FA-3 (L3 perception, absorbing the Perception Arc phases, now on a verified ground)** → **FA-4 (L4)** → **FA-5 (L5)** → close with the full articulable map: every component, four articulations, evidence-linked.
|
||||
|
||||
The Perception Arc is **not** discarded — its Phase 0 (read-rate ratchet) runs immediately as instrumentation, its Phase 4 deletions stand, and its Phase 1 is corrected (below) and re-gated behind FA-1's holonomy decision, because the honest experiment order is: *first establish whether the semantic ground carries meaning as designed; then ask what a reader over that ground can do.*
|
||||
|
||||
## 4. Phase 1 correction, imported from the systematic sweep
|
||||
|
||||
The operator hypothesis ("relations as geometric operators") is **not untested — it has three negative results**, found only when reachability was enumerated instead of grepped:
|
||||
|
||||
| Date | Experiment | Encoding | Verdict |
|
||||
|---|---|---|---|
|
||||
| 2026-06-04 | Field-reasoner wedge (`docs/analysis/field-wedge-ablation-result-2026-06-04.md`) | relations as **translator versors**, e1 line | **C3 — decoration**: wrong=0 held, caught zero symbolic errors, lost one case, diversity 0 |
|
||||
| 2026-07-19 | Relational operator ablation (`docs/analysis/relational-operator-ablation-dossier-2026-07-19.md`) | relations as **CGA dilation** (`VersorBinding`) | **Identical to rational baseline** (2/8 both); depth metadata inert; enrich=decoration |
|
||||
| 2026-07-28 | ADR-0252 §5 (`docs/research/sme-experiment-verdict-797ebad5.md`) | structure as **point configurations**, Procrustes | **NO-GO** at every attribute weight |
|
||||
|
||||
**Convergent mechanism:** wherever the geometric operation is isomorphic to the arithmetic it replaces, it reproduces it exactly and adds nothing — zero diversity is near-tautological, not disappointing. And all three hit the same binding constraint: **the reader** (6/8 refused; 5/500 decided; 1.28% read). **Reopening criterion, stated so this is never re-proposed blind:** a domain where the geometric encoding is *not* isomorphic to a trivial symbolic computation — which is precisely what FA-1's holonomy-resonance question is, and why it, not a fourth operator-arithmetic ablation, is the next experiment.
|
||||
|
||||
---
|
||||
*Method note: every claim above carries a path. The audit inherits the standing philosophies — red-before-green for every new pin, measure before ruling, record the error next to the fix — and the census, read-rate, and seam maps become pinned lanes so the audit's own instruments cannot go stale the way the ones it replaces did.*
|
||||
|
|
@ -4,6 +4,8 @@
|
|||
**Supersedes:** Track B's "widen from 19 constructions" (G-3, as previously specified) — see §4.
|
||||
**Governed by:** `AGENTS.md` standing philosophy; the pre-registration discipline that produced the §5 NO-GO.
|
||||
|
||||
> **RE-GATED 2026-07-28 (same day, later):** the [Foundations Audit](2026-07-28-foundations-audit.md) now **precedes** this arc — bottom-up, logos-first (FA-1/G-25). Phase 0 (read-rate ratchet) and Phase 4 (deletions) stand and may run; **Phase 1 as written below is CORRECTED and superseded** — the operator hypothesis has *three* prior negative results (see the audit charter §4), found by reachability census after this document shipped. The next experiment is FA-1's holonomy-resonance question, which tests the geometry on the job ADR-0015 actually assigned it.
|
||||
|
||||
---
|
||||
|
||||
## 0. The question this answers
|
||||
|
|
@ -78,7 +80,11 @@ Sequenced per `docs/conceptualizing_engineering_mastery.md`: measure → decide
|
|||
- **G-24** in the gap register: findings F-A…F-E, Tier A (it re-scopes Track B and Track C's entry conditions).
|
||||
- **The read-rate lane** (discharges G-21's open ask): `evals/perception/read_rate.py` measuring sentence-level read rate on `holdout_dev/v1` — logic reader and math reader, refusal taxonomy included. Pinned as a **floor ratchet** (1.28% / 1.0% at `0e1be8be`): the number may only rise, and a rise is a reviewed decision. This is the metric the arc is accountable to — capability, not machinery.
|
||||
|
||||
### Phase 1 · The §2 experiment — compositional core-primitive reading · **M** · *pre-registered, GO/NO-GO*
|
||||
### Phase 1 · ~~The §2 experiment~~ — **CORRECTED 2026-07-28: substantially already run, three negatives** · see Foundations Audit §4
|
||||
|
||||
*(Original text preserved below for the record. The wedge ablation (2026-06-04, verdict C3), the relational operator ablation (2026-07-19, identical-to-baseline), and §5 (NO-GO) jointly cover the operator-encoding branch this phase proposed. Reopening requires a domain where the geometry is not isomorphic to trivial symbolic computation — FA-1's holonomy gate is that domain candidate.)*
|
||||
|
||||
#### Original Phase 1 text (superseded)
|
||||
|
||||
The §5-shaped experiment for the untested branch. **Hypothesis H2:** a reader that (a) maps text onto the **four ratified primitives** — containment, transfer, accumulation, comparison — plus the Number and Object core-systems; (b) encodes each relation as an **operator** (the `relation_compiler` pattern, generalized); and (c) derives sentence meaning by **composition** of operators — generalizes **multiplicatively**: built from *k* atomic constructions, it reads novel compositions it was never shown.
|
||||
|
||||
|
|
|
|||
Loading…
Reference in a new issue