30-gap-register.md — CORE's first LIVE gap register since docs/gaps.md closed its 26th entry (supersession proposed for ruling). 20 entries in four tiers, each with evidence, deciding authority, and leverage rank: - Tier A (frontier-blocking): G-1 the unrun ADR-0252 §5 experiment (leverage 1 — authorized, scaffolded, verdict-less; the paradigm governs ALL future comprehension); G-2 the #138 fabrications (pre-labeled measured-and-pinned, held for ADR + ratification); G-3 reader 19 vs writer 1739; G-4 CR-2 no-chooser (the AGI-grade conceptual absence); G-5 L10 proof debt (no artifact, no suite, no cadence); G-6 the half-gated lived learning loop. - Tier B (enforcement debt): G-7 orphaned-pin meta-check (the highest- leverage single mechanical change); G-8 flag-default register; G-9 three doctrine laws without verified failing pins; G-10 curriculum scoping + absent ledger; G-11 identity authorization bar unstated. - Tier C (one-line rulings): CR-3 efferent, CR-4 temporal stance, CR-1 attention ADR, the daemon's owning ADR. - Tier D (latent/carried): aspect-arm defect class (generate/templates.py :79, unreachable today), sensorium entry criterion, curriculum-formation bypass (unverified carry), Wilson evidence debt, refusal materialisation. - Plus what is NOT in the register and why (deferred-with-ruling is the scripture model; benchmarks excluded by the completeness criterion). 31-hindrance-audit.md — 11 entries ranked by leverage, each with fitness verdict, evidence, proposed better home, deciding authority; decides nothing. Headliners: H-1 Wilson/replay counting basis (wrong-solution in the counting, not the gating — 21/25 bands short; the re-count may demote licenses and that is the mechanism working); H-2 decoration as testimony (DriveGradientMap, InhibitionMask — deletion per mastery step 2); H-3 the typed refusal discarded at the public boundary; H-4 extend the resolver's declared-precedence pattern upstream; H-8 three record/code contradictions (each one paragraph to fix); H-9 dead instruments standing as if live. Plus five examined-and-CLEARED candidates (the 18 organs' continued service is governance working; pure-Python-by-default is measured-correct; flag-gated conservatism is not hindrance — unregistered darkness is).
6.5 KiB
CORE Holistic Assessment — Scope and Method
Status: Ratified approach (Shay, 2026-07-27). Phase 0 complete.
Branch: docs/holistic-assessment (worktree core-wt-assess, based on forgejo/main @ 8927c563).
Nature: Read-only investigation. This assessment changes no runtime behavior, fixes no defect, and decides nothing. It produces evidence and judgments for ruling.
1. The question being answered
Four questions, in order of dependency:
- Where does CORE actually stand on its cognitive cycle — design articulated versus implementation fulfilled — from the telos down to individual components?
- Is the layer model itself complete? Are there layers, sublayers, or components missing from the design, not merely from the implementation — things an AGI/ASI-grade system requires that nothing in the current architecture accounts for?
- What is the metadata for every layer, sublayer, and component? Philosophical intent, functional contract, design shape, implementation status, evidence, capacity, and role — such that any future dive begins with a clear target rather than a grep.
- What is hindering us? Implemented ADRs or designs that are the wrong solution for their underlying problem; responsibilities lodged in the wrong subsystem; trade-offs tuned rather than dissolved.
Question 4 is not a courtesy pass. It is the question with the highest expected value, because a wrong component that works is more expensive than a missing component that is known to be missing.
2. Governing method
The assessment is conducted under docs/conceptualizing_engineering_mastery.md, applied to the assessment itself and not only to its subject.
Pillar I — Semantic Rigor. Completeness criteria and the metadata schema are defined before any component is judged (Phase 1), so that "missing" and "complete" have fixed meanings rather than per-component ones. Every implemented verdict requires an evidence pointer: a test, an eval lane, a pinned SHA, or an acceptance packet. "The module exists" is not evidence that it executes; "the flag is threaded" is not evidence that it changes behavior.
Pillar II — Mechanical Sympathy. Components are judged against CORE's own doctrine — deterministic decoding, exact recall, field-as-substrate with intelligence in the wiring, replay-gated learning, wrong=0-or-refuse — and not against a generic AGI checklist. A capability only counts as a gap if CORE's own telos requires it. Importing an external architecture's expectations would manufacture false gaps.
Pillar III — The Third Door. The hindrance audit's explicit charter: find the places where a bad trade-off was split rather than dissolved, and the places where deletion beats addition. Per the execution algorithm, deletion (step 2) precedes optimization (step 3); a component that should not exist is never a performance problem.
Two standing discipline rules carried in from prior work:
- The sabotage test. For every claim that a mechanism is live and load-bearing, ask what the measurement would look like with the mechanism removed. If it would look identical, the claim is decoration and is recorded as such. This repository has produced exactly that failure before — a rate reported as evidence of a reader that was
0.0throughout. - Identity, not value. When measuring whether two things are the same thing, measure identity rather than equal-looking values. Source-scanning metrics can move the wrong way on success.
3. Phases and division of labor
| Phase | Deliverable | Executor |
|---|---|---|
| 0 — Ground truth | Canonical-document ingestion; ADR triage; system-map recovery; raw material and open tensions for the taxonomy | Opus 5 — complete |
| 1 — Taxonomy & schema | The macro→micro layer taxonomy and the metadata card schema every card must fill | Fable 5 |
| 2 — Macro layer cards | One card per top-level layer; layer-level verdicts; cross-cutting concerns | Opus 5 |
| 3 — Micro component cards | Per-subsystem descent, depth allocated by load-bearing-ness | Fable 5 |
| 4 — Gap register + hindrance audit | Two separate registers; evidence-carrying | Fable 5 (reassigned from Opus 5 by Shay, 2026-07-27) |
| 5 — Synthesis | Executive assessment; ranked gaps; ranked hindrances; recommended R&D attack order | Fable 5 (same reassignment) |
Phase 1 is the keystone. A wrong taxonomy miscategorizes everything downstream, and it is the cheapest phase to correct.
4. Deliverables
docs/assessment/
00-scope-and-method.md # this file
01-phase0-ground-truth.md # Phase 0 findings + Phase 1 handoff
02-layer-taxonomy.md # Phase 1
03-card-schema.md # Phase 1
10-layer-cards/ # Phase 2
20-component-cards/ # Phase 3
30-gap-register.md # Phase 4
31-hindrance-audit.md # Phase 4
40-assessment.md # Phase 5
5. Rules of engagement
- Read-only. No runtime code is modified. No defect is fixed. Where a fix is obvious, it is recorded as a finding with its evidence, not applied.
- Verify against code, not against documents. Documentation in this repository has been measurably wrong about the repository before — most recently
docs/research/architecture-assessment-verification-2026-07-25.mdfalsified roughly a third of an external blueprint's work items by reading the implicated code. A claim sourced only from a document is labeled as such. - Settled rulings are constraints, not subjects. The deduction pivot, the scripture-content deferral, the no-merge-automation rule, and ratified ADRs enter as given. An ADR is reopened only on evidence of hindrance, and only as a flag for ruling — never as a unilateral recommendation to reverse.
- Known-and-held findings are recorded as such. The PR #138 fabrication findings (
every dog is a mammal→member(every_dog, mammal);Given: furthermore; p implies q; p.reaching served output) are measured and pinned, not fixed. The fixes are known — two of the 13 mutations — and are deliberately held out pending ADR and ratification because they change what CORE comprehends from user input, which is serving-path truth behavior. They enter the gap register pre-labeled and are never re-presented as newly discovered. - No timelines. Scope size, phase, and priority only.
- All wrinkles surfaced. Technical truths are volunteered unprompted, including ones that complicate the picture or reflect badly on prior work.