Phase A of the curriculum-license-loop arc. Design only; no production code, no flag, no default changed. Rules R1-R9: polarity is a row-level field (absent => affirmative), negatives compile to the sentential-not form and MUST reuse the affirmative connective (atom identity), contradiction is rejected at ratification, premise compilation becomes query-scoped, an empty scope is UNKNOWN, premise_count keeps reporting family size, the oracle learns polarity independently, and a band's honest ceiling must be reported before anyone authors for it. Three findings, each measured in-tree and each reversing something: 1. `polarity` is read by nothing. A row authored to refute a question today serves a confident wrong "Yes" -- and the independent oracle ignores the field for the same reason, so gold agrees and wrong=0 stays green. This is the silent-failure class Phase A existed to prevent, now demonstrated. 2. The 16-premise honesty cap holds a family to <=16 chains; at 17 the band declines everything, including the taught positive that works today. So ADR-0262 §5.1's own remedy (author ~219 relations) destroys the capability at row 17, tripping the fix §6 deferred as a future contingency. Jointly: a band has <=16 entailed cases against a 657-committed threshold, so no curriculum band can earn SERVE until scoping lands. The blocker is engineering, not content -- inverting the conclusion two arcs closed on. 3. Honest volume is bounded by taught vocabulary, not corpus size. 8 of 11 bands cannot reach 657 at all, including the plan's chosen first target physics·modal (ceiling 480, counting every possible question). Retargets Phase F to philosophy_theology·modal (ceiling 44,104, same 8-chain start). R5 is the only rule that could change an existing answer, so it is the only one carrying evidence: 8,520 routable questions, 0 verdict mismatches for both candidate scopes. Verdict-identity holds for any scope that is a superset of the query-atom rows, since each premise mints one independent atom. That is stronger than ADR-0262 §6's soundness argument, and it is why this narrowing is not the ADR-0261 §5.1 premise-dropping failure. ADR-0262 §5.1/§5.2/§6 carry forward-pointers; original text preserved. [Verification]: uv sync --locked on canonical CPython 3.12.13; in-worktree `core test --suite smoke -q` 556 passed (555 baseline + 1: the derived ratify-on-merge pin picked up ADR-0264 and discharged it), `--suite deductive -q` 285 passed. Four probes reproducible from docs/research/curriculum-premise-scope-2026-07-25.md; the corpus-mutating one restores the file and asserts byte-identity. |
||
|---|---|---|
| .. | ||
| adr | ||
| agents/grok | ||
| analysis | ||
| architecture | ||
| audit | ||
| benchmarks | ||
| briefs | ||
| curriculum | ||
| decisions | ||
| evals | ||
| examples | ||
| handoff | ||
| handoffs | ||
| implementation | ||
| issues | ||
| lab | ||
| outreach | ||
| paradigm-archive | ||
| plans | ||
| research | ||
| sessions | ||
| specs | ||
| workbench | ||
| zig | ||
| 3lang-depth-pr-plan.md | ||
| admissibility-exemplars.md | ||
| ci-optimization.md | ||
| core-rd-base-prompts.md | ||
| ethics_packs.md | ||
| EVAL_AUDIT_2026-05-20.md | ||
| eval_methodology.md | ||
| frontier_baselines.md | ||
| gaps.md | ||
| handoff_template.md | ||
| hitl-backpressure.md | ||
| holdout_recipients.txt | ||
| identity_packs.md | ||
| master-plan-post-substrate-audit.md | ||
| memo.html | ||
| model_dependency_size_tally.md | ||
| pack_inventory_2026-05-21.md | ||
| position_paper.md | ||
| PROGRESS.md | ||
| README.md | ||
| recognizer-registry.md | ||
| refusal-taxonomy.md | ||
| reviewers.yaml | ||
| RUST.md | ||
| safety_packs.md | ||
| sponsors.md | ||
| teaching_order.md | ||
| test-debt-quarantine.md | ||
| testing-lanes.md | ||
| Whitepaper.md | ||
| Yellowpaper.md | ||
CORE Documentation Index
This is the central index for all documentation in the CORE project.
Canonical Root Documents
- Whitepaper.md - The CORE architectural and philosophical whitepaper.
- Yellowpaper.md - Technical specifications and mathematical formulation of the CORE engine.
- PROGRESS.md - High-level project progress tracking.
- specs/runtime_contracts.md - Critical invariants and bounds for the runtime execution flow.
Directories
Architecture & Design
- adr/ - Architecture Decision Records (ADRs). The canonical history of all ratified engineering choices.
- architecture/ - High-level architectural documents (e.g., pipelines, schemas).
- analysis/ - Deep dives and master plans for structural changes.
- specs/ - Detailed technical specifications and invariants.
Planning & Progress
- plans/ - Capability roadmaps and implementation plans.
- briefs/ - Project briefs and scoping documents for upcoming work.
- issues/ - Detailed issue analyses and technical reproductions (not standard trackers).
- audit/ - Audit reports and claims ledgers.
Operations & Usage
- examples/ - Usage examples and reference integrations.
- workbench/ - Documentation for the CORE workbench UI and related tooling.
- agents/ - Agent-specific operational guides.
Note regarding
agents/grok/: Operational guide for using Grok 4.3 + Grok Build with CORE. This is not architecture documentation.
Learning & Evaluation
- curriculum/ - Documentation on the teaching/learning order and knowledge progression.
- evals/ - Evaluation methodology and performance criteria documentation.
- benchmarks/ - Benchmark evidence and performance tracking logs.
Historical & Experimental
- sessions/ - Chronological session logs documenting the "decision trail" and intellectual history of major choices.
- handoffs/ - Legacy brief, audit, and investigation notes (historical; the formal HANDOFF mechanism is retired — see AGENTS.md for the current lightweight
session-break-summary-<DATETIME>.mdconvention). - research/ - Raw research notes and preliminary findings.
- lab/ - Experimental content (Warning: Not ratified; must not be referenced as authoritative in production PRs).