docs: archive historical Confidence Engine evidence

This commit is contained in:
2026-08-19 12:07:17 +01:00
parent a12f9555af
commit e6d0327641
79 changed files with 27 additions and 9 deletions
+17
View File
@@ -15,6 +15,23 @@ All files below were moved from `docs/` on 2026-08-06 by Experiment 29 to reduce
| `docs/v0.7-observation-report.md` (136 lines) | `docs/archive/v0.7-observation-report.md` | Experimental observation snapshot from v0.7 UX work. | Useful as a reference but not a current working document. UX work is paused. | When reviewing past UX observations that may inform future interface design decisions. |
| `docs/archive/deferred-ux-backlog.md` (376 lines) | `docs/archive/deferred-ux-backlog.md` | Deferred and exploratory UX ideas from original `docs/backlog info.md` (lines 21390). Retained for historical reference. Not commitments, priorities or active tasks. | Superseded `docs/backlog info.md`. Deferred UX planning separated from mock reference in Experiment 31. | When a named past UX idea from the deferred backlog is being reviewed; not loaded by default. |
## Phase 2B Experiment Archives (2026-08-19)
All files below were classified `HISTORICAL_EVIDENCE + SAFE` during the Phase 1B/2B context audit and moved to reduce default reading burden while preserving full traceability. They are preserved evidence — not discarded, obsolete, or invalidated. Load only when a specific historical question requires them.
| Subdirectory | What Was Moved | Count |
|---|---|---|
| `docs/archive/experiments/reasoning-fidelity-v0.8/` | Experiment 56 family (reasoning-fidelity v0.8 pass) | 11 files (experiment-56am, excluding c) |
| `docs/archive/experiments/semantic-action-contract/` | Experiment 58 family (semantic action contract) | 8 files (experiment-58a1a6, b1b2) |
| `docs/archive/experiments/question-formulation/` | Experiment 59 family (question formulation) | 7 files (experiment-59a1a3, b1b4) |
| `docs/archive/experiments/decision-options/` | Experiment 60A family (decision options analysis) | 7 files (experiment-60a18, excluding a3) |
| `docs/archive/experiments/decision-closure-integration/` | Experiment 60B subfamilies {1015}, {5582}, {95,97,100} | 35 files (experiment-60b{10-15}, {55-56,58-82}, {95,97,100}) |
| `docs/archive/experiments/knowledge-mgmt/` | Cold-start validation historical evidence | 1 file (cold-start-validation.md) |
| `docs/archive/experiments/context-routing/` | Document-role review (classification/routing analysis) | 1 file (document-role-review.md) |
| `docs/archive/experiments/pre-RTO/` | Pre-Return-to-Origin experiments and version-specific docs: v0.5v0.7 | 7 files (pre-RTO experiments + release notes/UX pass) |
**Not moved in Phase 2B:** experiment-60b{18}, experiment-60b{1948}, checkpoint-60b93.md, experiment-57* family, experiment-60b{18} (carried forward), docs/design-evolution-log.md, docs/investigation-state-assessment*.md, architectural-principles.md, v0.6-reasoning-architecture.md, success-signals.md, failure-modes.md, investigation-narrative.md, behaviour-selection.md, orchestrator-contract.md, reasoning-contract-backlog.md, reasoning-refinement-requirements.md, reasoning-production-path-map.md. These will be reviewed in a later Phase 2C pass.
## Superseded Files
The following files were superseded by a structured split in Experiment 31 and are no longer in use. Their contents remain fully represented in the documents below.
+10 -9
View File
@@ -81,7 +81,7 @@ Grouped by task domain. Only load the group relevant to your work.
| Document | Purpose | Loaded When |
|---|---|---|
| `.claude/ux-guidelines.md` (97 lines) | Main user view, developer view priorities; "calm workspace" principle | Always loading the UX guidelines for any UI task |
| `docs/v0.7-user-workspace-ux-first-pass.md` (157 lines) | First UX pass: card layout, progress summary, developer details boundary | When continuing v0.7 UX work or reviewing layout decisions |
| `docs/archive/experiments/pre-RTO/v0.7-user-workspace-ux-first-pass.md` (157 lines) | First UX pass: card layout, progress summary, developer details boundary — archived Phase 2B (2026-08-19). Preserved evidence, not current context. | When reviewing v0.7 UX layout decisions from archived historical context |
| `docs/ui-mock-reference.md` (≈62 lines) | Task-specific reference for UI mock scenarios and fixture data locations. Contains available scenarios, purposes, where fixture data lives, when to use each, and warnings against treating mock behaviour as live-engine evidence. | When working on UI mock development or testing scenarios |
| `docs/v0.7-ui-mock-mode.md` (291 lines) | Mock investigation mode for UI development without Ollama | When developing UI features that need fixture-driven testing |
@@ -96,10 +96,10 @@ Grouped by task domain. Only load the group relevant to your work.
### Evidence, decision conditions, and passive classifiers
| Document | Purpose | Loaded When |
|---|---|---|
| `docs/v0.6-comparability-experiment.md` (48 lines) | Comparability hypothesis — observations must be comparable before contradiction | When working with contradiction detection or comparability assessment |
| `docs/v0.6-atomicity-experiment.md` (211 lines) | Atomic unknown decomposition experiment results | When reviewing how decomposed unknowns are handled |
| `docs/v0.6-selection-influence-experiment.md` (50 lines) | Selection influence — graph structure vs semantic keyword analysis | When investigating question selection drivers |
| `docs/v0.5-question-priority-generalisation.md` (48 lines) | Question priority generalisation from v0.5 | When reviewing historical prioritisation decisions |
| `docs/archive/experiments/pre-RTO/v0.6-comparability-experiment.md` (48 lines) | Comparability hypothesis — observations must be comparable before contradiction — archived Phase 2B (2026-08-19). Preserved evidence, not current context. | When reviewing historical comparability assessment from archived context |
| `docs/archive/experiments/pre-RTO/v0.6-atomicity-experiment.md` (211 lines) | Atomic unknown decomposition experiment results — archived Phase 2B (2026-08-19). Preserved evidence, not current context. | When reviewing historical atomicity decisions from archived context |
| `docs/archive/experiments/pre-RTO/v0.6-selection-influence-experiment.md` (50 lines) | Selection influence — graph structure vs semantic keyword analysis — archived Phase 2B (2026-08-19). Preserved evidence, not current context. | When reviewing historical selection driver analysis from archived context |
| `docs/archive/experiments/pre-RTO/v0.5-question-priority-generalisation.md` (48 lines) | Question priority generalisation from v0.5 — archived Phase 2B (2026-08-19). Preserved evidence, not current context. | When reviewing historical prioritisation decisions from archived context |
### Facilitator behaviour and investigation state
| Document | Purpose | Loaded When |
@@ -118,7 +118,7 @@ Grouped by task domain. Only load the group relevant to your work.
### Knowledge management
| Document | Purpose | Loaded When |
|---|---|---|
| `docs/document-role-review.md` (140 lines) | Classification of deferred documents; practical routing test for UI mock and reasoning tasks | When reviewing which documents to load; when a task involves architectural guidance or mock fixture reference |
| `docs/archive/experiments/context-routing/document-role-review.md` (140 lines) | Classification of deferred documents; practical routing test for UI mock and reasoning tasks — archived Phase 2B (2026-08-19). Preserved evidence, not current context. | When reviewing historical document classification decisions from archived context |
| `docs/task-context-packs.md` (~110 lines) | Task-routing entry point — four minimal packs for engine, UI, architecture review, and knowledge management work | Always for any new task — determines which pack to follow first |
### Testing and contracts
@@ -152,8 +152,9 @@ Documents or sections that are primarily historical evidence from past experimen
### Review Before Archive (may have future value; do not load by default now)
| Document | Size | Why review before archive |
|---|---|---|
| `docs/architectural-principles.md` (306 lines) | medium | Broader and aspirational architectural principles from experiments. Experiment 30 confirmed: 6 current, 4 aspirational targets, 3 overlap guardrails but add context. Role: task-specific reference for reasoning architecture work — not a statement of current implementation. Use `docs/current-working-principles.md` for default guidance instead. See `docs/document-role-review.md` §2 and Experiment 32 entry. |
| `docs/architectural-principles.md` (306 lines) | medium | Broader and aspirational architectural principles from experiments. Experiment 30 confirmed: 6 current, 4 aspirational targets, 3 overlap guardrails but add context. Role: task-specific reference for reasoning architecture work — not a statement of current implementation. Use `docs/current-working-principles.md` for default guidance instead. See `docs/archive/experiments/context-routing/document-role-review.md` §2 and Experiment 32 entry. |
| `docs/backlog info.md` (390 lines) | large | **Superseded by Experiment 31.** Content split into `docs/ui-mock-reference.md` (mock fixtures reference, ~62 lines) and `docs/archive/deferred-ux-backlog.md` (deferred UX planning, ~376 lines). See archive index for provenance. |
| Phase 2B experiment archives (see below) | various | Archived 2026-08-19 under Phase 2B context audit. All classified HISTORICAL_EVIDENCE + SAFE in Phase 1B. Moved to `docs/archive/experiments/` subdirectories. See docs/archive/README.md for updated index. |
### Do Not Move or Delete (evidence of design evolution)
These documents document the path from Phase 1 through Experiment 25B. Archiving them separately without review would lose the rationale behind later decisions.
@@ -170,7 +171,7 @@ These documents document the path from Phase 1 through Experiment 25B. Archiving
### Same principle appears in several documents
- The principle "The engine owns the complexity / user sees only the next step" appears in `01_Confidence_Engine_Founding_Principles.md`, `project-context.md`, `ux-guidelines.md`, and implicitly in `architecture-guardrails.md`. Consider consolidating or cross-referencing.
- The "calm workspace / minimal cognitive load" principle appears in `product-story.md`, `project-context.md`, `ux-guidelines.md`, and `v0.7-user-workspace-ux-first-pass.md`.
- The "calm workspace / minimal cognitive load" principle appears in `product-story.md`, `project-context.md`, `ux-guidelines.md`, and `docs/archive/experiments/pre-RTO/v0.7-user-workspace-ux-first-pass.md`.
### Current state is buried inside a long chronological log
- Experiment 25B (the most recent engine experiment) is at line ~1,483 of a 1,542-line document. A developer joining the project must scroll past 14+ phases to find the active state. Consider a "Current State" header near the top of `design-evolution-log.md`.
@@ -211,4 +212,4 @@ None. The five questions were answered accurately from the minimum context set.
## Return-to-Work Note
Engine experiments paused after Experiment 25B, which established scope-aware condition status classification — distinguishing direct evidence from relevant-but-different claims by checking subject, timeframe, and claim type. Present-state evidence does not settle future-feasibility conditions. The passive classifier layers remain isolated; no active integration yet. Knowledge-management experiments continue: five historical documents archived per Experiment 29; backlog info.md split in Experiment 31 into `docs/ui-mock-reference.md` (mock scenarios reference) and `docs/archive/deferred-ux-backlog.md` (deferred UX planning). Neither backlog item deleted or promoted. First file to inspect when resuming: `.claude/project-context.md`, then Experiments 2325B in `docs/design-evolution-log.md` (lines 12181520).
Engine experiments paused after Experiment 25B, which established scope-aware condition status classification — distinguishing direct evidence from relevant-but-different claims by checking subject, timeframe, and claim type. Present-state evidence does not settle future-feasibility conditions. The passive classifier layers remain isolated; no active integration yet. Knowledge-management experiments continue: five historical documents archived per Experiment 29; backlog info.md split in Experiment 31 into `docs/ui-mock-reference.md` (mock scenarios reference) and `docs/archive/deferred-ux-backlog.md` (deferred UX planning). Neither backlog item deleted or promoted. **Phase 2B context audit (2026-08-19):** ~76 historical experiment files moved to `docs/archive/experiments/` under eight subdirectories — preserved as evidence, removed from default context loading paths. No methodology or current operational paths changed. First file to inspect when resuming: `.claude/project-context.md`, then Experiments 2325B in `docs/design-evolution-log.md` (lines 12181520).