Outcome
Measure whether agents actually need a compiled historical retrieval layer, and admit only the smallest layer that demonstrably outperforms live GitHub/Git/spec search.
Baseline
Collect 20 representative historical questions from dev-backlog, dev-relay, and consuming-repository work. Compare:
- raw task-mirror grep,
gh + Git history + existing spec/ADR/context sources,
- a non-committed, on-demand compiled report.
Guardrails
- No committed topic graph or automatic per-PR memory update in the initial experiment.
- Derived output contains no authoritative current status.
- Every material claim links to an Issue, PR, commit, ADR, or spec source.
- Freshness distinguishes
compiled_at, sources_through, and human_verified_at.
- Partial source/API failure preserves existing output and performs no write.
Go criteria
Decision
- If the baseline sources are sufficient, close with no compiler.
- If the on-demand experiment wins, propose a separate
project-memory skill and human-gated charter amendment.
- A committed artifact requires a separate follow-up decision with a single integration writer, source cursor/CAS, dry-run diff, and regenerate-not-merge semantics.
Dependencies
The baseline can start immediately after the authority contract. Any productization decision waits until the mirror-off pilot has produced real usage evidence.
Outcome
Measure whether agents actually need a compiled historical retrieval layer, and admit only the smallest layer that demonstrably outperforms live GitHub/Git/spec search.
Baseline
Collect 20 representative historical questions from dev-backlog, dev-relay, and consuming-repository work. Compare:
gh+ Git history + existing spec/ADR/context sources,Guardrails
compiled_at,sources_through, andhuman_verified_at.Go criteria
Decision
project-memoryskill and human-gated charter amendment.Dependencies
The baseline can start immediately after the authority contract. Any productization decision waits until the mirror-off pilot has produced real usage evidence.