Repository navigation
Migrate specsmith → spec-kit + AEE; core on AEE library; GDELT Web Ngrams default - #54
Merged
Merged
Conversation
- specify init 1.0.10 (copilot/sh); aee + evaluator extensions vendored as real files - constitution ratified at .specify/memory/constitution.md from existing governance - specs/001-glossa-lab-baseline: as-built baseline spec/plan/tasks - remove specsmith skills, scaffold.yml, tracked backend/.specsmith state - runtime rate-limit state: .specsmith/ -> .glossa-state/ (base.py, model_intelligence.py), gitignored - AGENTS.md + LIFECYCLE.md updated to spec-kit flow; LEDGER entry appended (history untouched)
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Summary
Three-stage migration, executed per the repo's governance (read-first:
AGENTS.md,docs/governance/rules.md,LIFECYCLE.md) and ledgeredper stage in
LEDGER.md(append-only, AI-disclosed).Stage 1 — specsmith → spec-kit + AEE (
2e603341)specify init(v1.0.10, copilot integration) + aee & evaluatorextensions installed from local clones as real files (verified: no
gitlinks under
.specify/)..specify/memory/constitution.mdpreserves Glossa's real governance:citation/provenance, append-only LEDGER, foundation-check gate,
public/private correspondence boundary, AI disclosure, falsifiable
claims, spec-driven flow.
specs/001-glossa-lab-baseline/documents the as-built platform(FastAPI + React, discovery engine, Evidence Graph, Indus status per
README: 161 H+M readings, 90.96% coverage, preprint v4 DOI
10.5281/zenodo.20414696).
scaffold.yml, trackedrate-limit state removed); runtime persistence
.specsmith/→.glossa-state/infetchers/base.py+model_intelligence.py.Historical mentions in LEDGER/CHANGELOG/ledger-archive untouched.
Stage 2 — core wired to the AEE library (
a915eca5)applied-epistemic-engineering>=1.0.4,<2.backend/glossa_lab/aee_core.pyadapts extracted-claims JSON ontoAEE
Claim/Evidence/ClaimGraph+ScoringEngine.falsification_conditionmaps natively toClaim.falsification_tests;untested→ AEEDRAFT(original keptin
Claim.metadata); status enters scoring through attached evidencestrength (the engine takes no status input — documented in the
docstring).
GET /api/v1/indus-evidence/claims/aee-scoresand
?aee=trueonGET /claims. Verified live: 31 claims, meanpropagated score 0.485; default response shapes unchanged.
Stage 3 — GDELT guidance implemented (
48c1ad8f)gdelt_ngrams(Web Ngrams dataset):~5-min-ago request, bounded 15-min-heartbeat walk-back (24 marks),
quadgram matching with tri/bi/unigram reduction (>4-word keywords via
constituent windows), topic exclusions, TOC cross-reference,
.glossa-state/watermark.per GDELT's request during the Spanner migration.
specs/002-gdelt-ngrams-and-frontier-methods/also records (notimplemented): a future fully-cited daily briefing over the discovery
corpus, and manuscript-method design notes for future seal/tablet
vision work (physical metadata up front; discrete focused passes).
Sync note
Branch includes a clean merge of origin/main after Dependabot
#50–#53 landed (workflow actions v7, Pillow bump) — no conflicts; AEE
dependency and Dependabot changes coexist.
Test results
(
test_indus_evidence_api.py25 passed run separately; the rest 508passed / 9 skipped).
test_aee_core.py18 passed,test_gdelt_ngrams.py11passed (synthetic fixtures, no network). Ruff clean on changed files.
backend/scripts/foundation_check.pycannot run off the Windows devbox (hardcoded
C:\Users\trist\...path — pre-existing, untouched);noted in the stage-4 ledger entry.
Live smoke test (GDELT ngrams)
pair (2026-06-30 20:16 UTC): 1,063,647 quadgrams, 1,977 TOC entries;
end-to-end matching on real data works ("disease" → 16 items;
"Indus script" → 0 in that single minute).
marks over prior weeks all return 404 — the dataset appears not to be
publishing at test time. The fetcher handles this correctly (bounded
walk, no items, watermark untouched) and picks files up when
publication resumes.
🤖 Executed by an AI agent (Muse Spark, via Muse) at the direction of
Tristen Pierson, per constitution §VI. Not merged — awaiting review.