Skip to content
Draft
Show file tree
Hide file tree
Changes from all commits
Commits
File filter

Filter by extension

Filter by extension

Conversations
Failed to load comments.
Loading
Jump to
Jump to file
Failed to load files.
Loading
Diff view
Diff view
2 changes: 1 addition & 1 deletion .claude-plugin/marketplace.json
Original file line number Diff line number Diff line change
Expand Up @@ -19,7 +19,7 @@
{
"name": "aidd-dev",
"source": "./plugins/aidd-dev",
"description": "Code transformation: plan, implement, assert, audit, review, test, refactor, debug, for-sure. Hosts engineering agents.",
"description": "Code transformation: plan, implement, assert, audit, review, test, refactor, debug, goalify. Hosts engineering agents.",
"strict": true,
"metadata": {
"recommended": true
Expand Down
4 changes: 2 additions & 2 deletions aidd_docs/README.md
Original file line number Diff line number Diff line change
Expand Up @@ -40,7 +40,7 @@ Skills are grouped into plugins by domain. Install only the plugins you need.
| aidd-context | Bootstrap, project init, generation of context artifacts (skills, agents, rules, commands, hooks, plugins, marketplaces), mermaid diagrams, learn, discovery | `02-project-memory`, `03-context-generate`, `09-mermaid` |
| aidd-refine | Meta-cognition: brainstorm, challenge prior work, blind-spot scan, fact-check | `01-brainstorm`, `02-challenge`, `03-shadow-areas` |
| aidd-pm | Product management: backlog artifacts, refinement, Product Briefs, PRD, spec | `02-user-stories`, `05-spike`, `07-epic`, `09-defect`, `10-task` |
| aidd-dev | Code transformation: plan, implement, assert, audit, review, test, refactor, debug, for-sure | `01-plan`, `02-implement`, `05-review`, `06-test` |
| aidd-dev | Code transformation: plan, implement, assert, audit, review, test, refactor, debug, goalify | `01-plan`, `02-implement`, `05-review`, `06-test` |
| aidd-vcs | VCS workflows: commit, pull/merge request, release tag, issue creation | `01-commit`, `02-pull-request`, `04-issue-create` |
| aidd-orchestrator | Synchronous SDLC, async issue-to-PR automation, and product backlog | `00-async-dev`, `01-sdlc`, `02-backlog` |

Expand Down Expand Up @@ -104,7 +104,7 @@ AIDD is delivered as a plugin marketplace. Pick what you need; do not install ev
| ------------ | ------------------------------------------------------------------------------------------------------------------- |
| aidd-context | 00-onboard, 01-bootstrap, 02-project-memory, 03-context-generate, 09-mermaid, 10-learn, 11-explore |
| aidd-refine | 01-brainstorm, 02-challenge, 03-shadow-areas, 04-fact-check |
| aidd-dev | 01-plan, 02-implement, 03-assert, 04-audit, 05-review, 06-test, 07-refactor, 08-debug, 09-for-sure, 10-todo |
| aidd-dev | 01-plan, 02-implement, 03-assert, 04-audit, 05-review, 06-test, 07-refactor, 08-debug, 09-goalify, 10-todo |
| aidd-orchestrator | 00-async-dev, 01-sdlc |
| aidd-vcs | 01-commit, 02-pull-request, 03-release-tag, 04-issue-create |
| aidd-pm | 01-ticket-info, 02-user-stories, 03-prd, 04-spec, 05-spike, 06-product-brief, 07-epic, 08-three-amigos, 09-defect, 10-task |
Expand Down
4 changes: 2 additions & 2 deletions docs/CATALOG.md
Original file line number Diff line number Diff line change
Expand Up @@ -35,7 +35,7 @@ Bootstrap, project init, context-artifact generation, diagrams, learning, and ex

## 💻 aidd-dev

Code transformation: plan, implement, assert, audit, review, test, refactor, debug, for-sure, todo. Standalone Browser QA records short web evidence.
Code transformation: plan, implement, assert, audit, review, test, refactor, debug, goalify, todo. Standalone Browser QA records short web evidence.

| Skill | Role | Actions |
| --------------- | -------------------------------------------------------------------------- | ------------------------------------------------------------------------------- |
Expand All @@ -47,7 +47,7 @@ Code transformation: plan, implement, assert, audit, review, test, refactor, deb
| `06-test` | Write and iterate tests, validate user journeys in the browser | `01-test`, `02-test-journey` |
| `07-refactor` | Improve code without changing behavior across four axes | `01-performance`, `02-security`, `03-cleanup`, `04-architecture` |
| `08-debug` | Reproduce and fix bugs with a test-driven workflow | `01-reproduce`, `02-debug`, `03-reflect-issue` |
| `09-for-sure` | Iterative loop that retries until a success condition is met | `01-init-tracking`, `02-auto-accept`, `03-autonomous-loop` |
| `09-goalify` | Autonomous loop that replans and retries until a runnable success condition passes | `01-init-tracking`, `02-auto-accept`, `03-autonomous-loop` |
| `10-todo` | Split the prompt into independent todos, run one implementer agent per todo in parallel | `01-todo` |
| `11-browser-qa` | Record short reviewer videos for browser-scoped happy and edge cases | `00-prerequisites`, `01-load-scope`, `02-prepare-run`, `03-run-scenarios` |

Expand Down
Original file line number Diff line number Diff line change
Expand Up @@ -16,7 +16,7 @@ The contract every generated skill satisfies. `skill-generate` obeys it too.
- **R7.** The mermaid flow shows every path a run can take: the nominal chain, one entry node per case, a back-edge per loop, a terminal node per outcome. A branch stated in prose is a branch missing from the flow.
- **R8.** The action table is `| Action | Does |`, one row per action file, in run order. `Action` is the bare slug, no backticks and no number. `Does` is a lowercase imperative half-line with no final period. Above the table, one sentence: what to read next, nothing more.
- **R9.** `## Transversal rules` holds the rules no single action or reference owns. A rule stated there is stated nowhere else.
- **R10.** The router is loaded on every call, an action only when its turn comes: the router carries nothing an action or a reference could carry.
- **R10.** The router is loaded on every call, an action only when its turn comes: the router carries nothing an action or a reference could carry. Never link to `SKILL.md` from a skill file.

## An action

Expand Down
4 changes: 2 additions & 2 deletions plugins/aidd-dev/.claude-plugin/plugin.json
Original file line number Diff line number Diff line change
Expand Up @@ -2,7 +2,7 @@
"$schema": "https://json.schemastore.org/claude-code-plugin-manifest.json",
"name": "aidd-dev",
"version": "2.5.0",
"description": "Code transformation: plan, implement, assert, audit, review, test, refactor, debug, for-sure, plus short standalone Browser QA evidence. Hosts engineering agents.",
"description": "Code transformation: plan, implement, assert, audit, review, test, refactor, debug, goalify, plus short standalone Browser QA evidence. Hosts engineering agents.",
"author": {
"name": "AI-Driven Dev",
"url": "https://github.com/ai-driven-dev"
Expand All @@ -16,7 +16,7 @@
"./skills/06-test",
"./skills/07-refactor",
"./skills/08-debug",
"./skills/09-for-sure",
"./skills/09-goalify",
"./skills/10-todo",
"./skills/11-browser-qa"
],
Expand Down
18 changes: 9 additions & 9 deletions plugins/aidd-dev/CATALOG.md
Original file line number Diff line number Diff line change
Expand Up @@ -17,7 +17,7 @@ Auto-generated index of skills, agents, references and assets shipped by the `ai
- [`skills/06-test`](#skills06-test)
- [`skills/07-refactor`](#skills07-refactor)
- [`skills/08-debug`](#skills08-debug)
- [`skills/09-for-sure`](#skills09-for-sure)
- [`skills/09-goalify`](#skills09-goalify)
- [`skills/10-todo`](#skills10-todo)
- [`skills/11-browser-qa`](#skills11-browser-qa)

Expand Down Expand Up @@ -126,17 +126,17 @@ Auto-generated index of skills, agents, references and assets shipped by the `ai
| `references` | [mermaid-conventions.md](skills/08-debug/references/mermaid-conventions.md) | `Rules for generating valid, high-quality Mermaid diagrams. Apply when creating or reviewing any Mermaid diagram (flowchart, state, ER, sequence, gantt).` |
| `-` | [SKILL.md](skills/08-debug/SKILL.md) | `Reproduce and fix a known bug, or find an unknown root cause by hypothesis validation. Use when the user wants to fix a bug, find why something breaks, or reopen a stuck investigation. Not for building a feature or reviewing a diff.` |

#### `skills/09-for-sure`
#### `skills/09-goalify`

| Group | File | Description |
|-------|------|---|
| `actions` | [01-init-tracking.md](skills/09-for-sure/actions/01-init-tracking.md) | - |
| `actions` | [02-auto-accept.md](skills/09-for-sure/actions/02-auto-accept.md) | - |
| `actions` | [03-autonomous-loop.md](skills/09-for-sure/actions/03-autonomous-loop.md) | - |
| `assets` | [autonomous-loop-worker-prompt.md](skills/09-for-sure/assets/autonomous-loop-worker-prompt.md) | - |
| `assets` | [plan-template.md](skills/09-for-sure/assets/plan-template.md) | - |
| `references` | [autonomous-loop-log-format.md](skills/09-for-sure/references/autonomous-loop-log-format.md) | - |
| `-` | [SKILL.md](skills/09-for-sure/SKILL.md) | `Run an iterative agent loop that retries until a runnable success condition passes. Use when the user says "for sure", "keep trying until", or wants guaranteed completion against a success command. Not for one-shot tasks or uncheckable goals.` |
| `actions` | [01-init-tracking.md](skills/09-goalify/actions/01-init-tracking.md) | - |
| `actions` | [02-auto-accept.md](skills/09-goalify/actions/02-auto-accept.md) | - |
| `actions` | [03-autonomous-loop.md](skills/09-goalify/actions/03-autonomous-loop.md) | - |
| `assets` | [autonomous-loop-worker-prompt.md](skills/09-goalify/assets/autonomous-loop-worker-prompt.md) | - |
| `assets` | [plan-template.md](skills/09-goalify/assets/plan-template.md) | - |
| `references` | [autonomous-loop-log-format.md](skills/09-goalify/references/autonomous-loop-log-format.md) | - |
| `-` | [SKILL.md](skills/09-goalify/SKILL.md) | `Turn a goal into an autonomous loop that replans and retries until a runnable success condition passes. Use when the user says "goalify", "keep trying until", or wants a goal verified by a command. Not for one-shot tasks or uncheckable goals.` |

#### `skills/10-todo`

Expand Down
4 changes: 2 additions & 2 deletions plugins/aidd-dev/README.md
Original file line number Diff line number Diff line change
Expand Up @@ -8,7 +8,7 @@ Code transformation plugin for the AI-Driven Development framework.

First time? Install with `/plugin install aidd-dev@aidd-framework`, then run `aidd-dev:01-plan`.

Covers code transformation: planning, implementation, assertions, audits, code review, testing, refactoring, debugging, for-sure, and parallel todo fan-out. Standalone Browser QA records short web evidence. Also hosts AI agents.
Covers code transformation: planning, implementation, assertions, audits, code review, testing, refactoring, debugging, goalify, and parallel todo fan-out. Standalone Browser QA records short web evidence. Also hosts AI agents.

## Skills

Expand All @@ -22,7 +22,7 @@ Covers code transformation: planning, implementation, assertions, audits, code r
| [2.6] | [test](skills/06-test/SKILL.md) | Write and iterate on tests until they pass, and validate user journeys end-to-end in the browser. |
| [2.7] | [refactor](skills/07-refactor/SKILL.md) | Optimize code for performance and fix security vulnerabilities following OWASP guidelines. |
| [2.8] | [debug](skills/08-debug/SKILL.md) | Reproduce and fix bugs systematically using test-driven workflow, root cause analysis, and hypothesis validation. |
| [2.9] | [for-sure](skills/09-for-sure/SKILL.md) | Iterative agent loop that tracks attempts and retries until a success condition is met. |
| [2.9] | [goalify](skills/09-goalify/SKILL.md) | Autonomous loop that replans and retries until a runnable success condition passes. |
| [2.10] | [todo](skills/10-todo/SKILL.md) | Split the prompt into independent todos, run one executor agent per todo in parallel, then report a minimal table. |
| [2.11] | [browser-qa](skills/11-browser-qa/SKILL.md) | Record one short named video for a locked browser happy path and each sourced browser edge case. |

Expand Down
37 changes: 0 additions & 37 deletions plugins/aidd-dev/skills/09-for-sure/SKILL.md

This file was deleted.

This file was deleted.

55 changes: 55 additions & 0 deletions plugins/aidd-dev/skills/09-goalify/SKILL.md
Original file line number Diff line number Diff line change
@@ -0,0 +1,55 @@
---
name: 09-goalify
description: Turn a goal into an autonomous loop that replans and retries until a runnable success condition passes. Use when the user says "goalify", "keep trying until", or wants a goal verified by a command. Not for one-shot tasks or uncheckable goals.
argument-hint: task | command
---

# Skill: goalify

Frame a checkable goal interactively, then run unattended until its success condition passes or a safety stop is reached.

```mermaid
flowchart TD
Start[Setup or resume] --> Init[init-tracking]
Init -->|ready| Loop[autonomous-loop under auto-accept]
Init -->|already implemented| Done[Complete]
Init -->|unresolved prerequisite| Stop[Stop and report]
Loop --> Worker[Execute next step]
Worker -->|safety stop| Stop
Worker --> Verify[Verify evidence]
Verify -->|failure| Replan[Analyze and replan]
Replan -->|retry| Loop
Verify -->|more steps| Loop
Verify -->|all steps checked| Success{success_condition passes?}
Success -->|no| Replan
Success -->|yes| Done
```

## Actions

| Action | Does |
| --- | --- |
| [init-tracking](actions/01-init-tracking.md) | frame the goal, create or resume tracking, launch the loop |
| [auto-accept](actions/02-auto-accept.md) | decide and act within the task's safety limits |
| [autonomous-loop](actions/03-autonomous-loop.md) | dispatch workers, verify evidence, replan failures, check completion |

Run `init-tracking` interactively; it launches or resumes `autonomous-loop` under `auto-accept`.
Before running an action, read its file in `actions/`, not only the table or assets.

## Transversal rules

- Single source of truth: all task state lives in `aidd_docs/tasks/<task-name>.md` and nowhere else.
- No repeated failures: never retry a failed approach without a meaningful change.
- Honesty over escape: never set `status: implemented` until the success condition genuinely passes.
- Auto-accept: follow the action's rules within the original task; stop on payment or destructive actions.
- The loop spawns one worker agent per step and never does the work itself.
- Model policy: use a powerful available model for framing, planning, verification, and replanning, including every orchestrator launch or resume. At every worker launch or relaunch, use the smallest available model with its highest supported reasoning effort. Workers execute only their assigned step and return evidence.

## Assets

- `assets/plan-template.md`: the tracking file format (frontmatter, phases, acceptance criteria, Log).
- `assets/autonomous-loop-worker-prompt.md`: the prompt the loop spawns each per-step worker with.

## References

- `references/autonomous-loop-log-format.md`: the Log entry format the loop appends per attempt.
Loading