nw-ab-critique-dimensi…
Review dimensions for validating agent quality - template compliance, safety, testing, and priority validation
Runs a timeboxed PROBE to validate one core assumption, then optionally PROMOTES the probe into a walking skeleton — the first e2e thin slice of the feature, committed and demo-able. Use after DISCUSS when the feature involves a new mechanism, performance requirement, or
$ npx -y skills add nWave-ai/nWave --skill nw-spike --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/nw-spikeContext preview
The summary Claude sees to decide when to auto-load this skill.
Runs a timeboxed PROBE to validate one core assumption, then optionally PROMOTES the probe into a walking skeleton — the first e2e thin slice of the feature, committed and demo-able. Use after DISCUSS when the feature involves a new mechanism, performance requirement, or
name: nw-spike description: "Runs a timeboxed PROBE to validate one core assumption, then optionally PROMOTES the probe into a walking skeleton — the first e2e thin slice of the feature, committed and demo-able. Use after DISCUSS when the feature involves a new mechanism, performance requirement, or external integration." user-invocable: true argument-hint: "[feature-description] - Example: \"wave-matrix -- derive feature status from pytest + filesystem\""
**Wave**: SPIKE (between DISCUSS and DESIGN) | **Agent**: Attila (nw-software-crafter) | **Command**: `/nw-spike`
Execute a two-phase wave that turns a risky assumption into visible, iterable value as fast as possible:
1. **PROBE** — quick throwaway validation of one core assumption (30-60 min, code in `/tmp/`) 2. **PROMOTION GATE** (interactive) — ask the user whether to promote the probe 3. **WALKING SKELETON** — refactor the probe into an end-to-end thin slice committed to the repository (1-3 h, code in `src/` + 1 acceptance test)
The PROBE answers "does the mechanism work?". The WALKING SKELETON answers "can a user see it working end-to-end?". You never throw away working validated code — you promote it and iterate.
The spike is needed when the feature introduces:
If none of the above apply, skip SPIKE and go to DESIGN.
1. **DISCUSS artifacts**: Read `docs/feature/{feature-id}/discuss/` (required)
2. **DIVERGE artifacts**: Read `docs/feature/{feature-id}/diverge/recommendation.md` (if present)
**Question**: What is the ONE assumption you need to validate? **Examples**:
**Question**: What is the timing constraint? (Enter "none" if mechanism validation only) **Examples**:
**Question**: If this probe works, what would the thinnest end-to-end slice look like? Capture the rough path: `user-facing entry → business logic → persistence/services → user-visible output`. This is **not** a commitment — it's context for the promotion gate later.
Throwaway validation of the assumption.
@nw-software-crafter
Execute PROBE for "{feature-description}".
**Probe question**: {Decision 1 answer} **Performance budget**: {Decision 2 answer} **Target e2e path (for later)**: {Decision 3 answer}
**Rules**:
**After probe completes**: 1. Write findings to `docs/feature/{feature-id}/spike/findings.md` — binary verdict (WORKS / DOESN'T WORK), timing, edge cases, design implications. 2. Do **not** delete the probe code yet — wait for the promotion gate. 3. Report verdict and ask the orchestrator to run the promotion gate.
Run this gate **only after** the probe completes and findings.md is written.
Present the user with three choices:
| Choice | When to pick | Outcome | |---|---|---| | **PROMOTE** | Probe verdict is WORKS and the mechanism is worth building on | Proceed to Phase 3 — walking skeleton | | **DISCARD** | Probe verdict is WORKS but not worth pursuing (findings are enough) | Delete `/tmp/spike_{feature_id}/`. Commit `findings.md`. Hand off to DESIGN. | | **PIVOT** | Probe verdict is DOESN'T WORK or revealed a better approach | Delete probe code. Annotate `findings.md` with the pivot. Either loop back to DISCUSS or run a second probe. |
**Default**: if the probe verdict is WORKS and no reason to stop, recommend PROMOTE but let the user override.
Record the promotion decision in `docs/feature/{feature-id}/spike/wave-decisions.md` as an explicit wave decision with rationale.
Refactor the probe into the thinnest end-to-end slice that is committed, tested, and demo-able.
1. **End-to-end path**: the slice enters from a real user-facing entry point (CLI command, HTTP endpoint, UI action, hook) and exits at a real user-visible output (stdout, HTTP response, rendered screen, persisted file). Every layer in between is exercised — **no layer is mocked** unless that layer is an external paid service classified as costly in DISTILL's Walking Skeleton Strategy (then use the fake/contract test pattern). 2. **One acceptance test**: a `@walking_skeleton @driving_port` tagged scenario in `tests/{test-type-path}/{feature-id}/acceptance/walking-skeleton.feature`. The scenario MUST be green before hand-off. 3. **Production location**: code lives under `src/{production-path}/`, not in `/tmp/`. Minimal module skeleton is fine — no premature abstractions, no features beyond the walking skeleton. 4. **Committed**: the walking skeleton commit message is `feat({feature-id}): walking skeleton — {one-line description}`. 5. **Demo-able**: running the single acceptance test (or the real entry-point command) produces visible output that matches the user story from DISCUSS. 6. **Back-propagation**: if building the skeleton reveals a contradiction with DISCUSS or DESIGN, write the contradiction to `docs/feature/{feature-id}/spike/upstream-issues.md` and stop
AI agents that guide you from idea to working code, with human judgment at every gate. nWave runs inside Claude Code. It breaks feature delivery into seven waves (discover, diverge, discuss, design, devops, distill, deliver).
Repo: nWave-ai/nWave
Review dimensions for validating agent quality - template compliance, safety, testing, and priority validation
Review dimensions for validating agent quality - template compliance, safety, testing, and priority validation
Review dimensions for acceptance test quality - happy path bias, GWT compliance, business language purity, coverage completeness, walking skeleton…
Detailed 5-phase workflow for creating agents - from requirements analysis through validation and iterative refinement
5-layer testing approach for agent validation including adversarial testing, security validation, and prompt injection resistance
Architectural style selection decision matrices, trade-off analysis, structural enforcement rules, and combination patterns. Load when choosing or evaluating…