backprop
Bug → spec protocol. When a bug is found or a test fails, trace the cause, decide whether a new §V invariant would catch recurrence, append to §B. This is the…
Calibrated interrogation of a fuzzy idea before it becomes a spec. Asks one question at a time, recommends an answer, and lands each answer in §G (goal) or §C (constraints) — unknowns parked as `?` items, never guessed. The cheapest place to kill a bad idea is before §T exists.
$ npx -y skills add JuliusBrussee/cavekit --skill grill --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/grillContext preview
The summary Claude sees to decide when to auto-load this skill.
Calibrated interrogation of a fuzzy idea before it becomes a spec. Asks one question at a time, recommends an answer, and lands each answer in §G (goal) or §C (constraints) — unknowns parked as `?` items, never guessed. The cheapest place to kill a bad idea is before §T exists.
name: grill description: | Calibrated interrogation of a fuzzy idea before it becomes a spec. Asks one question at a time, recommends an answer, and lands each answer in §G (goal) or §C (constraints) — unknowns parked as `?` items, never guessed. The cheapest place to kill a bad idea is before §T exists. Triggers when the user has a vague idea, says "grill me", "stress-test this", "challenge my plan", "interview me before I spec", or invokes /ck:grill. Defers the actual write to the spec skill.
**One question at a time. Every answer lands in a § or gets parked `?`. Never guess a constraint into existence.**
Plan-then-execute guesses the fuzzy parts & builds the wrong thing. Grill drags the fuzz into §G/§C *before* a single §T row exists. A bad assumption caught here costs one question. Caught in §B it costs a bug.
Skip for a typo or a one-line fix. Grill scales to uncertainty, ⊥ to ego.
One opening read, not a quiz: 1. How well does user know this domain? (sets question depth) 2. How locked is the idea? (exploring vs committed) 3. Pressure wanted: light / normal / brutal.
Match it. Brutal grilling on a half-formed idea just demoralizes. Light grilling on a committed plan misses the load-bearing flaw.
Climb in order. Each rung, ask **one** question, **recommend** an answer, wait.
1. **Goal** — what must the code *do*, in one line? (→ §G) 2. **Done** — how do we know it works? name the observable. (→ §C / future §V) 3. **Boundary** — what is explicitly out of scope? (→ §C) 4. **Lock** — what tech/lib/pattern is non-negotiable? what is forbidden? (→ §C) 5. **Surface** — what does the outside world touch — cmd, api, file, env? (→ §I) 6. **Edge** — the one input that breaks the happy path? (→ future §V) 7. **Unknown** — what do we *not* know yet? (→ park as `?` §C bullet)
Stop climbing the moment the spec would be unambiguous. Do not ask all seven by reflex.
Each question carries a recommended answer so the user can grunt "yes" & move:
> Q: auth — session cookie or JWT? > rec: JWT — stateless, you named horizontal scaling as a §C. > (a) JWT (b) cookie (c) something else?
When done, emit a compact block — goal line, constraint bullets, surfaced unknowns as `?` — and hand to the **spec** skill to write §G/§C. Grill proposes; spec is the sole mutator. Never write SPEC.md directly.
Done when ALL hold:
Unresolved blocking unknown that needs the outside world → recommend `/research`, not a guess.
Frozen — compressed spec-driven development plugin for Claude Code. Still works; active development moved to JuliusBrussee/caveman.
Bug → spec protocol. When a bug is found or a test fails, trace the cause, decide whether a new §V invariant would catch recurrence, append to §B. This is the…
Plan-then-execute implementation against SPEC.md. Native single-thread loop, no sub-agents. On test or build failure, auto-invokes the backprop skill before…
Caveman encoding for SPEC.md and spec-adjacent writes. Loaded by /spec, /build, /check. Cuts tokens ~75% vs prose while staying precise. Triggers on any write…
Read-only drift detector. Diffs SPEC.md against current code and reports violations grouped by severity. Writes nothing — suggests remedies via the spec or…
Optional design-improvement pass for when you have spare usage to drain. Finds the shallowest modules in the code the spec touches, researches a deeper design,…
Gather external knowledge the spec needs and distill it into §R — the durable research log — so build grounds in facts instead of hallucinating library…