battle-test
Deep audit of a skills directory against the Skill Creator standard. Produces a scored report and phased remediation plan.
Convert delivery findings into skill, eval, workflow, and documentation improvements.
$ npx -y skills add hoangnguyen0403/agent-skills-standard --agent claude-codeHow it fires
How this command gets triggered: by you, by Claude, or both.
/retro-learnContext preview
What this command does when you run it.
Convert delivery findings into skill, eval, workflow, and documentation improvements.
Convert delivery findings into skill, eval, workflow, and documentation improvements.
**Input:** $ARGUMENTS
Optional args: slug=<feature>, ticket=<id/url>, mode=interactive|autonomous|channel, channel=<id>, auto_continue=true|false, profile=business|hybrid|technical.
Execute the following steps for **$ARGUMENTS**.
Goal: Turn defects, missed expectations, and delivery friction into durable standards improvements.
1. Gather evidence:
2. Classify:
3. Propose one targeted action per root cause:
4. Implement authorized candidates:
5. Evaluate and review:
6. Release and persist:
# Retro: [Name]
## Evidence
## Root Causes
| Finding | Category | Action |
| --- | --- | --- |
| [finding] | [category] | [action] |
## Skill Or Eval Updates
## Outcome Report
{schema_version: 1, run_id: "[run-id]", slug: "[slug]", workflow: retro-learn, feature_status: implemented, started_at: "[timestamp]", completed_at: "[timestamp]", requirement_trace: {brd_objectives: [], requirements: [], acceptance_criteria: [], srs: []}, completed_evidence: [], missing_evidence: [], decision_needed: [], recommended_next_workflow: null, cost: {source: unavailable}, agent: {identity: "[agent-identity]", model: "[model]"}}
## Next Workflow
## Follow-Ups
## Cost Report
Call `get_session_cost(workflow="retro-learn")` before final handoff.The portable SDLC standards layer for AI coding agents. Sync once, then work in your own runtime.
Repo: hoangnguyen0403/agent-skills-standard
Deep audit of a skills directory against the Skill Creator standard. Produces a scored report and phased remediation plan.
Clarify a rough product or engineering idea into a BRD-lite brief (Why) with measurable business value.
Run an AI-assisted PR code review using multi-layer lenses with confidence scoring.
Review an entire codebase for architecture, engineering health, and exploitable risk; generate a prioritized remediation plan, an evidence-anchored system…
Run a bounded cybersecurity exercise workflow with authorization, runtime, control, evidence, and independent adjudication gates; supports safe offline…
Controlled purple validation using paired action-observation evidence and explicit defensive outcomes.