Skip to content
Development
Command

/battle-test

Deep audit of a skills directory against the Skill Creator standard. Produces a scored report and phased remediation plan.

From plugin
agent-skills-standard
57033 skills21 agents33 commands1 MCP
Install
$ npx -y skills add hoangnguyen0403/agent-skills-standard --agent claude-code

How it fires

How this command gets triggered: by you, by Claude, or both.

  • Fires itselfClaude auto-loads it when your prompt matches the work.
  • You can call itInvoke it directly when you want it.
  • Slash command/battle-test

Context preview

What this command does when you run it.

Deep audit of a skills directory against the Skill Creator standard. Produces a scored report and phased remediation plan.

Command definition

battle-test.md

Battle Test

Deep audit of a skills directory against the Skill Creator standard. Produces a scored report and phased remediation plan.

**Input:** $ARGUMENTS

Optional args: slug=<feature>, ticket=<id/url>, mode=interactive|autonomous|channel, channel=<id>, auto_continue=true|false, profile=business|hybrid|technical.

Instructions

Execute the following steps for **$ARGUMENTS**.

⚔️ Battle Test Orchestrator

> **Goal**: Evaluate every `SKILL.md` in the target directory against `common-skill-creator`. Deliver a quantified health report and prioritized remediation plan.

---

Step 1 — Target Discovery & Tech Stack

Identify the tech stack and all skill files.

# Count total skills per category
find . -name "SKILL.md" | sed 's|/[^/]*/SKILL.md||' | sort | uniq -c

---

Step 2 — Frontmatter Audit (Breadth Scan)

Run scans to detect format and structure violations.

1. **Check for missing mandatory sections**: `grep -rL "triggers:\|priority:\|Anti-Patterns" <SKILLS>/` 2. **Check for broad glob triggers**: `grep -r "src/\*\*" <SKILLS>/` 3. **Check for length limits**: `find . -name "SKILL.md" -exec awk 'END{if(NR>100) print FILENAME": "NR" lines"}' {} \;`

---

Step 3 — Deep Audit & Scoring

Pick every P0 (CRITICAL) and a random sample of P1/P2 skills. Evaluate them against the **Grading Rubric** in: `<SKILLS>/common/common-skill-creator/references/rubric.md` when synced.

1. **Trigger Accuracy**: File patterns + keywords? 2. **Format Quality**: `**No X**: Do Y.` anti-patterns? 3. **Verification**: Mandatory checklists? 4. **Token Efficiency**: Under 100 lines? Imperative mood?

---

Step 4 — Scored Report

**Scoring Algorithm**: Start at 100 points for each category. Apply deductions for findings (🔴-15 / 🟠-8 / 🟡-3 / 🔵-1).

📊 Report Format

Output the report using the **Battle Test Report** and **Phased Plan** templates in: `<SKILLS>/common/common-skill-creator/references/rubric.md` when synced.

---

Step 5 — Interactive Follow-up

1. "Generate a `task.md` for Phase 1 remediation?" 2. "Fix the worst offender in [category] now?" 3. "Deep-dive audit on a specific category (e.g., `security`)?"

Read more
Ships withagent-skills-standard

The portable SDLC standards layer for AI coding agents. Sync once, then work in your own runtime.

Get the whole plugin

Other commands on agent-skills-standard.