prompt-evaluation-runn…
Use when evaluating prompts, LLM outputs, red-team suites, or model behavior with local eval configs and safe provider/cost controls.
Use when facing 2+ independent tasks that can be worked on without shared state or sequential dependencies
$ npx -y skills add yeaight7/agent-powerups --skill dispatching-parallel-agents --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/dispatching-parallel-agentsContext preview
The summary Claude sees to decide when to auto-load this skill.
Use when facing 2+ independent tasks that can be worked on without shared state or sequential dependencies
name: dispatching-parallel-agents description: Use when facing 2+ independent tasks that can be worked on without shared state or sequential dependencies
You delegate tasks to specialized agents with isolated context. By precisely crafting their instructions and context, you ensure they stay focused and succeed at their task. They should never inherit your session's context or history — you construct exactly what they need. This also preserves your own context for coordination work.
When you have multiple unrelated failures (different test files, different subsystems, different bugs), investigating them sequentially wastes time. Each investigation is independent and can happen in parallel.
**Core principle:** Dispatch one agent per independent problem domain. Let them work concurrently.
graph TD
MultipleFailures{"Multiple failures?"}
AreIndependent{"Are they independent?"}
SingleAgent["Single agent investigates all"]
OneAgent["One agent per problem domain"]
CanParallel{"Can they work in parallel?"}
Sequential["Sequential agents"]
Parallel["Parallel dispatch"]
MultipleFailures -->|yes| AreIndependent
AreIndependent -->|"no - related"| SingleAgent
AreIndependent -->|yes| CanParallel
CanParallel -->|yes| Parallel
CanParallel -->|"no - shared state"| Sequential**Use when:**
**Don't use when:**
Group failures by what's broken:
Each domain is independent — fixing tool approval doesn't affect abort tests.
Each agent gets:
Send ALL agent calls in a single message. This is the only way they run concurrently. Making one Agent call, waiting, then another is sequential.
When agents return:
Good agent prompts are: 1. **Focused** - One clear problem domain 2. **Self-contained** - All context needed to understand the problem 3. **Specific about output** - What should the agent return?
Fix the 3 failing tests in src/agents/agent-tool-abort.test.ts: 1. "should abort tool with partial output capture" - expects 'interrupted at' in message 2. "should handle mixed completed and aborted tools" - fast tool aborted instead of completed 3. "should properly track pendingToolCount" - expects 3 results but gets 0 These are timing/race condition issues. Your task: 1. Read the test file and understand what each test verifies 2. Identify root cause - timing issues or actual bugs? 3. Fix by replacing arbitrary timeouts with event-based waiting Do NOT just increase timeouts - find the real issue. Return: Summary of what you found and what you fixed.
**Too broad:** "Fix all the tests" — agent gets lost **Specific:** "Fix agent-tool-abort.test.ts" — focused scope
**No context:** "Fix the race condition" — agent doesn't know where **Context:** Paste the error messages and test names
**No constraints:** Agent might refactor everything **Constraints:** "Do NOT change production code"
**Vague output:** "Fix it" — you don't know what changed **Specific:** "Return summary of root cause and changes"
After agents return: 1. **Review each summary** — Understand what changed 2. **Check for conflicts** — Did agents edit same code? 3. **Run full suite** — Verify all fixes work together 4. **Spot check** — Agents can make systematic errors
Curated power-ups for coding agents: skills, slash commands, MCP configs, hooks, AGENTS.md templates, and workflows for serious software engineering. Claude Code, Codex, Antigravity CLI, Cursor and more
Repo: yeaight7/agent-powerups
Use when evaluating prompts, LLM outputs, red-team suites, or model behavior with local eval configs and safe provider/cost controls.
Use when creating or reviewing red-team eval plugins, attack templates, grader rubrics, safety fixtures, or model-risk test metadata.
Use when designing, running, debugging, or hardening deterministic eval suites for agent skills, prompts, tool workflows, or MCP-backed cases.
Use when designing tool definitions for a new agent or subagent, an agent shows high retry rates, ambiguous tool invocations, or silent failures, or an…
Use when routing a prompt to a local provider CLI for a second opinion, review, or plan -- you are about to call a provider directly, need the response saved…
Use when starting work in an unfamiliar area of a codebase, spawning a subagent that needs targeted file context, a first search pass missed the relevant file,…