prompt-evaluation-runn…
Use when evaluating prompts, LLM outputs, red-team suites, or model behavior with local eval configs and safe provider/cost controls.
Design high-quality MCP servers around workflows, narrow schemas, context-aware outputs, and actionable errors. Use when building or reviewing MCP tools for real agent tasks.
$ npx -y skills add yeaight7/agent-powerups --skill mcp-server-builder --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/mcp-server-builderContext preview
The summary Claude sees to decide when to auto-load this skill.
Design high-quality MCP servers around workflows, narrow schemas, context-aware outputs, and actionable errors. Use when building or reviewing MCP tools for real agent tasks.
name: mcp-server-builder description: Design high-quality MCP servers around workflows, narrow schemas, context-aware outputs, and actionable errors. Use when building or reviewing MCP tools for real agent tasks.
Use this skill when designing or implementing an MCP server.
tool name: stable, verb-noun, describes the workflow step input schema: typed, narrow, required fields only + optional detail flags output shape: consistent structure across all tools in the server failure modes: named error codes + correction hint
Bad error: `"Error: 404 Not Found"`
Good error: `"Resource 'project-123' not found. Use list_projects to see available project IDs."`
Every error should tell the agent its next valid action.
Use the Python and Node references only for the stack you are actually shipping.
Use them as optional helpers, not mandatory runtime requirements.
Curated power-ups for coding agents: skills, slash commands, MCP configs, hooks, AGENTS.md templates, and workflows for serious software engineering. Claude Code, Codex, Antigravity CLI, Cursor and more
Repo: yeaight7/agent-powerups
Use when evaluating prompts, LLM outputs, red-team suites, or model behavior with local eval configs and safe provider/cost controls.
Use when creating or reviewing red-team eval plugins, attack templates, grader rubrics, safety fixtures, or model-risk test metadata.
Use when designing, running, debugging, or hardening deterministic eval suites for agent skills, prompts, tool workflows, or MCP-backed cases.
Use when designing tool definitions for a new agent or subagent, an agent shows high retry rates, ambiguous tool invocations, or silent failures, or an…
Use when routing a prompt to a local provider CLI for a second opinion, review, or plan -- you are about to call a provider directly, need the response saved…
Use when starting work in an unfamiliar area of a codebase, spawning a subagent that needs targeted file context, a first search pass missed the relevant file,…