/external-model-delegation
A `UserPromptSubmit` hook classifies every user prompt into one of six cost/performance tiers. The hook injects `additionalContext` instructing Claude to run a specific delegation script and return the output.
$ npx -y skills add alinaqi/claude-bootstrap --skill external-model-delegation --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
- Fires itselfAuto-invocation. Claude auto-loads it when your prompt matches the work.Auto-invocation is when the right skill fires by itself at the right moment, driven by a FLOW.md router and a hook, instead of you invoking it by name. It is the difference between a skill being installed and a skill actually getting used.Read the full definition →
- You can call itInvoke it directly when you want it.
- Slash command
/external-model-delegation
Context preview
The summary Claude sees to decide when to auto-load this skill.
A `UserPromptSubmit` hook classifies every user prompt into one of six cost/performance tiers. The hook injects `additionalContext` instructing Claude to run a specific delegation script and return the output.
SKILL.md
external-model-delegation.SKILL.mdExternal Model Delegation Pattern
A `UserPromptSubmit` hook classifies every user prompt into one of six cost/performance tiers. The hook injects `additionalContext` instructing Claude to run a specific delegation script and return the output.
Tier routing table
| Tier | Delegation command | Cost | |------|-------------------|------| | QWEN | `qwen3 "prompt"` | $0 (local Ollama) | | DEEPSEEK_FLASH | `deepseek --flash "prompt"` | $0.14 / $0.28 per M tokens | | DEEPSEEK_PRO | `deepseek --pro "prompt"` | $0.44 / $0.87 per M tokens | | KIMI | `kimi --quiet -p "prompt"` | $0.60 / $2.50 per M tokens | | CODEX | `codex exec` | varies | | CLAUDE | handle natively | $3-5 / $15-25 per M tokens |
Delegation script pattern
Each script is a self-contained executable in `~/bin/` that accepts a prompt and writes the response to stdout:
~/bin/
├── qwen3 # Shell: curl to local Ollama API
├── kimi # Shell: execs Kimi CLI binary
├── deepseek # Python: httpx to DeepSeek Anthropic-compat API
└── route-task # Shell + qwen3: classifies prompt into tier
Script contract
1. Accept prompt as first argument: `qwen3 "what is 2+2"` 2. Support `--flash` / `--pro` model flags (deepseek) 3. Support `--quiet` mode flag (kimi) 4. Write response to stdout, errors to stderr 5. Exit 0 on success, non-zero on error
Writing a new delegation script
#!/bin/bash
# Minimal delegator template
PROMPT="$1"
API_KEY="${EXTERNAL_API_KEY:-}"
# Call external API, write result to stdout
curl -s https://api.example.com/chat \
-H "Authorization: Bearer $API_KEY" \
-d "$(jq -n --arg p "$PROMPT" '{prompt: $p}')" \
| jq -r '.response'Routing hook flow
User types prompt
↓
UserPromptSubmit hook fires
↓
qwen3 classifies into tier (QWEN|DEEPSEEK_FLASH|DEEPSEEK_PRO|KIMI|CODEX|CLAUDE)
↓
Hook injects additionalContext: "Run: <delegation-command>"
↓
Claude reads context, spawns delegation script, returns output
↓
User sees response from the delegated modelClassification tiers
| Tier | Task types | |------|-----------| | QWEN | grep, find, regex, shell, syntax lookups, log reading, short summaries | | DEEPSEEK_FLASH | Simple code, boilerplate, CRUD, test writing, small fixes, config | | DEEPSEEK_PRO | Multi-file features, refactors, debugging, medium coding, docs | | KIMI | Single-file review, medium reasoning, commit messages, diff summaries | | CODEX | Bulk generation, mechanical changes across many files | | CLAUDE | Architecture, security, complex debugging, system design, quality-critical |
Environment
# Required env vars (set in ~/.zshrc)
export DEEPSEEK_API_KEY="sk-..." # For deepseek delegator
export OPENAI_API_KEY="sk-..." # For codex CLI
# Ollama must be running locally for qwen3 classification + delegation
Read more
External Model Delegation Pattern
A `UserPromptSubmit` hook classifies every user prompt into one of six cost/performance tiers. The hook injects `additionalContext` instructing Claude to run a specific delegation script and return the output.
Tier routing table
| Tier | Delegation command | Cost | |------|-------------------|------| | QWEN | `qwen3 "prompt"` | $0 (local Ollama) | | DEEPSEEK_FLASH | `deepseek --flash "prompt"` | $0.14 / $0.28 per M tokens | | DEEPSEEK_PRO | `deepseek --pro "prompt"` | $0.44 / $0.87 per M tokens | | KIMI | `kimi --quiet -p "prompt"` | $0.60 / $2.50 per M tokens | | CODEX | `codex exec` | varies | | CLAUDE | handle natively | $3-5 / $15-25 per M tokens |
Delegation script pattern
Each script is a self-contained executable in `~/bin/` that accepts a prompt and writes the response to stdout:
~/bin/ ├── qwen3 # Shell: curl to local Ollama API ├── kimi # Shell: execs Kimi CLI binary ├── deepseek # Python: httpx to DeepSeek Anthropic-compat API └── route-task # Shell + qwen3: classifies prompt into tier
Script contract
1. Accept prompt as first argument: `qwen3 "what is 2+2"` 2. Support `--flash` / `--pro` model flags (deepseek) 3. Support `--quiet` mode flag (kimi) 4. Write response to stdout, errors to stderr 5. Exit 0 on success, non-zero on error
Writing a new delegation script
#!/bin/bash
# Minimal delegator template
PROMPT="$1"
API_KEY="${EXTERNAL_API_KEY:-}"
# Call external API, write result to stdout
curl -s https://api.example.com/chat \
-H "Authorization: Bearer $API_KEY" \
-d "$(jq -n --arg p "$PROMPT" '{prompt: $p}')" \
| jq -r '.response'Routing hook flow
User types prompt
↓
UserPromptSubmit hook fires
↓
qwen3 classifies into tier (QWEN|DEEPSEEK_FLASH|DEEPSEEK_PRO|KIMI|CODEX|CLAUDE)
↓
Hook injects additionalContext: "Run: <delegation-command>"
↓
Claude reads context, spawns delegation script, returns output
↓
User sees response from the delegated modelClassification tiers
| Tier | Task types | |------|-----------| | QWEN | grep, find, regex, shell, syntax lookups, log reading, short summaries | | DEEPSEEK_FLASH | Simple code, boilerplate, CRUD, test writing, small fixes, config | | DEEPSEEK_PRO | Multi-file features, refactors, debugging, medium coding, docs | | KIMI | Single-file review, medium reasoning, commit messages, diff summaries | | CODEX | Bulk generation, mechanical changes across many files | | CLAUDE | Architecture, security, complex debugging, system design, quality-critical |
Environment
# Required env vars (set in ~/.zshrc) export DEEPSEEK_API_KEY="sk-..." # For deepseek delegator export OPENAI_API_KEY="sk-..." # For codex CLI # Ollama must be running locally for qwen3 classification + delegation
Turn Claude Code into a self-reviewing, test-enforced engineering system that remembers context across sessions — then route work across 13 models from a single dashboard.
Repo: alinaqi/claude-bootstrap
Other skills on maggy.
- /aeo-optimization
AI Engine Optimization - semantic triples, page templates, content clusters for AI citations
Open skill - /agent-teams
Claude Code Agent Teams - default team-based development with strict TDD pipeline enforcement
Open skill - /agentic-development
Build AI agents with Pydantic AI (Python) and Claude SDK (Node.js)
Open skill - /ai-models
Latest AI models reference - Claude, OpenAI, Gemini, Eleven Labs, Replicate
Open skill - /android-java
Android Java development with MVVM, ViewBinding, and Espresso testing
Open skill - /android-kotlin
Android Kotlin development with Coroutines, Jetpack Compose, Hilt, and MockK testing
Open skill

