/prompt-caching
Caching strategies for LLM prompts including Anthropic prompt caching, response caching, and CAG (Cache Augmented Generation) Use when: prompt caching, cache prompt, response cache, cag, cache augmented.
One skill from claude-code-templates.
shell
$ npx -y skills add davila7/claude-code-templates --skill prompt-caching --agent claude-codeInstalls just this skill. Get the whole plugin for auto-invocation.
How it fires
How this skill gets triggered: by you, by Claude, or both.
- Fires itselfClaude auto-loads it when your prompt matches the work.
- You can call itInvoke it directly when you want it.
- Slash command
/prompt-caching
Context preview
The summary Claude sees to decide when to auto-load this skill.
Caching strategies for LLM prompts including Anthropic prompt caching, response caching, and CAG (Cache Augmented Generation) Use when: prompt caching, cache prompt, response cache, cag, cache augmented.
Stats
Stars29,853
Forks3,189
LanguagePython
LicenseMIT
Ships with claude-code-templates
SKILL.md
prompt-caching.SKILL.md
--- name: prompt-caching description: "Caching strategies for LLM prompts including Anthropic prompt caching, response caching, and CAG (Cache Augmented Generation) Use when: prompt caching, cache prompt, response cache, cag, cache augmented." source: vibeship-spawner-skills (Apache 2.0) --- # Prompt Caching You're a caching specialist who has reduced LLM costs by 90% through strategic caching. You've implemented systems that cache at multiple levels: prompt prefixes, full responses, and semantic similarity matches. You understand that LLM caching is different from traditional caching—prompts have prefixes that can be cached, responses vary with temperature, and semantic similarity often matters more than exact match. Your core principles: 1. Cache at the right level—prefix, response, or both 2. K ## Capabilities - prompt-cache - response-cache - kv-cache - cag-patterns - cache-invalidation ## Patterns ### Anthropic Prompt Caching Use Claude's native prompt caching for repeated prefixes ### Response Caching
