/common-llm-security
OWASP LLM Top 10 (2025) audit checklist for AI applications, agent tools, RAG pipelines, and prompt construction. Use when performing any security review touching LLM client code, prompt templates, agent tools, or vector stores.
$ npx -y skills add hoangnguyen0403/agent-skills-standard --skill common-llm-security --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
- Fires itselfAuto-invocation. Claude auto-loads it when your prompt matches the work.Auto-invocation is when the right skill fires by itself at the right moment, driven by a FLOW.md router and a hook, instead of you invoking it by name. It is the difference between a skill being installed and a skill actually getting used.Read the full definition →
- You can call itInvoke it directly when you want it.
- Slash command
/common-llm-security
Context preview
The summary Claude sees to decide when to auto-load this skill.
OWASP LLM Top 10 (2025) audit checklist for AI applications, agent tools, RAG pipelines, and prompt construction. Use when performing any security review touching LLM client code, prompt templates, agent tools, or vector stores.
SKILL.md
common-llm-security.SKILL.mdname: common-llm-security
description: OWASP LLM Top 10 (2025) audit checklist for AI applications, agent tools, RAG pipelines, and prompt construction. Use when performing any security review touching LLM client code, prompt templates, agent tools, or vector stores.
metadata:
triggers:
keywords:
- LLM security
- prompt injection
- agent security
- RAG security
- AI security
- openai
- anthropic
- langchain
- LLM reviewOWASP LLM Top 10 Security Checklist (2025)
**Priority: P0 (CRITICAL)**
Implementation Guidelines
- **Check LLM01 first**: Prompt injection #1 LLM finding — any user input concatenated directly into prompt string immediate P0.
- **Check LLM06 next**: Agent tools with write/delete/execute capabilities without confirmation P0.
- **Mark each item**: ✅ not affected | ⚠️ needs review | 🔴 confirmed finding.
- **P0 finding caps Security score at 40/100** — not skip any item.
- See [references/owasp-llm.md](references/owasp-llm.md) for full detection signals.
OWASP LLM Top 10 (2025)
| ID | Risk | Key Detection Signal | | ----- | ---- | -------------------- | | LLM01 | Prompt Injection | User input string-concatenated into prompt. Retrieved docs inserted into system turn. | | LLM02 | Sensitive Information Disclosure | PII or credentials passed into prompt context. LLM response logged without redaction. | | LLM03 | Supply Chain | Unverified model weights or plugins. Third-party agent added without trust review. | | LLM04 | Data & Model Poisoning | User-controlled data written to training sets or embedding stores without validation. | | LLM05 | Improper Output Handling | LLM output used directly in DOM sink, SQL query, shell command, or redirect URL. | | LLM06 | Excessive Agency | Agent tool with write/delete/network access — no human-in--loop confirmation. | | LLM07 | System Prompt Leakage | System prompt content returned via tool output, error message, or API response. | | LLM08 | Vector & Embedding Weaknesses | User text injected into vector store without sanitization. No tenant namespace isolation. | | LLM09 | Misinformation | LLM output used for critical decisions (medical, financial, legal) without verification. | | LLM10 | Unbounded Consumption | No `max_tokens` on LLM call. No rate limit on invocations. Agent loop without depth cap. |
Anti-Patterns
- **No prompt concat**: Pass user input as separate `user` turn, never interpolated into system prompts.
- **No raw LLM output in sinks**: Sanitize LLM responses before writing to DOM, queries, or shell.
- **No uncapped agent loops**: Every agentic recursion must enforce max iteration/depth limit.
References
- [OWASP LLM — Full Detection Signals](references/owasp-llm.md) — load when auditing any LLM client code
Canonical response anchors
When this skill applies, preserve the following domain terminology or equivalent concrete examples in the answer when relevant:
- sanitize
Read more
name: common-llm-security
description: OWASP LLM Top 10 (2025) audit checklist for AI applications, agent tools, RAG pipelines, and prompt construction. Use when performing any security review touching LLM client code, prompt templates, agent tools, or vector stores.
metadata:
triggers:
keywords:
- LLM security
- prompt injection
- agent security
- RAG security
- AI security
- openai
- anthropic
- langchain
- LLM reviewOWASP LLM Top 10 Security Checklist (2025)
**Priority: P0 (CRITICAL)**
Implementation Guidelines
- **Check LLM01 first**: Prompt injection #1 LLM finding — any user input concatenated directly into prompt string immediate P0.
- **Check LLM06 next**: Agent tools with write/delete/execute capabilities without confirmation P0.
- **Mark each item**: ✅ not affected | ⚠️ needs review | 🔴 confirmed finding.
- **P0 finding caps Security score at 40/100** — not skip any item.
- See [references/owasp-llm.md](references/owasp-llm.md) for full detection signals.
OWASP LLM Top 10 (2025)
| ID | Risk | Key Detection Signal | | ----- | ---- | -------------------- | | LLM01 | Prompt Injection | User input string-concatenated into prompt. Retrieved docs inserted into system turn. | | LLM02 | Sensitive Information Disclosure | PII or credentials passed into prompt context. LLM response logged without redaction. | | LLM03 | Supply Chain | Unverified model weights or plugins. Third-party agent added without trust review. | | LLM04 | Data & Model Poisoning | User-controlled data written to training sets or embedding stores without validation. | | LLM05 | Improper Output Handling | LLM output used directly in DOM sink, SQL query, shell command, or redirect URL. | | LLM06 | Excessive Agency | Agent tool with write/delete/network access — no human-in--loop confirmation. | | LLM07 | System Prompt Leakage | System prompt content returned via tool output, error message, or API response. | | LLM08 | Vector & Embedding Weaknesses | User text injected into vector store without sanitization. No tenant namespace isolation. | | LLM09 | Misinformation | LLM output used for critical decisions (medical, financial, legal) without verification. | | LLM10 | Unbounded Consumption | No `max_tokens` on LLM call. No rate limit on invocations. Agent loop without depth cap. |
Anti-Patterns
- **No prompt concat**: Pass user input as separate `user` turn, never interpolated into system prompts.
- **No raw LLM output in sinks**: Sanitize LLM responses before writing to DOM, queries, or shell.
- **No uncapped agent loops**: Every agentic recursion must enforce max iteration/depth limit.
References
- [OWASP LLM — Full Detection Signals](references/owasp-llm.md) — load when auditing any LLM client code
Canonical response anchors
When this skill applies, preserve the following domain terminology or equivalent concrete examples in the answer when relevant:
- sanitize
The portable SDLC standards layer for AI coding agents. Sync once, then work in your own runtime.
Repo: hoangnguyen0403/agent-skills-standard
Other skills on agent-skills-standard.
- /android-agp-upgrade
Upgrade an Android project to Android Gradle Plugin (AGP) 9. Use when migrating to AGP 9, updating Gradle build files, migrating to built-in Kotlin, or adopting the new AGP DSL.
Open skill - /android-architecture
Apply Clean Architecture layering, modularization, and Unidirectional Data Flow in Android projects. Use when setting up project structure, placing code in layers, configuring feature/core modules, or implementing UDF patterns; defer Compose state and ViewModel/StateFlow
Open skill - /android-background-work
Implement WorkManager and background processing correctly on Android. Use when creating Worker classes, scheduling tasks, choosing between WorkManager and Foreground Services, or setting up Hilt in workers; defer FCM and notification delivery to android-notifications.
Open skill - /android-compose-migration
Migrate an Android XML View to Jetpack Compose following a structured 10-step workflow. Use when converting XML layouts to Compose, setting up Compose in an existing View-based project, or incrementally adopting Compose.
Open skill - /android-compose
Build high-performance declarative UI with Jetpack Compose. Use when writing Composable functions, optimizing recomposition, hoisting state, or working with LazyColumn and side effects; defer deep-link and navigation routing to android-navigation.
Open skill - /android-concurrency
Write correct coroutine scopes, lifecycle collection, and dispatcher injection in Android production code. Use for suspend functions, coroutine scopes, and dispatcher mechanics; defer ViewModel StateFlow/LiveData architecture, Fragment lifecycle recipes,
Open skill

