/pr-review
Diff-based PR review across code quality, test coverage, silent failures, type design, and comment quality with severity-ranked findings. Triggers on: "review my PR", "review this code", "check my changes", "audit this PR", "code review". NOT for pre-landing gate, use
$ npx -y skills add Mathews-Tom/armory --skill pr-review --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
- Fires itselfAuto-invocation. Claude auto-loads it when your prompt matches the work.Auto-invocation is when the right skill fires by itself at the right moment, driven by a FLOW.md router and a hook, instead of you invoking it by name. It is the difference between a skill being installed and a skill actually getting used.Read the full definition →
- You can call itInvoke it directly when you want it.
- Slash command
/pr-review
Context preview
The summary Claude sees to decide when to auto-load this skill.
Diff-based PR review across code quality, test coverage, silent failures, type design, and comment quality with severity-ranked findings. Triggers on: "review my PR", "review this code", "check my changes", "audit this PR", "code review". NOT for pre-landing gate, use
SKILL.md
pr-review.SKILL.mdname: pr-review
description: 'Diff-based PR review across code quality, test coverage, silent failures, type design, and comment quality with severity-ranked findings. Triggers on: "review my PR", "review this code", "check my changes", "audit this PR", "code review". NOT for pre-landing gate, use pre-landing-review.'
metadata:
version: 1.1.1
category: review
tags: [code-review, pull-request, quality, diff-analysis]
difficulty: intermediate
phase: review
PR Review
Diff-based code review across five dimensions. Reads the changed files, selects applicable review methodologies, and produces an aggregated report with severity-ranked findings.
> **Native alternative:** Claude Code's `/ultrareview` runs a lightweight native bug-focused review (three free per month on Pro/Max plans at Opus 4.7's launch). Use this skill for five-dimension severity-ranked analysis (code quality + tests + error handling + types + comments) with file:line references; use `/ultrareview` for a quick bug-hunting pass on a diff.
Reference Files
| File | Contents | Load When | | ------------------------------- | ------------------------------------------------------- | ------------------------------- | | `references/code-review.md` | Guideline compliance, bug detection, confidence scoring | Always | | `references/test-analysis.md` | Behavioral test coverage, criticality rating | Test files changed | | `references/error-handling.md` | Silent failure patterns, catch block analysis | Error handling changed | | `references/type-design.md` | Invariant analysis, 4-dimension rating rubric | Type definitions added/modified | | `references/comment-quality.md` | Comment accuracy, long-term value, rot detection | Comments/docstrings added |
---
Workflow
Phase 1: Scope
1. Determine the review target:
- Default: `git diff` (unstaged changes)
- If user specifies a PR: `git diff main...HEAD` or `gh pr diff <number>`
- If user specifies files: review those files directly
2. List all changed files with `git diff --name-only` 3. Read the project's CLAUDE.md (if present) for project-specific rules
Phase 2: Route
Classify changed files and select applicable dimensions:
| Condition | Dimension | Reference to Load | | ---------------------------------------------------------------------------- | --------------- | ------------------------------- | | Always | Code review | `references/code-review.md` | | Files matching `*test*`, `*spec*`, `*_test.*`, `test_*` | Test analysis | `references/test-analysis.md` | | Files containing try/catch, except, .catch, Result, error callbacks | Error handling | `references/error-handling.md` | | Files containing class, interface, type, struct, enum, dataclass definitions | Type design | `references/type-design.md` | | Files with new/modified docstrings, JSDoc, or block comments | Comment quality | `references/comment-quality.md` |
Load only the reference files that apply. Skip dimensions with no matching files.
Phase 3: Review
For each applicable dimension, analyze the diff using the loaded methodology:
1. **Code review** — scan every changed file for guideline violations and bugs. Apply confidence scoring (0-100). Only report issues >= 80. 2. **Test analysis** — map test coverage to changed code paths. Rate gaps 1-10. Only report gaps >= 5. 3. **Error handling** — examine every error handler in the diff for silent failures. Classify CRITICAL/HIGH/MEDIUM. 4. **Type design** — evaluate new or modified types on 4 dimensions (encapsulation, invariant expression, usefulness, enforcement). Rate each 1-10. 5. **Comment quality** — verify accuracy, assess long-term value, flag comment rot.
Phase 4: Aggregate
Merge all findings into a single report, deduplicated and severity-ranked.
**Deduplication rules:**
- If two dimensions flag the same file:line, keep the higher-severity finding
- If code-review and error-handling both flag an empty catch block, merge into one
finding with the error-handling severity (it's the specialist)
**Severity mapping across dimensions:**
| Dimension | Maps to Critical | Maps to Important | Maps to Suggestion | | --------------- | ------------------- | ------------------------ | --------------------- | | Code review | Confidence 90-100 | Confidence 80-89 | — | | Test analysis | Rating 9-10 | Rating 7-8 | Rating 5-6 | | Error handling | CRITICAL | HIGH | MEDIUM | | Type design | Any rating <= 3/10 | Any rating 4-6/10 | Rating 7-8/10 | | Comment quality | Factually incorrect | Misleading or incomplete | Restates obvious code |
---
Output Format
# PR Review Summary
**Scope:** [X files changed, Y dimensions applied]
**Dimensions:** [list of active dimensions]
## Critical Issues (must fix before merge)
- **[dimension]** `file:line` — Description. Fix suggestion.
## Important Issues (should fix)
- **[dimension]** `file:line` — Description. Fix suggestion.
## Suggestions (consider)
- **[dimension]** `file:line` — Description.
## Strengths
- What's well-done in this changeset.
## Recommended Action
1. Fix critical issues
2. Address important issues
3. Consider suggestions
4. Re-run review after fixes
If no issues are found at any severity level, confirm the code meets standards with a brief summary of what was reviewed and which dimensions were applied.
---
Aspect Selection
Users can request specific dimensions ins
Read more
name: pr-review description: 'Diff-based PR review across code quality, test coverage, silent failures, type design, and comment quality with severity-ranked findings. Triggers on: "review my PR", "review this code", "check my changes", "audit this PR", "code review". NOT for pre-landing gate, use pre-landing-review.' metadata: version: 1.1.1 category: review tags: [code-review, pull-request, quality, diff-analysis] difficulty: intermediate phase: review
PR Review
Diff-based code review across five dimensions. Reads the changed files, selects applicable review methodologies, and produces an aggregated report with severity-ranked findings.
> **Native alternative:** Claude Code's `/ultrareview` runs a lightweight native bug-focused review (three free per month on Pro/Max plans at Opus 4.7's launch). Use this skill for five-dimension severity-ranked analysis (code quality + tests + error handling + types + comments) with file:line references; use `/ultrareview` for a quick bug-hunting pass on a diff.
Reference Files
| File | Contents | Load When | | ------------------------------- | ------------------------------------------------------- | ------------------------------- | | `references/code-review.md` | Guideline compliance, bug detection, confidence scoring | Always | | `references/test-analysis.md` | Behavioral test coverage, criticality rating | Test files changed | | `references/error-handling.md` | Silent failure patterns, catch block analysis | Error handling changed | | `references/type-design.md` | Invariant analysis, 4-dimension rating rubric | Type definitions added/modified | | `references/comment-quality.md` | Comment accuracy, long-term value, rot detection | Comments/docstrings added |
---
Workflow
Phase 1: Scope
1. Determine the review target:
- Default: `git diff` (unstaged changes)
- If user specifies a PR: `git diff main...HEAD` or `gh pr diff <number>`
- If user specifies files: review those files directly
2. List all changed files with `git diff --name-only` 3. Read the project's CLAUDE.md (if present) for project-specific rules
Phase 2: Route
Classify changed files and select applicable dimensions:
| Condition | Dimension | Reference to Load | | ---------------------------------------------------------------------------- | --------------- | ------------------------------- | | Always | Code review | `references/code-review.md` | | Files matching `*test*`, `*spec*`, `*_test.*`, `test_*` | Test analysis | `references/test-analysis.md` | | Files containing try/catch, except, .catch, Result, error callbacks | Error handling | `references/error-handling.md` | | Files containing class, interface, type, struct, enum, dataclass definitions | Type design | `references/type-design.md` | | Files with new/modified docstrings, JSDoc, or block comments | Comment quality | `references/comment-quality.md` |
Load only the reference files that apply. Skip dimensions with no matching files.
Phase 3: Review
For each applicable dimension, analyze the diff using the loaded methodology:
1. **Code review** — scan every changed file for guideline violations and bugs. Apply confidence scoring (0-100). Only report issues >= 80. 2. **Test analysis** — map test coverage to changed code paths. Rate gaps 1-10. Only report gaps >= 5. 3. **Error handling** — examine every error handler in the diff for silent failures. Classify CRITICAL/HIGH/MEDIUM. 4. **Type design** — evaluate new or modified types on 4 dimensions (encapsulation, invariant expression, usefulness, enforcement). Rate each 1-10. 5. **Comment quality** — verify accuracy, assess long-term value, flag comment rot.
Phase 4: Aggregate
Merge all findings into a single report, deduplicated and severity-ranked.
**Deduplication rules:**
- If two dimensions flag the same file:line, keep the higher-severity finding
- If code-review and error-handling both flag an empty catch block, merge into one
finding with the error-handling severity (it's the specialist)
**Severity mapping across dimensions:**
| Dimension | Maps to Critical | Maps to Important | Maps to Suggestion | | --------------- | ------------------- | ------------------------ | --------------------- | | Code review | Confidence 90-100 | Confidence 80-89 | — | | Test analysis | Rating 9-10 | Rating 7-8 | Rating 5-6 | | Error handling | CRITICAL | HIGH | MEDIUM | | Type design | Any rating <= 3/10 | Any rating 4-6/10 | Rating 7-8/10 | | Comment quality | Factually incorrect | Misleading or incomplete | Restates obvious code |
---
Output Format
# PR Review Summary **Scope:** [X files changed, Y dimensions applied] **Dimensions:** [list of active dimensions] ## Critical Issues (must fix before merge) - **[dimension]** `file:line` — Description. Fix suggestion. ## Important Issues (should fix) - **[dimension]** `file:line` — Description. Fix suggestion. ## Suggestions (consider) - **[dimension]** `file:line` — Description. ## Strengths - What's well-done in this changeset. ## Recommended Action 1. Fix critical issues 2. Address important issues 3. Consider suggestions 4. Re-run review after fixes
If no issues are found at any severity level, confirm the code meets standards with a brief summary of what was reviewed and which dimensions were applied.
---
Aspect Selection
Users can request specific dimensions ins
Curated, production-grade skills, agents, hooks, rules, commands, utilities, and presets for AI coding agents. No magic, no demos — battle-tested workflows built for developers who use AI seriously.
Repo: Mathews-Tom/armory
Other skills on armory.
- /adr-writer
Generates Architecture Decision Records capturing context, rationale, alternatives, and consequences in numbered status-tracked format. Triggers on: "write an ADR", "document this decision", "architecture decision record", "decision record", "design decision", "ADR for".
Open skill - /agent-builder
Build AI agents and automate Claude Code programmatically via the Claude Agent SDK and headless CLI mode. Covers Python SDK, claude -p, SDK MCP servers, hooks, sessions. Triggers on: "build an agent", "agent SDK", "headless mode", "automate Claude", "programmatic agent".
Open skill - /api-docs-generator
Audits and enhances FastAPI and REST API documentation: missing descriptions, response codes, examples, docstrings, Pydantic models, OpenAPI spec. Triggers on: "generate API docs", "document this API", "OpenAPI for", "FastAPI docs", "document endpoints", "swagger docs".
Open skill - /architecture-diagram
Generate layered architecture diagrams as self-contained HTML with inline SVG icons, CSS Grid containers, and connection overlays. Triggers on: "architecture diagram", "infra diagram", "system diagram", "deployment diagram", "topology", "draw architecture". NOT for architecture
Open skill - /architecture-reviewer
Architecture reviews across 7 dimensions (structural, scalability, enterprise readiness, performance, security, ops, data) with scored reports. Triggers on: "review architecture", "critique design", "audit system", "assess scalability", "enterprise readiness", "technical due
Open skill - /arxiv-figures
Optimize and prepare figures for arXiv submission: format conversion (EPS/PDF/PNG/JPG), size reduction, metadata stripping, processor compatibility (DVI vs PDFLaTeX). Triggers on: "optimize figures for arXiv", "reduce figure size", "convert figures for arXiv", "fix arXiv
Open skill

