agent-loop-ext
Crash-resilient external agent loop with state persistence and CI/CD integration
Validate documentation for unsupported claims, made-up metrics, and unverifiable statements
$ npx -y skills add jmagly/aiwg --skill claims-validator --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/claims-validatorContext preview
The summary Claude sees to decide when to auto-load this skill.
Validate documentation for unsupported claims, made-up metrics, and unverifiable statements
namespace: aiwg name: claims-validator platforms: [all] description: Validate documentation for unsupported claims, made-up metrics, and unverifiable statements
Validate documentation for unsupported claims, made-up metrics, and unverifiable statements.
Alternate expressions and non-obvious activations (primary phrases are matched automatically from the skill description):
This skill identifies statements that make claims without evidence, including:
When triggered, this skill:
1. **Scans for metric claims**:
2. **Identifies unsupported comparatives**:
3. **Checks for feature claims**:
4. **Validates citations**:
5. **Generates report**:
# Flagged "Time Saved: 92-96% (9-15 hours → 45-60 minutes)" "99x faster routing" "45x cache speedup" # Problem No benchmark data, methodology, or reproducible test # Fix Remove claim, or add: "Based on [benchmark/test], measured [how]"
# Flagged "Budget $20-50/month for moderate use" "Light usage: ~$10-20/month" "Enterprise teams may see $100-500+/month" # Problem No actual usage data, varies wildly by use case # Fix Remove specific numbers, or link to pricing calculator/methodology
# Flagged "Deploy Full SDLC Framework (2 Minutes)" "5 minutes, replaces 2-4 hours manual work" "campaign setup from 2-3 weeks → 1 week" # Problem No measurement, varies by project complexity # Fix Remove time claims, describe what it does instead
# Flagged "faster than manual processes" "more efficient than traditional approaches" "better than existing solutions" # Problem No specific comparison, no baseline defined # Fix Remove comparison, or specify exactly what's being compared
# Flagged "aiwg -migrate-workspace # Optional migration tool" "Run 'config-validator --fix' to apply automated fixes" # Problem Command doesn't exist in codebase # Fix Remove until implemented, or mark as "Planned:"
# Flagged "comprehensive", "revolutionary", "game-changing" "best-in-class", "industry-leading", "cutting-edge" "seamless", "effortless", "zero-friction" # Problem Subjective claims that can't be verified # Fix Replace with specific, factual descriptions
# Claims Validation Report **Document**: README.md **Date**: 2025-12-09 **Claims Found**: 12 **Issues**: 8 ## Summary | Category | Found | Unsupported | Action Needed | |----------|-------|-------------|---------------| | Metrics | 5 | 4 | Remove or cite | | Time estimates | 3 | 3 | Remove | | Cost estimates | 2 | 2 | Remove | | Comparatives | 1 | 1 | Rephrase | | Features | 1 | 1 | Remove (not implemented) | ## Issues ### 1. Unsupported Metric **Line 204**: "Time Savings: 20-98% reduction across 5 core use cases" **Problem**: No benchmark data or methodology **Action**: Remove claim ### 2. Cost Estimate Without Data **Line 1180**: "Light usage: ~$10-20/month" **Problem**: No actual usage data to support estimate **Action**: Remove specific numbers ### 3. Non-existent Feature **Line 583**: "aiwg -migrate-workspace" **Problem**: Command not implemented **Action**: Remove until implemented ## Recommendations 1. **Remove all percentage-based time savings claims** 2. **Remove cost estimates** - direct users to measure their own usage 3. **Remove non-existent commands** - only document what exists 4. **Replace superlatives with specifics** - say what it does, not how good it is
User: "Check README.md for unsupported claims" Output: "Claims Validation: README.md Found 8 unsupported claims: Metrics (4): - Line 204: '20-98% reduction' - no data - Line 362: '56-63% time saved' - no data - Line 588: '99x faster routing' - no benchmark - Line 589: '45x cache speedup' - no benchmark Time Estimates (2): - Line 187: '5 minutes, replaces 2-4 hours' - unmeasured - Line 366: '2 Minutes' in heading - varies by project Non-existent Features (2): - Line 583: 'aiwg -migrate-workspace' - not implemented - Line 651: 'aiwg -rollback-workspace' - not implemented Recommendation: Remove these claims. Describe what features do, not how much time/money they save."
User: "Any BS in this section?" Output: "Found 2 issues: 1. 'Revolutionary approach' - marketing speak, be specific 2. '10x productivity gain' - no measurement Suggest: Replace with factual descriptions of functionality."
This skill complements:
Reusable project context and specialist workflows for the AI tools you already use. Plan software, coordinate specialist reviews, prepare campaigns, investigate incidents, organize research, curate media, and maintain operational knowledge.
Repo: jmagly/aiwg
Crash-resilient external agent loop with state persistence and CI/CD integration
Detect requests for iterative autonomous agent loops and route to the appropriate loop executor
Automatically execute tests when code-generating agents modify source files, enforcing the execute-before-return pattern
Enable agent loops to learn from similar past tasks and share patterns across loops
Query and manage the executable feedback debug memory
Execute tests on generated code and iterate until passing