/anti-benchmark
Challenge industry best practices' hidden assumptions. Deconstruct benchmarks
$ npx -y skills add yogsoth-ai/de-anthropocentric-research-engine --skill anti-benchmark --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
- Fires itselfAuto-invocation. Claude auto-loads it when your prompt matches the work.Auto-invocation is when the right skill fires by itself at the right moment, driven by a FLOW.md router and a hook, instead of you invoking it by name. It is the difference between a skill being installed and a skill actually getting used.Read the full definition →
- You can call itInvoke it directly when you want it.
- Slash command
/anti-benchmark
Context preview
The summary Claude sees to decide when to auto-load this skill.
Challenge industry best practices' hidden assumptions. Deconstruct benchmarks
SKILL.md
anti-benchmark.SKILL.mdname: anti-benchmark
description: Challenge industry best practices' hidden assumptions. Deconstruct benchmarks
to reveal unexamined constraints.
execution: strategy
dependencies:
tactics:
- creative-ideation-assumption-enumeration
sops:
- benchmark-challenge
- constructive-rebellion
- creative-ideation-assumption-perturbation
- destruction-synthesis
- sacred-cow-identification
Anti-Benchmark
Challenge industry best practices' hidden assumptions.
State Ledger
| Resource | Target | Current | % | |----------|--------|---------|---| | web-search | 25 | 0 | 0% | | web-research | 10 | 0 | 0% | | paper-overview | 25 | 0 | 0% | | paper-search | 15 | 0 | 0% | | paper-research | 5 | 0 | 0% |
HARD-GATE
Cannot exit strategy until ≥80% of each budget line is consumed OR yield targets are met with justification for remaining budget.
Available Tactics
| Tactic | Role | |--------|------| | assumption-enumeration | Surface assumptions hidden in benchmarks |
Available SOPs
| SOP | Role | |-----|------| | benchmark-challenge | Identify and negate benchmark assumptions | | sacred-cow-identification | Find unquestioned beliefs behind best practices | | assumption-perturbation | Test what happens when benchmark assumptions fail | | constructive-rebellion | Build alternatives that violate benchmarks constructively | | destruction-synthesis | Synthesize anti-benchmark outputs |
Execution Guidance
1. **Identify benchmarks**: Catalog the industry best practices and standards in the domain 2. **Deconstruct**: Use benchmark-challenge to expose hidden assumptions in each 3. **Surface sacred cows**: Use sacred-cow-identification for deeper unquestioned beliefs 4. **Perturb**: Use assumption-perturbation to test "what if this standard is wrong?" 5. **Research alternatives**: Search for domains that succeed WITHOUT these benchmarks 6. **Build**: Use constructive-rebellion to form solutions that violate conventions productively 7. **Synthesize**: Produce structured output via destruction-synthesis
<!-- BEGIN available-tables (generated) -->
Available Tactics
Optional, no fixed order; the final leaf is always a sop.
| Tactic | When to use | | --- | --- | | creative-ideation-assumption-enumeration | Surface, perturb, and prioritize assumptions by disruption potential. Orchestrates assumption surfacing → perturbation → sacred cow identification → prioritization. |
Available SOPs
Optional, no fixed order; the final leaf is always a sop.
| SOP | When to use | | --- | --- | | benchmark-challenge | Identify and negate benchmark assumptions. Deconstruct best practices to reveal hidden constraints and open new spaces. | | constructive-rebellion | Build constructive alternatives from destructive negation. Transform violated assumptions into viable innovation directions. | | creative-ideation-assumption-perturbation | Perturb each assumption, observe system response. Systematic stress-testing of assumptions to reveal fragility and opportunity. | | destruction-synthesis | Synthesize all assumption destruction outputs into structured destructive innovation report. | | sacred-cow-identification | Find domain's unquestioned beliefs. Systematic identification of dogma that constrains innovation. |
<!-- END available-tables (generated) -->
Read more
name: anti-benchmark description: Challenge industry best practices' hidden assumptions. Deconstruct benchmarks to reveal unexamined constraints. execution: strategy dependencies: tactics: - creative-ideation-assumption-enumeration sops: - benchmark-challenge - constructive-rebellion - creative-ideation-assumption-perturbation - destruction-synthesis - sacred-cow-identification
Anti-Benchmark
Challenge industry best practices' hidden assumptions.
State Ledger
| Resource | Target | Current | % | |----------|--------|---------|---| | web-search | 25 | 0 | 0% | | web-research | 10 | 0 | 0% | | paper-overview | 25 | 0 | 0% | | paper-search | 15 | 0 | 0% | | paper-research | 5 | 0 | 0% |
HARD-GATE
Cannot exit strategy until ≥80% of each budget line is consumed OR yield targets are met with justification for remaining budget.
Available Tactics
| Tactic | Role | |--------|------| | assumption-enumeration | Surface assumptions hidden in benchmarks |
Available SOPs
| SOP | Role | |-----|------| | benchmark-challenge | Identify and negate benchmark assumptions | | sacred-cow-identification | Find unquestioned beliefs behind best practices | | assumption-perturbation | Test what happens when benchmark assumptions fail | | constructive-rebellion | Build alternatives that violate benchmarks constructively | | destruction-synthesis | Synthesize anti-benchmark outputs |
Execution Guidance
1. **Identify benchmarks**: Catalog the industry best practices and standards in the domain 2. **Deconstruct**: Use benchmark-challenge to expose hidden assumptions in each 3. **Surface sacred cows**: Use sacred-cow-identification for deeper unquestioned beliefs 4. **Perturb**: Use assumption-perturbation to test "what if this standard is wrong?" 5. **Research alternatives**: Search for domains that succeed WITHOUT these benchmarks 6. **Build**: Use constructive-rebellion to form solutions that violate conventions productively 7. **Synthesize**: Produce structured output via destruction-synthesis
<!-- BEGIN available-tables (generated) -->
Available Tactics
Optional, no fixed order; the final leaf is always a sop.
| Tactic | When to use | | --- | --- | | creative-ideation-assumption-enumeration | Surface, perturb, and prioritize assumptions by disruption potential. Orchestrates assumption surfacing → perturbation → sacred cow identification → prioritization. |
Available SOPs
Optional, no fixed order; the final leaf is always a sop.
| SOP | When to use | | --- | --- | | benchmark-challenge | Identify and negate benchmark assumptions. Deconstruct best practices to reveal hidden constraints and open new spaces. | | constructive-rebellion | Build constructive alternatives from destructive negation. Transform violated assumptions into viable innovation directions. | | creative-ideation-assumption-perturbation | Perturb each assumption, observe system response. Systematic stress-testing of assumptions to reveal fragility and opportunity. | | destruction-synthesis | Synthesize all assumption destruction outputs into structured destructive innovation report. | | sacred-cow-identification | Find domain's unquestioned beliefs. Systematic identification of dogma that constrains innovation. |
<!-- END available-tables (generated) -->
The complete research orchestration system for AI-native science. What It Does Design Philosophy Architecture (v3.2.2) Quick Start Configuration Roadmap License DARE is not a tool that helps you do research. It is the researcher.
Repo: yogsoth-ai/de-anthropocentric-research-engine
Other skills on de-anthropocentric-research-engine.
- /formated-results
Closing skill for the research-executor, loaded as the last step of formated-specs. Summarize the design just produced into one research-result JSON fenced block in your reply. Do not execute the research.
Open skill - /formated-specs
Spec-slot skill for the research-executor. Emit the 4-layer DARE orchestration of the assigned topic as one research-graph JSON fenced block in your reply. Replaces the generic spec-writing step.
Open skill - /injection-fidelity
Loss-1 judge (codex role). Given one sample's de-identified dialogue and its PolicyCard, decide axis-by-axis whether the user-simulator enacted the card's per-axis pressure. Judge enactment of the card, never whether the research is good.
Open skill - /ladder-quality-order
Loss-2 judge (codex role). Over one topic's 6 shuffled research-design samples, pairwise-rank by quality using the D1–D5 standard. Emit the pairwise log; the harness computes the order and the ladder verdicts. Judge quality difference, never against academic standards.
Open skill - /optimization-loop
The optimizer brain for the ladder-foundry pretraining loop. Runs the two-level nested batch loop, delegates gating to gate_eval, attributes a failing batch to one weight (attribute-first), and recovers from disk after compaction. Control flow is fully scripted; only the
Open skill - /acu-nugget-recall
Tactic: Extract atomic units from one paper and score how much of a caller-supplied summary covers. Use for ACU-style binary or Nugget-style ternary recall checks; cannot run without a target summary.
Open skill

