Skip to content
Development
Skill

/storm-research

Use when researching any topic, question, market, company, technology, or decision — before writing reports or proposals, making business or investment decisions, preparing for negotiations, interviews, or presentations, or learning a new field. This is the default research

From plugin
operating-core
37 skills1 agent
Install
$ npx -y skills add josherau/claude-operating-core --skill storm-research --agent claude-code

How it fires

How this skill gets triggered: by you, by Claude, or both.

  • Fires itselfAuto-invocation. Claude auto-loads it when your prompt matches the work.Auto-invocation is when the right skill fires by itself at the right moment, driven by a FLOW.md router and a hook, instead of you invoking it by name. It is the difference between a skill being installed and a skill actually getting used.Read the full definition →
  • You can call itInvoke it directly when you want it.
  • Slash command/storm-research

Context preview

The summary Claude sees to decide when to auto-load this skill.

Use when researching any topic, question, market, company, technology, or decision — before writing reports or proposals, making business or investment decisions, preparing for negotiations, interviews, or presentations, or learning a new field. This is the default research

SKILL.md

storm-research.SKILL.md
name: storm-research
description: Use when researching any topic, question, market, company, technology, or decision — before writing reports or proposals, making business or investment decisions, preparing for negotiations, interviews, or presentations, or learning a new field. This is the default research method for all research tasks. Skip only for single-fact lookups, or when the user explicitly asks for a quick answer or names a different method.

STORM Research

Overview

Multi-perspective research method adapted from Stanford's STORM (Shao et al., NAACL 2024, github.com/stanford-oval/storm). Core principle: **one query returns the majority view; five grounded perspectives plus a contradiction map returns understanding.** The STORM paper reports ~25% better organization and ~10% broader coverage from multi-perspective question asking versus its single-pass baseline.

**The Iron Rule: every perspective must be grounded in real retrieval.** Persona-prompting without sources is five flavors of one model's blind spots hallucinating in harmony. Each perspective's claims come from actual searches (WebSearch, WebFetch, or whatever search MCPs you have wired up) with sources noted. No search → no claim.

The Four Phases

Run all four, in order. Do not stop after Phase 1 or 3.

Phase 1 — Multi-Perspective Scan

Pick 5 perspectives suited to the topic. Default set:

| Perspective | Asks | |---|---| | Practitioner | What actually happens day-to-day? What breaks in reality? | | Skeptic | What's overhyped or wrong? What could fail? | | Economist | Who profits? What incentives shape the claims? | | Historian | Has this pattern happened before? How did it end? | | Academic | What does the actual evidence/data say? |

Swap perspectives when the topic demands (e.g. regulator, customer, competitor, clinician). For each: generate that persona's 2-3 sharpest questions, then **search to answer them**. Note the source for each finding.

Phase 2 — Contradiction Map

Find where the perspectives fight and why. Rules of evidence:

  • All 5 agree → probably true.
  • Perspectives conflict → the real understanding lives here; name the disagreement and the reason (incentive, timeframe, evidence quality).
  • Nobody addressed it → you found a gap; say so explicitly.

Phase 3 — Synthesis

One briefing: what's established (with sources), the named contradictions, reliability ranking of major claims, and a **specific recommended action** — not a summary that hedges everything.

Phase 4 — Peer Review

Adversarially grade your own output: strongest claim, weakest claim, likely source bias, missing perspective, and a confidence score per key conclusion. The original STORM pipeline has no self-critique step, so this phase is our addition — and it is not optional. (And per the review-panel rule: Phase 4 is still self-critique — the finished deliverable gets an independent reviewer too.)

Scaling

| Mode | When | How | |---|---|---| | Quick | Casual question, minutes matter | 3 perspectives, ~5 searches, all inline | | Standard (default) | Most research requests | 5 perspectives, 8-15 searches, inline or 2-3 parallel subagents | | Deep | "Thorough/comprehensive/audit" | One subagent per perspective in parallel, then synthesize; consider a dedicated deep-research harness if you run one |

Output Shape

Deliver: **Briefing** (synthesis + recommended action) → **Contradiction map** → **Perspective findings with sources** → **Peer review with confidence scores**. Briefing first — the reader gets the answer before the apparatus.

Common Mistakes

  • **Answering from training data with persona flavoring.** The perspectives structure the questions; retrieval answers them.
  • **Perspectives that converge.** If the skeptic sounds like the academic, you picked lazy personas — force genuine disagreement in framing.
  • **Skipping Phase 2.** The contradiction map is the step that separates this from a book report.
  • **Skipping Phase 4 to save time.** 60 seconds of self-critique is the method's error bar.
Read more
Ships withoperating-core

Quality gates for Claude Code. An agent should never grade its own work, outbound content should be pretested before it costs you money, research should be grounded in more than one perspective, and the reasoning behind a hard solve should outlive the session

Get the whole plugin
Stats
3
Stars
2
Forks
Maintained
Maintenance
Shell
Language
MIT
License
1mo ago
Last commit
1mo ago
Created

Repo: josherau/claude-operating-core

Other skills on operating-core.