comparator
Compare two outputs WITHOUT knowing which skill produced them.
An agent is a specialist Claude hands a whole job to, with its own tools and its own context.
4,474 agents across 361 plugins.
Compare two outputs WITHOUT knowing which skill produced them.
Evaluate expectations against an execution transcript and outputs.
Analyze Logic-Lens benchmark/eval failures. Use after running content-evals, or when pointed at a `skills-workspace/iteration-*` directory or a…
The verify gate of the Logic-Lens iteration loop. Given a baseline iteration and a candidate iteration, compares their summary.json (overall, logic vs format…
Applies a single, minimal, generalized edit to a Logic-Lens skill (SKILL.md / guide / _shared file) given a concrete failure diagnosis. Use inside the…
Use this agent when you need to validate design implementations in real browsers for Drupal or WordPress projects. This agent should be used proactively after…
Meta-agent that adopts external personalities and adapts them to LETS modes. Loads identity from personality text provided in prompt, then operates as that…
System design expert for architecture reviews, pattern analysis, SOLID principles evaluation, and coupling/abstraction assessments. Use when reviewing…
Backend development expert for API design review, business logic analysis, error handling assessment, and performance evaluation. Use when reviewing…
Creative research brainstormer that generates ideas researchers working within a single field would miss — cross-field connections, challenges to conventional…
Adversarial but constructive research sparring partner that stress-tests research ideas BEFORE the researcher invests months of effort. Evaluates ideas along 7…
AUTONOMOUS PAPER PRODUCTION AGENT. This is the primary agent of the plugin. Takes a paper title, topic, or research question and AUTONOMOUSLY executes the…
Use this agent after creating or modifying a CLAUDE.md file. Reviews quality including instruction specificity, token efficiency, correct separation of…
Use this agent after creating or modifying a hook. Reviews quality including exit code contract, performance, file filtering, settings.json registration, and…
Reviews a single refactor phase diff before commit. Called by applying-refactors at each phase checkpoint. Enforces 400 LOC cap, phase-type discipline, and…
Specialized agent for creating comprehensive infrastructure and architecture documentation.
Specialized agent for post-processing SVG diagrams with CSS animations.
Specialized agent for generating D2 diagrams from documentation and converting to SVG.
Use when asked to analyze BigQuery SQL files across a project for cost optimization opportunities, estimate total query costs, or audit a codebase for…
Use when asked to review all SQL files in a project or directory for BigQuery anti-patterns, scan a codebase for SQL performance issues, or audit BigQuery…
Use when asked to analyze table schemas across a project, recommend partitioning and clustering strategies for existing tables, audit schema design, or…
Requirements analyst. MUST BE USED for ambiguous requests, requirements gathering, scope assessment, feasibility checks, and proposal writing. NOT for small,…
Technical architect. MUST BE USED for system design, API design, implementation planning, multi-file refactoring, and agent orchestration. Receives proposal.md…
Analyzes tasks and decomposes them into a sequence of agent steps for execution.
Reviews code changes for architectural consistency and patterns. Use PROACTIVELY after any structural changes, new services, or API modifications. Ensures…
Design RESTful APIs, microservice boundaries, and database schemas. Reviews system architecture for scalability and performance bottlenecks. Use PROACTIVELY…
Validates review findings against full source context to remove false positives. Runs after synthesis, before user approval. Does NOT add new findings.
Use this agent when bare-metal code, firmware, or an FPGA design will not come up on real hardware: it boots but hangs, faults early, gives no output, or…
Use this agent to review HDL/RTL changes (ROHD, Chisel, SpinalHDL, Verilog, VHDL) before merge or tapeout. It reviews a diff or a set of modules against…
Use this agent before a tapeout or MPW shuttle submission to run the precheck gate and decide whether the design is genuinely ready. It checks DRC/LVS status,…
Deep analysis agent for reference materials, transcripts, and knowledge bases. Extracts insights, themes, quotes, and patterns to inform content creation. Use…
Execution agent that creates content based on approved plans. Invokes appropriate specialist skills and ensures quality output. Only operates after planning…
Planning agent for content creation projects. Analyzes user requests, reviews reference materials, and creates detailed content plans before any creation…
Business operations specialist covering market analysis, pricing strategy, go-to-market planning, and competitive intelligence. Use proactively when evaluating…
Content strategist specializing in messaging, copywriting, tutorials, documentation, and launch communications. Use proactively when creating content, writing…
Product designer specializing in UI/UX, user flows, HTML mockups, and design systems. Use proactively when creating user interfaces, planning user experiences,…
Impl-blind test/oracle author — writes conformance tests from a spec-only brief. Tool-restricted by definition (no Read/Grep/Glob/Edit), so "authored blind" is…
Implementer — writes production code, tests, and migrations. The "generic engineer" fallback when no narrower specialist exists. Activate only when the…
Log and metrics analyst — reads .cladding/audit.log.jsonl, perf/baseline.json, and drift reports; surfaces patterns the human can act on. Activate only when…
Action-specific Auto-Harness Evaluator subagent for contract review. Use only when the current legal action is evaluator_review.
Action-specific Auto-Harness Evaluator subagent for parallel contract review. Use only when the current legal action is evaluator_review_parallel.
Action-specific Auto-Harness Evaluator subagent for final report synthesis. Use only when the current legal action is evaluator_final.
Use proactively in the background after creating or modifying multiple vault notes, or when asked about vault link health. Lowest priority of the PKM agents —…
Use proactively in the background after completing significant work blocks, after git commits (triggered automatically by PreToolUse hook), or before session…
Use proactively when researching what the vault knows about a topic, before creating new notes, or when exploring existing knowledge and connections. Run in…
© 2026 Flowy · Free and open source
Built for Claude Code · Not affiliated with Anthropic