Skip to content

Claude Code Monitoring skills :)

Flowy lists 271 skills in the Monitoring category for Claude Code, across 19 plugins. A skill is a folder of instructions an agent loads while you work. Every one here shows what is inside before you install, who wrote it, and whether it is Auto-invoked, meaning a FLOW.md router fires it as you prompt.

Browse all skills →Monitoring plugins →

Showing 96 of 271

haiku
Skill
Monitoring

haiku

When writing a haiku for this bot, follow these conventions:

@agno-agi@agno-agiView Skill
analyze_current
Skill
Monitoring

analyze_current

Read and understand the current baseline implementation. Extract all relevant information about the existing approach without modifying anything, and record…

@upsonic@upsonicView Skill
benchmark
Skill
Monitoring

benchmark

Define the comparison metrics and extract baseline values from the current implementation. Record them as a structured JSON entry so downstream phases and…

@upsonic@upsonicView Skill
code-review
Skill
Monitoring

code-review

Perform structured code reviews with actionable feedback. Use when a user asks to review code, check code quality, find bugs, audit security, improve…

@upsonic@upsonicView Skill
observal
Skill
Monitoring

observal

Core Observal CLI operations: pull agents into your harness, scan installed components, diagnose and patch harness configs, authenticate, manage CLI settings,…

@observal@observalView Skill
observal-admin
Skill
Monitoring

observal-admin

Observal admin operations including user management, server settings, submission review queue, security events, audit logs, and SSO configuration. Use when the…

@observal@observalView Skill
observal-advanced
Skill
Monitoring

observal-advanced

Advanced Observal operations including session reconciliation, CLI upgrades and downgrades, complete uninstallation, and local fallback mode for offline use.…

@observal@observalView Skill
evolution-engine
Skill
Monitoring

evolution-engine

Domain knowledge for the Evolution Engine — LLM-powered autonomous strategy discovery from raw OHLCV data. Covers the generate-backtest-select-evolve loop,…

risk-management
Skill
Monitoring

risk-management

Risk management domain knowledge for trading agents — affective state monitoring, position sizing, drawdown management, tilt detection, and behavioral…

trade-memory
Skill
Monitoring

trade-memory

Compliance-grade decision audit trail for AI trading agents. Records every trading decision with full context (conditions, filters, indicators, risk state),…

criterium
Skill
Monitoring

criterium

Use this skill when users ask about benchmarking Clojure code, measuring performance, profiling execution time, or using the criterium library. Covers the…

@hugoduncan@hugoduncanView Skill
dd-apm
Skill
Monitoring

dd-apm

APM - traces, services, dependencies, performance analysis.

pup3d975
@datadog@datadogView Skill
dd-code-generation
Skill
Monitoring

dd-code-generation

Use pup CLI for immediate Datadog operations or generate code for integration into applications

pup3d975
@datadog@datadogView Skill
dd-debugger
Skill
Monitoring

dd-debugger

Live Debugger - inspect runtime argument/variable values in production by placing log probes on methods. Use when asked what values a function receives, what…

pup3d975
@datadog@datadogView Skill
00-meta-eval
Skill
Monitoring

00-meta-eval

Use when the user wants to build an evaluation system for an LLM/agent application but doesn't know where to start — they have traces, prompts, RAG pipelines,…

@agentscope-ai@agentscope-aiView Skill
01-eval-design
Skill
Monitoring

01-eval-design

Use when the user needs to design evaluation datasets, create test cases, stratify samples, generate adversarial examples, extract eval dimensions from…

@agentscope-ai@agentscope-aiView Skill
02-metric-design
Skill
Monitoring

02-metric-design

Use when the user has evaluation principles or a dataset but needs help choosing the right graders, designing evaluation metrics, creating LLM-as-judge…

@agentscope-ai@agentscope-aiView Skill
add-agent-support
Skill
Monitoring

add-agent-support

Create and ship AgentSessions support for a new or changed local AI agent/provider. Use when adding, reviewing, testing, documenting, or marketing a provider…

agent-support-matrix
Skill
Monitoring

agent-support-matrix

Maintain Agent Sessions agent support matrix and JSON/JSONL parsing compatibility. Use when checking upstream agent releases for session format changes,…

bug-audit
Skill
Monitoring

bug-audit

Weekly multi-agent audit for serious bugs (data integrity, silent caps, staleness, timestamp math, trust boundaries). Fans out Sonnet scanners + Opus deep…

issue-brief
Skill
Monitoring

issue-brief

Explain a GitHub issue, discussion, or feature request in plain language before deciding whether to build it. Covers what the reporter actually wants, a…

context-modes
Skill
Monitoring

context-modes

Structured work modes for agent sessions. Set LACP_CONTEXT_MODE to activate: tdd (red-green-refactor), debugging (4-phase root cause), sprint (pre-agreed…

@0xnyk@0xnykView Skill
quality-gate
Skill
Monitoring

quality-gate

Production quality gate for agent sessions. Activates on session stop to evaluate work quality using 4-dimension weighted scoring (completeness, honesty,…

@0xnyk@0xnykView Skill
session-hardening
Skill
Monitoring

session-hardening

Production hardening for agent sessions. Includes pretool guards (blocks rm -rf, co-author injection, publishing without approval, data exfiltration),…

@0xnyk@0xnykView Skill
releasing
Skill
Monitoring

releasing

Full release lifecycle — version bump, CHANGELOG, rich release notes, tag, publish. Use when user says "release", "tag and release", "publish version", "cut a…

@alexei-led@alexei-ledView Skill
SkillCompass
Skill
Monitoring

SkillCompass

Evaluate skill quality, find the weakest dimension, and apply directed improvements. Also tracks usage to spot idle or risky skills. Use when: first session…

@evol-ai@evol-aiView Skill
session-monitoring
Skill
Monitoring

session-monitoring

Provides awareness of claudectl session state, health checks, and cost tracking. Activated when the user asks about session health, spending, brain decisions,…

@mercurialsolo@mercurialsoloView Skill
platform-skills
Skill
Monitoring

platform-skills

Use when troubleshooting, implementing, reviewing, or auditing platform infrastructure as a system — where Kubernetes, GitOps, CI/CD, and security concerns…

tokenscope
Skill
Monitoring

tokenscope

Judgment-layer review of Claude Code token usage. Use when the user asks to audit token costs, review Claude Code spending, check context waste, or interpret a…

@avivavi@avivaviView Skill
data-analysis
Skill
Monitoring

data-analysis

Analyze, explore, clean, and visualize datasets with statistical rigor. Use when user asks to analyze data, find patterns, compute statistics, create…

@upsonic@upsonicView Skill
evaluate
Skill
Monitoring

evaluate

Compare baseline and new implementation results. Produce the machine-readable final report `result.json`, update `experiments.json`, and append a row to…

@upsonic@upsonicView Skill
experiment_management
Skill
Monitoring

experiment_management

Set up and manage the experiment folder structure. This is Phase 0 — it runs before any analysis begins. All bookkeeping files are JSON (never markdown).

@upsonic@upsonicView Skill
observal-agents
Skill
Monitoring

observal-agents

Create, update, version, and manage Observal agents. Use when the user wants to create a new agent, update an existing one, release a new version, scaffold a…

@observal@observalView Skill
observal-ops
Skill
Monitoring

observal-ops

View traces, spans, metrics, feedback, telemetry health, and agent insight reports, including suggestions that reuse components already in the registry. Use…

@observal@observalView Skill
observal-registry
Skill
Monitoring

observal-registry

Submit, browse, install, edit, archive, restore, transfer, and version MCPs, skills, hooks, prompts, and sandboxes in the Observal registry, and get components…

@observal@observalView Skill
tradememory-bridge
Skill
Monitoring

tradememory-bridge

Bridge between Binance trading events and TradeMemory Protocol. Automatically journals trades, recalls similar past setups, detects behavioral biases, and…

trading-memory
Skill
Monitoring

trading-memory

Domain knowledge for AI trading memory — Outcome-Weighted Memory (OWM) architecture, 5 memory types, recall scoring, and behavioral analysis. Use when…

dd-docs
Skill
Monitoring

dd-docs

Datadog docs lookup using docs.datadoghq.com/llms.txt and linked Markdown pages.

pup3d975
@datadog@datadogView Skill
dd-logs
Skill
Monitoring

dd-logs

Log management - search, pipelines, archives, and cost control.

pup3d975
@datadog@datadogView Skill
03-align-human
Skill
Monitoring

03-align-human

Use when the user has a judge/grader and human-labeled data, and wants to measure how well the judge agrees with humans, detect systematic biases, determine…

@agentscope-ai@agentscope-aiView Skill
04-eval-report
Skill
Monitoring

04-eval-report

Use when the user has run multiple evaluation skills and wants a comprehensive analysis — maturity assessment, cross-skill signals, trends, prioritized…

@agentscope-ai@agentscope-aiView Skill
05-rag-eval
Skill
Monitoring

05-rag-eval

Use when the user has a RAG (Retrieval-Augmented Generation) system and wants to evaluate its quality — separating retrieval issues from generation issues.…

@agentscope-ai@agentscope-aiView Skill
deploy
Skill
Monitoring

deploy

Use when shipping a release of Agent Sessions — bumping version, updating CHANGELOG, building, signing, notarizing, publishing appcast, and creating a GitHub…

release-notes
Skill
Monitoring

release-notes

Use when writing or curating the user-facing release copy for an Agent Sessions release — README "What's New", GitHub release notes, Sparkle release notes, or…

sc-skill
Skill
Monitoring

sc-skill

Capture deterministic macOS screenshots for testing, docs, release notes, and marketing assets. Use when asked to automate app screenshots, batch-generate…

implement
Skill
Monitoring

implement

Create a new Jupyter notebook implementing the method from the research paper, using the same data as the baseline. Record implementation details and measured…

@upsonic@upsonicView Skill
progress
Skill
Monitoring

progress

Maintain a **machine-readable** progress file so dashboards, CLIs, and notebooks can poll the experiment's state at any time. The file is a JSON document —…

@upsonic@upsonicView Skill
research
Skill
Monitoring

research

Read the materialized research source and extract actionable information needed to implement the proposed method. Record the findings as a structured JSON…

@upsonic@upsonicView Skill
dd-monitors
Skill
Monitoring

dd-monitors

Monitor management - create, update, mute, and alerting best practices.

pup3d975
@datadog@datadogView Skill
dd-pup
Skill
Monitoring

dd-pup

Datadog CLI (pup). OAuth2 auth with token refresh.

pup3d975
@datadog@datadogView Skill
dd-symdb
Skill
Monitoring

dd-symdb

Symbol Database - search service symbols, find probe-able methods.

pup3d975
@datadog@datadogView Skill
06-prompt-regression
Skill
Monitoring

06-prompt-regression

Use when the user has changed a prompt (system prompt, RAG template, agent instruction, etc.) and wants to know whether the candidate is better or worse than…

@agentscope-ai@agentscope-aiView Skill
07-redteam
Skill
Monitoring

07-redteam

Use when the user wants to test their LLM/agent application for safety and security vulnerabilities — jailbreaks, prompt injection, PII extraction, harmful…

@agentscope-ai@agentscope-aiView Skill
08-bootstrap
Skill
Monitoring

08-bootstrap

Use when the user has nothing — no traces, no labels, no eval set — and needs to build a v0 evaluation from scratch. Also use when the user says "I need to…

@agentscope-ai@agentscope-aiView Skill
summarization
Skill
Monitoring

summarization

Summarize documents, articles, conversations, code, and technical content into concise, accurate summaries. Use when user asks to summarize, condense, create a…

@upsonic@upsonicView Skill
dd-triage-flaky-test
Skill
Monitoring

dd-triage-flaky-test

Load when investigating a specific flaky test. Gets history, failure pattern, and category, then recommends fix, quarantine, or escalate.

pup3d975
@datadog@datadogView Skill
dd-unblock-pr
Skill
Monitoring

dd-unblock-pr

Load when investigating a failing PR CI pipeline or checking PR health. Attributes each CI failure as flaky, infra, or regression, proposes a targeted action,…

pup3d975
@datadog@datadogView Skill
auto-arena
Skill
Monitoring

auto-arena

Automatically evaluate and compare multiple AI models or agents without pre-existing test data. Generates test queries from a task description, collects…

@agentscope-ai@agentscope-aiView Skill
bib-verify
Skill
Monitoring

bib-verify

Verify a BibTeX file for hallucinated or fabricated references by cross-checking every entry against CrossRef, arXiv, and DBLP. Reports each reference as…

@agentscope-ai@agentscope-aiView Skill
claude-authenticity
Skill
Monitoring

claude-authenticity

Detect whether an API endpoint is backed by genuine Claude (not a wrapper, proxy, or impersonator) using 9 weighted rule-based checks that mirror the…

@agentscope-ai@agentscope-aiView Skill
find-skills-combo
Skill
Monitoring

find-skills-combo

Discover and recommend **combinations** of agent skills to complete complex, multi-faceted tasks. Provides two recommendation strategies — **Maximum Quality**…

@agentscope-ai@agentscope-aiView Skill
mmx-cli
Skill
Monitoring

mmx-cli

Generate text, images, video, speech, and music via the MiniMax AI platform. Covers text generation (MiniMax-M3 model), image generation (image-01), video…

@agentscope-ai@agentscope-aiView Skill
openjudge
Skill
Monitoring

openjudge

Build custom LLM evaluation pipelines using the OpenJudge framework. Covers selecting and configuring graders (LLM-based, function-based, agentic), running…

@agentscope-ai@agentscope-aiView Skill
paper-review
Skill
Monitoring

paper-review

Review academic papers for correctness, quality, and novelty using OpenJudge's multi-stage pipeline. Supports PDF files and LaTeX source packages…

@agentscope-ai@agentscope-aiView Skill
ref-hallucination-arena
Skill
Monitoring

ref-hallucination-aren

Benchmark LLM reference recommendation capabilities by verifying every cited paper against Crossref, PubMed, arXiv, and DBLP. Measures hallucination rate,…

@agentscope-ai@agentscope-aiView Skill
rl-reward
Skill
Monitoring

rl-reward

Build RL reward signals using the OpenJudge framework. Covers choosing between pointwise and pairwise reward strategies based on RL algorithm, task type, and…

@agentscope-ai@agentscope-aiView Skill

175 more in Monitoring. See them all →