00-andruia-consultant
Arquitecto de Soluciones Principal y Consultor Tecnológico de Andru.ia. Diagnostica y traza la hoja de ruta óptima para proyectos de IA en español.
Use when an agent, harness, gateway, MCP workflow, or multi-step automation claims completion and the available traces, checkpoints, approvals, tool calls, or deployment records must be judged without trusting self-reported success.
$ npx -y skills add sickn33/antigravity-awesome-skills --skill audit-agent-run-evidence --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/audit-agent-run-evidenceContext preview
The summary Claude sees to decide when to auto-load this skill.
Use when an agent, harness, gateway, MCP workflow, or multi-step automation claims completion and the available traces, checkpoints, approvals, tool calls, or deployment records must be judged without trusting self-reported success.
name: audit-agent-run-evidence description: "Use when an agent, harness, gateway, MCP workflow, or multi-step automation claims completion and the available traces, checkpoints, approvals, tool calls, or deployment records must be judged without trusting self-reported success." risk: safe source: self date_added: "2026-08-19"
Turn an end-to-end success statement into independently decidable claims. Reconstruct what happened from available records, grade each claim against the strongest witness, and keep missing evidence distinct from failure.
This is a read-only audit. Do not rerun tools, approve actions, resume workers, deploy artifacts, or modify evidence unless the user separately authorizes those actions.
Do not use this skill to design instrumentation for a future run or to perform the missing actions. It evaluates evidence that already exists.
Record these inputs before judging the run:
Do not silently strengthen the original success criteria. Do not weaken them to match the evidence that happens to exist.
Split the overall claim into atomic predicates. Give every row a stable claim ID.
| Field | Required content | |---|---| | `claim_id` | Stable identifier | | `predicate` | One falsifiable statement | | `required_witness` | Source that can independently prove it | | `evidence_refs` | Exact event, log, artifact, or record IDs | | `counterevidence_refs` | Conflicting records | | `coverage` | Required instances versus observed instances | | `verdict` | `proven`, `partially_proven`, `contradicted`, or `not_proven` | | `gap` | Missing field, actor, interval, or verification |
Typical predicates include:
Preserve original records and create a normalized event view with:
{
"run_id": "run-123",
"event_id": "evt-42",
"sequence": 42,
"observed_at": "RFC3339 timestamp",
"actor": {"type": "worker", "id": "worker-2"},
"operation": "mcp.search",
"state_before": "researching",
"state_after": "researching",
"attempt": 2,
"request_id": "req-9",
"idempotency_key": "task-7:search:2",
"input_digest": "sha256:...",
"output_digest": "sha256:...",
"checkpoint_seq": 3,
"parent_event_id": "evt-41",
"status": "succeeded",
"evidence_ref": "tool-log:991"
}Use `null` or `unknown` for absent values. Never synthesize IDs, timestamps, digests, costs, approvals, or outcomes.
Verify bundle hashes or signatures when supplied. Check duplicate IDs, broken parent links, non-monotonic per-source sequences, impossible state transitions, unaccounted clock skew, and unexplained trace gaps. Treat an integrity failure as counterevidence for claims that depend on the affected records.
Prefer the witness closest to the effect:
| Claim | Strong witness | Insufficient alone | |---|---|---| | Code changed | Commit/tree and diff | Agent narration | | Test passed | Complete test result bound to revision | Command invocation | | MCP effect occurred | Server or provider audit record | Client request | | Checkpoint resumed | Durable checkpoint plus verified load event | Checkpoint file exists | | Human approved | Authorization-system decision bound to artifact and target | Approval requested | | Deployment succeeded | Platform record plus required health checks | Deployment started | | Memory grounded a decision | Versioned memory read and citation | Final answer resembles memory |
An orchestrator and its child worker are not independent witnesses when they repeat the same unverified result. A cryptographic digest proves byte identity, not semantic correctness.
1. Order events by causal links and per-source sequence; use timestamps only as supporting evidence. 2. Build the state-transition path and mark every gap or illegal transition. 3. Link each retry chain by logical operation, request ID, and idempotency key. 4. Link checkpoints to the state they contain and the resume event that consumes them. 5. Preserve every parallel branch outcome; apply the declared `all_required`, `quorum`, `first_success`, or other join rule. 6. Track remaining budgets at each transition. A late success after budget exhaustion is a budget violation. 7. Bind approvals and deployment records to exact artifact digests and targets.
Do not infer successful completion from a final state label when required intermediate predicates are missing.
Find reusable instructions for your project, inspect their complete files, and keep an exact skill set you can review and reuse. Codex or Claude inspects your project and chooses exact skills from the complete local AAS catalog.
Repo: sickn33/antigravity-awesome-skills
Arquitecto de Soluciones Principal y Consultor Tecnológico de Andru.ia. Diagnostica y traza la hoja de ruta óptima para proyectos de IA en español.
Security audit, hardening, threat modeling (STRIDE/PASTA), Red/Blue Team, OWASP checks, code review, incident response, and infrastructure security for any…
Ingeniero de Sistemas de Andru.ia. Diseña, redacta y despliega nuevas habilidades (skills) dentro del repositorio siguiendo el Estándar de Diamante.
Estratega de Inteligencia de Dominio de Andru.ia. Analiza el nicho específico de un proyecto para inyectar conocimientos, regulaciones y estándares únicos del…
AI-powered presentation generation via the 2slides API — create slides from text, match a reference image style, summarize documents into decks, add AI voice…