account-rotation
Switch coding-agent accounts and verify runtime identity. Use when: the caller requests an account change; never rotate automatically to evade a quota.
Write behavioral tests, practice TDD or inspect important coverage gaps. Use when: test design or missing proof needs work; running an existing suite needs no skill.
$ npx -y skills add boshu2/agentops --skill test --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/testContext preview
The summary Claude sees to decide when to auto-load this skill.
Write behavioral tests, practice TDD or inspect important coverage gaps. Use when: test design or missing proof needs work; running an existing suite needs no skill.
name: test
description: 'Write behavioral tests, practice TDD or inspect important coverage gaps. Use when: test design or missing proof needs work; running an existing suite needs no skill.'
practices:
- tdd
- property-based-testing
- bdd-gherkin
hexagonal_role: supporting
consumes:
- standards
- repo-context
produces:
- test-evidence
context_rel: []
skill_api_version: 1
user-invocable: true
context:
window: fork
intent:
mode: task
sections:
exclude:
- HISTORY
metadata:
capabilities: [test]
effects: [write_test_files, write_test_evidence, modify_source_files]
canonical_status: canonical
disposition: keep_specialist
tier: execution
dependencies: []
output_contract: behavioral tests and reproducible check facts; coverage results when requestedWrite or strengthen tests for a named behavior. Use existing tests directly when the task is only to run a known suite; this skill is not a required wrapper. A test is useful when it distinguishes an accepted outcome from a plausible failure, not merely when it executes the implementation.
| Mode | Use when | Result | |---|---|---| | `generate` | Existing behavior needs tests | Useful tests and focused/suite results | | `coverage` | The caller asks to find or fill gaps | Before/after coverage, valuable tests and remaining risks | | `tdd` | New behavior is being developed test first | Real expected RED, implementation, green and refactor | | `strategy` | The caller wants test design only | Prioritized risks and proposed checks in the existing discussion |
Default to `generate`; mode and scope are skill prompt choices, not invented CLI flags. Coverage thresholds come from the caller or repository.
conversation, bead, specification or existing contract before inventing new ones.
bounded contexts may use different terms; do not unify them by renaming tests.
from accidental timing, ordering and mutable shared-state dependencies.
manufacture a RED claim or alter acceptance to excuse a product defect.
reproducer and finding. Do not mask it by deleting or weakening a test.
Prefer exact observable values or errors when known. Use properties or invariants when they express the contract more faithfully than one example. Differential agreement needs an independently credible reference. A smoke check proves only what it observes; it cannot establish an exact behavior by itself. Explain a material oracle limit in the native handoff, without creating a worksheet or mandatory report.
Establish that an important new behavioral check can catch the defect it claims to guard. An authentic pre-fix RED or reproduction is usually sufficient. If a regression test was written after the fix, run it against the pre-fix version or use a safe, targeted negative control in an isolated copy. Mutate only when that would resolve real doubt about the oracle, then restore and verify the candidate. Do not demand one mutation experiment per table row or new test.
Confirm the runner completed, the intended tests actually ran, and assertions observe the promised behavior. Report crashes, truncation, unexpected skips or exclusions as gaps. When runner discovery or failure reporting changed, use a negative control through that same path before trusting green. No need to re-prove an unchanged healthy runner on each edit.
1. Read the accepted examples and relevant public interface. For a small change, one discriminating example may suffice; add consequential error/boundary cases where they could falsify acceptance. A `.feature` file is optional. If the repository already uses scenario-to-test annotations, maintain them and use its scenario coverage checker. Do not add a feature file just to satisfy this skill. 2. Find the owning suite, applicable repository standards and a narrow baseline. Use [Domain's standards](../domain/references/standards/test-pyramid.md) only if additional guidance would affect the test choice. Measure broad coverage only for `coverage` mode or an existing repository requirement. 3. Write the smallest test that observes the promised result through a stable interface. In `tdd` mode run it before implementation and require the expected missing-behavior failure, then implement and refactor under green. In other modes use evidence appropriate to existing versus newly fixed behavior. 4. Run the focused checks during editing, then the relevant integration recipe before handoff. Broaden only for changed risk, a failure or repository policy; avoid replaying the full suite after every small edit. 5. Return test changes, literal commands and results, discovered defects and material unchecked behavior. Compare against the original accepted examples. New tests added after implementation may supplement but never replace them.
Load only the guidance needed by the subject:
Tests belong in the repository's language-native locations. Check facts and limits belong in the existing handoff. Persist coverage or ot
Agent work you can verify and build on. AgentOps means agent operations: applying years of DevOps experience to how coding agents plan, implement, validate, and hand off work.
Switch coding-agent accounts and verify runtime identity. Use when: the caller requests an account change; never rotate automatically to evade a quota.
Coordinate selected writers with Agent Mail messages and advisory file reservations. Use when: this adapter is requested; mail does not own tracker status.
Dispatch independent tasks to parallel workers or selected persistent roles. Use when: delegation is authorized with disjoint scopes; execution does not…
Run a supplied task in AGY Antigravity and collect its result. Use when: the caller selects AGY; never a fallback for native coding.
Search agent session logs and cited episodes with CASS. Use when: past prompts, decisions or failures may answer a question; repeated text is not a proven…