coordinate-external-ag…
Coordinate independently operated external agents through durable handoffs. Use when work crosses hosts, sessions, accounts, services, queues, boards, pull…
Implement and evaluate activation steering, representation engineering, logit changes, or weight-space interventions. Use when changing model behavior without ordinary fine-tuning or sweeping layer, strength, and persistence choices.
$ npx -y skills add gaelic-ghost/socket --skill steer-language-model-behavior --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/steer-language-model-behaviorContext preview
The summary Claude sees to decide when to auto-load this skill.
Implement and evaluate activation steering, representation engineering, logit changes, or weight-space interventions. Use when changing model behavior without ordinary fine-tuning or sweeping layer, strength, and persistence choices.
name: steer-language-model-behavior description: Implement and evaluate activation steering, representation engineering, logit changes, or weight-space interventions. Use when changing model behavior without ordinary fine-tuning or sweeping layer, strength, and persistence choices.
1. Invoke `research-model-representations` to define or validate the steering signal. 2. Freeze target and guardrail evaluation sets before selecting layers or strengths. 3. Record layer, hook point, token position, normalization, sign, magnitude, schedule, and generation settings. 4. Sweep a bounded strength range including zero and negative controls. 5. Compare with prompt-only, random-direction, and norm-matched controls. 6. Measure target success, capability regressions, fluency, calibration, diversity, and off-target behavioral changes. 7. For persistent changes, preserve the base checkpoint as immutable, write a new artifact, record both checksums and the transformation, and never edit weights in place. 8. Re-run the exact packaged or merged artifact if the intervention becomes persistent. 9. Report the smallest effective intervention and the operating range where the claim holds.
A successful steering vector demonstrates controllability under tested conditions. It does not by itself establish a unique representation, a complete mechanism, or safe generalization. Stronger target behavior with broad unrelated regressions is not a clean success.
Use `references/steering-controls.md` as the minimum comparison matrix.
Stuff for Agents on macOS Promo audio: Socket Codex Marketplace Promo
Coordinate independently operated external agents through durable handoffs. Use when work crosses hosts, sessions, accounts, services, queues, boards, pull…
Assign worktree, branch, write, validation, integration, and cleanup ownership before parallel repository work. Use when a worker will inspect or modify…
Design framework-neutral agent and automation workflows before implementation. Use when choosing between Codex app automations, codex exec, Codex subagents,…
Design evaluation workflows for agent, skill, prompt, and automation behavior before implementation. Use when choosing eval cases, graders, thresholds,…
Design safe n8n workflows with deterministic routing, credentials, idempotency, recovery, local-model checks, drafts, and exact approval gates.
Coordinate bounded worker tasks with a launch envelope, report-back, escalation, and synthesis contract. Use before spawning, resuming, steering, cancelling,…