coral-debug
Verify and debug changes to CORAL itself — smallest reproduce loop per area (grader / daemon / CLI / hooks / manager / workspace / hub / template / config /…
Run and manage CORAL experiments from the operator side — launch agents with `coral start` (dotlist overrides, model/count, tmux vs local), monitor with `coral status` / `coral log` / `coral show` / the web dashboard, and drive the loop with `coral resume` (inject instructions,
$ npx -y skills add Human-Agent-Society/CORAL --skill running-coral-experiments --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/running-coral-experimentsContext preview
The summary Claude sees to decide when to auto-load this skill.
Run and manage CORAL experiments from the operator side — launch agents with `coral start` (dotlist overrides, model/count, tmux vs local), monitor with `coral status` / `coral log` / `coral show` / the web dashboard, and drive the loop with `coral resume` (inject instructions,
name: running-coral-experiments description: Run and manage CORAL experiments from the operator side — launch agents with `coral start` (dotlist overrides, model/count, tmux vs local), monitor with `coral status` / `coral log` / `coral show` / the web dashboard, and drive the loop with `coral resume` (inject instructions, fork from an attempt), `coral heartbeat` (tune reflection cadence), and `coral stop`. Use whenever the user wants to start a CORAL run, check on agents, read scores/leaderboard, steer or resume a run, diagnose agents that keep restarting or fail every eval, scale to more agents or islands, or stop a run. Deep references for steering/heartbeat tuning and scaling/troubleshooting live alongside this skill.
You drive a run with five verbs: **start → status → log/show → resume → stop**. Everything else is a flag on those or a deeper topic in the references. Prefer `coral <cmd> --help` over guessing flags.
**Prereq:** a task (`task.yaml` + `seed/` + grader package) that passes `coral validate .`. No task yet → that's the `creating-a-coral-task` skill. Each runtime CLI must be installed and authenticated → the `setting-up-coral` skill.
coral start -c task.yaml # auto-tmux session coral start -c task.yaml agents.count=4 agents.model=opus # dotlist overrides (no quotes needed) coral start -c task.yaml run.verbose=true run.ui=true # verbose logs + web dashboard coral start -c task.yaml run.session=local # foreground, no tmux
coral status # agent health + leaderboard snapshot (the quick pulse) coral runs # active runs across tasks; --all includes finished coral ui --port 8420 # web dashboard: live leaderboard, logs, DAG
`coral status` answers "who's alive, how many evals, current best". If it looks healthy but scores never move, jump to budget classes + troubleshooting in [references/scaling-and-ops.md](references/scaling-and-ops.md).
coral log # top 20 real attempts by score coral log -n 5 --recent # most recent instead of best coral log --search "kernel" --agent agent-1 coral log --class grader_error # surface crashing graders (first stop when unhealthy) coral show <hash> # one attempt: score, explanation, files changed coral show <hash> --diff # full diff — see exactly what the leader did
`<hash>` comes from `coral log`/`coral status`. By default `coral log` hides `tune` and `grader_error` attempts; `--all` shows them, `--class {real|tune|grader_error}` filters to one. What the classes mean → [references/scaling-and-ops.md](references/scaling-and-ops.md).
coral resume # resume latest run, sessions restored coral resume -i "Try greedy approaches first" # inject guidance agents read next loop coral resume --from <hash> -i "Continue this fork" # reset an agent to an attempt, then steer coral export <hash> -b winning-idea # export an attempt's commit as a git branch
`resume -i` is how you nudge a run without restarting from scratch (stop → resume with an instruction). `--from` forks a promising line that later regressed. You can also retune the reflection cadence — `coral heartbeat set/remove/reset` — to make agents reflect less, pivot sooner, etc. Both topics, with worked examples: [references/steering.md](references/steering.md).
coral stop # stop the current/latest run (picker if several) coral stop --all # stop every active run
Stopping leaves all results, notes, and the leaderboard on disk — `coral resume` later, or just inspect with `coral log`/`coral show`.
coral validate . # grader scores the seed (once) coral start -c task.yaml agents.count=2 # launch coral status # ... check periodically coral log -n 5 --recent # see what agents are trying coral show <best-hash> --diff # inspect the leader coral resume -i "Focus on the inner loop" # steer if they plateau coral stop # done
Note: `coral eval / diff / revert / checkout / wait` are **agent-side** commands run *inside* a worktree during a run — agents already know them from the generated `CORAL.md`. As the operator you rarely touch them; you drive the verbs above. Full CLI reference: https://coral.compounding-intelligence.ai/docs/cli/reference
Robust, lightweight infrastructure for multi-agent self-evolution, built for autoresearch. CORAL is infrastructure for autonomous AI agent organizations that run experiments, share knowledge, and continuously improve solutions.
Verify and debug changes to CORAL itself — smallest reproduce loop per area (grader / daemon / CLI / hooks / manager / workspace / hub / template / config /…
Add a new component to the CORAL framework itself — a new agent runtime under `coral/agent/builtin/` (claude_code/codex/cursor_agent style), a new CLI command…
End-to-end recipe for adding a new task under `examples/` — the three pieces that have to line up (`task.yaml`, `seed/`, and `grader/`), what to put in each,…
Use when preparing, reviewing, resolving conflicts for, or merging a CORAL release pull request from the long-lived dev branch into main.
Write a note to {shared_dir}/notes/ that future agents can actually act on. Use after every coral eval, when a heartbeat (reflect / consolidate / pivot) asks…
Research the problem domain before coding. Web search for techniques, save raw sources, write structured findings, update the index.