Skip to content

/data-rigor-and-leakage

Use BEFORE training any model, to build correct train/val/test splits and hunt data leakage - the #1 cause of fake-high accuracy. Covers group/patient/subject splits, temporal splits, official-benchmark splits, label correctness, class balance, and preprocessing parity. Triggers

From plugin
823 skills1 agents1 commands
shell
$ npx -y skills add mxslr/mlcraft --skill data-rigor-and-leakage --agent claude-code

How it fires

How this skill gets triggered: by you, by Claude, or both.

  • Fires itselfAuto-invocation. Claude auto-loads it when your prompt matches the work.
  • You can call itInvoke it directly when you want it.
  • Slash command/data-rigor-and-leakage
How auto-invocation works

Context preview

The summary Claude sees to decide when to auto-load this skill.

Use BEFORE training any model, to build correct train/val/test splits and hunt data leakage - the #1 cause of fake-high accuracy. Covers group/patient/subject splits, temporal splits, official-benchmark splits, label correctness, class balance, and preprocessing parity. Triggers
Ships withmlcraft

A research-first AI/ML research-engineer workflow for Claude Code

Get the whole plugin, auto-invoked

Other skills on mlcraft.