/higgsfield-models
Use when the user asks which model to use, wants to compare models, or needs guidance on selecting between Kling, Sora 2, Wan, Seedance, Veo 3, Minimax Hailuo, Soul, Nano Banana, or other Higgsfield engines.
$ npx -y skills add OSideMedia/higgsfield-ai-prompt-skill --skill higgsfield-models --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
- Fires itselfAuto-invocation. Claude auto-loads it when your prompt matches the work.Auto-invocation is when the right skill fires by itself at the right moment, driven by a FLOW.md router and a hook, instead of you invoking it by name. It is the difference between a skill being installed and a skill actually getting used.Read the full definition →
- You can call itInvoke it directly when you want it.
- Slash command
/higgsfield-models
Context preview
The summary Claude sees to decide when to auto-load this skill.
Use when the user asks which model to use, wants to compare models, or needs guidance on selecting between Kling, Sora 2, Wan, Seedance, Veo 3, Minimax Hailuo, Soul, Nano Banana, or other Higgsfield engines.
SKILL.md
higgsfield-models.SKILL.mdname: higgsfield-models
description: >
Use when the user asks which model to use, wants to compare models,
or needs guidance on selecting between Kling, Sora 2, Wan, Seedance,
Veo 3, Minimax Hailuo, Soul, Nano Banana, or other Higgsfield engines.
user-invocable: true
metadata:
references:
- MODELS-DEEP-REFERENCE.md
tags: [higgsfield, models, Kling, Sora, Wan, Seedance, Veo, Soul, NanoBanana]
version: 3.2.0
updated: 2026-07-05
parent: higgsfieldHiggsfield Model Selection Guide
Choosing the right model is the single biggest factor in output quality after the prompt. This file handles most selection questions. For deep per-model documentation (prompting specifics, parameters, edge cases, API details) → read `MODELS-DEEP-REFERENCE.md`.
---
Quick Decision Flowchart
Fast lookup — for detailed comparisons see the full tables below.
| Need | Recommended Model | Tier | |------|-------------------|------| | Top-tier cinematic video + audio | Kling 3.0 | Premium | | Epic scale / spectacle | Sora 2 | Premium | | Nature / landscapes + ref images | Veo 3.1 | Premium | | Artistic / stylized video | Wan 2.6 | Mid | | Fast video iteration | Seedance 2.0 Pro | Mid | | VFX / fluid motion | Minimax Hailuo 2.3 | Mid | | Budget-friendly video | Kling 2.5 Turbo / Higgsfield DoP Lite | Free–Low | | Fashion / aesthetic images | Soul 2.0 | Free | | Photorealistic sharp images | Nano Banana Pro | Low | | AI actor generation | Soul Cast | Low | | Native 4K images | Kling Image 3.0 | Mid | | Photo style transformation | Photodump (29 presets) | Low |
**Pricing tiers:** Free (Soul 2.0, DoP Lite) · Low (0.1–2 credits) · Mid (2–10 credits) · Premium (10+ credits). See the Credit Cost Reference below for exact per-model costs.
---
Video Models — Comparison
| Model | Realism | Character | Motion | Style | Duration | Audio | Best for | |-------|---------|-----------|--------|-------|----------|-------|----------| | Kling 3.0 | ★★★★★ | ★★★★★ | ★★★★★ | ★★★★☆ | 3–15s | ✅ | Cinematic, long, audio, multi-shot | | Kling 3.0 Omni | ★★★★★ | ★★★★★ | ★★★★★ | ★★★★☆ | 3–15s | ✅ | Video clone, storyboard control | | Kling 3.0 Omni Edit | ★★★★★ | ★★★★★ | — | ★★★★☆ | 3–10s | ✅ | Edit footage at 3.0 quality | | Kling O1 Video (legacy) | ★★★★★ | ★★★★★ | ★★★★☆ | ★★★☆☆ | 5–10s | ❌ | Multi-ref (7), start/end frame | | Kling O1 Video Edit (legacy) | ★★★★☆ | ★★★★★ | — | ★★★★★ | 3–10s | ❌ | Relight, restyle, swap, remove | | Kling 3.0 Motion Control | ★★★★★ | ★★★★☆ | ★★★★★ | ★★★☆☆ | 3–30s | Optional | Motion transfer from reference video | | Kling 2.6 (legacy) | ★★★★★ | ★★★★★ | ★★★★☆ | ★★★☆☆ | 5/10s | ✅ | Character drama, realism; native audio via `sound` toggle (default on) | | Kling 2.5 Turbo | ★★★★☆ | ★★★★☆ | ★★★★☆ | ★★★☆☆ | 5–10s | ❌ | Fast Kling iteration | | Sora 2 | ★★★★☆ | ★★★☆☆ | ★★★★★ | ★★★★☆ | — | ❌ | Epic scale, physics, action | | Wan 2.7 | ★★★★★ | ★★★★☆ | ★★★★★ | ★★★★★ | 2–15s | ✅ | 60fps, T2V/I2V/R2V/edit, first+last frame | | Wan 2.6 | ★★★★☆ | ★★★☆☆ | ★★★★☆ | ★★★★★ | 5/10/15s | ❌ | Artistic, stylized, improved physics | | Wan 2.5 | ★★★★☆ | ★★★☆☆ | ★★★★☆ | ★★★★★ | 5–10s | ✅ | Native audio, artistic, fantasy | | Seedance 2.0 | ★★★★★ | ★★★★★ | ★★★★★ | ★★★★☆ | 4–15s | ✅ | 12-asset multimodal, complex motion | | Seedance 1.5 Pro | ★★★★☆ | ★★★★☆ | ★★★★☆ | ★★★★☆ | 4/8/12s | ✅ | Best lip-sync, multilingual audio | | Seedance Pro | ★★★☆☆ | ★★★☆☆ | ★★★☆☆ | ★★★☆☆ | 10s | ❌ | Fast iteration, no audio needed | | Veo 3.1 | ★★★★★ | ★★★★☆ | ★★★★☆ | ★★★★☆ | 4/6/8s | ✅ | Ref images, first/last frame, 4K | | Veo 3.1 Lite | ★★★★☆ | ★★★★☆ | ★★★★☆ | ★★★★☆ | 4/6/8s | ✅ | Budget 3.1 quality, 1080p, I2V, volume | | Veo 3 | ★★★★☆ | ★★★☆☆ | ★★★★☆ | ★★★☆☆ | 4–8s | ✅ | Nature, environment, stable model | | Gemini Omni Flash | ★★★★☆ | ★★★★☆ | ★★★☆☆ | ★★★☆☆ | 4–10s | ✅ | Reference-driven video (image + video refs), native audio, 720p | | Grok Imagine Video | ★★★★☆ | ★★★☆☆ | ★★★★☆ | ★★★★☆ | 1–15s | ✅ | Video editing, animate images | | Minimax Hailuo 2.3 | ★★★★★ | ★★★★☆ | ★★★★★ | ★★★★☆ | 6–10s | ❌ | VFX, fluid motion, anime, physics | | Minimax Hailuo 02 | ★★★★☆ | ★★★☆☆ | ★★★★★ | ★★★☆☆ | 6–10s | ❌ | Dance, sports, fluid motion | | Higgsfield DoP (Lite/Standard/Turbo) | ★★★☆☆ | ★★★☆☆ | ★★★★☆ | ★★★☆☆ | 3–5s | ❌ | I2V specialist, 50+ presets, optical physics |
---
Decision Flowchart
Is this image or video?
├── IMAGE
│ ├── Person / portrait? → Soul 2.0
│ ├── Cinematic keyframe for I2V pipeline? → Soul Cinema Preview
│ ├── Native 4K / image series / storyboarding? → Kling Image 3.0
│ ├── Maximum sharpness / 4K? → Nano Banana Pro
│ ├── Fast pro-quality / text rendering? → Nano Banana 2
│ ├── Reference consistency or dense text? → Seedream 4.5
│ ├── Complex layout / multi-panel? → Seedream 5.0 Lite
│ ├── Text/logo in image? → GPT Image 1.5
│ └── Edit an existing image? → Flux Kontext
│
└── VIDEO
├── EDIT existing footage?
│ ├── Relight, restyle, swap, remove → Kling O1 Video Edit
│ └── Higher quality 3.0 edit → Kling 3.0 Omni Edit
│
├── Is a human character the focus?
│ ├── Need audio, long clip (15s), multi-shot → Kling 3.0
│ ├── Need to clone from reference video → Kling 3.0 Omni
│ ├── Best lip-sync + multilingual → Seedance 1.5 Pro
│ ├── Legacy-tier great character (audio togglable via `sound`) → Kling 2.6
│ └── Fast iteration → Kling 2.5 Turbo
│
├── Need motion transfer from reference video?
│ └── → Kling 3.0 Motion Control
│
├── Animate a still image with cinematic camera?
│ └── → Higgsfield DoP (Lite/Standard/Turbo)
│
├── Is the environment/phenomenon the hero?
│ ├── Nature, documentary, stable → Veo 3
│ ├── Need ref image consistency → Veo 3.1
│ ├── Budget Veo 3.1 quality / volume → Veo 3.1 Lite
│ ├── 60fps, first+last frame, ref images → Wan 2.7
│ └── Artistic, painterly, fantasy → Wan 2.5/2.6
│
├── Is it action/spectacle?
│ ├── Epic scale, crowds, physRead more
name: higgsfield-models
description: >
Use when the user asks which model to use, wants to compare models,
or needs guidance on selecting between Kling, Sora 2, Wan, Seedance,
Veo 3, Minimax Hailuo, Soul, Nano Banana, or other Higgsfield engines.
user-invocable: true
metadata:
references:
- MODELS-DEEP-REFERENCE.md
tags: [higgsfield, models, Kling, Sora, Wan, Seedance, Veo, Soul, NanoBanana]
version: 3.2.0
updated: 2026-07-05
parent: higgsfieldHiggsfield Model Selection Guide
Choosing the right model is the single biggest factor in output quality after the prompt. This file handles most selection questions. For deep per-model documentation (prompting specifics, parameters, edge cases, API details) → read `MODELS-DEEP-REFERENCE.md`.
---
Quick Decision Flowchart
Fast lookup — for detailed comparisons see the full tables below.
| Need | Recommended Model | Tier | |------|-------------------|------| | Top-tier cinematic video + audio | Kling 3.0 | Premium | | Epic scale / spectacle | Sora 2 | Premium | | Nature / landscapes + ref images | Veo 3.1 | Premium | | Artistic / stylized video | Wan 2.6 | Mid | | Fast video iteration | Seedance 2.0 Pro | Mid | | VFX / fluid motion | Minimax Hailuo 2.3 | Mid | | Budget-friendly video | Kling 2.5 Turbo / Higgsfield DoP Lite | Free–Low | | Fashion / aesthetic images | Soul 2.0 | Free | | Photorealistic sharp images | Nano Banana Pro | Low | | AI actor generation | Soul Cast | Low | | Native 4K images | Kling Image 3.0 | Mid | | Photo style transformation | Photodump (29 presets) | Low |
**Pricing tiers:** Free (Soul 2.0, DoP Lite) · Low (0.1–2 credits) · Mid (2–10 credits) · Premium (10+ credits). See the Credit Cost Reference below for exact per-model costs.
---
Video Models — Comparison
| Model | Realism | Character | Motion | Style | Duration | Audio | Best for | |-------|---------|-----------|--------|-------|----------|-------|----------| | Kling 3.0 | ★★★★★ | ★★★★★ | ★★★★★ | ★★★★☆ | 3–15s | ✅ | Cinematic, long, audio, multi-shot | | Kling 3.0 Omni | ★★★★★ | ★★★★★ | ★★★★★ | ★★★★☆ | 3–15s | ✅ | Video clone, storyboard control | | Kling 3.0 Omni Edit | ★★★★★ | ★★★★★ | — | ★★★★☆ | 3–10s | ✅ | Edit footage at 3.0 quality | | Kling O1 Video (legacy) | ★★★★★ | ★★★★★ | ★★★★☆ | ★★★☆☆ | 5–10s | ❌ | Multi-ref (7), start/end frame | | Kling O1 Video Edit (legacy) | ★★★★☆ | ★★★★★ | — | ★★★★★ | 3–10s | ❌ | Relight, restyle, swap, remove | | Kling 3.0 Motion Control | ★★★★★ | ★★★★☆ | ★★★★★ | ★★★☆☆ | 3–30s | Optional | Motion transfer from reference video | | Kling 2.6 (legacy) | ★★★★★ | ★★★★★ | ★★★★☆ | ★★★☆☆ | 5/10s | ✅ | Character drama, realism; native audio via `sound` toggle (default on) | | Kling 2.5 Turbo | ★★★★☆ | ★★★★☆ | ★★★★☆ | ★★★☆☆ | 5–10s | ❌ | Fast Kling iteration | | Sora 2 | ★★★★☆ | ★★★☆☆ | ★★★★★ | ★★★★☆ | — | ❌ | Epic scale, physics, action | | Wan 2.7 | ★★★★★ | ★★★★☆ | ★★★★★ | ★★★★★ | 2–15s | ✅ | 60fps, T2V/I2V/R2V/edit, first+last frame | | Wan 2.6 | ★★★★☆ | ★★★☆☆ | ★★★★☆ | ★★★★★ | 5/10/15s | ❌ | Artistic, stylized, improved physics | | Wan 2.5 | ★★★★☆ | ★★★☆☆ | ★★★★☆ | ★★★★★ | 5–10s | ✅ | Native audio, artistic, fantasy | | Seedance 2.0 | ★★★★★ | ★★★★★ | ★★★★★ | ★★★★☆ | 4–15s | ✅ | 12-asset multimodal, complex motion | | Seedance 1.5 Pro | ★★★★☆ | ★★★★☆ | ★★★★☆ | ★★★★☆ | 4/8/12s | ✅ | Best lip-sync, multilingual audio | | Seedance Pro | ★★★☆☆ | ★★★☆☆ | ★★★☆☆ | ★★★☆☆ | 10s | ❌ | Fast iteration, no audio needed | | Veo 3.1 | ★★★★★ | ★★★★☆ | ★★★★☆ | ★★★★☆ | 4/6/8s | ✅ | Ref images, first/last frame, 4K | | Veo 3.1 Lite | ★★★★☆ | ★★★★☆ | ★★★★☆ | ★★★★☆ | 4/6/8s | ✅ | Budget 3.1 quality, 1080p, I2V, volume | | Veo 3 | ★★★★☆ | ★★★☆☆ | ★★★★☆ | ★★★☆☆ | 4–8s | ✅ | Nature, environment, stable model | | Gemini Omni Flash | ★★★★☆ | ★★★★☆ | ★★★☆☆ | ★★★☆☆ | 4–10s | ✅ | Reference-driven video (image + video refs), native audio, 720p | | Grok Imagine Video | ★★★★☆ | ★★★☆☆ | ★★★★☆ | ★★★★☆ | 1–15s | ✅ | Video editing, animate images | | Minimax Hailuo 2.3 | ★★★★★ | ★★★★☆ | ★★★★★ | ★★★★☆ | 6–10s | ❌ | VFX, fluid motion, anime, physics | | Minimax Hailuo 02 | ★★★★☆ | ★★★☆☆ | ★★★★★ | ★★★☆☆ | 6–10s | ❌ | Dance, sports, fluid motion | | Higgsfield DoP (Lite/Standard/Turbo) | ★★★☆☆ | ★★★☆☆ | ★★★★☆ | ★★★☆☆ | 3–5s | ❌ | I2V specialist, 50+ presets, optical physics |
---
Decision Flowchart
Is this image or video?
├── IMAGE
│ ├── Person / portrait? → Soul 2.0
│ ├── Cinematic keyframe for I2V pipeline? → Soul Cinema Preview
│ ├── Native 4K / image series / storyboarding? → Kling Image 3.0
│ ├── Maximum sharpness / 4K? → Nano Banana Pro
│ ├── Fast pro-quality / text rendering? → Nano Banana 2
│ ├── Reference consistency or dense text? → Seedream 4.5
│ ├── Complex layout / multi-panel? → Seedream 5.0 Lite
│ ├── Text/logo in image? → GPT Image 1.5
│ └── Edit an existing image? → Flux Kontext
│
└── VIDEO
├── EDIT existing footage?
│ ├── Relight, restyle, swap, remove → Kling O1 Video Edit
│ └── Higher quality 3.0 edit → Kling 3.0 Omni Edit
│
├── Is a human character the focus?
│ ├── Need audio, long clip (15s), multi-shot → Kling 3.0
│ ├── Need to clone from reference video → Kling 3.0 Omni
│ ├── Best lip-sync + multilingual → Seedance 1.5 Pro
│ ├── Legacy-tier great character (audio togglable via `sound`) → Kling 2.6
│ └── Fast iteration → Kling 2.5 Turbo
│
├── Need motion transfer from reference video?
│ └── → Kling 3.0 Motion Control
│
├── Animate a still image with cinematic camera?
│ └── → Higgsfield DoP (Lite/Standard/Turbo)
│
├── Is the environment/phenomenon the hero?
│ ├── Nature, documentary, stable → Veo 3
│ ├── Need ref image consistency → Veo 3.1
│ ├── Budget Veo 3.1 quality / volume → Veo 3.1 Lite
│ ├── 60fps, first+last frame, ref images → Wan 2.7
│ └── Artistic, painterly, fantasy → Wan 2.5/2.6
│
├── Is it action/spectacle?
│ ├── Epic scale, crowds, physA comprehensive Claude skill library for generating high-quality prompts on Higgsfield AI — the cinematic video and image generation platform.
Other skills on higgsfield-ai-prompt-skill.
- /higgsfield-acting
Writes the character-performance layer of a video prompt as behavior under pressure, not displayed emotion — objective, obstacle, tactics, beats, subtext, listening, body/status/proxemics, and mandatory eye life. Produces a reusable 150–220-word acting master profile per
Open skill - /higgsfield-apps
Use when the user asks about Higgsfield's one-click Apps, wants to know which app to use for a specific output, or needs guidance on the Apps workflow.
Open skill - /higgsfield-assist
Use when the user asks about Higgsfield Assist (the built-in GPT-5 copilot), how to use the platform's native AI assistant, credit optimization strategies, plan selection, how to get more from fewer credits, or platform efficiency tips.
Open skill - /higgsfield-audio
Use when the user asks about audio in Higgsfield videos, needs to add dialogue or lip-sync, wants sound effects or ambient sound in generated video, asks about music or BGM in output, or is using any audio-capable model (Kling 3.0, Seedance 1.5 Pro, Seedance 2.0, Veo 3/3.1, Grok
Open skill - /higgsfield-camera
Use when the user asks about camera movements, shot types, or how to describe camera behavior in a Higgsfield prompt. Contains all named camera controls with descriptions, best use cases, and example prompt phrases.
Open skill - /higgsfield-canvas
Use when the user mentions Higgsfield Canvas, a node-based or node graph workspace, an infinite board/canvas, chaining generations into a pipeline, or wants to wire prompts → images → videos across models on one surface. Covers what Canvas is, the node categories, the seven
Open skill

