/avatar-video
Create AI avatar videos with precise control over avatars, voices, scripts, scenes, and backgrounds using HeyGen's v2 API. Use when: (1) Choosing a specific avatar and voice for a video, (2) Writing exact scripts for an avatar to speak, (3) Building multi-scene videos with
$ npx -y skills add calesthio/OpenMontage --skill avatar-video --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
- Fires itselfAuto-invocation. Claude auto-loads it when your prompt matches the work.Auto-invocation is when the right skill fires by itself at the right moment, driven by a FLOW.md router and a hook, instead of you invoking it by name. It is the difference between a skill being installed and a skill actually getting used.Read the full definition →
- You can call itInvoke it directly when you want it.
- Slash command
/avatar-video
Context preview
The summary Claude sees to decide when to auto-load this skill.
Create AI avatar videos with precise control over avatars, voices, scripts, scenes, and backgrounds using HeyGen's v2 API. Use when: (1) Choosing a specific avatar and voice for a video, (2) Writing exact scripts for an avatar to speak, (3) Building multi-scene videos with
SKILL.md
avatar-video.SKILL.mdname: avatar-video
description: |
Create AI avatar videos with precise control over avatars, voices, scripts, scenes, and backgrounds using HeyGen's v2 API. Use when: (1) Choosing a specific avatar and voice for a video, (2) Writing exact scripts for an avatar to speak, (3) Building multi-scene videos with different backgrounds per scene, (4) Creating transparent WebM videos for compositing, (5) Using talking photos as video presenters, (6) Integrating HeyGen avatars with Remotion, (7) Batch video generation with exact specs, (8) Brand-consistent production videos with precise control.
homepage: https://docs.heygen.com/reference/create-a-video
allowed-tools: mcp__heygen__*
metadata:
openclaw:
requires:
env:
- HEYGEN_API_KEY
primaryEnv: HEYGEN_API_KEYAvatar Video
Create AI avatar videos with full control over avatars, voices, scripts, scenes, and backgrounds. Build single or multi-scene videos with exact configuration using HeyGen's `/v2/video/generate` API.
Authentication
All requests require the `X-Api-Key` header. Set the `HEYGEN_API_KEY` environment variable.
curl -X GET "https://api.heygen.com/v2/avatars" \
-H "X-Api-Key: $HEYGEN_API_KEY"
Tool Selection
If HeyGen MCP tools are available (`mcp__heygen__*`), **prefer them** over direct HTTP API calls — they handle authentication and request formatting automatically.
| Task | MCP Tool | Fallback (Direct API) | |------|----------|----------------------| | Check video status / get URL | `mcp__heygen__get_video` | `GET /v2/videos/{video_id}` | | List account videos | `mcp__heygen__list_videos` | `GET /v2/videos` | | Delete a video | `mcp__heygen__delete_video` | `DELETE /v2/videos/{video_id}` |
Video generation (`POST /v2/video/generate`) and avatar/voice listing are done via direct API calls — see reference files below.
Default Workflow
1. **List avatars** — `GET /v2/avatars` → pick an avatar, preview it, note `avatar_id` and `default_voice_id`. See [avatars.md](references/avatars.md) 2. **List voices** (if needed) — `GET /v2/voices` → pick a voice matching the avatar's gender/language. See [voices.md](references/voices.md) 3. **Write the script** — Structure scenes with one concept each. See [scripts.md](references/scripts.md) 4. **Generate the video** — `POST /v2/video/generate` with avatar, voice, script, and background per scene. See [video-generation.md](references/video-generation.md) 5. **Poll for completion** — `GET /v2/videos/{video_id}` until status is `completed`. See [video-status.md](references/video-status.md)
Quick Reference
| Task | Read | |------|------| | List and preview avatars | [avatars.md](references/avatars.md) | | List and select voices | [voices.md](references/voices.md) | | Write and structure scripts | [scripts.md](references/scripts.md) | | Generate video (single or multi-scene) | [video-generation.md](references/video-generation.md) | | Add custom backgrounds | [backgrounds.md](references/backgrounds.md) | | Add captions / subtitles | [captions.md](references/captions.md) | | Add text overlays | [text-overlays.md](references/text-overlays.md) | | Create transparent WebM video | [video-generation.md](references/video-generation.md) (WebM section) | | Use templates | [templates.md](references/templates.md) | | Create avatar from photo | [photo-avatars.md](references/photo-avatars.md) | | Check video status / download | [video-status.md](references/video-status.md) | | Upload assets (images, audio) | [assets.md](references/assets.md) | | Use with Remotion | [remotion-integration.md](references/remotion-integration.md) | | Set up webhooks | [webhooks.md](references/webhooks.md) |
When to Use This Skill vs Create Video
This skill is for **precise control** — you choose the avatar, write the exact script, configure each scene.
If the user just wants to **describe a video idea** and let AI handle the rest (script, avatar, visuals), use the **create-video** skill instead.
| User Says | Create Video Skill | This Skill | |-----------|:------------------:|:----------:| | "Make me a video about X" | ✓ | | | "Create a product demo" | ✓ | | | "I want avatar Y to say exactly Z" | | ✓ | | "Multi-scene video with different backgrounds" | | ✓ | | "Transparent WebM for compositing" | | ✓ | | "Use this specific voice for my script" | | ✓ | | "Batch generate videos with exact specs" | | ✓ |
Reference Files
Core Video Creation
- [references/avatars.md](references/avatars.md) - Listing avatars, styles, avatar_id selection
- [references/voices.md](references/voices.md) - Listing voices, locales, speed/pitch
- [references/scripts.md](references/scripts.md) - Writing scripts, pauses, pacing
- [references/video-generation.md](references/video-generation.md) - POST /v2/video/generate and multi-scene videos
Video Customization
- [references/backgrounds.md](references/backgrounds.md) - Solid colors, images, video backgrounds
- [references/text-overlays.md](references/text-overlays.md) - Adding text with fonts and positioning
- [references/captions.md](references/captions.md) - Auto-generated captions and subtitles
Advanced Features
- [references/templates.md](references/templates.md) - Template listing and variable replacement
- [references/photo-avatars.md](references/photo-avatars.md) - Creating avatars from photos
- [references/webhooks.md](references/webhooks.md) - Webhook endpoints and events
Integration
- [references/remotion-integration.md](references/remotion-integration.md) - Using HeyGen in Remotion compositions
Foundation
- [references/video-status.md](references/video-status.md) - Polling patterns and download URLs
- [references/assets.md](references/assets.md) - Uploading images, videos, audio
- [references/dimensions.md](references/dimensions.md) - Resolution and aspect ratios
- [references/quota.md](references/quota.md) - Credit system and usage limits
Best Practices
1. **Preview avatars before generating** —
Read more
name: avatar-video
description: |
Create AI avatar videos with precise control over avatars, voices, scripts, scenes, and backgrounds using HeyGen's v2 API. Use when: (1) Choosing a specific avatar and voice for a video, (2) Writing exact scripts for an avatar to speak, (3) Building multi-scene videos with different backgrounds per scene, (4) Creating transparent WebM videos for compositing, (5) Using talking photos as video presenters, (6) Integrating HeyGen avatars with Remotion, (7) Batch video generation with exact specs, (8) Brand-consistent production videos with precise control.
homepage: https://docs.heygen.com/reference/create-a-video
allowed-tools: mcp__heygen__*
metadata:
openclaw:
requires:
env:
- HEYGEN_API_KEY
primaryEnv: HEYGEN_API_KEYAvatar Video
Create AI avatar videos with full control over avatars, voices, scripts, scenes, and backgrounds. Build single or multi-scene videos with exact configuration using HeyGen's `/v2/video/generate` API.
Authentication
All requests require the `X-Api-Key` header. Set the `HEYGEN_API_KEY` environment variable.
curl -X GET "https://api.heygen.com/v2/avatars" \ -H "X-Api-Key: $HEYGEN_API_KEY"
Tool Selection
If HeyGen MCP tools are available (`mcp__heygen__*`), **prefer them** over direct HTTP API calls — they handle authentication and request formatting automatically.
| Task | MCP Tool | Fallback (Direct API) | |------|----------|----------------------| | Check video status / get URL | `mcp__heygen__get_video` | `GET /v2/videos/{video_id}` | | List account videos | `mcp__heygen__list_videos` | `GET /v2/videos` | | Delete a video | `mcp__heygen__delete_video` | `DELETE /v2/videos/{video_id}` |
Video generation (`POST /v2/video/generate`) and avatar/voice listing are done via direct API calls — see reference files below.
Default Workflow
1. **List avatars** — `GET /v2/avatars` → pick an avatar, preview it, note `avatar_id` and `default_voice_id`. See [avatars.md](references/avatars.md) 2. **List voices** (if needed) — `GET /v2/voices` → pick a voice matching the avatar's gender/language. See [voices.md](references/voices.md) 3. **Write the script** — Structure scenes with one concept each. See [scripts.md](references/scripts.md) 4. **Generate the video** — `POST /v2/video/generate` with avatar, voice, script, and background per scene. See [video-generation.md](references/video-generation.md) 5. **Poll for completion** — `GET /v2/videos/{video_id}` until status is `completed`. See [video-status.md](references/video-status.md)
Quick Reference
| Task | Read | |------|------| | List and preview avatars | [avatars.md](references/avatars.md) | | List and select voices | [voices.md](references/voices.md) | | Write and structure scripts | [scripts.md](references/scripts.md) | | Generate video (single or multi-scene) | [video-generation.md](references/video-generation.md) | | Add custom backgrounds | [backgrounds.md](references/backgrounds.md) | | Add captions / subtitles | [captions.md](references/captions.md) | | Add text overlays | [text-overlays.md](references/text-overlays.md) | | Create transparent WebM video | [video-generation.md](references/video-generation.md) (WebM section) | | Use templates | [templates.md](references/templates.md) | | Create avatar from photo | [photo-avatars.md](references/photo-avatars.md) | | Check video status / download | [video-status.md](references/video-status.md) | | Upload assets (images, audio) | [assets.md](references/assets.md) | | Use with Remotion | [remotion-integration.md](references/remotion-integration.md) | | Set up webhooks | [webhooks.md](references/webhooks.md) |
When to Use This Skill vs Create Video
This skill is for **precise control** — you choose the avatar, write the exact script, configure each scene.
If the user just wants to **describe a video idea** and let AI handle the rest (script, avatar, visuals), use the **create-video** skill instead.
| User Says | Create Video Skill | This Skill | |-----------|:------------------:|:----------:| | "Make me a video about X" | ✓ | | | "Create a product demo" | ✓ | | | "I want avatar Y to say exactly Z" | | ✓ | | "Multi-scene video with different backgrounds" | | ✓ | | "Transparent WebM for compositing" | | ✓ | | "Use this specific voice for my script" | | ✓ | | "Batch generate videos with exact specs" | | ✓ |
Reference Files
Core Video Creation
- [references/avatars.md](references/avatars.md) - Listing avatars, styles, avatar_id selection
- [references/voices.md](references/voices.md) - Listing voices, locales, speed/pitch
- [references/scripts.md](references/scripts.md) - Writing scripts, pauses, pacing
- [references/video-generation.md](references/video-generation.md) - POST /v2/video/generate and multi-scene videos
Video Customization
- [references/backgrounds.md](references/backgrounds.md) - Solid colors, images, video backgrounds
- [references/text-overlays.md](references/text-overlays.md) - Adding text with fonts and positioning
- [references/captions.md](references/captions.md) - Auto-generated captions and subtitles
Advanced Features
- [references/templates.md](references/templates.md) - Template listing and variable replacement
- [references/photo-avatars.md](references/photo-avatars.md) - Creating avatars from photos
- [references/webhooks.md](references/webhooks.md) - Webhook endpoints and events
Integration
- [references/remotion-integration.md](references/remotion-integration.md) - Using HeyGen in Remotion compositions
Foundation
- [references/video-status.md](references/video-status.md) - Polling patterns and download URLs
- [references/assets.md](references/assets.md) - Uploading images, videos, audio
- [references/dimensions.md](references/dimensions.md) - Resolution and aspect ratios
- [references/quota.md](references/quota.md) - Credit system and usage limits
Best Practices
1. **Preview avatars before generating** —
World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI coding assistant into a full video production studio.
Repo: calesthio/OpenMontage
Other skills on openmontage.
- /acestep
AI music generation with ACE-Step 1.5 — background music, vocal tracks, covers, stem extraction for video production. Use when generating music, soundtracks, jingles, or working with audio stems. Triggers include background music, soundtrack, jingle, music generation, stem
Open skill - /agents
Build voice AI agents with ElevenLabs. Use when creating voice assistants, customer service bots, interactive voice characters, or any real-time voice conversation experience.
Open skill - /ai-video-gen
Generate AI videos from text prompts using multiple provider gateways. Use when: (1) Generating videos from text descriptions, (2) Creating AI-generated video clips for content production, (3) Image-to-video generation with a reference image, (4) Choosing between video
Open skill - /azure-speech-to-text
Transcribe audio to text using Azure AI Speech (Fast Transcription REST API). Use when converting audio/video to text, generating subtitles, or processing spoken content in OpenMontage. Optional cloud STT provider — preferred when AZURE_SPEECH_KEY is configured; the local
Open skill - /beautiful-mermaid
Render Mermaid diagrams as SVG and PNG using the Beautiful Mermaid library. Use when the user asks to render a Mermaid diagram.
Open skill - /bfl-api
BFL FLUX API integration guide covering endpoints, async polling patterns, rate limiting, error handling, webhooks, and regional endpoints with Python and TypeScript code examples.
Open skill

