create-image-fal
Generate or edit an image via any FAL image model (nano-banana edit, gpt-image, flux, ...), ROUTED THROUGH THE fal-proxy so it bills the Ads agent. image_urls…
Assemble an absurdist animated-explainer video ad (~38s, 9:16) from per-scene i2v clips + their measured VO windows — retime each clip to its VO, re-encode every segment to identical 30fps/libx264/yuv420p so the concat demuxer never drops frames, concat, build a REAL-product PIL
$ npx -y skills add gooseworks-ai/goose-skills --skill render-absurdist-explainer --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/render-absurdist-explainerContext preview
The summary Claude sees to decide when to auto-load this skill.
Assemble an absurdist animated-explainer video ad (~38s, 9:16) from per-scene i2v clips + their measured VO windows — retime each clip to its VO, re-encode every segment to identical 30fps/libx264/yuv420p so the concat demuxer never drops frames, concat, build a REAL-product PIL
name: render-absurdist-explainer description: Assemble an absurdist animated-explainer video ad (~38s, 9:16) from per-scene i2v clips + their measured VO windows — retime each clip to its VO, re-encode every segment to identical 30fps/libx264/yuv420p so the concat demuxer never drops frames, concat, build a REAL-product PIL end card (never AI) with a slow Ken-Burns, mix VO (loudnorm I=-14) under music (loudnorm I=-26, volume 0.62, amix normalize=0), and burn libass captions last. FREE deterministic assembly (bash-free, Python + ffmpeg + PIL); the recipe supplies the clips, VO, music, product photo, palette, and caption table and gates the paid keyframe/clip/VO/music calls to their own capabilities. Use for the absurdist-explainer format. status: active
The free, deterministic renderer for the **absurdist-explainer** video ad format — the bright Pixar/Disney 3D spot where a personified villain (the problem) narrates the whole ad in one voice, teaches the product's ownable mechanism through cartoon biology, lists the damage, then watches its own scheme collapse when the product arrives. This capability is the **FREE assembly stage only**. All generative work (nano-banana keyframes, Seedance i2v clips, ElevenLabs VO + music) happens upstream in the recipe and is handed to this capability as files.
It ports the validated compose recipe from two reference runs (HUM "Big Chill" cortisol absurdism and Soteri "Eczema, the pH villain"). The recipe is deterministic — iterate the cut for free, re-roll only the offending paid beat.
1. **Per-scene retime.** Each i2v clip is retimed to its **measured** VO window (`scale=1080:1920:force_original_aspect_ratio=increase,crop=1080:1920,fps=30,setsar=1`, then `tpad=stop_mode=clone` if the VO is longer than the clip, else `-t` trim). 2. **Identical re-encode.** Every segment is re-encoded `libx264 -crf 18 -pix_fmt yuv420p -r 30` even if already correct — a framerate mismatch makes the concat demuxer silently drop frames. 3. **Concat** all scene segments + the end card via the concat demuxer (`-c copy`). 4. **Real-product end card.** `build_endcard.py` composites the REAL retail product photo over the brand palette (flat, or sampled from the photo's own edge pixel) with a typeset wordmark + claim rows + CTA pill in PIL `ImageDraw.text` — **never an AI cartoon bottle, never AI-rendered brand text**. `compose.py` Ken-Burnses it 1.00 → 1.04 over the dwell. 5. **Mix.** VO bus `loudnorm I=-14 TP=-1.5`, music bus `loudnorm I=-26 TP=-3` then `volume=0.62`, `amix inputs=2 duration=first normalize=0` → master lands at -14.5..-13.5 LUFS with the music ducked under the VO. 6. **Captions last.** `make_captions.py` emits a libass `.ass` (one cue per scene, Arial 64 white / 6px outline / MarginV=330, `start = scene_start + 0.08s`, suppressed on the end card). `compose.py` burns it as the final filter so captions sit on top.
layer (wordmark / product line / claim rows / accent CTA pill). Reads the same `config.json`. Run this FIRST so `end_card.image` exists before `compose.py`.
compose reads, so caption windows stay in lockstep with the cut. Run before `compose.py` (or point `config.captions_ass` at nothing to skip captions).
concat → Ken-Burns end card → VO/music loudnorm mix → burn captions → master mp4.
brand-neutralised Soteri values as a worked reference).
`config.json` carries: `scenes[]` (each `{id, clip, target_sec, vo, caption, atempo?}` where `target_sec` is the **measured** VO window), `end_card{product_image, image, dwell_sec, zoom_to, wordmark, product_line, claims[], cta, background?}`, `brand_palette {primary, primary_lite, accent, grey}`, `music_bed`, `music_volume` (default 0.62), `atempo` (compose-stage VO speed-up, default off; the reference runs used 1.3 when the VO read slow), `captions_ass`, and `caption_style`. See `config.example.json`.
Both reference runs shipped an AI bottle first and had to re-shoot with the real photo.
`ImageDraw.text`. AI draws the world + characters only.
concat demuxer silently drops frames.
count — VO drives the per-scene timing.
to `0.70` (Big Chill), `amix normalize=0`. Master target -14.5..-13.5 LUFS, true-peak ≤ -1.5 dBFS.
carries the message — two text layers at one spot are both unreadable).
`watch` (QC the final master — confirm the villain silhouette holds, the single voice carries the whole spot, the motif lands ≥3×, no AI brand text leaked into a cartoon background, the end card is the real product, and duration is within ±0.1s of the summed windows). The recipe gates the paid `create-image-fal` (keyframes), `create-video-fal` (Seedance i2v), `create-vo-elevenlabs`, and `create-music-elevenlabs` calls to their own capabilities — this capability itself makes NO paid calls.
Put your AI agent on the growth team. Research customers and competitors, analyze what is working, create the next campaign, and learn from the result.
Repo: gooseworks-ai/goose-skills
Generate or edit an image via any FAL image model (nano-banana edit, gpt-image, flux, ...), ROUTED THROUGH THE fal-proxy so it bills the Ads agent. image_urls…
Generate a single photoreal or designed image with OpenAI gpt-image via fal.ai. Supports gpt-image-1 (default, fixed sizes — the FAL fallback for Higgsfield's…
Generate an instrumental music bed via ElevenLabs Music, ROUTED THROUGH THE elevenlabs-proxy so it bills the Ads agent. Trims any sparse intro, loudnorm, fades…
Image-to-video (or text-to-video) via any FAL video model (Kling, Seedance, Veo), ROUTED THROUGH THE GooseWorks fal-proxy so the call bills the Ads agent. The…
Generate a voiceover (VO) clip via ElevenLabs text-to-speech, ROUTED THROUGH THE elevenlabs-proxy so it bills the Ads agent. Voice id + script text come from…
Scrape competitor ads from Google Ads by domain. Returns ad creatives, formats, and campaign details. Use for competitive ad research and messaging analysis.