minimax-h3listed
Install: claude install-skill teskor-hub/minimax-h3-skill
# MiniMax H3 — prompting and ComfyUI setup
MiniMax H3 is an open-weight omni-modal video model: it generates video **and native stereo audio in a single forward pass**, at 24 fps, with a trained clip length of roughly 5–15 s. It ships as two checkpoints — `fl2va` (frame-conditioned) and `ref2va` (reference-conditioned) — which are different weights, not modes of one model.
This skill follows MiniMax's own prompt-writing guides and adds the failure modes those guides do not cover.
- `references/prompting.md` — the official output format for T2VA / I2VA / FL2VA / L2VA
- `references/reference-mode.md` — the official six-section format for full reference (Ref2VA)
- `references/templates.md` — fill-in templates for every mode
- `references/troubleshooting.md` — symptom → cause → fix, from real failures
- `references/comfyui.md` — checkpoints, quants, VRAM, node-by-node settings
- `references/reel-to-prompt.md` — rebuilding a reference clip: measure its cuts, read its frames, write the prompt from what is there
- `references/reel-modes.md` — ControlNet / Motion Strip / Hybrid selection, source dialogue and mode-specific readiness
**When the user supplies a reference clip or a link to one**, do not describe it from memory. Run `tools/reel_shots.py` first — it downloads the clip, detects every cut, writes frames at each shot's head, middle and tail, and lists the valid `17k+5` lengths bracketing the source duration. Then read those frames. Beat timings taken from a measured cut