10 music video
Generates music video and beat-synced visual content prompts for Higgsfield. An audio track is always required. Routes to veo3_1 (primary, native audio) or veo3_1_lite (fallback). Use when the user wants a music video, lyric video, beat-synced visuals, performance video, concert visual, album art animation, or any music-driven visual content.From its SKILL.md
npx -y skills add pixelab-ch/higgsfield-skills --skill 10-music-videoAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
2 things to look at
- 2 stars2 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.
- runs commandsInstructs the agent to run 4 commands, including `higgsfield:generate_video` and 3 more.
SKILL.md
6.4 KB, ~1.4k tokens by cl100k_base, as published. Nobody here has run it
Higgsfield Music-Video Skill
What this skill does
Crafts production-ready music video prompts and routes them to the best Higgsfield model for beat-synced, audio-driven output. Handles performance videos, narrative videos, abstract visualizers, lyric videos, and multi-segment assembly for full-length songs.
An audio track is always required. This is a mandatory-audio skill: every generation
needs a user-supplied audio file uploaded and confirmed via the media branch before calling
higgsfield:generate_video. The confirmed audio is attached with role audio.
Model routing
Primary and fallback video models
| veo3_1 (primary) | veo3_1_lite (fallback) | |
|---|---|---|
| Rationale | Native audio generation; best audio-visual sync | Lower cost; use when primary is unavailable |
| Aspect ratios | 16:9, 9:16 only — no 1:1 | 16:9, 9:16, auto |
| Duration | 4, 6, 8 s — discrete set only | 4, 6, 8 s (same discrete set) |
| Tunable params | quality {basic, high, ultra}; model {veo-3-1-preview, veo-3-1-fast} | resolution {720p, 1080p}; generate_audio {true, false, def false} |
| Media roles | start_image (max 1) | start_image, end_image |
| Native audio | Yes | Basic (via generate_audio param) |
Routing rule: Use veo3_1 by default. Fall back to veo3_1_lite only when the primary
is unavailable. When falling back, inform the user of what they give up: native audio
quality and the veo-3-1-preview quality tier. Never switch silently.
Aspect and duration constraints:
- Neither model supports
1:1aspect ratio. - Durations are the discrete set
[4, 6, 8]— do not pass 5, 7, 10, or 15. - A full music video requires multiple segments (see model-specs.md for planning guidance).
MODEL-06 directive: If a parameter is rejected at generation time, call
higgsfield:models_explore with the target model name to re-verify the live schema. Full
parameter tables: references/model-specs.md.
Prompt-building workflow
-
Gather intent — Confirm: song genre, visual strategy (performance / narrative / abstract), target platform, aspect ratio, mood/aesthetic, desired clip duration.
-
Select model — Apply the routing table above. Use
veo3_1by default. -
Build the prompt — Use the craft references below:
- Beat-sync techniques, energy mapping, hook framework, camera techniques: references/beat-sync.md
- Genre-specific visual language, colors, lighting, trigger words: references/genres.md
- Master template and worked example prompts: references/examples.md
-
Plan segments — Because veo3_1 clips are 4–8 s, map the song's structure to segments before generating. See references/model-specs.md for multi-segment planning.
-
Present for review — Show the assembled prompt and all parameters to the user for review and refinement before any generation call.
Opt-in generation
Generation costs Higgsfield credits and requires explicit user confirmation before any generate call. This skill never auto-generates.
Full step-by-step flow (confirmation gate, balance/cost surface, generate → poll →
display): ../../shared/generation-flow.md
This skill's primary model: veo3_1
Media upload — MANDATORY (GEN-04):
An audio track is always required for video generation. Before calling
higgsfield:generate_video:
- Ask the user to provide their audio file.
- Upload it: run
higgsfield:media_upload→ returns apending_id. - Confirm it: run
higgsfield:media_confirm→ returns aconfirmed_id. - Attach it in
input_fileswith roleaudio:[{ "id": "<confirmed_id>", "role": "audio" }]
If the user also provides a reference image for the video, add it with role start_image
in the same input_files array:
[
{ "id": "<audio_confirmed_id>", "role": "audio" },
{ "id": "<image_confirmed_id>", "role": "start_image" }
]
Never pass a pending_id directly to input_files — it will be rejected. See
../../shared/generation-flow.md Step 2b for the full
atomic-pair detail.
Tool signatures: ../../shared/mcp-tools.md
Reference materials
| File | Contents |
|---|---|
| references/model-specs.md | Per-model parameter tables for veo3_1 and veo3_1_lite; per-platform recommendations; multi-segment assembly note; verification annotation |
| references/beat-sync.md | Beat-visual synchronization philosophy, 12 music-video hook styles, 7 beat-sync techniques (timing references, energy mapping, frequency-responsive effects, syncopation visuals), energy arc table, performance/narrative/abstract strategies, camera techniques, multi-segment strategy |
| references/genres.md | Genre visual language guide for 10 genres (hip-hop, pop, rock/metal, EDM, R&B, lo-fi, classical, jazz, country, K-pop) with color palettes, lighting approaches, trigger words, visual hooks, and example prompts |
| references/examples.md | Master template, 5 worked music-video prompts (hip-hop performance, pop ballad, EDM abstract, R&B narrative, K-pop choreography), lyric video technique table |
What ships with it: 4 files
41.9 KB alongside SKILL.md
references/
- beat-sync.md11.3 KB
- examples.md13.3 KB
- genres.md11.9 KB
- model-specs.md5.3 KB