Seedance audio
OpenGod - 14-module AI super skill. Fuses 18 of Top 50 GitHub AI projects.
npx -y skills add 381487190/opengod --skill seedance-audioAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
One thing to look at
- 0 stars0 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.
What its author says it does
Copied from the file, not written here
This skill should be used when the user asks for Seedance 2.0 audio, dialogue, lip-sync, music, sound effects, ambience, beat-sync, audio-reference mapping, desync troubleshooting, or sound-driven visual timing.
The file declares its own license as MIT. That is the author’s claim about this one file, and it is not the same thing as the license GitHub reports for the repository, which is listed with the other numbers below.
SKILL.md
4.6 KB, 875 tokens by cl100k_base, as published. Nobody here has run it
seedance-audio
Use this for dialogue, lip-sync, sound layers, music, ambience, beat-sync, audio-reference mapping, desync troubleshooting, or sound-driven visual timing. Audio should support the visible beat instead of becoming a second competing prompt.
Load [ref:audio-guide] for how the audio model behaves, per-language dialogue capacity, the voice-reference lip-sync path, beat-sync, desync repair, audio-reference conflicts, and multi-character workarounds. Load [ref:audio-post-delivery] when the user needs stems, M&E, dubbing, loudness, sync, mix, or delivery guidance.
Intent
Half of every emotion enters through the ears, and users almost always forget sound until its absence makes the clip feel dead. The soul here is giving every scene its sound before being asked - the room's breath, the action's evidence, the line that lands. When they hear it, they realize it was always part of what they meant.
Core Rules
Keep dialogue short, quote spoken lines, and assign every line to a named speaker. Prefer locked or stable framing for lip-sync. Remove head-turning, large face motion, extreme camera moves, or busy hand gestures while mouth accuracy matters. Treat @Audio1 as a rhythm, pacing, mood, voice-tone, or ambience reference unless the active platform documents exact playback behavior; on surfaces that accept a spoken-voice reference, field reports indicate an attached voice clip can drive lip-sync directly: the model syncs to your audio instead of synthesizing speech - the most reliable field-reported path for non-English dialogue. Use only rights-cleared voices.
Reliability is probabilistic and language-dependent: field reports rank Mandarin strongest for lip-sync, English a close second, with Japanese, Korean, Russian, and others weaker. Keep non-English lines very short or use a voice reference, and budget retakes rather than promising a clean voiced take. See [ref:audio-guide] for the field-observed per-language dialogue-capacity table.
Sound Layer Pattern
Use compact layers: Dialogue: ... Sound: ... SFX: ... Music: ... Silence: .... Include only the layers that matter. Silence is valid when it sharpens drama or avoids confusing lip-sync.
| Need | Stable audio direction |
|---|---|
| Lip-sync | Character A, locked medium close-up, says "I found it." Clear dry dialogue, no head turn. |
| Product ad | Sound: low room tone. SFX: magnetic click on lid open, soft glass chime at final frame. |
| Beat sync | @Audio1 provides tempo only; light pulses and foot taps match the downbeat. |
| Drama | Distant rain and refrigerator hum; no music during the line. |
| Action | Breathing grows louder, shoe squeak at landing, metal door buzzer at endpoint. |
Multi-Character Dialogue
Use one speaker per short clip when reliability matters. If two characters must speak, separate turns and keep the camera stable: Character A says... pause. Character B answers.... For complex exchanges, recommend generating controlled single-speaker clips and compositing in post.
Failure Fixes
If dialogue desyncs, shorten the line, lock the camera, remove head turns, clean the audio role, and reduce competing SFX. If the wrong speaker talks, assign tags and split lines by speaker. If audio is ignored, remove extra music/SFX instructions and make the reference role explicit.
If audio and video references fight each other, mute the reference video before upload when possible, or make the priority explicit: @Video1 controls camera only; @Audio1 controls tempo and energy.
Sequence State
When sequence state is present, inherit completed dialogue, active dialogue, ambience, music phase, SFX phase, current clip scope, continuity locks, exact reference tags, and reserved future beats. Do not repeat completed dialogue unless the user explicitly asks for a reprise. Continue or intentionally change the audio phase instead of restarting it by accident.
Output Contract
Return speaker map, quoted dialogue, sound layers, audio reference role, lip-sync constraints, post/delivery notes if needed, and a compact prompt-ready audio block.
What ships with it
Read from the repository
Just SKILL.md. No reference files, no scripts.
Gives 0 of the 12 instructions most video audio skills give in 875 tokens
Counted across 622 of the 795 authors here whose files we hold, read 2026-08-07
- Read individual rule files for detailed explanationsin 21 of 622, across 10 files
- Render final videoin 13 of 622, across 6 files
- Use WAV PCM 16kHz mono audio formatin 12 of 622, across 3 files
- Use this skill when dealing with Remotion codein 11 of 622, across 4 files
- Save generated audio to a WAV filein 11 of 622, across 4 files
- Handle conversion errors gracefullyin 10 of 622, across 6 files
- Add captions to videos alwaysin 10 of 622, across 4 files
- Generate music from text descriptions using MusicGenin 9 of 622, across 2 files
- Do not skip pipeline layersin 9 of 622, across 3 files
- Do not make one tool do everythingin 9 of 622, across 3 files
- Use Azure Document Intelligence for complex PDFsin 9 of 622, across 4 files
- Never ask the user to paste their full API keyin 9 of 622, across 3 files
Said here and by no other author read
- keep dialogue short
- quote spoken dialogue lines
- assign every line to a named speaker
- prefer locked framing for lip-sync
- remove head turns during mouth motion
- use only rights-cleared voices
Grouped from the skills themselves: near-identical wordings counted once, and counted by distinct author, so one author publishing three of these counts once. Length counted with cl100k_base; the agent that loads this file may tokenize it differently.