seedance-audio
GitHub用于Seedance 2.0音视频同步、对话生成、音效及节拍同步的提示词构建技能,涵盖音频参考映射与不同步故障排查。
Trigger Scenarios
Install
npx skills add Emily2040/seedance-2.0 --skill seedance-audio -g -y
SKILL.md
Frontmatter
{
"name": "seedance-audio",
"tags": [
"audio",
"lip-sync",
"dialogue",
"seedance-20"
],
"license": "MIT",
"metadata": {
"author": "Iamemily2050 (@iamemily2050)",
"parent": "seedance-20",
"updated": "2026-08-01",
"version": "6.7.0",
"openclaw": {
"emoji": "🎬",
"homepage": "https:\/\/github.com\/Emily2040\/seedance-2.0"
},
"repository": "https:\/\/github.com\/Emily2040\/seedance-2.0"
},
"description": "This skill should be used when the user asks for Seedance 2.0 audio, dialogue, lip-sync, music, sound effects, ambience, beat-sync, audio-reference mapping, desync troubleshooting, or sound-driven visual timing.",
"user-invocable": true
}
seedance-audio
Before producing prompt text, a prompt-ready block, a rewrite, an example, or a compiled clip, load the Director's Read, classify the brief, and complete its canonical narrative or non-narrative record. Translate that record into visible or audible carriers and keep its internal labels out of final generation prose.
Use this for dialogue, lip-sync, sound layers, music, ambience, beat-sync, audio-reference mapping, desync troubleshooting, or sound-driven visual timing. Audio should support the visible beat instead of becoming a second competing prompt.
Load audio-guide for the audio evidence boundary, dialogue timing, supported voice-reference options, beat-sync, desync repair, audio-reference conflicts, and multi-character workarounds. Load audio-post-delivery when the user needs stems, M&E, dubbing, loudness, sync, mix, or delivery guidance.
Intent
Half of every emotion enters through the ears, and users almost always forget sound until its absence makes the clip feel dead. The soul here is giving every scene its sound before being asked - the room's breath, the action's evidence, the line that lands. When they hear it, they realize it was always part of what they meant.
Core Rules
Preserve the user's exact dialogue, language, script, and register. Quote each line and name its speaker. Stable framing and one speaker turn are useful starting heuristics when mouth accuracy matters; do not remove required performance or rewrite words without authorization. Check spoken duration and inspect the returned words, performance, and visible sync separately.
Treat @Audio1 as a reference for its assigned role. Use a spoken-voice workflow only when the active operation supports it, with rights-cleared material; do not promise exact playback or guaranteed lip-sync. Offer native generation, a supported reference workflow, and post dubbing with their tradeoffs. No universal Mandarin-first ranking or per-language line ceiling is established by this repository. See audio-guide for the scoped model-card evidence and dialogue calibration for an optional pilot within the user's existing take/spend authorization.
Sound Layer Pattern
Use compact layers: Dialogue: ... Sound: ... SFX: ... Music: ... Silence: .... Include only the layers that matter. Silence is valid when it sharpens drama or avoids confusing lip-sync.
| Need | Stable audio direction |
|---|---|
| Lip-sync | Character A, locked medium close-up, says "I found it." Clear dry dialogue, no head turn. |
| Product ad | Sound: low room tone. SFX: magnetic click on lid open, soft glass chime at final frame. |
| Beat sync | @Audio1 provides tempo only; light pulses and foot taps match the downbeat. |
| Drama | Distant rain and refrigerator hum; no music during the line. |
| Action | Breathing grows louder, shoe squeak at landing, metal door buzzer at endpoint. |
Multi-Character Dialogue
Use one speaker per short clip when reliability matters. If two characters must speak, separate turns and keep the camera stable: Character A says... pause. Character B answers.... For complex exchanges, recommend generating controlled single-speaker clips and compositing in post.
Failure Fixes
If dialogue desyncs, offer a shorter line if wording can change, lock the camera, remove head turns, clean the audio role, and reduce competing SFX. If the wrong speaker talks, assign tags and split lines by speaker. If audio is ignored, remove extra music/SFX instructions and make the reference role explicit.
If audio and video references fight each other, mute the reference video before upload when possible, or make the priority explicit: @Video1 controls camera only; @Audio1 controls tempo and energy.
Sequence State
When sequence state is present, inherit completed dialogue, active dialogue, ambience, music phase, SFX phase, current clip scope, continuity locks, exact reference tags, and reserved future beats. Do not repeat completed dialogue unless the user explicitly asks for a reprise. Continue or intentionally change the audio phase instead of restarting it by accident.
Output Contract
Return speaker map, quoted dialogue, sound layers, audio reference role, lip-sync constraints, post/delivery notes if needed, and a compact prompt-ready audio block.
For worked examples with a chosen approach, copyable prompt, evidence boundary and failure checks, load performance and dialogue cards. These are authored concepts, not observed generation results.
Version History
-
9ea203f
Current 2026-09-08 20:18
新增具体的表演和精确对话教学示例;校准关于对话语言限制的说明,明确不虚构语言限制。
-
b8cdc5b
2026-08-04 19:05
新增强制要求在处理提示前加载Director's Read并分类简报的规则;修复嵌套运行时路由的可移植性问题。
- 7f22eac 2026-08-02 22:00
- 6c51262 2026-07-30 20:23


