Agent Skills
› SamurAIGPT/Generative-Media-Skills
› muapi-music-video
muapi-music-video
GitHub根据歌曲主题生成包含关键帧、动画片段和匹配音乐的短视频。通过并行调用图像生成、音频创建及视频动画工具,实现从文本到多媒体内容的自动化编排与输出。
Trigger Scenarios
music video
mv
video story
song visualization
Install
npx skills add SamurAIGPT/Generative-Media-Skills --skill muapi-music-video -g -y
SKILL.md
Frontmatter
{
"name": "muapi-music-video",
"slug": "muapi-music-video",
"version": "1.0.0",
"description": "Build a short music video from a song theme — N keyframes, animate each, generate matching music.",
"acceptLicenseTerms": true
}
Music Video
Build a short music video from a song theme — N keyframes, animate each, generate matching music.
Inputs
| Name | Type | Required | Default | Description |
|---|---|---|---|---|
theme |
text | yes | — | Song / video theme (e.g. "lonely robot finds a friend, hopeful"). |
scenes |
int | no | 3 | Number of scenes (each becomes a 5s clip). |
music_style |
text | no | ambient cinematic, instrumental, slow tempo, warm | Suno-style tags for the soundtrack. |
visual_style |
text | no | cinematic, photoreal, soft volumetric light, 16:9 |
Steps
Build one the plan covering:
- Layer A (parallel) — N keyframes + 1 music track all at once.
- For each scene 1..N:
muapi image generatewith a beat-specific prompt +{{visual_style}}, model=nano-banana-pro (these feed video gen). - One
muapi audio create(kind=music) using{{music_style}}, duration = N × 5 + a 2s tail.
- For each scene 1..N:
- Layer B (parallel, depends on Layer A) — animate each keyframe.
- For each scene:
muapi video from-imagewithimage=$nX.url, model=veo3.1-image-to-video, duration=5, prompt=scene-specific motion direction.
- For each scene:
- Return:
- The scene keyframes (asset ids in order).
- The animation clips (asset ids in order).
- The music track asset id.
- A short summary describing the cut order.
Notes
- Keep character continuity by repeating the character description in every scene prompt verbatim.
- Don't auto-confirm any single video call > 50 cr — those need the user's nod (the loop will prompt automatically).
- If a scene's
muapi video from-imagefails after failover, fall back tomuapi video generate(text-to-video) for that scene only.
Trigger Keywords
music video, mv, video story, song visualization
Notes for the Executing Agent
- This recipe is LLM-orchestrated: read each phase, gather any missing inputs from the user, then call
muapiCLI commands. Usemuapi auth configurefirst ifMUAPI_API_KEYis unset. - For model IDs without a CLI alias yet, fall back to the raw endpoint via
curl -X POST https://api.muapi.ai/api/v1/<endpoint> -H "x-api-key: $MUAPI_API_KEY" -H 'content-type: application/json' -d '{...}'and poll withmuapi predict wait <request_id>. - Substitute
{{input_name}}placeholders with the user's actual inputs before issuing each call.
Version History
- d4ce293 Current 2026-07-25 11:20


