Agent Skillsbytedance/deer-flow › video-generation

video-generation

GitHub

根据用户请求生成高质量视频,支持结构化JSON提示词及参考图片引导。通过调用Python脚本执行AIGC视频生成流程,涵盖需求分析、提示词构建及自动执行环节。

skills/public/video-generation/SKILL.md bytedance/deer-flow

Trigger Scenarios

用户请求生成或创建视频 需要基于参考图片或特定风格进行视频创作

Install

npx skills add bytedance/deer-flow --skill video-generation -g -y
More Options

Non-standard path

npx skills add https://github.com/bytedance/deer-flow/tree/main/skills/public/video-generation -g -y

Use without installing

npx skills use bytedance/deer-flow@video-generation

指定 Agent (Claude Code)

npx skills add bytedance/deer-flow --skill video-generation -a claude-code -g -y

安装 repo 全部 skill

npx skills add bytedance/deer-flow --all -g -y

预览 repo 内 skill

npx skills add bytedance/deer-flow --list

SKILL.md

Frontmatter
{
    "name": "video-generation",
    "description": "Use this skill when the user requests to generate, create, or imagine videos. Supports structured prompts and reference image for guided generation."
}

Video Generation Skill

Overview

This skill generates high-quality videos using structured prompts and a Python script. The workflow includes creating JSON-formatted prompts and executing video generation with optional reference image.

Core Capabilities

  • Create structured JSON prompts for AIGC video generation
  • Support reference image as guidance or the first/last frame of the video
  • Generate videos through automated Python script execution

Workflow

Step 1: Understand Requirements

When a user requests video generation, identify:

  • Subject/content: What should be in the image
  • Style preferences: Art style, mood, color palette
  • Technical specs: Aspect ratio, composition, lighting
  • Reference image: Any image to guide generation
  • You don't need to check the folder under /mnt/user-data

Step 2: Create Structured Prompt

Generate a structured JSON file in /mnt/user-data/workspace/ with naming pattern: {descriptive-name}.json

Step 3: Create Reference Image (Optional when image-generation skill is available)

Generate reference image for the video generation.

  • If only 1 image is provided, use it as the guided frame of the video

Step 3: Execute Generation

Call the Python script:

python /mnt/skills/public/video-generation/scripts/generate.py \
  --prompt-file /mnt/user-data/workspace/prompt-file.json \
  --reference-images /path/to/ref1.jpg \
  --output-file /mnt/user-data/outputs/generated-video.mp4 \
  --aspect-ratio 16:9

Parameters:

  • --prompt-file: Absolute path to JSON prompt file (required)
  • --reference-images: Absolute paths to reference image (optional)
  • --output-file: Absolute path to output image file (required)
  • --aspect-ratio: Aspect ratio of the generated image (optional, default: 16:9)

[!NOTE] Do NOT read the python file, instead just call it with the parameters.

Video Generation Example

User request: "Generate a short video clip depicting the opening scene from "The Chronicles of Narnia: The Lion, the Witch and the Wardrobe"

Step 1: Search for the opening scene of "The Chronicles of Narnia: The Lion, the Witch and the Wardrobe" online

Step 2: Create a JSON prompt file with the following content:

{
  "title": "The Chronicles of Narnia - Train Station Farewell",
  "background": {
    "description": "World War II evacuation scene at a crowded London train station. Steam and smoke fill the air as children are being sent to the countryside to escape the Blitz.",
    "era": "1940s wartime Britain",
    "location": "London railway station platform"
  },
  "characters": ["Mrs. Pevensie", "Lucy Pevensie"],
  "camera": {
    "type": "Close-up two-shot",
    "movement": "Static with subtle handheld movement",
    "angle": "Profile view, intimate framing",
    "focus": "Both faces in focus, background soft bokeh"
  },
  "dialogue": [
    {
      "character": "Mrs. Pevensie",
      "text": "You must be brave for me, darling. I'll come for you... I promise."
    },
    {
      "character": "Lucy Pevensie",
      "text": "I will be, mother. I promise."
    }
  ],
  "audio": [
    {
      "type": "Train whistle blows (signaling departure)",
      "volume": 1
    },
    {
      "type": "Strings swell emotionally, then fade",
      "volume": 0.5
    },
    {
      "type": "Ambient sound of the train station",
      "volume": 0.5
    }
  ]
}

Step 3: Use the image-generation skill to generate the reference image

Load the image-generation skill and generate a single reference image narnia-farewell-scene-01.jpg according to the skill.

Step 4: Use the generate.py script to generate the video

python /mnt/skills/public/video-generation/scripts/generate.py \
  --prompt-file /mnt/user-data/workspace/narnia-farewell-scene.json \
  --reference-images /mnt/user-data/outputs/narnia-farewell-scene-01.jpg \
  --output-file /mnt/user-data/outputs/narnia-farewell-scene-01.mp4 \
  --aspect-ratio 16:9

Do NOT read the python file, just call it with the parameters.

Output Handling

After generation:

  • Videos are typically saved in /mnt/user-data/outputs/
  • Share generated videos (come first) with user as well as generated image if applicable, using present_files tool
  • Provide brief description of the generation result
  • Offer to iterate if adjustments needed

Notes

  • Always use English for prompts regardless of user's language
  • JSON format ensures structured, parsable prompts
  • Reference image enhance generation quality significantly
  • Iterative refinement is normal for optimal results

Providers (Gemini / MiniMax)

Auto-selected by environment variables (CLI unchanged):

  • GEMINI_API_KEY set → Gemini Veo (default, unchanged).
  • Only MINIMAX_API_KEY set → MiniMax video (/v1/video_generation, async 3-step poll/download).
  • Force with VIDEO_GENERATION_PROVIDER=gemini|minimax.

MiniMax overrides: MINIMAX_API_HOST (default https://api.minimaxi.com), MINIMAX_VIDEO_MODEL (default MiniMax-Hailuo-2.3). The first reference image is used as MiniMax first_frame_image. MiniMax ignores --aspect-ratio (it uses resolution/duration).

Version History

  • a5acc25 Current 2026-08-20 17:01

Same Skill Collection

.agent/skills/blocking-io-guard/SKILL.md
.agent/skills/deerflow-maintainer-orchestrator/SKILL.md
skills/public/academic-paper-review/SKILL.md
skills/public/bootstrap/SKILL.md
skills/public/chart-visualization/SKILL.md
skills/public/claude-to-deerflow/SKILL.md
skills/public/code-documentation/SKILL.md
skills/public/data-analysis/SKILL.md
skills/public/deep-research/SKILL.md
skills/public/find-skills/SKILL.md
skills/public/frontend-design/SKILL.md
skills/public/github-deep-research/SKILL.md
skills/public/image-generation/SKILL.md
skills/public/music-generation/SKILL.md
skills/public/newsletter-generation/SKILL.md
skills/public/podcast-generation/SKILL.md
skills/public/ppt-generation/SKILL.md
skills/public/skill-creator/SKILL.md
skills/public/skill-reviewer/SKILL.md
skills/public/surprise-me/SKILL.md
skills/public/systematic-literature-review/SKILL.md
skills/public/vercel-deploy-claimable/SKILL.md
skills/public/web-design-guidelines/SKILL.md
.agent/skills/engineer-system-change/SKILL.md
skills/public/consulting-analysis/SKILL.md

Metadata

Files
0
Version
a5acc25
Hash
7a511606
Indexed
2026-08-20 17:01

- 위키
Copyright © 2011-2026 iteam. Current version is 2.155.2. UTC+08:00, 2026-08-22 11:01
浙ICP备14020137号-1 $방문자$