Agent Skillsvllm-project/vllm-omni › handdrawn-live-video-generator

handdrawn-live-video-generator

GitHub

用于生成手绘风格与实景融合的超现实短视频提示词,包含分镜结构、视觉风格定义及渲染指令,适用于温和奇幻短片创作。

.agents/skills/handdrawn-live-video-generator/SKILL.md vllm-project/vllm-omni

触发场景

需要生成手绘与真人结合的视频脚本 请求超现实混合媒体短片创意 定制特定时长的视频分镜描述

安装

npx skills add vllm-project/vllm-omni --skill handdrawn-live-video-generator -g -y
更多选项

非标准路径

npx skills add https://github.com/vllm-project/vllm-omni/tree/main/.agents/skills/handdrawn-live-video-generator -g -y

不安装直接使用

npx skills use vllm-project/vllm-omni@handdrawn-live-video-generator

指定 Agent (Claude Code)

npx skills add vllm-project/vllm-omni --skill handdrawn-live-video-generator -a claude-code -g -y

安装 repo 全部 skill

npx skills add vllm-project/vllm-omni --all -g -y

预览 repo 内 skill

npx skills add vllm-project/vllm-omni --list

SKILL.md

Frontmatter
{
    "name": "handdrawn-live-video-generator",
    "metadata": {
        "compatibility": "Portable prompt writing with sibling h3-prompt-writing; actual generation requires an available video backend and any referenced media."
    },
    "description": "Write or generate a continuous live-action scene with a rough luminous hand-drawn entity that contacts real objects, morphs, and escapes a slightly delayed handheld camera. Use for gentle surreal mixed-media shorts, not flat-vector-only animation or horror clips."
}

Handdrawn Live-Action Fusion

Read the portable workflow. Keep the user's requested scope and language. A creative-prompt request delivers the prompt, without generating media. For a requested finished video, prepare the prompt and proceed under existing authorization using available tools.

Creative contract

The default is a 15-second, 16:9 continuous scene in an everyday real space. A flat luminous drawn entity visibly contacts a real hand/object early, transforms as one traceable being, escapes through connected space, and the handheld camera reacts slightly late. Duration and ratio follow the user's explicit choices.

Use rough crayon, chalk, colored pencil, pastel, or brush strokes with uneven fill, frayed edges, line jitter, and frame-by-frame redraw. Keep the live-action setting physically grounded and the drawn presence visibly planar. Uniform neon tubes, plush characters, polished CG, and clean vector strokes would change this style.

Choose a fresh setting, entity, palette, contact mechanism, transformation chain, escape route, camera reaction, and emotional ending within the user's constraints. Do not ban an otherwise requested motif simply because an upstream example used it. The tone is warm, playful, and gently surprising rather than threatening.

Timed action structure

For a 15-second brief, use these intervals as a starting structure:

  • 0–3s: establish obvious real/drawn contact, such as landing on a palm or slipping around a finger. The real participant reacts.
  • 3–6s: begin a continuous morph while retaining a color smear, line, curve, or tail from the previous shape. Give it a playful action.
  • 6–10s: let it escape into an adjacent part of the same space. The camera pans, tilts, or moves only after the entity starts leaving its view.
  • 10–13s: add another readable interaction or surprise, keeping the same entity.
  • 13–15s: extend an earlier drawn motif across the space into a gentle larger transformation, then resolve with a small emotional/comic detail.

Retiming changes the intervals proportionally or reorganizes beats; do not leave 15-second timestamps in a shorter prompt. Preserve a user's explicit final form instead of forcing the spatial finale when it conflicts with their ending.

Prompt delivery and rendering

For a same-language creative prompt, begin with duration, ratio, real space, and the live-action/drawing fusion. Then describe phone-camera texture, the timed action intervals, drawn material, camera lag, specific exclusions, and ambient sound. Avoid unrelated headings or an essay when the user asked for copy-ready text.

For actual H3 generation, create a separate structured rendering prompt using the appropriate base/Ref2VA guide: English prose with literal dialogue/visible text in the requested language. Preserve the creative intent and actual media bindings. Do not send unsupported model settings or promise exact first-frame conditioning merely because an image is present.

Review

Check clear early physical contact, one continuous entity, connected geography, retained shape motifs, delayed camera reaction, rough planar drawing, and a non-threatening ending. Avoid jump scares, sudden blackouts, menacing anatomy, and unrequested extra characters. Deliver the actual prompt/video and state any remaining contact, continuity, texture, or camera-timing defects.

版本历史

  • f7e2834 当前 2026-09-22 16:47

同 Skill 集合

.agents/skills/3d-animation-short-generator/SKILL.md
.agents/skills/brand-promo-video-generator/SKILL.md
.agents/skills/co-op-game-intro-generator/SKILL.md
.agents/skills/h3-prompt-writing/SKILL.md
.agents/skills/minimalist-product-ad-generator/SKILL.md
.agents/skills/music-video-subtitle-generator/SKILL.md
.agents/skills/paper-collage-explainer-generator/SKILL.md
.agents/skills/papercraft-stop-motion-explainer/SKILL.md
.claude/skills/add-diffusion-model/SKILL.md
.claude/skills/add-recipe/SKILL.md
.claude/skills/add-tts-model/SKILL.md
.claude/skills/diffusion-perf-opt/SKILL.md
.claude/skills/find-simplifications/SKILL.md
.claude/skills/precheck-pr/SKILL.md
.claude/skills/quantization/SKILL.md
.claude/skills/vllm-omni-npu-upgrade/SKILL.md
.claude/skills/vllm-omni-test/SKILL.md
.claude/skills/production-add-diffusion-model/SKILL.md
.claude/skills/review-pr/SKILL.md

元信息

文件数
0
版本
f7e2834
Hash
6c94484f
收录时间
2026-09-22 16:47

首页 - Wiki
Copyright © 2011-2026 iteam. Current version is 2.155.2. UTC+08:00, 2026-09-22 18:52
浙ICP备14020137号-1