Agent Skills › vllm-project/vllm-omni › handdrawn-live-video-generator

handdrawn-live-video-generator

GitHub

生成手绘风格与实拍融合的短视频提示词或视频,通过发光手绘实体接触、变形及逃脱镜头的创意流程,打造温暖超现实的混合媒体短片。

.agents/skills/handdrawn-live-video-generator/SKILL.md vllm-project/vllm-omni

Trigger Scenarios

需要生成手绘与实拍结合的视频 请求创作超现实混合媒体短片

Install

npx skills add vllm-project/vllm-omni --skill handdrawn-live-video-generator -g -y
More Options

Non-standard path

npx skills add https://github.com/vllm-project/vllm-omni/tree/main/.agents/skills/handdrawn-live-video-generator -g -y

Use without installing

npx skills use vllm-project/vllm-omni@handdrawn-live-video-generator

指定 Agent (Claude Code)

npx skills add vllm-project/vllm-omni --skill handdrawn-live-video-generator -a claude-code -g -y

安装 repo 全部 skill

npx skills add vllm-project/vllm-omni --all -g -y

预览 repo 内 skill

npx skills add vllm-project/vllm-omni --list

SKILL.md

Frontmatter
{
    "name": "handdrawn-live-video-generator",
    "metadata": {
        "compatibility": "Portable prompt writing with sibling h3-prompt-writing; actual generation requires an available video backend and any referenced media."
    },
    "description": "Write or generate a continuous live-action scene with a rough luminous hand-drawn entity that contacts real objects, morphs, and escapes a slightly delayed handheld camera. Use for gentle surreal mixed-media shorts, not flat-vector-only animation or horror clips."
}

Handdrawn Live-Action Fusion

Read the portable workflow. Keep the user's requested scope and language. A creative-prompt request delivers the prompt, without generating media. For a requested finished video, prepare the prompt and proceed under existing authorization using available tools.

Creative contract

The default is a 15-second, 16:9 continuous scene in an everyday real space. A flat luminous drawn entity visibly contacts a real hand/object early, transforms as one traceable being, escapes through connected space, and the handheld camera reacts slightly late. Duration and ratio follow the user's explicit choices.

Use rough crayon, chalk, colored pencil, pastel, or brush strokes with uneven fill, frayed edges, line jitter, and frame-by-frame redraw. Keep the live-action setting physically grounded and the drawn presence visibly planar. Uniform neon tubes, plush characters, polished CG, and clean vector strokes would change this style.

Choose a fresh setting, entity, palette, contact mechanism, transformation chain, escape route, camera reaction, and emotional ending within the user's constraints. Do not ban an otherwise requested motif simply because an upstream example used it. The tone is warm, playful, and gently surprising rather than threatening.

Timed action structure

For a 15-second brief, use these intervals as a starting structure:

  • 0–3s: establish obvious real/drawn contact, such as landing on a palm or slipping around a finger. The real participant reacts.
  • 3–6s: begin a continuous morph while retaining a color smear, line, curve, or tail from the previous shape. Give it a playful action.
  • 6–10s: let it escape into an adjacent part of the same space. The camera pans, tilts, or moves only after the entity starts leaving its view.
  • 10–13s: add another readable interaction or surprise, keeping the same entity.
  • 13–15s: extend an earlier drawn motif across the space into a gentle larger transformation, then resolve with a small emotional/comic detail.

Retiming changes the intervals proportionally or reorganizes beats; do not leave 15-second timestamps in a shorter prompt. Preserve a user's explicit final form instead of forcing the spatial finale when it conflicts with their ending.

Prompt delivery and rendering

For a same-language creative prompt, begin with duration, ratio, real space, and the live-action/drawing fusion. Then describe phone-camera texture, the timed action intervals, drawn material, camera lag, specific exclusions, and ambient sound. Avoid unrelated headings or an essay when the user asked for copy-ready text.

For actual H3 generation, create a separate structured rendering prompt using the appropriate base/Ref2VA guide: English prose with literal dialogue/visible text in the requested language. Preserve the creative intent and actual media bindings. Do not send unsupported model settings or promise exact first-frame conditioning merely because an image is present.

Review

Check clear early physical contact, one continuous entity, connected geography, retained shape motifs, delayed camera reaction, rough planar drawing, and a non-threatening ending. Avoid jump scares, sudden blackouts, menacing anatomy, and unrequested extra characters. Deliver the actual prompt/video and state any remaining contact, continuity, texture, or camera-timing defects.

Version History

  • f7e2834 Current 2026-09-22 16:47

Same Skill Collection

.agents/skills/3d-animation-short-generator/SKILL.md
.agents/skills/brand-promo-video-generator/SKILL.md
.agents/skills/co-op-game-intro-generator/SKILL.md
.agents/skills/h3-prompt-writing/SKILL.md
.agents/skills/minimalist-product-ad-generator/SKILL.md
.agents/skills/music-video-subtitle-generator/SKILL.md
.agents/skills/paper-collage-explainer-generator/SKILL.md
.agents/skills/papercraft-stop-motion-explainer/SKILL.md
.claude/skills/add-diffusion-model/SKILL.md
.claude/skills/add-recipe/SKILL.md
.claude/skills/add-tts-model/SKILL.md
.claude/skills/diffusion-perf-opt/SKILL.md
.claude/skills/find-simplifications/SKILL.md
.claude/skills/precheck-pr/SKILL.md
.claude/skills/quantization/SKILL.md
.claude/skills/vllm-omni-npu-upgrade/SKILL.md
.claude/skills/vllm-omni-test/SKILL.md
.claude/skills/production-add-diffusion-model/SKILL.md
.claude/skills/review-pr/SKILL.md

Metadata

Files
0
Version
817f5d0
Hash
6c94484f
Indexed
2026-09-22 16:47

ホーム - Wiki
Copyright © 2011-2026 iteam. Current version is 2.155.2. UTC+08:00, 2026-09-29 19:04
浙ICP备14020137号-1