image-gen

GitHub

用于通过后端 API 生成和编辑 AI 图像的技能,支持多种模型选择、参数配置及进度追踪,适用于需要创建或修改图片的场景。

claude/skills/image-gen/SKILL.md ChatCut-Inc/agent-plugin

触发场景

用户请求生成图片或照片 用户请求基于参考图进行图像编辑

安装

npx skills add ChatCut-Inc/agent-plugin --skill image-gen -g -y
更多选项

非标准路径

npx skills add https://github.com/ChatCut-Inc/agent-plugin/tree/main/claude/skills/image-gen -g -y

不安装直接使用

npx skills use ChatCut-Inc/agent-plugin@image-gen

指定 Agent (Claude Code)

npx skills add ChatCut-Inc/agent-plugin --skill image-gen -a claude-code -g -y

安装 repo 全部 skill

npx skills add ChatCut-Inc/agent-plugin --all -g -y

预览 repo 内 skill

npx skills add ChatCut-Inc/agent-plugin --list

SKILL.md

Frontmatter
{
    "name": "image-gen",
    "description": "AI image generation and reference-image editing via GPT Image 2.5 Flare\/Sunburst, GPT Image 2, and Nano Banana. Use when the user wants to generate or create an image \/ picture \/ still through the backend image-generation jobs.\n",
    "user-invocable": true
}

Image Gen

ChatCut image generation and editing require an active paid Pro subscription and use credits for every model. If submit_image returns FEATURE_NOT_INCLUDED, surface the upgrade requirement; do not retry another ChatCut model to bypass it. Codex native image generation keeps its own host entitlement and is not governed by this ChatCut requirement.

Generate AI images through the backend generation API. Submit-only: creates a generation job and returns a jobId.

After submission, use track_progress tool to check status or wait for completion.

Model Selection

Model Strengths Max refs
gpt-image-2.5-flare Default: fast everyday generation and editing 10
gpt-image-2.5-sunburst Precision edits and detailed creative work 10
gpt-image-2 Previous GPT image model 10
nano-banana Strongest reference-image fidelity 14
  • gpt-image-2.5-flare is the default. Honor the image model attached to the user prompt or explicitly requested.
  • Use gpt-image-2.5-sunburst when the user selects it for precise editing.
  • nano-banana is the reference-heavy choice. Use it when reference-image fidelity matters more than text rendering, or when the user needs more than 10 reference images.

IMPORTANT: Before generating, READ the model's reference document for params, limits, and prompt tips:

Tool Params

Param Values Default
aspectRatio 1:1, 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, 4:5, 5:4, 21:9 16:9
imageSize 1K, 2K, 4K 1K
quality low, medium, high, auto; Image 2.5 also xhigh, max high
background auto, opaque, transparent (Image 2.5 only; PNG output) auto
referenceAssetIds Array of project asset ids — backend resolves bytes server-side
name Short descriptive asset name shown in the library
count Number of images to generate (1–10, each becomes a separate job) 1

Defaults

  • Aspect ratio: 16:9. If the project composition is not 16:9, ASK the user which aspect ratio they want before generating.
  • Size: 1K.

Ask Before Submit

  • Never auto-upgrade size.
  • Only pass imageSize: "2K" or "4K" when the user explicitly asks. Warn that 2K/4K are EXPERIMENTAL and may be slower.

Reference Images

Use when the user provides source material to edit, blend, or use as visual guidance (e.g. "change the background", "combine these into a poster").

  • Pass project asset ids via referenceAssetIds. The backend fetches and encodes them server-side — never pull the asset bytes yourself.
  • When the user @-references an image asset, pass its id directly in referenceAssetIds.
  • Formats accepted by backend: png, jpeg, webp, svg (auto-rasterized to png), heic, heif. Each ≤ 50MB.

Run

// Basic generation
submit_image({
  model: "gpt-image-2.5-flare",
  prompt: "a cute orange cat",
  name: "Cat",
});

// With quality (OpenAI models)
submit_image({
  model: "gpt-image-2.5-flare",
  prompt: "hero poster with bold title",
  quality: "high",
  name: "Hero Poster",
});

// With reference images — pass project asset ids; backend resolves bytes
submit_image({
  model: "gpt-image-2.5-flare",
  prompt: "change background to beach",
  referenceAssetIds: ["<assetId>"],
  name: "Beach Edit",
});

// Reference-heavy with nano-banana
submit_image({
  model: "nano-banana",
  prompt: "composite poster",
  referenceAssetIds: ["<id1>", "<id2>"],
  name: "Composite",
});

// Multiple images
submit_image({
  model: "gpt-image-2.5-flare",
  prompt: "product shots",
  count: 3,
  name: "Product",
});

After submission, call track_progress with action=status jobIds=<jobId>. If the current task depends on a non-terminal result, follow its checkBackAfterSeconds sleep guidance before each later status check; action=wait is only a non-blocking compatibility alias.

Rules

  • Always provide name with a short descriptive asset name.
  • Wait for completion unless the user only asked to queue. Use returned asset ids for further edits; use edit_item to place an image on the timeline when requested.
  • Generation costs credits. Before submitting, briefly tell the user what you're about to generate — especially when generating multiple images.
  • Do not use this skill for job management. Use track_progress tool for that.

版本历史

  • 506b5b0 当前 2026-09-11 19:17

    新增 gpt-image-2.5-flare 和 gpt-image-2.5-sunburst 模型支持,增加背景透明等参数选项,更新默认模型为 flare,并补充订阅限制说明。

  • f39cdae 2026-07-30 21:53
  • 5e9afe0 2026-07-22 10:57

同 Skill 集合

chatcut-desktop-codex-plugin/skills/connect-chatcut-desktop/SKILL.md
chatcut/skills/asset-import/SKILL.md
chatcut/skills/chatcut-plugin-basics-claude/SKILL.md
chatcut/skills/chatcut-plugin-basics/SKILL.md
chatcut/skills/create-motion-graphics/SKILL.md
chatcut/skills/export/SKILL.md
chatcut/skills/image-gen/SKILL.md
chatcut/skills/known-errors/SKILL.md
chatcut/skills/music/SKILL.md
chatcut/skills/shader-gen/SKILL.md
chatcut/skills/transcription/SKILL.md
chatcut/skills/verification/SKILL.md
chatcut/skills/video-gen/SKILL.md
chatcut/skills/voice/SKILL.md
chatcut/skills/widget-forms/SKILL.md
claude/skills/create-motion-graphics/SKILL.md
claude/skills/export/SKILL.md
claude/skills/known-errors/SKILL.md
claude/skills/motion-graphic-gen/SKILL.md
claude/skills/music/SKILL.md
claude/skills/shader-gen/SKILL.md
claude/skills/transcription/SKILL.md
claude/skills/verification/SKILL.md
claude/skills/video-gen/SKILL.md
claude/skills/widget-forms/SKILL.md
codex/skills/asset-import/SKILL.md
codex/skills/chatcut-plugin-basics/SKILL.md
codex/skills/create-motion-graphics/SKILL.md
codex/skills/export/SKILL.md
codex/skills/image-gen/SKILL.md
codex/skills/known-errors/SKILL.md
codex/skills/music/SKILL.md
codex/skills/shader-gen/SKILL.md
codex/skills/transcription/SKILL.md
codex/skills/verification/SKILL.md
codex/skills/video-gen/SKILL.md
codex/skills/widget-forms/SKILL.md
chatcut/skills/product-help/SKILL.md
chatcut/skills/talking-head-guide/SKILL.md
claude/skills/asset-import/SKILL.md
claude/skills/chatcut-plugin-basics-claude/SKILL.md
claude/skills/digital-human/SKILL.md
claude/skills/multicam-sync/SKILL.md
claude/skills/product-help/SKILL.md
claude/skills/talking-head-guide/SKILL.md
claude/skills/video-translation/SKILL.md
claude/skills/voice/SKILL.md
codex/skills/digital-human/SKILL.md
codex/skills/multicam-sync/SKILL.md

元信息

文件数
0
版本
506b5b0
Hash
cf300f5b
收录时间
2026-07-22 10:57

首页 - Wiki
Copyright © 2011-2026 iteam. Current version is 2.155.2. UTC+08:00, 2026-09-12 01:49
浙ICP备14020137号-1 $访客地图$