Agent SkillsYaoApp/yao › yao-image

yao-image

GitHub

图像专家技能,支持读取、分析及生成图片。提供image_read用于视觉模型描述,image_generate用于文生图,涵盖本地文件、URL及多参数配置。

tools/skills/yao-image/SKILL.md YaoApp/yao

Trigger Scenarios

需要读取或分析图片内容时 需要根据文本提示生成新图片时

Install

npx skills add YaoApp/yao --skill yao-image -g -y
More Options

Non-standard path

npx skills add https://github.com/YaoApp/yao/tree/main/tools/skills/yao-image -g -y

Use without installing

npx skills use YaoApp/yao@yao-image

指定 Agent (Claude Code)

npx skills add YaoApp/yao --skill yao-image -a claude-code -g -y

安装 repo 全部 skill

npx skills add YaoApp/yao --all -g -y

预览 repo 内 skill

npx skills add YaoApp/yao --list

SKILL.md

Frontmatter
{
    "name": "yao-image",
    "description": "Image expert. ALWAYS invoke this skill when you need to read, analyze, describe, or generate images. Use for screenshots, photos, charts, diagrams, AI-generated images, or any visual content."
}

Image Tools

Use these tools when you encounter images you cannot read natively, or when you need to generate new images.

image_read

Send an image to a vision-capable model and get a text description.

Local file (most common):

tai tool image_read --image_path /path/to/image.png --prompt "Describe this image"

URL:

tai tool image_read --image_path https://example.com/photo.jpg --prompt "What is shown?"

With a specific vision provider:

tai tool image_read --image_path /path/to/image.png --prompt "Describe" --provider llm.my-openai:gpt-4o
Parameter Type Required Description
image_path string yes Image file path or URL
prompt string no Analysis instruction (default: describe in detail)
max_size integer no Max dimension in pixels for longest edge (default: 1080)
provider string no Vision provider connector ID. If omitted, uses default vision model

Images are automatically resized (preserving aspect ratio) before sending to the vision model. Supported formats: PNG, JPEG, GIF, WebP.

image_generate

Generate a new image from a text prompt (text-to-image). For editing an existing image, use image_edit instead.

Basic usage (always specify output):

tai tool image_generate --prompt "A serene mountain landscape at sunset" --output landscape.png

With specific provider, model and size:

tai tool image_generate --prompt "A futuristic city skyline" --provider llm.my-openai --model gpt-image-1 --dimensions 1792x1024 --output output/city.png

Transparent background (for icons, stickers, product shots):

tai tool image_generate --prompt "A cute fox mascot" --background transparent --output_format png --output fox.png

WebP output with quality control:

tai tool image_generate --prompt "A landscape painting" --output_format webp --quality high --output painting.webp
Parameter Type Required Description
prompt string yes Text description of the image to generate
output string yes Output file path for the generated image
provider string no Provider connector ID (use image_providers to list). Auto-selects if omitted
dimensions string no Image dimensions (default: 1024x1024). Common: 1024x1024, 1024x1792, 1792x1024
model string no Model name to use. Overrides the provider's default model
background string no transparent, opaque, or auto. Use transparent for PNG/WebP with no background
output_format string no png, jpeg, or webp. Default: png
output_compression integer no Compression level 0-100 for jpeg/webp. Higher = better quality. Default: 100
quality string no low, medium, high, or auto. Higher takes longer. Default: auto
extra JSON no Provider-specific parameters as a JSON object (e.g. --extra '{"moderation":"low"}')

If output is omitted, the image is saved to a default path in the working directory.

image_edit

Edit or transform an existing image based on a text prompt (image-to-image). Use for style transfer, background replacement, adding/removing elements, or any modification that requires a reference image.

Basic usage:

tai tool image_edit --image_path /path/to/photo.png --prompt "Change the background to a beach scene" --output edited.png

With URL image:

tai tool image_edit --image_path https://example.com/photo.jpg --prompt "Make it look like a watercolor painting" --output watercolor.png

With specific provider and model:

tai tool image_edit --image_path /path/to/original.png --prompt "Remove the person in the foreground" --provider llm.my-openai --model gpt-image-1 --dimensions 1024x1024 --output result.png

With mask (edit only the masked region):

tai tool image_edit --image_path /path/to/photo.png --mask /path/to/mask.png --prompt "Replace with a garden" --output edited.png

Transparent background edit:

tai tool image_edit --image_path /path/to/product.png --prompt "Remove background" --background transparent --output_format png --output cutout.png
Parameter Type Required Description
image_path string yes Reference image file path or URL
prompt string yes Text description of the desired edit or transformation
output string yes Output file path for the edited image
provider string no Provider connector ID (use image_providers with capability=image_editing). Auto-selects if omitted
dimensions string no Output dimensions (default: 1024x1024). Common: 1024x1024, 1024x1792, 1792x1024
model string no Model name to use. Overrides the provider's default model
mask string no Mask image path/URL. Transparent areas in the mask define the editable region
background string no transparent, opaque, or auto. Use transparent for PNG/WebP with no background
output_format string no png, jpeg, or webp. Default: png
output_compression integer no Compression level 0-100 for jpeg/webp. Higher = better quality. Default: 100
quality string no low, medium, high, or auto. Higher takes longer. Default: auto
extra JSON no Provider-specific parameters as a JSON object (e.g. --extra '{"input_fidelity":"high"}')

If output is omitted, the image is saved to a default path in the working directory.

image_providers

List available image providers filtered by capability.

List image generation providers (default):

tai tool image_providers

List image editing providers:

tai tool image_providers --capability image_editing

List vision (image reading) providers:

tai tool image_providers --capability vision
Parameter Type Required Description
capability string no image_generation (default), image_editing, or vision

Returns a list of providers with their available models and connector IDs that can be passed to image_generate, image_edit, or image_read.

Constraints

Use only the parameters listed above for each tool. The supported first-class parameters are: prompt, output, provider, dimensions, model, background, output_format, output_compression, quality, mask (edit only), and extra.

Do not pass n, style, or response_format — they are unsupported and will be ignored or cause errors.

For provider-specific parameters not covered above (e.g. moderation, input_fidelity), pass them through the extra JSON object.

Version History

  • d90d41d Current 2026-09-23 10:26

    增强图像生成与编辑功能,新增背景处理、输出格式、质量设置等选项,支持透明背景和遮罩编辑。

  • 94ab88a 2026-08-20 19:13

Same Skill Collection

tools/skills/yao-agent/SKILL.md
tools/skills/yao-audio/SKILL.md
tools/skills/yao-board/SKILL.md
tools/skills/yao-decision/SKILL.md
tools/skills/yao-doc/SKILL.md
tools/skills/yao-ocr/SKILL.md
tools/skills/yao-process/SKILL.md
tools/skills/yao-secret/SKILL.md
tools/skills/yao-web/SKILL.md
tools/skills/yao-workspace-config/SKILL.md
tools/skills/yao-workspace/SKILL.md

Metadata

Files
0
Version
d90d41d
Hash
153dd2ae
Indexed
2026-08-20 19:13

Home - Wiki
Copyright © 2011-2026 iteam. Current version is 2.155.2. UTC+08:00, 2026-09-24 19:05
浙ICP备14020137号-1