moderate

GitHub

用于对文本进行安全审核,检测是否包含违规或敏感内容。基于OpenAI的13项分类标准返回安全判定结果及置信度分数,适用于发布前内容过滤等场景。

agents/claude/skills/moderate/SKILL.md zerogpu/zerogpu-router

Trigger Scenarios

用户要求审核文本安全性 检查用户生成内容是否适合发布

Install

npx skills add zerogpu/zerogpu-router --skill moderate -g -y
More Options

Non-standard path

npx skills add https://github.com/zerogpu/zerogpu-router/tree/main/agents/claude/skills/moderate -g -y

Use without installing

npx skills use zerogpu/zerogpu-router@moderate

指定 Agent (Claude Code)

npx skills add zerogpu/zerogpu-router --skill moderate -a claude-code -g -y

安装 repo 全部 skill

npx skills add zerogpu/zerogpu-router --all -g -y

预览 repo 内 skill

npx skills add zerogpu/zerogpu-router --list

SKILL.md

Frontmatter
{
    "name": "moderate",
    "description": "Screen text for unsafe, harmful, or policy-sensitive content and return a safety verdict across OpenAI's 13 moderation categories. Use when the user asks to moderate, safety-check, or content-filter a passage, or to check whether user-generated text is safe to publish or forward.",
    "allowed-tools": "Bash(zerogpu moderations *)",
    "argument-hint": "<text>"
}

Run moderation. $ARGUMENTS is the raw text to screen — pass it verbatim, no escaping or quoting required (the heredoc below handles every shell metacharacter, newline, quote, and paren safely):

zerogpu moderations -m zlm-v1-moderation-edge <<'ZGPU_END_OF_INPUT'
$ARGUMENTS
ZGPU_END_OF_INPUT

Output is OpenAI's moderations envelope as JSON: for each entry in results, a flagged verdict, per-category booleans in categories, and calibrated confidence scores in category_scores, across all 13 safety categories. Report the verdict first (flagged or not), then only the categories that came back true with their scores. Do not restate the flagged text itself.

Backed by zlm-v1-moderation-edge (86M parameters, $0.02 / $0.05 per 1M input/output tokens), an edge model that beats omni-moderation-latest on the binary safe/unsafe decision and on 9 of 13 categories. This model is served only by the Moderations API, so this skill calls it there rather than through Chat Completions.

Screening text is a safety check, not an endorsement. Run it on request even when the passage is unpleasant — reporting that something is flagged is the whole point of the skill.

Savings note: only if the command output literally contains a line starting with 💰 ZeroGPU savings, append that exact line, unchanged, as the last line of your reply. If no such line is present, say nothing about savings and do not mention or suggest /zerogpu-router:cost-savings — this note is intentionally occasional, not shown every time.

Version History

  • 4e1b070 Current 2026-09-22 08:07

    升级至v2.3.0:将命令调用方式改为模型无关的CLI端点(zerogpu moderations),使插件独立于CLI模型变更,需依赖zerogpu-cli >= 3.8.0。

  • 87b63e9 2026-08-27 16:48

Same Skill Collection

agents/claude/skills/chat-deepseek-v4-1-flash/SKILL.md
agents/claude/skills/chat-deepseek/SKILL.md
agents/claude/skills/chat-glm/SKILL.md
agents/claude/skills/chat-liquid/SKILL.md
agents/claude/skills/chat-qwen/SKILL.md
agents/claude/skills/chat-thinking/SKILL.md
agents/claude/skills/chat/SKILL.md
agents/claude/skills/classify-domain/SKILL.md
agents/claude/skills/classify-iab-enriched/SKILL.md
agents/claude/skills/classify-iab/SKILL.md
agents/claude/skills/classify-structured/SKILL.md
agents/claude/skills/classify-zero-shot/SKILL.md
agents/claude/skills/cost-savings/SKILL.md
agents/claude/skills/embed/SKILL.md
agents/claude/skills/extract-entities/SKILL.md
agents/claude/skills/extract-json/SKILL.md
agents/claude/skills/extract-pii/SKILL.md
agents/claude/skills/extract-signals/SKILL.md
agents/claude/skills/generate-followups/SKILL.md
agents/claude/skills/moderate-llama/SKILL.md
agents/claude/skills/redact-pii/SKILL.md
agents/claude/skills/signin/SKILL.md
agents/claude/skills/summarize/SKILL.md
agents/openclaw/plugin/skills/chat-deepseek-v4-1-flash/SKILL.md
agents/openclaw/plugin/skills/chat-deepseek/SKILL.md
agents/openclaw/plugin/skills/chat-glm/SKILL.md
agents/openclaw/plugin/skills/chat-liquid/SKILL.md
agents/openclaw/plugin/skills/chat-qwen/SKILL.md
agents/openclaw/plugin/skills/chat-thinking/SKILL.md
agents/openclaw/plugin/skills/chat/SKILL.md
agents/openclaw/plugin/skills/classify-domain/SKILL.md
agents/openclaw/plugin/skills/classify-iab-enriched/SKILL.md
agents/openclaw/plugin/skills/classify-iab/SKILL.md
agents/openclaw/plugin/skills/classify-structured/SKILL.md
agents/openclaw/plugin/skills/classify-zero-shot/SKILL.md
agents/openclaw/plugin/skills/cost-savings/SKILL.md
agents/openclaw/plugin/skills/embed/SKILL.md
agents/openclaw/plugin/skills/extract-entities/SKILL.md
agents/openclaw/plugin/skills/extract-json/SKILL.md
agents/openclaw/plugin/skills/extract-pii/SKILL.md
agents/openclaw/plugin/skills/extract-signals/SKILL.md
agents/openclaw/plugin/skills/generate-followups/SKILL.md
agents/openclaw/plugin/skills/moderate-llama/SKILL.md
agents/openclaw/plugin/skills/moderate/SKILL.md
agents/openclaw/plugin/skills/redact-pii/SKILL.md
agents/openclaw/plugin/skills/signin/SKILL.md
agents/openclaw/plugin/skills/status/SKILL.md
agents/openclaw/plugin/skills/summarize/SKILL.md
agents/openclaw/plugin/skills/zerogpu-summarize/SKILL.md

Metadata

Files
0
Version
4e1b070
Hash
7f61f10f
Indexed
2026-08-27 16:48

Home - Wiki
Copyright © 2011-2026 iteam. Current version is 2.155.2. UTC+08:00, 2026-09-27 07:32
浙ICP备14020137号-1