moderate

GitHub

调用 ZeroGPU API 对文本进行内容安全审核,基于 OpenAI 13个分类判断是否违规,返回标记结果、类别及置信度分数。

agents/openclaw/plugin/skills/moderate/SKILL.md zerogpu/zerogpu-router

触发场景

用户要求审查文本安全性 检查内容是否符合发布政策

安装

npx skills add zerogpu/zerogpu-router --skill moderate -g -y
更多选项

非标准路径

npx skills add https://github.com/zerogpu/zerogpu-router/tree/main/agents/openclaw/plugin/skills/moderate -g -y

不安装直接使用

npx skills use zerogpu/zerogpu-router@moderate

指定 Agent (Claude Code)

npx skills add zerogpu/zerogpu-router --skill moderate -a claude-code -g -y

安装 repo 全部 skill

npx skills add zerogpu/zerogpu-router --all -g -y

预览 repo 内 skill

npx skills add zerogpu/zerogpu-router --list

SKILL.md

Frontmatter
{
    "name": "moderate",
    "metadata": {
        "openclaw": {
            "install": [
                {
                    "bins": [
                        "zerogpu"
                    ],
                    "kind": "node",
                    "package": "zerogpu-cli"
                }
            ],
            "requires": {
                "bins": [
                    "zerogpu"
                ]
            }
        }
    },
    "description": "Screen text for unsafe, harmful, or policy-sensitive content and return a safety verdict across OpenAI's 13 moderation categories. Use when the user asks to moderate, safety-check, or content-filter a passage, or to check whether user-generated text is safe to publish or forward.",
    "allowed-tools": "Bash(zerogpu moderate *)",
    "argument-hint": "<text>"
}

Sends your input to ZeroGPU's hosted API for inference — this is not local processing. Don't pass secrets, credentials, or regulated data you aren't cleared to share with a third party. See the plugin README's "Data & privacy" section.

Run moderation. $ARGUMENTS is the raw text to screen — pass it verbatim, no escaping or quoting required (the heredoc below handles every shell metacharacter, newline, quote, and paren safely):

ZGPU_TEXT=$(cat <<'ZGPU_END_OF_INPUT'
$ARGUMENTS
ZGPU_END_OF_INPUT
)
zerogpu moderate "$ZGPU_TEXT"

Output is OpenAI's moderations envelope as JSON: a flagged verdict, per-category booleans in categories, and calibrated confidence scores in category_scores, across all 13 safety categories. Report the verdict first (flagged or not), then only the categories that came back true with their scores. Do not restate the flagged text itself.

Backed by zlm-v1-moderation-edge (86M parameters, $0.02 / $0.05 per 1M input/output tokens), an edge model that beats omni-moderation-latest on the binary safe/unsafe decision and on 9 of 13 categories. This model is served by the Moderations API rather than the Responses API; the CLI routes it automatically.

Screening text is a safety check, not an endorsement. Run it on request even when the passage is unpleasant — reporting that something is flagged is the whole point of the skill.

Savings note: only if the command output literally contains a line starting with 💰 ZeroGPU savings, append that exact line, unchanged, as the last line of your reply. If no such line is present, say nothing about savings and do not mention or suggest the cost-savings skill — this note is intentionally occasional, not shown every time.

版本历史

  • 87b63e9 当前 2026-08-27 16:49

同 Skill 集合

agents/claude/skills/chat-deepseek/SKILL.md
agents/claude/skills/chat-glm/SKILL.md
agents/claude/skills/chat-liquid/SKILL.md
agents/claude/skills/chat-qwen/SKILL.md
agents/claude/skills/chat-thinking/SKILL.md
agents/claude/skills/chat/SKILL.md
agents/claude/skills/classify-domain/SKILL.md
agents/claude/skills/classify-iab-enriched/SKILL.md
agents/claude/skills/classify-iab/SKILL.md
agents/claude/skills/classify-structured/SKILL.md
agents/claude/skills/classify-zero-shot/SKILL.md
agents/claude/skills/cost-savings/SKILL.md
agents/claude/skills/embed/SKILL.md
agents/claude/skills/extract-entities/SKILL.md
agents/claude/skills/extract-json/SKILL.md
agents/claude/skills/extract-pii/SKILL.md
agents/claude/skills/generate-followups/SKILL.md
agents/claude/skills/moderate/SKILL.md
agents/claude/skills/redact-pii/SKILL.md
agents/claude/skills/signin/SKILL.md
agents/claude/skills/summarize/SKILL.md
agents/openclaw/plugin/skills/chat-deepseek/SKILL.md
agents/openclaw/plugin/skills/chat-glm/SKILL.md
agents/openclaw/plugin/skills/chat-liquid/SKILL.md
agents/openclaw/plugin/skills/chat-qwen/SKILL.md
agents/openclaw/plugin/skills/chat-thinking/SKILL.md
agents/openclaw/plugin/skills/chat/SKILL.md
agents/openclaw/plugin/skills/classify-domain/SKILL.md
agents/openclaw/plugin/skills/classify-iab-enriched/SKILL.md
agents/openclaw/plugin/skills/classify-iab/SKILL.md
agents/openclaw/plugin/skills/classify-structured/SKILL.md
agents/openclaw/plugin/skills/classify-zero-shot/SKILL.md
agents/openclaw/plugin/skills/cost-savings/SKILL.md
agents/openclaw/plugin/skills/embed/SKILL.md
agents/openclaw/plugin/skills/extract-entities/SKILL.md
agents/openclaw/plugin/skills/extract-json/SKILL.md
agents/openclaw/plugin/skills/extract-pii/SKILL.md
agents/openclaw/plugin/skills/generate-followups/SKILL.md
agents/openclaw/plugin/skills/redact-pii/SKILL.md
agents/openclaw/plugin/skills/signin/SKILL.md
agents/openclaw/plugin/skills/summarize/SKILL.md
agents/openclaw/plugin/skills/zerogpu-summarize/SKILL.md
agents/claude/skills/status/SKILL.md
agents/openclaw/plugin/skills/status/SKILL.md

元信息

文件数
0
版本
87b63e9
Hash
a87189bc
收录时间
2026-08-27 16:49

首页 - Wiki
Copyright © 2011-2026 iteam. Current version is 2.155.2. UTC+08:00, 2026-09-06 16:31
浙ICP备14020137号-1 $访客地图$