moderate

GitHub

对文本进行安全审查,检测是否包含不安全、有害或违反政策的内容。支持OpenAI的13项分类标准,返回判定结果及置信度分数,适用于内容过滤和发布前安全检查。

agents/openclaw/plugin/skills/moderate/SKILL.md zerogpu/zerogpu-router

Trigger Scenarios

用户要求审核文本安全性 检查内容是否符合发布政策

Install

npx skills add zerogpu/zerogpu-router --skill moderate -g -y
More Options

Non-standard path

npx skills add https://github.com/zerogpu/zerogpu-router/tree/main/agents/openclaw/plugin/skills/moderate -g -y

Use without installing

npx skills use zerogpu/zerogpu-router@moderate

指定 Agent (Claude Code)

npx skills add zerogpu/zerogpu-router --skill moderate -a claude-code -g -y

安装 repo 全部 skill

npx skills add zerogpu/zerogpu-router --all -g -y

预览 repo 内 skill

npx skills add zerogpu/zerogpu-router --list

SKILL.md

Frontmatter
{
    "name": "moderate",
    "metadata": {
        "openclaw": {
            "install": [
                {
                    "bins": [
                        "zerogpu"
                    ],
                    "kind": "node",
                    "package": "zerogpu-cli"
                }
            ],
            "requires": {
                "bins": [
                    "zerogpu"
                ]
            }
        }
    },
    "description": "Screen text for unsafe, harmful, or policy-sensitive content and return a safety verdict across OpenAI's 13 moderation categories. Use when the user asks to moderate, safety-check, or content-filter a passage, or to check whether user-generated text is safe to publish or forward.",
    "allowed-tools": "Bash(zerogpu moderate *)",
    "argument-hint": "<text>"
}

Sends your input to ZeroGPU's hosted API for inference — this is not local processing. Don't pass secrets, credentials, or regulated data you aren't cleared to share with a third party. See the plugin README's "Data & privacy" section.

Run moderation. $ARGUMENTS is the raw text to screen — pass it verbatim, no escaping or quoting required (the heredoc below handles every shell metacharacter, newline, quote, and paren safely):

ZGPU_TEXT=$(cat <<'ZGPU_END_OF_INPUT'
$ARGUMENTS
ZGPU_END_OF_INPUT
)
zerogpu moderate "$ZGPU_TEXT"

Output is OpenAI's moderations envelope as JSON: a flagged verdict, per-category booleans in categories, and calibrated confidence scores in category_scores, across all 13 safety categories. Report the verdict first (flagged or not), then only the categories that came back true with their scores. Do not restate the flagged text itself.

Backed by zlm-v1-moderation-edge (86M parameters, $0.02 / $0.05 per 1M input/output tokens), an edge model that beats omni-moderation-latest on the binary safe/unsafe decision and on 9 of 13 categories. This model is served by the Moderations API rather than the Responses API; the CLI routes it automatically.

Screening text is a safety check, not an endorsement. Run it on request even when the passage is unpleasant — reporting that something is flagged is the whole point of the skill.

Savings note: only if the command output literally contains a line starting with 💰 ZeroGPU savings, append that exact line, unchanged, as the last line of your reply. If no such line is present, say nothing about savings and do not mention or suggest the cost-savings skill — this note is intentionally occasional, not shown every time.

Version History

  • 87b63e9 Current 2026-08-27 16:49

Same Skill Collection

agents/claude/skills/chat-deepseek/SKILL.md
agents/claude/skills/chat-glm/SKILL.md
agents/claude/skills/chat-liquid/SKILL.md
agents/claude/skills/chat-qwen/SKILL.md
agents/claude/skills/chat-thinking/SKILL.md
agents/claude/skills/chat/SKILL.md
agents/claude/skills/classify-domain/SKILL.md
agents/claude/skills/classify-iab-enriched/SKILL.md
agents/claude/skills/classify-iab/SKILL.md
agents/claude/skills/classify-structured/SKILL.md
agents/claude/skills/classify-zero-shot/SKILL.md
agents/claude/skills/cost-savings/SKILL.md
agents/claude/skills/embed/SKILL.md
agents/claude/skills/extract-entities/SKILL.md
agents/claude/skills/extract-json/SKILL.md
agents/claude/skills/extract-pii/SKILL.md
agents/claude/skills/generate-followups/SKILL.md
agents/claude/skills/moderate/SKILL.md
agents/claude/skills/redact-pii/SKILL.md
agents/claude/skills/signin/SKILL.md
agents/claude/skills/summarize/SKILL.md
agents/openclaw/plugin/skills/chat-deepseek/SKILL.md
agents/openclaw/plugin/skills/chat-glm/SKILL.md
agents/openclaw/plugin/skills/chat-liquid/SKILL.md
agents/openclaw/plugin/skills/chat-qwen/SKILL.md
agents/openclaw/plugin/skills/chat-thinking/SKILL.md
agents/openclaw/plugin/skills/chat/SKILL.md
agents/openclaw/plugin/skills/classify-domain/SKILL.md
agents/openclaw/plugin/skills/classify-iab-enriched/SKILL.md
agents/openclaw/plugin/skills/classify-iab/SKILL.md
agents/openclaw/plugin/skills/classify-structured/SKILL.md
agents/openclaw/plugin/skills/classify-zero-shot/SKILL.md
agents/openclaw/plugin/skills/cost-savings/SKILL.md
agents/openclaw/plugin/skills/embed/SKILL.md
agents/openclaw/plugin/skills/extract-entities/SKILL.md
agents/openclaw/plugin/skills/extract-json/SKILL.md
agents/openclaw/plugin/skills/extract-pii/SKILL.md
agents/openclaw/plugin/skills/generate-followups/SKILL.md
agents/openclaw/plugin/skills/redact-pii/SKILL.md
agents/openclaw/plugin/skills/signin/SKILL.md
agents/openclaw/plugin/skills/summarize/SKILL.md
agents/openclaw/plugin/skills/zerogpu-summarize/SKILL.md
agents/claude/skills/status/SKILL.md
agents/openclaw/plugin/skills/status/SKILL.md

Metadata

Files
0
Version
87b63e9
Hash
a87189bc
Indexed
2026-08-27 16:49

Accueil - Wiki
Copyright © 2011-2026 iteam. Current version is 2.155.2. UTC+08:00, 2026-08-27 21:59
浙ICP备14020137号-1 $Carte des visiteurs$