Agent Skills
› Hmbown/Codewhale
› pdf
处理PDF文件的通用技能,支持读取、提取文本、OCR识别、拆分、合并、旋转、水印、填表及创建等功能。强调工具选择策略、结果验证及数据安全,适用于以PDF为核心输入输出的各类任务场景。
触发场景
需要读取或解析PDF内容
对PDF进行拆分、合并或格式转换
从扫描版PDF中提取文字
生成或修改PDF文档
安装
npx skills add Hmbown/Codewhale --skill pdf -g -y
SKILL.md
Frontmatter
{
"name": "pdf",
"description": "Read, extract, split, merge, rotate, watermark, fill, OCR, or create PDF files with verification of page counts and text extraction."
}
Use this skill for any task where a PDF is the primary input or output.
Workflow
- Identify the PDF operation: read, extract, OCR, split, merge, rotate, watermark, redact, fill forms, encrypt/decrypt, or create.
- Preserve originals. Write outputs with explicit names.
- Use the most reliable available tool:
- the built-in
Filetool (action: "read") for basic text extraction from PDFs pdftotext,pdfinfo,qpdf, ormutoolwhen installed- Python libraries such as
pypdf,pdfplumber,PyMuPDF, orreportlabwhen available - OCR tools only for scanned pages
- the built-in
- For extraction, report page coverage and note when layout, tables, or OCR quality may affect accuracy.
- For generated or modified PDFs, verify page count, text extraction where possible, and file size. For redaction, confirm removed text is not extractable from the output.
Ask before installing dependencies or running OCR over large documents. Do not represent a visually scanned PDF as fully accurate text unless OCR quality has been checked.
版本历史
-
0fe366b
当前 2026-08-16 09:03
将过时的内部工具名(如read_file)替换为正确的模型可见名称(File action: read),修复提示词中的错误引导。
- b0e4926 2026-07-24 17:42


