Agent Skillsfirecrawl/cli › firecrawl-crawl

firecrawl-crawl

GitHub

用于批量抓取网站或特定目录下的所有页面内容,支持深度限制、路径过滤和并发提取。适用于需要获取大量同一站点内容的场景,如文档库爬取。

skills/firecrawl-crawl/SKILL.md firecrawl/cli

Trigger Scenarios

用户要求爬取整个网站或特定部分 用户提及 'crawl', 'get all the pages', 'bulk extract' 等关键词 需要从多个页面提取内容

Install

npx skills add firecrawl/cli --skill firecrawl-crawl -g -y
More Options

Use without installing

npx skills use firecrawl/cli@firecrawl-crawl

指定 Agent (Claude Code)

npx skills add firecrawl/cli --skill firecrawl-crawl -a claude-code -g -y

安装 repo 全部 skill

npx skills add firecrawl/cli --all -g -y

预览 repo 内 skill

npx skills add firecrawl/cli --list

SKILL.md

Frontmatter
{
    "name": "firecrawl-crawl",
    "description": "Bulk extract content from an entire website or site section. Use this skill when the user wants to crawl a site, extract all pages from a docs section, bulk-scrape multiple pages following links, or says \"crawl\", \"get all the pages\", \"extract everything under \/docs\", \"bulk extract\", or needs content from many pages on the same site. Handles depth limits, path filtering, and concurrent extraction.\n",
    "allowed-tools": [
        "Bash(firecrawl *)",
        "Bash(npx firecrawl-cli *)"
    ]
}

firecrawl crawl

Bulk extract content from a website. Crawls pages following links up to a depth/limit.

Prerequisite: crawl requires authentication (no keyless free tier); without credentials the CLI prompts an interactive login.

When to use

  • You need content from many pages on a site (e.g., all /docs/)
  • You want to extract an entire site section
  • Step 4 in the workflow escalation pattern: search → scrape → map + scrape → crawl → monitor → interact

Quick start

# Crawl a docs section
firecrawl crawl "<url>" --include-paths /docs --limit 50 --wait -o .firecrawl/crawl.json

# Full crawl with depth limit
firecrawl crawl "<url>" --max-depth 3 --wait --progress -o .firecrawl/crawl.json

# Check status of a running crawl
firecrawl crawl <job-id>

Options

Option Description
--wait Wait for crawl to complete before returning
--progress Show progress while waiting
--limit <n> Max pages to crawl
--max-depth <n> Max link depth to follow
--include-paths <paths> Only crawl URLs matching these paths
--exclude-paths <paths> Skip URLs matching these paths
--delay <ms> Delay between requests
--max-concurrency <n> Max parallel crawl workers
--pretty Pretty print JSON output
-o, --output <path> Output file path

Tips

  • Always use --wait when you need the results immediately. It has no default timeout; use --timeout <seconds> to bound polling. Without --wait, crawl returns a job ID for async polling.
  • Use --include-paths to scope the crawl — don't crawl an entire site when you only need one section.
  • Crawl consumes credits per page. Check firecrawl credit-usage before large crawls (credit-usage requires authentication).

See also

Version History

  • 253abde Current 2026-08-19 19:06

    新增认证前置说明;更新工作流步骤描述;修复链接并优化文档结构。

  • 6c50c5d 2026-07-24 11:49

Same Skill Collection

skills/firecrawl-agent/SKILL.md
skills/firecrawl-download/SKILL.md
skills/firecrawl-map/SKILL.md
skills/firecrawl-scrape/SKILL.md
skills/firecrawl-cli/SKILL.md
skills/firecrawl-interact/SKILL.md
skills/firecrawl-monitor/SKILL.md
skills/firecrawl-parse/SKILL.md
skills/firecrawl-search/SKILL.md
skills/firecrawl/SKILL.md

Metadata

Files
0
Version
253abde
Hash
41f01edb
Indexed
2026-07-24 11:49

inicio - Wiki
Copyright © 2011-2026 iteam. Current version is 2.155.2. UTC+08:00, 2026-08-22 07:27
浙ICP备14020137号-1 $mapa de visitantes$