firecrawl-agent
GitHub用于从网站自主提取结构化数据(如价格、产品列表),支持多页导航和JSON Schema约束,适用于需要跨页面获取特定模式数据的场景。
Trigger Scenarios
Install
npx skills add firecrawl/cli --skill firecrawl-agent -g -y
SKILL.md
Frontmatter
{
"name": "firecrawl-agent",
"description": "Autonomous multi-page extraction into structured JSON. Use when the user wants website data matching a schema — pricing tiers, product listings — beyond a single-page scrape.\n",
"allowed-tools": [
"Bash(firecrawl *)",
"Bash(npx firecrawl-cli *)"
]
}
firecrawl agent
AI-powered autonomous extraction. The agent navigates sites and extracts structured data (takes 2-5 minutes).
Quick start
# Extract structured data
firecrawl agent "extract all pricing tiers" --wait --json -o .firecrawl/pricing.json
# With a JSON schema for structured output
firecrawl agent "extract products" --schema '{"type":"object","properties":{"name":{"type":"string"},"price":{"type":"number"}}}' --wait --json -o .firecrawl/products.json
# Focus on specific pages
firecrawl agent "get feature list" --urls "<url>" --wait --json -o .firecrawl/features.json
Run firecrawl agent --help for the full option list.
Done when: the output file contains valid JSON answering the request — or a job ID was intentionally returned for later polling.
Job IDs
Omitting --wait returns a job ID. A UUID positional argument is auto-detected as a status check:
# Check once (equivalent to adding --status)
firecrawl agent "<job-id>"
# Wait on an existing job, polling every 10 seconds for up to 5 minutes
firecrawl agent "<job-id>" --wait --poll-interval 10 --timeout 300
# Cancel an active job
firecrawl agent "<job-id>" --cancel
Tips
- Use
--waitfor inline results; omit it only when you want a job ID to poll later (see Job IDs). - Use
--schemafor predictable, structured output — otherwise the agent returns freeform data. - Agent runs consume more credits than simple scrapes. Use
--max-creditsto cap spending. - For simple single-page extraction, prefer
scrape— it's faster and cheaper.
See also
- firecrawl-scrape — simpler single-page extraction
- firecrawl-interact — scrape + interact for manual page interaction (more control)
- firecrawl-crawl — bulk extraction without AI
- firecrawl-build-scrape — building structured extraction into an app instead of running it here
Version History
-
86aaf06
Current 2026-08-27 16:50
重构文档结构,移除过时选项表与重复说明,优化触发描述以减少Token消耗,新增完成标准。
-
253abde
2026-08-19 19:06
修复了之前版本中关于CLI命令、参数及行为描述的错误声明和示例。
- 6c50c5d 2026-07-24 11:49


