firecrawl-agent

作者: firecrawl

AI驅動的自動化結構化資料提取,適用於複雜的多頁面網站。能智慧導航網站以定位並提取資料,以JSON格式回傳結果,並可選擇進行結構驗證。支援自訂JSON結構以產生可預測的結構化輸出,若未提供結構則可進行自由格式提取。提供兩種模型層級(spark-1-mini和spark-1-pro),各有信用額度限制,並可選擇等待內嵌結果。最適合多頁面提取任務;若為較簡單的抓取,請使用...

npx skills add https://github.com/firecrawl/cli --skill firecrawl-agent

firecrawl agent

AI-powered autonomous extraction. The agent navigates sites and extracts structured data (takes 2-5 minutes).

When to use

  • You need structured data from complex multi-page sites
  • Manual scraping would require navigating many pages
  • You want the AI to figure out where the data lives

Quick start

# Extract structured data
firecrawl agent "extract all pricing tiers" --wait -o .firecrawl/pricing.json

# With a JSON schema for structured output
firecrawl agent "extract products" --schema '{"type":"object","properties":{"name":{"type":"string"},"price":{"type":"number"}}}' --wait -o .firecrawl/products.json

# Focus on specific pages
firecrawl agent "get feature list" --urls "<url>" --wait -o .firecrawl/features.json

Options

OptionDescription
--urls <urls>Starting URLs for the agent
--model <model>Model to use: spark-1-mini or spark-1-pro
--schema <json>JSON schema for structured output
--schema-file <path>Path to JSON schema file
--max-credits <n>Credit limit for this agent run
--waitWait for agent to complete
--prettyPretty print JSON output
-o, --output <path>Output file path

Tips

  • Always use --wait to get results inline. Without it, returns a job ID.
  • Use --schema for predictable, structured output — otherwise the agent returns freeform data.
  • Agent runs consume more credits than simple scrapes. Use --max-credits to cap spending.
  • For simple single-page extraction, prefer scrape — it's faster and cheaper.

See also

來自 firecrawl 的更多技能