firecrawl-crawl

Bulk extract content from an entire website or site section. Use this skill when the user wants to crawl a site, extract all pages from a docs section,…

npx skills add https://github.com/firecrawl/firecrawl-codex-plugin --skill firecrawl-crawl

firecrawl crawl

Bulk extract content from a website. Crawls pages following links up to a depth/limit.

When to use

  • You need content from many pages on a site (e.g., all /docs/)
  • You want to extract an entire site section
  • Step 4 in the workflow escalation pattern: search → scrape → map → crawl → interact

Quick start

# Crawl a docs section
firecrawl crawl "<url>" --include-paths /docs --limit 50 --wait -o .firecrawl/crawl.json

# Full crawl with depth limit
firecrawl crawl "<url>" --max-depth 3 --wait --progress -o .firecrawl/crawl.json

# Check status of a running crawl
firecrawl crawl <job-id>

Options

OptionDescription
--waitWait for crawl to complete before returning
--progressShow progress while waiting
--limit <n>Max pages to crawl
--max-depth <n>Max link depth to follow
--include-paths <paths>Only crawl URLs matching these paths
--exclude-paths <paths>Skip URLs matching these paths
--delay <ms>Delay between requests
--max-concurrency <n>Max parallel crawl workers
--prettyPretty print JSON output
-o, --output <path>Output file path

Tips

  • Always use --wait when you need the results immediately. Without it, crawl returns a job ID for async polling.
  • Use --include-paths to scope the crawl — don't crawl an entire site when you only need one section.
  • Crawl consumes credits per page. Check firecrawl credit-usage before large crawls.

See also

Lebih banyak skill dari firecrawl

oracle
firecrawl
Praktik terbaik dalam menggunakan CLI oracle (penggabungan prompt dan file, mesin, sesi, dan pola lampiran file).
official
pinecone
firecrawl
Basis data vektor terkelola untuk aplikasi AI produksi. Sepenuhnya terkelola, penskalaan otomatis, dengan pencarian hibrida (padat + jarang), pemfilteran metadata, dan ruang nama.…
official
sentence-transformers
firecrawl
Kerangka kerja untuk embedding kalimat, teks, dan gambar terkini. Menyediakan lebih dari 5000 model terlatih untuk kesamaan semantik, pengelompokan, dan pengambilan data.…
official
wp-playground
firecrawl
Gunakan untuk alur kerja WordPress Playground: instance WP cepat yang dapat dibuang di browser atau secara lokal melalui @wp-playground/cli (server, run-blueprint, build-snapshot),…
official
wp-plugin-development
firecrawl
Gunakan saat mengembangkan plugin WordPress: arsitektur dan hooks, aktivasi/deaktivasi/pencopotan, antarmuka admin dan API Pengaturan, penyimpanan data, cron/tugas, keamanan…
official
wp-project-triage
firecrawl
Gunakan saat Anda memerlukan inspeksi deterministik terhadap repositori WordPress (plugin/tema/tema blok/WP inti/Gutenberg/situs penuh) termasuk perkakas/pengujian/versi…
official
wp-rest-api
firecrawl
Gunakan saat membangun, memperluas, atau men-debug endpoint/rute REST API WordPress: register_rest_route, kelas WP_REST_Controller/controller, skema/argumen…
official
wp-wpcli-and-ops
firecrawl
Gunakan saat bekerja dengan WP-CLI (wp) untuk operasi WordPress: pencarian-penggantian aman, ekspor/impor db, manajemen plugin/tema/pengguna/konten, cron, pembersihan cache,…
official