firecrawl-map
bởi firecrawl
Khám phá và lọc các URL trên một trang web, có tùy chọn tìm kiếm để xác định các trang cụ thể. Hỗ trợ lọc theo truy vấn tìm kiếm để tìm các trang khớp với từ khóa trong các trang web lớn. Bao gồm các chiến lược xử lý sơ đồ trang web (bao gồm, bỏ qua hoặc chỉ sử dụng) và tùy chọn bao gồm tên miền phụ. Đầu ra kết quả dưới dạng văn bản thuần hoặc JSON với giới hạn URL có thể cấu hình. Thường được kết hợp với firecrawl-scrape: sử dụng map kèm tìm kiếm để tìm URL mục tiêu, sau đó scrape nó.
npx skills add https://github.com/firecrawl/cli --skill firecrawl-mapfirecrawl map
Discover URLs on a site. Use --search to find a specific page within a large site.
When to use
- You need to find a specific subpage on a large site
- You want a list of all URLs on a site before scraping or crawling
- Step 3 in the workflow escalation pattern: search → scrape → map → crawl → interact
Quick start
# Find a specific page on a large site
firecrawl map "<url>" --search "authentication" -o .firecrawl/filtered.txt
# Get all URLs
firecrawl map "<url>" --limit 500 --json -o .firecrawl/urls.json
Options
| Option | Description |
|---|---|
--limit <n> | Max number of URLs to return |
--search <query> | Filter URLs by search query |
--sitemap <include|skip|only> | Sitemap handling strategy |
--include-subdomains | Include subdomain URLs |
--json | Output as JSON |
-o, --output <path> | Output file path |
Tips
- Map + scrape is a common pattern: use
map --searchto find the right URL, thenscrapeit. - Example:
map https://docs.example.com --search "auth"→ found/docs/api/authentication→scrapethat URL.
See also
- firecrawl-scrape — scrape the URLs you discover
- firecrawl-crawl — bulk extract instead of map + scrape
- firecrawl-download — download entire site (uses map internally)