firecrawl-map
作者: firecrawl
发现并过滤网站上的URL,支持通过搜索定位特定页面。可基于搜索查询过滤大型网站中匹配关键词的页面,包含站点地图处理策略(包含、跳过或仅使用)以及可选的子域名包含功能。输出结果支持纯文本或JSON格式,并可配置URL数量限制。通常与firecrawl-scrape配合使用:先通过搜索映射找到目标URL,再进行抓取。
npx skills add https://github.com/firecrawl/cli --skill firecrawl-mapfirecrawl map
Discover URLs on a site. Use --search to find a specific page within a large site.
When to use
- You need to find a specific subpage on a large site
- You want a list of all URLs on a site before scraping or crawling
- Step 3 in the workflow escalation pattern: search → scrape → map → crawl → interact
Quick start
# Find a specific page on a large site
firecrawl map "<url>" --search "authentication" -o .firecrawl/filtered.txt
# Get all URLs
firecrawl map "<url>" --limit 500 --json -o .firecrawl/urls.json
Options
| Option | Description |
|---|---|
--limit <n> | Max number of URLs to return |
--search <query> | Filter URLs by search query |
--sitemap <include|skip|only> | Sitemap handling strategy |
--include-subdomains | Include subdomain URLs |
--json | Output as JSON |
-o, --output <path> | Output file path |
Tips
- Map + scrape is a common pattern: use
map --searchto find the right URL, thenscrapeit. - Example:
map https://docs.example.com --search "auth"→ found/docs/api/authentication→scrapethat URL.
See also
- firecrawl-scrape — scrape the URLs you discover
- firecrawl-crawl — bulk extract instead of map + scrape
- firecrawl-download — download entire site (uses map internally)