firecrawl-map
作者: firecrawl
发现并过滤网站上的URL,支持通过搜索定位特定页面。可基于搜索查询过滤大型网站中匹配关键词的页面,包含站点地图处理策略(包含、跳过或仅使用)以及可选的子域名包含功能。输出结果支持纯文本或JSON格式,并可配置URL数量限制。通常与firecrawl-scrape配合使用:先通过搜索映射找到目标URL,再进行抓取。
npx skills add https://github.com/firecrawl/cli --skill firecrawl-mapfirecrawl map
Discover URLs on a site. Use --search to find a specific page within a large site.
Prerequisite: map requires authentication (no keyless free tier); without credentials the CLI prompts an interactive login.
Quick start
# Find a specific page on a large site
firecrawl map "<url>" --search "authentication" -o .firecrawl/filtered.txt
# Get all URLs
firecrawl map "<url>" --limit 500 --json -o .firecrawl/urls.json
Run firecrawl map --help for the full option list (sitemap handling, subdomains, etc.).
Done when: the URL list is saved under .firecrawl/ and you have selected the URLs to scrape or crawl next.
Tips
- Map + scrape is a common pattern: use
map --searchto find the right URL, thenscrapeit. - Example:
map https://docs.example.com --search "auth"→ found/docs/api/authentication→scrapethat URL.
See also
- firecrawl-scrape — scrape the URLs you discover
- firecrawl-crawl — bulk extract instead of map + scrape
- firecrawl-download — download entire site (uses map internally)
- firecrawl-build-search — building URL discovery into an app instead of running it here