firecrawl-scrape

Extraia markdown limpo de qualquer URL, incluindo aplicativos de página única renderizados por JavaScript. Lida tanto com páginas estáticas quanto com SPAs renderizados por JS, com tempos de espera configuráveis para renderização. Suporta scraping simultâneo de múltiplas URLs com opções de formato de saída, incluindo markdown, HTML, links e capturas de tela. Inclui opções de filtragem de conteúdo, como o modo apenas-conteúdo-principal para remover navegação e rodapés, além de inclusão/exclusão de tags. Opcionalmente, responde a perguntas inline via flag --query para consultas direcionadas...

npx skills add https://github.com/firecrawl/cli --skill firecrawl-scrape

firecrawl scrape

Scrape one or more URLs. Returns clean, LLM-optimized markdown. Multiple URLs are scraped concurrently.

When to use

  • You have a specific URL and want its content
  • The page is static or JS-rendered (SPA)
  • Step 2 in the workflow escalation pattern: search → scrape → map → crawl → interact

Quick start

# Basic markdown extraction
firecrawl scrape "<url>" -o .firecrawl/page.md

# Main content only, no nav/footer
firecrawl scrape "<url>" --only-main-content -o .firecrawl/page.md

# Wait for JS to render, then scrape
firecrawl scrape "<url>" --wait-for 3000 -o .firecrawl/page.md

# Multiple URLs (each saved to .firecrawl/)
firecrawl scrape https://example.com https://example.com/blog https://example.com/docs

# Get markdown and links together
firecrawl scrape "<url>" --format markdown,links -o .firecrawl/page.json

# Ask a question about the page
firecrawl scrape "https://example.com/pricing" --query "What is the enterprise plan price?"

Options

OptionDescription
-f, --format <formats>Output formats: markdown, html, rawHtml, links, screenshot, json
-Q, --query <prompt>Ask a question about the page content (5 credits)
-HInclude HTTP headers in output
--only-main-contentStrip nav, footer, sidebar — main content only
--wait-for <ms>Wait for JS rendering before scraping
--include-tags <tags>Only include these HTML tags
--exclude-tags <tags>Exclude these HTML tags
--redact-piiRedact personally identifiable information from output
-o, --output <path>Output file path

Tips

  • Prefer plain scrape over --query. Scrape to a file, then use grep, head, or read the markdown directly — you can search and reason over the full content yourself. Use --query only when you want a single targeted answer without saving the page (costs 5 extra credits).
  • Try scrape before interact. Scrape handles static pages and JS-rendered SPAs. Only escalate to interact when you need interaction (clicks, form fills, pagination).
  • Multiple URLs are scraped concurrently — check firecrawl --status for your concurrency limit.
  • Single format outputs raw content. Multiple formats (e.g., --format markdown,links) output JSON.
  • Always quote URLs — shell interprets ? and & as special characters.
  • Naming convention: .firecrawl/{site}-{path}.md

See also

Mais skills de firecrawl

oracle
firecrawl
Melhores práticas para usar a CLI do oracle (prompt + agrupamento de arquivos, engines, sessões e padrões de anexo de arquivos).
official
pinecone
firecrawl
Banco de dados vetorial gerenciado para aplicações de IA em produção. Totalmente gerenciado, com escalonamento automático, busca híbrida (densa + esparsa), filtragem por metadados e namespaces.…
official
sentence-transformers
firecrawl
Framework para embeddings de sentenças, textos e imagens de última geração. Fornece mais de 5000 modelos pré-treinados para similaridade semântica, clusterização e recuperação.
official
wp-playground
firecrawl
Use para fluxos de trabalho do WordPress Playground: instâncias WP descartáveis e rápidas no navegador ou localmente via @wp-playground/cli (server, run-blueprint, build-snapshot),…
official
wp-plugin-development
firecrawl
Use ao desenvolver plugins WordPress: arquitetura e hooks, ativação/desativação/desinstalação, interface administrativa e Settings API, armazenamento de dados, cron/tarefas, segurança…
official
wp-project-triage
firecrawl
Use quando precisar de uma inspeção determinística de um repositório WordPress (plugin/tema/tema de bloco/WP core/Gutenberg/site completo) incluindo ferramentas/testes/versão…
official
wp-rest-api
firecrawl
Use ao construir, estender ou depurar endpoints/rotas da REST API do WordPress: register_rest_route, classes WP_REST_Controller/controller, schema/argumentos…
official
wp-wpcli-and-ops
firecrawl
Use ao trabalhar com WP-CLI (wp) para operações no WordPress: substituição segura de texto, exportação/importação de banco de dados, gerenciamento de plugins/temas/usuários/conteúdo, cron, limpeza de cache,…
official