firecrawl-agent

Extração autônoma com inteligência artificial. O agente navega por sites e extrai dados estruturados (leva de 2 a 5 minutos).

npx skills add https://github.com/firecrawl/firecrawl-claude-plugin --skill firecrawl-agent

firecrawl agent

AI-powered autonomous extraction. The agent navigates sites and extracts structured data (takes 2-5 minutes).

Before starting autonomous extraction for structured records or listings, check firecrawl search alexandria '<data you need>' for a ready-made workflow or data provider. Inspect a matching contract with firecrawl list <provider> <capability> --pretty and execute with firecrawl scrape --alexandria <provider>/<capability> --options '<input JSON>' if it covers the task. Use the exact provider, capability, and input fields from that contract. Continue with Agent when no suitable tool exists or the task requires autonomous navigation.

Quick start

# Extract structured data
firecrawl agent "extract all pricing tiers" --wait --json -o .firecrawl/pricing.json

# With a JSON schema for structured output
firecrawl agent "extract products" --schema '{"type":"object","properties":{"name":{"type":"string"},"price":{"type":"number"}}}' --wait --json -o .firecrawl/products.json

# Focus on specific pages
firecrawl agent "get feature list" --urls "<url>" --wait --json -o .firecrawl/features.json

Run firecrawl agent --help for the full option list.

Done when: the output file contains valid JSON answering the request — or a job ID was intentionally returned for later polling.

Job IDs

Omitting --wait returns a job ID. A UUID positional argument is auto-detected as a status check:

# Check once (equivalent to adding --status)
firecrawl agent "<job-id>"

# Wait on an existing job, polling every 10 seconds for up to 5 minutes
firecrawl agent "<job-id>" --wait --poll-interval 10 --timeout 300

# Cancel an active job
firecrawl agent "<job-id>" --cancel

Tips

  • Use --wait for inline results; omit it only when you want a job ID to poll later (see Job IDs).
  • Use --schema for predictable, structured output — otherwise the agent returns freeform data.
  • Agent runs consume more credits than simple scrapes. Use --max-credits to cap spending.
  • For simple single-page extraction, prefer scrape — it's faster and cheaper.

See also

Alexandria session feedback

To report an Alexandria session outcome or a provider/capability gap, use firecrawl alexandria feedback --rating good|partial|bad --url <website> --requested-functionality '<what was needed>' --rationale '<what happened>' --json. Use observed results in the rationale. No job ID is needed; this session feedback has no job-age deadline and no credit refund. Optional --provider-feedback and --capability-feedback JSON arrays describe specific gaps; inspect firecrawl alexandria feedback --help for their fields. Use the capability issue missing_capability when a provider exists but lacks the needed capability, and new_capability_request (with requestedFunctionality) to ask for one.

Mais skills de firecrawl

firecrawl-research-index
firecrawl
Encontre os artigos que respondem a uma consulta de pesquisa com o Firecrawl Research, utilizando busca semântica, expansão semântica e estrutural, e verificação no corpo do texto. Use sempre esta habilidade para qualquer tarefa de localização de literatura ou recuperação de artigos — consultas de um único artigo ou conjuntos completos de múltiplos artigos.
data-analysisresearchweb-scraping
oracle
firecrawl
Melhores práticas para usar a CLI do oracle (prompt + agrupamento de arquivos, engines, sessões e padrões de anexo de arquivos).
pinecone
firecrawl
Banco de dados vetorial gerenciado para aplicações de IA em produção. Totalmente gerenciado, com escalonamento automático, busca híbrida (densa + esparsa), filtragem por metadados e namespaces.…
wpds
firecrawl
Use ao construir UIs que utilizam o WordPress Design System (WPDS) e seus componentes, tokens, padrões, etc.
audiocraft-audio-generation
firecrawl
Biblioteca PyTorch para geração de áudio, incluindo texto para música (MusicGen) e texto para som (AudioGen). Use quando precisar gerar música a partir de texto…
skypilot-multi-cloud-orchestration
firecrawl
Orquestração multi-cloud para cargas de trabalho de ML com otimização automática de custos. Use quando precisar executar treinamento ou jobs em lote em múltiplas nuvens, aproveitar…
firecrawl-seo-audit
firecrawl
Audite o SEO de um site com Firecrawl. Use quando o usuário solicitar uma auditoria de SEO, revisão de metadados e cabeçalhos, análise de sitemap/estrutura do site, oportunidades de palavras-chave, comparação de SERP de concorrentes ou recomendações priorizadas de otimização de busca.
data-analysisresearchweb-scraping
gh-issues
firecrawl
Buscar issues do GitHub, criar subagentes para implementar correções e abrir PRs, depois monitorar e tratar comentários de revisão dos PRs. Uso: /gh-issues [owner/repo] [--label…