firecrawl-agent

KI-gestützte autonome Datenextraktion, die komplexe Websites durchsucht und strukturierte JSON-Daten zurückgibt. Verwenden Sie diese Fähigkeit, wenn der Benutzer strukturierte Daten von…

npx skills add https://github.com/firecrawl/firecrawl-cursor-plugin --skill firecrawl-agent

firecrawl agent

AI-powered autonomous extraction. The agent navigates sites and extracts structured data (takes 2-5 minutes).

Before starting autonomous extraction for structured records or listings, check firecrawl search alexandria '<data you need>' for a ready-made workflow or data provider. Inspect a matching contract with firecrawl list <provider> <capability> --pretty and execute with firecrawl scrape --alexandria <provider>/<capability> --options '<input JSON>' if it covers the task. Use the exact provider, capability, and input fields from that contract. Continue with Agent when no suitable tool exists or the task requires autonomous navigation.

Quick start

# Extract structured data
firecrawl agent "extract all pricing tiers" --wait --json -o .firecrawl/pricing.json

# With a JSON schema for structured output
firecrawl agent "extract products" --schema '{"type":"object","properties":{"name":{"type":"string"},"price":{"type":"number"}}}' --wait --json -o .firecrawl/products.json

# Focus on specific pages
firecrawl agent "get feature list" --urls "<url>" --wait --json -o .firecrawl/features.json

Run firecrawl agent --help for the full option list.

Done when: the output file contains valid JSON answering the request — or a job ID was intentionally returned for later polling.

Job IDs

Omitting --wait returns a job ID. A UUID positional argument is auto-detected as a status check:

# Check once (equivalent to adding --status)
firecrawl agent "<job-id>"

# Wait on an existing job, polling every 10 seconds for up to 5 minutes
firecrawl agent "<job-id>" --wait --poll-interval 10 --timeout 300

# Cancel an active job
firecrawl agent "<job-id>" --cancel

Tips

  • Use --wait for inline results; omit it only when you want a job ID to poll later (see Job IDs).
  • Use --schema for predictable, structured output — otherwise the agent returns freeform data.
  • Agent runs consume more credits than simple scrapes. Use --max-credits to cap spending.
  • For simple single-page extraction, prefer scrape — it's faster and cheaper.

See also

Alexandria session feedback

To report an Alexandria session outcome or a provider/capability gap, use firecrawl alexandria feedback --rating good|partial|bad --url <website> --requested-functionality '<what was needed>' --rationale '<what happened>' --json. Use observed results in the rationale. No job ID is needed; this session feedback has no job-age deadline and no credit refund. Optional --provider-feedback and --capability-feedback JSON arrays describe specific gaps; inspect firecrawl alexandria feedback --help for their fields. Use the capability issue missing_capability when a provider exists but lacks the needed capability, and new_capability_request (with requestedFunctionality) to ask for one.

Mehr Skills von firecrawl

firecrawl-research-index
firecrawl
Finden Sie die Paper, die eine Forschungsfrage mit Firecrawl Research beantworten, unter Verwendung von semantischer Suche, semantischer und struktureller Erweiterung sowie In-Text-Verifikation. Verwenden Sie diese Fähigkeit stets für jede Aufgabe zur Literatursuche / Paper-Beschaffung – sowohl für einzelne Paper als auch für vollständige Multi-Paper-Sets.
data-analysisresearchweb-scraping
oracle
firecrawl
Bewährte Verfahren für die Verwendung der oracle CLI (Prompt- und Dateibündelung, Engines, Sitzungen und Dateianhänge-Muster).
pinecone
firecrawl
Verwaltete Vektordatenbank für KI-Produktionsanwendungen. Vollständig verwaltet, automatisch skalierend, mit Hybrid-Suche (dicht + dünnbesetzt), Metadatenfilterung und Namespaces.…
wpds
firecrawl
Verwenden beim Erstellen von Benutzeroberflächen, die das WordPress Design System (WPDS) und seine Komponenten, Tokens, Patterns usw. nutzen.
audiocraft-audio-generation
firecrawl
PyTorch-Bibliothek zur Audioerzeugung, einschließlich Text-zu-Musik (MusicGen) und Text-zu-Geräusch (AudioGen). Verwenden Sie diese, wenn Sie Musik aus Text generieren müssen…
skypilot-multi-cloud-orchestration
firecrawl
Multi-Cloud-Orchestrierung für ML-Workloads mit automatischer Kostenoptimierung. Verwenden Sie dies, wenn Sie Trainings- oder Batch-Jobs über mehrere Clouds hinweg ausführen müssen, nutzen Sie…
firecrawl-seo-audit
firecrawl
Führen Sie eine SEO-Überprüfung einer Website mit Firecrawl durch. Verwenden Sie dies, wenn der Benutzer nach einer SEO-Analyse, einer Überprüfung von Metadaten und Überschriften, einer Sitemap-/Seitenstrukturanalyse, Keyword-Möglichkeiten, einem Vergleich der SERP von Mitbewerbern oder priorisierten Suchoptimierungsempfehlungen fragt.
data-analysisresearchweb-scraping
gh-issues
firecrawl
GitHub-Issues abrufen, Unter-Agenten zur Implementierung von Korrekturen und zum Erstellen von PRs einsetzen, dann PR-Review-Kommentare überwachen und bearbeiten. Verwendung: /gh-issues [owner/repo] [--label…