firecrawl-scrape

作成者: firecrawl

任意のURLから、JavaScriptでレンダリングされたSPAを含む、クリーンなマークダウンを抽出します。ユーザーがURLを提供してその内容を希望する場合、「scrape」と言った場合などに、このスキルを使用してください。

npx skills add https://github.com/firecrawl/firecrawl-codex-plugin --skill firecrawl-scrape

firecrawl scrape

Scrape one or more URLs. Returns clean, LLM-optimized markdown. Multiple URLs are scraped concurrently.

When to use

  • You have a specific URL and want its content
  • The page is static or JS-rendered (SPA)
  • Step 2 in the workflow escalation pattern: search → scrape → map → crawl → interact

Quick start

# Basic markdown extraction
firecrawl scrape "<url>" -o .firecrawl/page.md

# Main content only, no nav/footer
firecrawl scrape "<url>" --only-main-content -o .firecrawl/page.md

# Wait for JS to render, then scrape
firecrawl scrape "<url>" --wait-for 3000 -o .firecrawl/page.md

# Multiple URLs (each saved to .firecrawl/)
firecrawl scrape https://example.com https://example.com/blog https://example.com/docs

# Get markdown and links together
firecrawl scrape "<url>" --format markdown,links -o .firecrawl/page.json

# Ask a question about the page
firecrawl scrape "https://example.com/pricing" --query "What is the enterprise plan price?"

Options

OptionDescription
-f, --format <formats>Output formats: markdown, html, rawHtml, links, screenshot, json
-Q, --query <prompt>Ask a question about the page content (5 credits)
-HInclude HTTP headers in output
--only-main-contentStrip nav, footer, sidebar — main content only
--wait-for <ms>Wait for JS rendering before scraping
--include-tags <tags>Only include these HTML tags
--exclude-tags <tags>Exclude these HTML tags
-o, --output <path>Output file path

Tips

  • Prefer plain scrape over --query. Scrape to a file, then use grep, head, or read the markdown directly — you can search and reason over the full content yourself. Use --query only when you want a single targeted answer without saving the page (costs 5 extra credits).
  • Try scrape before interact. Scrape handles static pages and JS-rendered SPAs. Only escalate to interact when you need interaction (clicks, form fills, pagination).
  • Multiple URLs are scraped concurrently — check firecrawl --status for your concurrency limit.
  • Single format outputs raw content. Multiple formats (e.g., --format markdown,links) output JSON.
  • Always quote URLs — shell interprets ? and & as special characters.
  • Naming convention: .firecrawl/{site}-{path}.md

See also

firecrawlのその他のスキル

firecrawl-research-index
firecrawl
Firecrawl Researchを使用して、研究クエリに回答する論文を、意味検索、意味的・構造的拡張、本文内検証により見つけます。文献検索や論文取得のタスク(単一論文の検索から複数論文のセットまで)には、常にこのスキルを使用してください。
data-analysisresearchweb-scraping
oracle
firecrawl
oracle CLIのベストプラクティス(プロンプトとファイルのバンドル、エンジン、セッション、ファイル添付パターン)
pinecone
firecrawl
プロダクションAIアプリケーション向けの管理型ベクトルデータベース。フルマネージド、自動スケーリング、ハイブリッド検索(高密度+スパース)、メタデータフィルタリング、名前空間を備えています。…
wpds
firecrawl
WordPress Design System(WPDS)とそのコンポーネント、トークン、パターンなどを活用してUIを構築する際に使用します。
audiocraft-audio-generation
firecrawl
音声生成のためのPyTorchライブラリ。テキストから音楽を生成するMusicGenと、テキストから効果音を生成するAudioGenを含む。テキストから音楽を生成する必要がある場合に使用する…
skypilot-multi-cloud-orchestration
firecrawl
MLワークロード向けのマルチクラウドオーケストレーションで、自動コスト最適化を実現。複数のクラウドにまたがってトレーニングやバッチジョブを実行する必要がある場合、または…を活用する場合に使用します。
firecrawl-seo-audit
firecrawl
Firecrawlを使用してウェブサイトのSEOを監査します。ユーザーがSEO監査、メタデータや見出しのレビュー、サイトマップやサイト構造の分析、キーワードの機会、競合のSERP比較、優先順位付けされた検索最適化の推奨事項を求めた場合に使用します。
data-analysisresearchweb-scraping
gh-issues
firecrawl
GitHubのIssueを取得し、修正を実装してPRを開くサブエージェントを起動し、その後PRレビューコメントを監視して対応します。使用方法: /gh-issues [owner/repo] [--label…]