firecrawl-parse

द्वारा firecrawl

किसी भी स्थानीय फ़ाइल—जैसे PDF, DOCX, DOC, ODT, RTF, XLSX, XLS, या HTML—की सामग्री को कुशलतापूर्वक निकालें और साफ, सुव्यवस्थित मार्कडाउन में परिवर्तित करें, जो सहेजी गई…

npx skills add https://github.com/firecrawl/firecrawl-codex-plugin --skill firecrawl-parse

firecrawl parse

Turn a local document into clean markdown on disk. Supports PDF, DOCX, DOC, ODT, RTF, XLSX, XLS, HTML/HTM/XHTML.

When to use

  • You have a file on disk (not a URL) and want its text as markdown
  • User drops a PDF/DOCX and asks what it says, or to summarize it
  • Use scrape instead when the source is a URL

Quick start

Always save to .firecrawl/ with -o — parsed docs can be hundreds of KB and blow up context if streamed to stdout. Add .firecrawl/ to .gitignore.

mkdir -p .firecrawl

# File → markdown
firecrawl parse ./paper.pdf -o .firecrawl/paper.md

# AI summary
firecrawl parse ./paper.pdf -S -o .firecrawl/paper-summary.md

# Ask a question about the doc
firecrawl parse ./paper.pdf -Q "What are the main conclusions?" \
  -o .firecrawl/paper-qa.md

Then head, grep, rg etc., or incrementally read the file - don't load the whole thing at once.

Options

OptionDescription
-S, --summaryAI-generated summary
-Q, --query <prompt>Ask a question about the parsed content
-o, --output <path>Output file path — always use this
-f, --format <fmt>markdown (default), html, summary
--timeout <ms>Timeout for the parse job
--timingShow request duration

Tips

  • Quote paths with spaces: firecrawl parse "./My Doc.pdf" -o .firecrawl/mydoc.md.
  • Max upload size: 50 MB per file.
  • Credits: ~1 per PDF page; HTML is 1 flat.
  • Check .firecrawl/ before re-parsing the same file.
  • To check your credit balance (recommended for batch processing and similar workflows), use the firecrawl credit-usage command.

See also

firecrawl की और Skills

firecrawl-research-index
firecrawl
Firecrawl Research के साथ किसी शोध प्रश्न का उत्तर देने वाले पेपर खोजें, जिसमें सिमैंटिक सर्च, सिमैंटिक और स्ट्रक्चरल विस्तार, और इन-बॉडी सत्यापन का उपयोग किया जाता है। किसी भी साहित्य-खोज / पेपर-प्राप्ति कार्य — एकल-पेपर खोज या पूर्ण मल्टी-पेपर सेट — के लिए हमेशा इस स्किल का उपयोग करें।
data-analysisresearchweb-scraping
oracle
firecrawl
ओरेकल CLI के उपयोग के लिए सर्वोत्तम अभ्यास (प्रॉम्प्ट + फ़ाइल बंडलिंग, इंजन, सत्र और फ़ाइल अटैचमेंट पैटर्न)।
pinecone
firecrawl
उत्पादन AI अनुप्रयोगों के लिए प्रबंधित वेक्टर डेटाबेस। पूरी तरह से प्रबंधित, स्वचालित स्केलिंग, हाइब्रिड खोज (डेंस + स्पार्स), मेटाडेटा फ़िल्टरिंग और नेमस्पेस के साथ।…
wpds
firecrawl
जब वर्डप्रेस डिज़ाइन सिस्टम (WPDS) और इसके घटकों, टोकन, पैटर्न आदि का उपयोग करके यूआई बनाया जा रहा हो, तब उपयोग करें।
audiocraft-audio-generation
firecrawl
ऑडियो जनरेशन के लिए PyTorch लाइब्रेरी जिसमें टेक्स्ट-टू-म्यूजिक (MusicGen) और टेक्स्ट-टू-साउंड (AudioGen) शामिल है। इसका उपयोग तब करें जब आपको टेक्स्ट से संगीत उत्पन्न करने की आवश्यकता हो…
skypilot-multi-cloud-orchestration
firecrawl
मल्टी-क्लाउड ऑर्केस्ट्रेशन एमएल वर्कलोड के लिए स्वचालित लागत अनुकूलन के साथ। उपयोग करें जब आपको कई क्लाउड्स पर प्रशिक्षण या बैच जॉब चलाने की आवश्यकता हो, लाभ उठाएं…
firecrawl-seo-audit
firecrawl
Firecrawl के साथ किसी वेबसाइट का SEO ऑडिट करें। इसका उपयोग तब करें जब उपयोगकर्ता SEO ऑडिट, मेटाडेटा और हेडिंग समीक्षा, साइटमैप/साइट-संरचना विश्लेषण, कीवर्ड अवसर, प्रतिस्पर्धी SERP तुलना, या प्राथमिकता वाली खोज अनुकूलन अनुशंसाओं का अनुरोध करे।
data-analysisresearchweb-scraping
gh-issues
firecrawl
GitHub की समस्याएँ लाएँ, सुधार लागू करने और PR खोलने के लिए उप-एजेंट बनाएँ, फिर PR समीक्षा टिप्पणियों की निगरानी करें और उनका समाधान करें। उपयोग: /gh-issues [owner/repo] [--label…