firecrawl-knowledge-base

โดย firecrawl

สร้างฐานความรู้จากเนื้อหาเว็บด้วย Firecrawl ใช้สำหรับเอกสารอ้างอิงในเครื่อง, ชิ้นส่วนที่พร้อมสำหรับ RAG, ชุดข้อมูลสำหรับปรับแต่งละเอียด, กระจกสะท้อนเอกสาร, คลังข้อมูลหัวข้อ, หรือมาร์กดาวน์ที่พร้อมสำหรับ LLM ซึ่งจัดระเบียบจากแหล่งข้อมูลเว็บ

npx skills add https://github.com/firecrawl/firecrawl-workflows --skill firecrawl-knowledge-base

Firecrawl Knowledge Base

Use this to turn URLs or topics into organized LLM-ready content.

Onboarding Interview

Infer the source, goal, depth, and output location from context. If the source and goal are clear, proceed immediately.

Ask at most 1-3 concise questions only if blocked, such as the source URL/topic, whether the output is reference/RAG/training/docs, or training format if training is requested.

Firecrawl Collection Plan

Use Firecrawl map for documentation sites, search for topic-based corpora, scrape pages into markdown, and preserve code examples and tables.

For files, follow the Firecrawl download-style convention:

.firecrawl/
  <hostname>/
    <path>/
      index.md

Parallel Work

If appropriate, use sub-agents or equivalent parallel task runners:

  • one docs section per researcher
  • official docs, tutorials, community discussions, and references by source type
  • source scraping vs chunk generation vs manifest generation

Output Modes

  • Reference: markdown files, index.md, and sources.json.
  • RAG: markdown files plus chunk files and manifest.json.
  • Training: scraped source files plus training-data.jsonl and training-metadata.json.
  • Docs mirror: complete markdown mirror with a table of contents.

Final Deliverable

# Knowledge Base: [Source]

## Summary
[What was collected and why]

## Output Structure
[Files/directories created]

## Coverage
[Sections, source types, counts]

## Usage Notes
[How to use in RAG, docs, training, or agent context]

## Sources
[URLs collected]

## Rerun Inputs
workflow: firecrawl-knowledge-base
source: [url/topic]
goal: [reference/rag/train/docs]
depth: [quick/thorough/exhaustive]
output_dir: [.firecrawl/]

Quality Bar

  • Preserve code examples and formatting.
  • Remove boilerplate navigation where possible.
  • Include source URLs in frontmatter or metadata.

Skills เพิ่มเติมจาก firecrawl

oracle
firecrawl
แนวทางปฏิบัติที่ดีที่สุดสำหรับการใช้ oracle CLI (การรวม prompt และไฟล์, เอ็นจิน, เซสชัน, และรูปแบบการแนบไฟล์)
official
firecrawl-demo-walkthrough
firecrawl
ดำเนินการตามขั้นตอนหลักของผลิตภัณฑ์ด้วยเบราว์เซอร์ Firecrawl และสร้างคู่มือการใช้งาน UX/ผลิตภัณฑ์ที่มีโครงสร้าง ใช้สำหรับการสมัครสมาชิก การเริ่มต้นใช้งาน ราคา เอกสาร แดชบอร์ด การเตรียมสาธิตผลิตภัณฑ์ การวิเคราะห์ UX และการวิเคราะห์ประสบการณ์การใช้งานครั้งแรก
browser-automationofficialresearch
sag
firecrawl
ElevenLabs แปลงข้อความเป็นเสียง พร้อมประสบการณ์การใช้งานแบบ say สไตล์ Mac
official
pinecone
firecrawl
ฐานข้อมูลเวกเตอร์ที่จัดการแล้วสำหรับแอปพลิเคชัน AI ในระบบผลิต จัดการเต็มรูปแบบ ปรับขนาดอัตโนมัติ พร้อมการค้นหาแบบไฮบริด (dense + sparse) การกรองเมตาดาต้า และเนมสเปซ…
official
sentence-transformers
firecrawl
เฟรมเวิร์กสำหรับการฝังประโยค ข้อความ และรูปภาพที่ทันสมัยที่สุด มีโมเดลที่ผ่านการฝึกอบรมล่วงหน้ามากกว่า 5000 โมเดลสำหรับความคล้ายคลึงทางความหมาย การจัดกลุ่ม และการดึงข้อมูล
official
model-merging
firecrawl
รวมโมเดลที่ปรับแต่งหลายตัวโดยใช้ mergekit เพื่อรวมความสามารถโดยไม่ต้องฝึกซ้ำ ใช้เมื่อสร้างโมเดลเฉพาะทางโดยการผสมผสานโมเดลที่เชี่ยวชาญเฉพาะด้าน...
official
deepspeed
firecrawl
คำแนะนำจากผู้เชี่ยวชาญสำหรับการฝึกอบรมแบบกระจายด้วย DeepSpeed - ขั้นตอนการปรับแต่ง ZeRO, การทำงานแบบขนานของไปป์ไลน์, FP16/BF16/FP8, 1-bit Adam, การสนใจแบบกระจัดกระจาย
official
discord
firecrawl
การดำเนินการ Discord ผ่านเครื่องมือข้อความ (channel=discord)
official