firecrawl-knowledge-base

작성자: firecrawl

Firecrawl을 사용하여 웹 콘텐츠로 지식 베이스를 구축하세요. 로컬 참조 문서, RAG 준비 청크, 파인튜닝 데이터셋, 문서 미러, 주제 코퍼스 또는 웹 소스에서 정리된 LLM 준비 마크다운에 사용할 수 있습니다.

npx skills add https://github.com/firecrawl/firecrawl-workflows --skill firecrawl-knowledge-base

Firecrawl Knowledge Base

Use this to turn URLs or topics into organized LLM-ready content.

Onboarding Interview

Infer the source, goal, depth, and output location from context. If the source and goal are clear, proceed immediately.

Ask at most 1-3 concise questions only if blocked, such as the source URL/topic, whether the output is reference/RAG/training/docs, or training format if training is requested.

Firecrawl Collection Plan

Use Firecrawl map for documentation sites, search for topic-based corpora, scrape pages into markdown, and preserve code examples and tables.

For files, follow the Firecrawl download-style convention:

.firecrawl/
  <hostname>/
    <path>/
      index.md

Parallel Work

If appropriate, use sub-agents or equivalent parallel task runners:

  • one docs section per researcher
  • official docs, tutorials, community discussions, and references by source type
  • source scraping vs chunk generation vs manifest generation

Output Modes

  • Reference: markdown files, index.md, and sources.json.
  • RAG: markdown files plus chunk files and manifest.json.
  • Training: scraped source files plus training-data.jsonl and training-metadata.json.
  • Docs mirror: complete markdown mirror with a table of contents.

Final Deliverable

# Knowledge Base: [Source]

## Summary
[What was collected and why]

## Output Structure
[Files/directories created]

## Coverage
[Sections, source types, counts]

## Usage Notes
[How to use in RAG, docs, training, or agent context]

## Sources
[URLs collected]

## Rerun Inputs
workflow: firecrawl-knowledge-base
source: [url/topic]
goal: [reference/rag/train/docs]
depth: [quick/thorough/exhaustive]
output_dir: [.firecrawl/]

Quality Bar

  • Preserve code examples and formatting.
  • Remove boilerplate navigation where possible.
  • Include source URLs in frontmatter or metadata.

firecrawl의 다른 스킬

oracle
firecrawl
oracle CLI 사용 모범 사례 (프롬프트 + 파일 번들링, 엔진, 세션 및 파일 첨부 패턴)
official
firecrawl-demo-walkthrough
firecrawl
Firecrawl 브라우저로 제품의 주요 흐름을 살펴보고 구조화된 UX/제품 워크스루를 생성합니다. 회원가입, 온보딩, 가격 정책, 문서, 대시보드, 제품 데모 준비, UX 분석, 첫 실행 경험 분석에 사용하세요.
browser-automationofficialresearch
sag
firecrawl
ElevenLabs 텍스트 음성 변환, Mac 스타일 say 사용자 경험.
official
pinecone
firecrawl
프로덕션 AI 애플리케이션을 위한 관리형 벡터 데이터베이스입니다. 완전 관리형, 자동 확장, 하이브리드 검색(밀집 + 희소), 메타데이터 필터링, 네임스페이스를 지원합니다.
official
sentence-transformers
firecrawl
최신 문장, 텍스트 및 이미지 임베딩을 위한 프레임워크입니다. 의미 유사도, 클러스터링, 검색을 위한 5000개 이상의 사전 훈련된 모델을 제공합니다.
official
model-merging
firecrawl
여러 미세 조정된 모델을 mergekit을 사용하여 병합함으로써 재학습 없이 기능을 결합합니다. 도메인 특화 모델을 혼합하여 특수 모델을 생성할 때 사용합니다.
official
deepspeed
firecrawl
분산 학습을 위한 DeepSpeed 전문 가이드 - ZeRO 최적화 단계, 파이프라인 병렬 처리, FP16/BF16/FP8, 1비트 Adam, 희소 어텐션
official
discord
firecrawl
메시지 도구를 통한 Discord 운영 (채널=discord).
official