gguf-quantization
von firecrawl
GGUF-Format und llama.cpp-Quantisierung für effiziente CPU/GPU-Inferenz. Verwenden Sie dies beim Bereitstellen von Modellen auf Consumer-Hardware, Apple Silicon oder wenn flexible…
npx skills add https://github.com/firecrawl/ai-research-skills --skill gguf-quantization