davila7/claude-code-templates
gguf-quantization
GGUF format and llama.cpp quantization for efficient CPU/GPU inference. Use when deploying models on consumer hardware, Apple Silicon, or when needing flexible…
Installation
npx skills add https://github.com/davila7/claude-code-templates
Similar popular skills
Related neighbors and high-traction skills in the same topics — useful to compare before installing.
GGUF format and llama.cpp quantization for efficient CPU/GPU inference. Use when deploying mode…
765 installsHalf-Quadratic Quantization for LLMs without calibration data. Use when quantizing models to 4/…
752 installsActivation-aware weight quantization for 4-bit LLM compression with 3x speedup and minimal accu…
750 installsExpert skill for AI model quantization and optimization. Covers 4-bit/8-bit quantization, GGUF …
191 installsAlso in this package
Other skills from davila7/claude-code-templates · top by installs.
npx skills add https://github.com/davila7/claude-code-templates
More details
Agent compatibility
Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.
Repository health
main
Skill metadata
Parsed from SKILL.md frontmatter.
History
- First seen on skills.sh
- First recorded snapshot · 385 installs