davila7/claude-code-templates
speculative-decoding
Accelerate LLM inference using speculative decoding, Medusa multiple heads, and lookahead decoding techniques. Use when optimizing inference speed (1.5-3.6×…
Installation
npx skills add https://github.com/davila7/claude-code-templates
Similar popular skills
Related neighbors and high-traction skills in the same topics — useful to compare before installing.
Add EAGLE-3 or draft-model speculative decoding to a Jetson vLLM server when TPOT is the bottle…
1.1K installsAccelerate LLM inference using speculative decoding, Medusa multiple heads, and lookahead decod…
750 installsAlso in this package
Other skills from davila7/claude-code-templates · top by installs.
npx skills add https://github.com/davila7/claude-code-templates
More details
Agent compatibility
Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.
Repository health
main
Skill metadata
Parsed from SKILL.md frontmatter.
History
- First seen on skills.sh
- First recorded snapshot · 344 installs