eyadsibai/ltk
llm-inference
Use when "LLM inference", "serving LLM", "vLLM", "llama.cpp", "GGUF", "text generation", "model serving", "inference optimization", "KV cache", "continuous…
Installation
npx skills add https://github.com/eyadsibai/ltk
Similar popular skills
Related neighbors and high-traction skills in the same topics — useful to compare before installing.
Start, query, and stop a network-specific TAO inference microservice ({network_arch}-inference-…
1.5K installsPick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM…
1.1K installsConnects to and performs inference with Google Cloud Agent Platform GenAI models, including Fir…
5.6K installsDeploys and optimizes AI/ML inference workloads on GKE, using GPUs, TPUs, and model servers. Us…
4K installs>- MANDATORY recipe for every Caffeine build that calls an LLM, chatbot, GPT, or ChatGPT **on C…
2.5K installsInspect the availability of model serving on a completed Itô compute booking and, when the cano…
1.1K installsAlso in this package
Other skills from eyadsibai/ltk · top by installs.
npx skills add https://github.com/eyadsibai/ltk
More details
Agent compatibility
Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.
Repository health
master
History
- First seen on skills.sh
- First recorded snapshot · 81 installs