google/skills
gke-inference
Deploys and optimizes AI/ML inference workloads on GKE, using GPUs, TPUs, and model servers. Use when deploying GKE inference servers, configuring GKE GPU…
Installation
npx skills add https://github.com/google/skills
Similar popular skills
Related neighbors and high-traction skills in the same topics — useful to compare before installing.
Start, query, and stop a network-specific TAO inference microservice ({network_arch}-inference-…
1.5K installsPick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM…
1.1K installs>- MANDATORY recipe for every Caffeine build that calls an LLM, chatbot, GPT, or ChatGPT **on C…
2.5K installsInspect the availability of model serving on a completed Itô compute booking and, when the cano…
1.1K installsConnects NemoClaw to a local inference server. Use when setting up Ollama, vLLM, TensorRT-LLM, …
846 installsDetermine cause-and-effect relationships using propensity scoring, instrumental variables, and …
501 installsAlso in this package
Other skills from google/skills · top by installs.
npx skills add https://github.com/google/skills
More details
Agent compatibility
Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.
Repository health
main
History
- First seen on skills.sh
- First recorded snapshot · 4,025 installs