togethercomputer/skills
together-evaluations
LLM-as-a-judge evaluation framework on Together AI. Classify, score, and compare model outputs, select judge models, use external-provider judges or targets,…
Installation
npx skills add https://github.com/togethercomputer/skills
Similar popular skills
Related neighbors and high-traction skills in the same topics — useful to compare before installing.
Investigate AI observability evaluations — `hog` (deterministic code-based), `llm_judge` (LLM-p…
293 installsInvestigate AI observability evaluations — `hog` (deterministic code-based), `llm_judge` (LLM-p…
128 installsCompatibility router for LangWatch evaluation requests.
120 installsAlso in this package
Other skills from togethercomputer/skills · top by installs.
npx skills add https://github.com/togethercomputer/skills
More details
Agent compatibility
Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.
Repository health
main
History
- First seen on skills.sh
- First recorded snapshot · 76 installs