yonatangross/orchestkit
llm-evaluation
LLM output evaluation and quality assessment. Use when implementing LLM-as-judge patterns, quality gates for AI outputs, or automated evaluation pipelines.
Installation
npx skills add https://github.com/yonatangross/orchestkit
Similar popular skills
Related neighbors and high-traction skills in the same topics — useful to compare before installing.
Implement comprehensive evaluation strategies for LLM applications using automated metrics, hum…
11K installsPluginEval quality methodology — dimensions, rubrics, statistical methods, and scoring formulas…
5.9K installsStructured scholarly-work evaluation for papers, proposals, literature reviews, methods section…
5.1K installsAutomates the end-to-end detection engineering workflow in Google SecOps using MCP tools. Use w…
4K installsUse after completing any non-trivial task. The agent self-rates its output on 5 axes — accuracy…
2.9K installsRun an expert review against Nielsen's heuristics and domain criteria, with severity ratings. U…
1.6K installsAlso in this package
Other skills from yonatangross/orchestkit · top by installs.
npx skills add https://github.com/yonatangross/orchestkit
More details
Agent compatibility
Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.
Repository health
main
History
- First seen on skills.sh
- First recorded snapshot · 14 installs