greyhaven-ai/claude-code-config
grey-haven-evaluation
Evaluate LLM outputs with multi-dimensional rubrics, handle non-determinism, and implement LLM-as-judge patterns. Essential for production LLM systems. Use…
Installation
npx skills add https://github.com/greyhaven-ai/claude-code-config
Similar popular skills
Related neighbors and high-traction skills in the same topics — useful to compare before installing.
Implement comprehensive evaluation strategies for LLM applications using automated metrics, hum…
11K installsPluginEval quality methodology — dimensions, rubrics, statistical methods, and scoring formulas…
5.9K installsStructured scholarly-work evaluation for papers, proposals, literature reviews, methods section…
5.1K installsAutomates the end-to-end detection engineering workflow in Google SecOps using MCP tools. Use w…
4K installsUse after completing any non-trivial task. The agent self-rates its output on 5 axes — accuracy…
2.9K installsRun an expert review against Nielsen's heuristics and domain criteria, with severity ratings. U…
1.6K installsAlso in this package
Other skills from greyhaven-ai/claude-code-config · top by installs.
npx skills add https://github.com/greyhaven-ai/claude-code-config
More details
Agent compatibility
Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.
Repository health
main
Skill metadata
Parsed from SKILL.md frontmatter.
History
- First seen on skills.sh
- First recorded snapshot · 9 installs