npx skills add sickn33/agentic-awesome-skills --skill advanced-evaluation
bastani-inc/atomic
advanced-evaluation
This skill should be used when the user asks to "implement LLM-as-judge", "compare model outputs", "create evaluation rubrics", "mitigate evaluation bias", or…
Installation
npx skills add https://github.com/bastani-inc/atomic
Similar popular skills
Related neighbors and high-traction skills in the same topics — useful to compare before installing.
Design and operate LLM-as-a-Judge evaluation systems using direct scoring, pairwise comparison,…
171 installsImplement comprehensive evaluation strategies for LLM applications using automated metrics, hum…
11K installsPluginEval quality methodology — dimensions, rubrics, statistical methods, and scoring formulas…
5.9K installsStructured scholarly-work evaluation for papers, proposals, literature reviews, methods section…
5.1K installsAutomates the end-to-end detection engineering workflow in Google SecOps using MCP tools. Use w…
4K installsUse after completing any non-trivial task. The agent self-rates its output on 5 axes — accuracy…
2.9K installsAlso in this package
Other skills from bastani-inc/atomic · top by installs.
npx skills add https://github.com/bastani-inc/atomic
More details
Agent compatibility
Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.
Also listed on
Alternate registries and mirrors of this skill.
npx skills add ranbot-ai/awesome-skills --skill advanced-evaluation
npx skills add smithery/neversight --skill advanced-evaluation
Repository health
main
History
- First seen on skills.sh
- First recorded snapshot · 163 installs