mathews-tom/armory
skill-evaluator
Evaluate Claude Code SKILL.md quality across 6 weighted dimensions: frontmatter quality, trigger coverage, structural completeness, content depth, consistency…
Installation
npx skills add https://github.com/mathews-tom/armory
Similar popular skills
Related neighbors and high-traction skills in the same topics — useful to compare before installing.
INVOKE THIS SKILL when building evaluation pipelines for LangSmith. Covers three core component…
4.2K installsEvaluate models, datasets, and agents with the NeMo Evaluator plugin. Use for metric selection,…
1.8K installsHandles LLM-as-judge evaluation workflows on Arize including creating/updating evaluators, runn…
1.1K installsHandles LLM-as-judge and code evaluator workflows on Arize including creating/updating evaluato…
2.5K installsTechnology stack evaluation and comparison with TCO analysis, security assessment, and ecosyste…
904 installsEvaluates LLMs across 100+ benchmarks from 18+ harnesses (MMLU, HumanEval, GSM8K, safety, VLM) …
741 installsAlso in this package
Other skills from mathews-tom/armory · top by installs.
npx skills add https://github.com/mathews-tom/armory
More details
Agent compatibility
Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.
Repository health
main
History
- First seen on skills.sh
- First recorded snapshot · 23 installs