latitude-dev/eval-skills · Archived
llm-evals-audit
Use this skill when a developer wants to check whether their existing evaluations are trustworthy and well-targeted. Triggers on: "audit my evals", "are my…
Installation
npx skills add https://github.com/latitude-dev/eval-skills
Stronger alternatives
This repository is archived — consider an actively maintained alternative.
Use this skill when a developer wants to validate how well their LLM judge aligns with human ju…
10 installsUse this skill when a developer wants to create LLM-as-a-judge evaluators for their AI system. …
9 installsUse this skill when a developer wants to annotate their LLM outputs, set up an annotation proce…
9 installsUse this skill when a developer wants to decide what type of evaluation to build for their AI s…
8 installsSimilar popular skills
Related neighbors and high-traction skills in the same topics — useful to compare before installing.
DEPRECATED redirect — this skill was renamed to online-evals. Do not use this skill; invoke onl…
2.6K installsAttach judges to config variations for automatic LLM-as-a-judge evaluation. Create custom judge…
1.7K installsScaffolds evaluation suites for the Axiom AI SDK. Generates eval files, scorers, flag schemas, …
1.6K installsRun evaluations for Hugging Face Hub models using inspect-ai and lighteval on local hardware.
1.5K installsBuild and run evaluators for AI/LLM applications using Phoenix.
1.1K installsHelp users build robust infrastructure for measuring, monitoring, and iterating on AI product p…
1.9K installsAlso in this package
Other skills from latitude-dev/eval-skills.
npx skills add https://github.com/latitude-dev/eval-skills
More details
Agent compatibility
Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.
Repository health
main
History
- First seen on skills.sh
- First recorded snapshot · 8 installs