Diagnose why an agent harness misbehaved by reading the local flight-recorder ledger (.vigiles/runs.jsonl) — which skills fired or got hijacked, which hooks blocked or wrongly allowed, which subagent tool-contract violations happened, and how a skill's trigger rate moved.
All-time #9042Trending #1991First seen Jul 2, 2026
Diagnose why an agent harness misbehaved by reading the local flight-recorder ledger (.vigiles/runs.jsonl) — which skills fired or got hijacked, which hooks blocked or wrongly allowed, which subagent tool-contract violations happened, and how a skill's trigger rate moved.
Use when asked why a skill stopped firing, why a hook didn't block, why the wrong skill ran, or to debug/investigate what the harness actually did.
NOT for writing new rules (use strengthen) or editing the spec (use edit-spec).
Similar popular skills
Related neighbors and high-traction skills in the same topics — useful to compare before installing.
Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.
Claude CodeNot declared
CursorNot declared
CodexNot declared
GitHub CopilotNot declared
WindsurfNot declared
Gemini CLINot declared
ClineNot declared
OpenCodeNot declared
Repository health
Stars15
LicenseLICENSE
Default branchmain
Open issues17
Status
Active
Skill metadata
Parsed from SKILL.md frontmatter.
Allowed toolsRead, Glob, Grep
Package contents
Files included with this skill beyond the listing page.
skill mdSKILL.md3,410 B
docsSUMMARY.md526 B
History
First seen on skills.sh
First recorded snapshot · 1,130 installs
SKILL.md
Diagnose harness misbehavior from the flight recorder — the local, append-only ledger at .vigiles/runs.jsonl that vigiles writes as your harness runs. It records what actually happened, so you debug from evidence instead of guessing.
capability-diff — a blast-radius change: {pr, added, removed, widened}.
Instructions
Step 1: Read the ledger
Read .vigiles/runs.jsonl (JSONL — one record per line; tolerate a torn last line). If it's absent or empty, say so — there's nothing recorded yet; suggest running the harness (or vigiles audit) first. Do NOT fabricate records.
Step 2: Answer the specific question, evidence-first
Match the user's question to the ledger:
"Why did skill X stop firing / why does the wrong one run?" — count skill fires by
name over time. If X's fire-rate dropped, look for a sibling that fired on the same kinds of prompts (a selection collision) and check their descriptions for overlap. Recommend differentiating or merging the descriptions.
"Why didn't my hook block that?" — find hook records for the event. A `decision:
allow on something that should be denied, or mode: observe (shadow, never blocks), or the absence of any record, tells you which. Recommend flipping observe→enforce` or fixing the gate logic.
"Did a subagent misbehave?" — list agent records with allowed: false: the agent
reached for a tool outside its declared contract. Point at the contract to tighten or widen.
"Is it getting worse?" — compare eval metric values (recall/precision) across runs;
a downward trend is drift (often after a harness/model upgrade).
Step 3: Recommend a fix, tied to the evidence
Prefer promoting an ignored-but-decidable rule from prose to a deterministic gate: a repeated agent violation or a rule the agent keeps breaking → a compiled hook or a tighter tool-contract (the strengthen skill can help). A description collision → differentiate the skill descriptions. Always cite the specific records you based the diagnosis on.
Step 4: Offer the next step
If the fix is a spec change, hand off to edit-spec. If it's promoting guidance to a linter rule, hand off to strengthen. If a behavioral claim needs measuring (does the skill fire now?), hand off to test-harness (measureTriggerRate).