smithery/willoscar

evidence-draft

Create per-subsection evidence packs (NO PROSE): claim candidates, concrete comparisons, evaluation protocol, limitations, plus citation-backed evidence snippets with provenance.

Installation

$ npx skills add smithery/willoscar --skill evidence-draft

Similar popular skills

Related neighbors and high-traction skills in the same topics — useful to compare before installing.

Also in this package

Other skills from smithery/willoscar · top by installs.

npx skills add smithery/willoscar

Browse all from smithery/willoscar

More details

Agent compatibility

Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.

Claude Code Not declared
Cursor Not declared
Codex Declared
GitHub Copilot Not declared
Windsurf Not declared
Gemini CLI Not declared
Cline Not declared
OpenCode Not declared

Skill metadata

Parsed from SKILL.md frontmatter.

Declared agents codex

Package contents

Files included with this skill beyond the listing page.

  • skill md SKILL.md 5,045 B
  • docs SUMMARY.md 857 B

History

  1. First recorded snapshot · 0 installs

SKILL.md

Evidence Draft

Triggers & routing

  • Trigger: evidence draft, evidence pack, claim candidates, concrete comparisons, evidence snippets, provenance, 证据草稿, 证据包, 可引用事实.

Build deterministic outline/evidence_drafts.jsonl packs from briefs + notes + optional evidence bindings.

Compatibility mode is active: this migration preserves the existing JSONL contract while moving evidence-quality policy, sparse-evidence routing, and evaluation-anchor rules into references/ and assets/.

Load Order

Always read:

  • references/overview.md
  • references/evidencequalitypolicy.md

Read by task:

  • references/blockvsdowngrade.md when deciding whether thin evidence should block drafting or only downgrade claim strength
  • references/evaluationanchorrules.md when evaluation tokens, protocol context, or numeric claims are weak
  • references/examplessparseevidence.md for evidence-thin pack calibration
  • references/sourcetexthygiene.md when paper self-narration or generic result wrappers are leaking into pack snippets / claim candidates

Machine-readable assets:

  • assets/evidencepackschema.json
  • assets/evidence_policy.json
  • assets/sourcetexthygiene.json
  • repo-wide assets/limitation-signals.json — shared polarity rules that keep

resolved failures and positive improvements out of limitation slots

Inputs

Required:

  • outline/subsection_briefs.jsonl
  • papers/paper_notes.jsonl
  • citations/ref.bib

Optional but recommended:

  • papers/evidence_bank.jsonl
  • outline/evidence_bindings.jsonl

Outputs

Keep the current output contract:

  • outline/evidence_drafts.jsonl
  • optional human-readable mirrors under outline/evidence_drafts/

Script Boundary

Use scripts/run.py only for:

  • deterministic joins across briefs / notes / evidence bank / bindings
  • snippet extraction and provenance assembly
  • policy-driven blockingmissing / downgradesignals / verify_fields materialization
  • pack validation and Markdown mirror generation

Do not treat run.py as the place for:

  • filler bullets that make thin evidence look complete
  • hidden sparse-evidence judgment that is not inspectable from references/ / assets/
  • reader-facing narrative prose

Output Shape Rules

Keep these stable:

  • preserve the existing top-level pack fields already used by downstream survey pipelines
  • claim_candidates must remain snippet-derived
  • concrete_comparisons must remain genuinely two-sided; if one cluster has no usable highlight, drop the card and surface thin evidence upstream instead of fabricating an A-vs-B contrast
  • snippet sampling should stay cluster-aware: when a subsection has explicit clusters, evidence selection should avoid collapsing onto one route just because its abstracts contain louder result sentences
  • sparse evidence should surface as explicit blockers / downgrade signals / verify fields, not filler bullets
  • citation keys must remain constrained to citations/ref.bib

Compatibility Notes

Current mode is reference-first with deterministic compatibility:

  • assets/evidence_policy.json defines pack thresholds and sparse-evidence routing
  • assets/evidencepackschema.json documents/validates the stable pack shape
  • assets/sourcetexthygiene.json owns this Skill's wrapper cleanup, while the

repo-wide assets/limitation-signals.json owns limitation polarity across paper-notes, evidence-draft, and writer-context-pack

  • scripts/run.py still materializes the existing JSONL + Markdown outputs, but no longer pads sparse sections with generic caution prose

Quick Start

  • uv run python .codex/skills/evidence-draft/scripts/run.py --workspace <workspace>

Execution Notes

When running in compatibility mode, scripts/run.py currently reads:

  • outline/subsection_briefs.jsonl
  • papers/paper_notes.jsonl
  • citations/ref.bib
  • optionally papers/evidencebank.jsonl and outline/evidencebindings.jsonl
  • assets/evidencepolicy.json and assets/evidencepack_schema.json

Script

Quick Start

  • uv run python .codex/skills/evidence-draft/scripts/run.py --workspace <workspace>

All Options

  • --workspace <dir>
  • --unit-id <id>
  • --inputs <path1;path2>
  • --outputs <path1;path2>
  • --checkpoint <C*>

Examples

  • uv run python .codex/skills/evidence-draft/scripts/run.py --workspace <workspace>

Troubleshooting

  • If packs look complete despite thin evidence, inspect assets/evidencepolicy.json and references/blockvs_downgrade.md before changing Python.
  • If evaluation bullets are generic, inspect references/evaluationanchorrules.md and the policy asset.
  • If claims are strong but evidence is abstract/title-only, downgrade via downgradesignals and verifyfields rather than adding narrative caveats.