PlaygroundCatalog › Evidence Grading
EG

Evidence Grading

🔵 Stable🕐 updated 2026-07-19 🔷 SkillSpec L3 pm-cowork

Grade the evidence behind a claim before betting on it — the hierarchy for business evidence (experiments > usage data > surveys > interviews > anecdotes > opinion), the fit-for-decision test, and the mixed-evidence verdicts that real questions produce. Use when asked how strong is our evidence for this, grade what we know before the decision, is this enough to bet on, or we have three anecdotes and a survey — now what. Produces the evidence inventory with grades, the sufficiency verdict against the decision's stakes, and the cheapest-upgrade path.

▶ Run it free — no key needed 📝 Grade your existing draft View SKILL.md ↗

What to give it

The claim and the decision riding on it — "users want X" feeding a backlog item vs. feeding a repositioning are different sufficiency bars; the decision's reversibility and cost set the bar
The evidence, itemized — every piece: the data pull, the survey, the five customer quotes, the competitor's move, the expert's opinion — including the inconvenient items (an inventory that omits contradicting evidence is advocacy)
The evidence's provenance — n, selection, dates, who collected it and with what incentive ([source-triangulation](../source-triangulation/SKILL.md) supplies the externals; internal evidence has incentives too)

✅ The bar it holds itself to

Every skill in this library self-verifies — these are this skill's own quality checks, straight from its definition.

Every item has a rung and within-rung quality notes
Contradicting evidence is in the inventory at its own grade
Echoes were collapsed before weighing
The verdict is stakes-relative, not absolute
The insufficient branch ends in a priced upgrade, not just a no

⚠️ What it refuses to do

Do not grade by vividness — the memorable anecdote outshines the boring dataset in every meeting; the hierarchy exists to resist exactly that
Do not conclude prevalence from existence — the most common grading felony
Do not omit the inconvenient items — an advocacy inventory grades the author, not the claim
Do not demand experimental grade for reversible bets — over-evidencing cheap decisions is its own waste
Do not end at "insufficient" — the upgrade path is the difference between rigor and obstruction

Install

npx pm-claude-skills add --agent claude   # or codex · cursor · gemini · hermes
# or one-line MCP (every skill, any client):
claude mcp add pm-skills -- npx -y pm-claude-skills-mcp

Related skills

🔌 Embed this skill

Drop this on your blog, docs, or site — it renders a "Run this skill" card:

<div data-pm-skill="evidence-grading"></div>
<script src="https://mohitagw15856.github.io/pm-claude-skills/embed.js" async></script>

💬 Discussion

Evidence Grading is one of 750 open-source professional AI agent skills — all SkillSpec L3. Try them all in the browser · ⭐ Star on GitHub · Browse the full catalog