pm-red-team — independently scanned and version-tracked by SaferSkills.
SaferSkills independently audited pm-red-team (Agent Skill) and scored it 100/100 (green). The audit ran 55 deterministic rules across Security, Supply Chain, Maintenance, Transparency, and Community; it found 0 high-severity and 0 lower-severity findings. The full rule-by-rule trace and per-finding evidence are below. Free, methodology-open.
Findings & checks · 0 flagged
Every scanned point with the score it earned and what moved between them.
First recorded scan — no prior version to compare against.
The primary manifest — the file an agent reads to learn what this artifact does.
Runs an adversarial second pass on something that has already been written or critiqued. The goal is not to be contrarian — it's to surface the parts of the analysis that wouldn't survive a hostile reviewer (an exec, a board member, a sharp peer PM, a competitor strategist). Adapted from the chief-of-staff pattern of critically evaluating AI output rather than deferring to it.
The skill is opinionated. It will disagree with the prior pass when it has reason to. It will keep what's strong. It will not generate noise just to look productive.
Use this when:
pm-evaluator, pm-prd-drafter, pm-decision-coach, pm-value-hypothesis-tester, pm-metrics-critic, pm-launch-reviewer) has produced an evaluation or draft, and the user wants it pressure-tested before sending up the chainDon't use this when:
pm-prd-drafter or pm-decision-coachTreat them as two distinct objects:
Most red-team failures come from skipping step one and just re-critiquing the prior analysis without going back to the source. Read both.
The first pass usually applied one of the repo's primary lenses (rubric scoring, problem framing, value-hypothesis stress test). The second pass should pick a different lens, not run the same one harder.
Useful lenses to switch to, depending on what the first pass used:
rubrics/pm-evaluation-rubric.md): run the stakeholder lens — whose perspective is missing? Procurement, support, sales reps, finance, regulatory, downstream teams. (See cross-functional/stakeholder-alignment.md § "Bundle thinking.")decision-making/value-hypothesis.md and decision-making/competitive-moat.md.)cross-functional/engineering-partnership.md and cross-functional/failure-management.md.)decision-making/metrics.md § "When the metrics are lying.")Volume is not value. The first pass already caught the obvious issues. The job here is to find the two or three most consequential gaps — the ones that, if exposed in an exec review, would derail the recommendation.
Useful filters:
Adversarial review is not blanket disagreement. If the first pass got the diagnosis right, say so. If a section is already tight, don't manufacture a critique. Calling out three things that are good enough to keep makes the three things that are not land harder.
A red-team review that ends with "consider revising" is useless. Specify:
If the right move is don't ship this, say so directly: "The load-bearing assumption is a guess; the recommendation is not yet ready. Defer until [specific test] runs."
If the prior analysis is itself wrong — too lenient, off-target, applying the wrong framework — name it. "The prior pass scored this 4/5 on Criterion 3, but the bundle effect on the primary product is unaddressed; on the rubric as written, this should be at most 2/5."
This is the chief-of-staff pattern made explicit: critically evaluate the prior AI output, do not defer to it.
## TL;DR
[One paragraph honest take. Does the artifact, after the prior critique, hold up under a hostile reviewer? Yes / partially / no — and the one thing that most determines the answer.]
## What the prior pass got right
[Two or three specific things that are already strong enough to keep. Be concrete — quote or cite.]
## What the prior pass missed
For each (max three; quality > quantity):
**[Hole 1 — short title]**
- *Where:* [the specific section, claim, or metric in the original artifact]
- *The miss:* [what's wrong or absent, in one or two sentences]
- *Why it matters:* [the consequence — what an exec / board member / competitor would do with it]
- *Re-write:* [the specific replacement line or the specific test that would close the gap]
## Where the prior pass itself went off
[Optional. If the prior critique is itself weak — wrong framework, lenient grading, missed lens — name it. One or two paragraphs.]
## What I'd do before sending this up the chain
[The 1-3 concrete actions. Each one must be doable. "Run the bundle-effect analysis on the primary product" beats "consider bundle effects." Pre-commit a kill criterion. Re-run the value-hypothesis test on a tighter segment. Etc.]
## Verdict
[Ship / revise / defer / kill, with one sentence of reasoning. Take a position.]cross-functional/stakeholder-alignment.md, decision-making/value-hypothesis.md, etc.) so the user can deepen the read.Common invocation patterns:
If the user invokes this skill without first having a primary analysis in scope, ask: "What's the prior analysis (the artifact and any existing critique)? I need both to run a useful second pass."
~30 seconds. Free. No account. Every finding cites a rule and a line of evidence.