adversary — independently scanned and version-tracked by SaferSkills.
SaferSkills independently audited adversary (Agent Skill) and scored it 100/100 (green). The audit ran 55 deterministic rules across Security, Supply Chain, Maintenance, Transparency, and Community; it found 0 high-severity and 0 lower-severity findings. The full rule-by-rule trace and per-finding evidence are below. Free, methodology-open.
Findings & checks · 0 flagged
Every scanned point with the score it earned and what moved between them.
First recorded scan — no prior version to compare against.
The primary manifest — the file an agent reads to learn what this artifact does.
You have an answer. Before you deliver it — write the case that it should be rejected.
Not a balanced review. Not "on the other hand." A prosecution. You are the most competent opponent this answer will ever face. Build the case for rejection with the same quality you used to build the answer.
You are about to generate a review of your own work that:
This is confirmation cascade — each supporting token makes the next supporting token more likely. The review becomes a rubber stamp. Breaking the cascade requires generating content that actively undermines your own conclusion.
One paragraph. What is the answer, recommendation, or plan you are about to deliver?
Write it clearly enough that an opponent could attack it. If you can't state it in one paragraph — the answer isn't coherent enough to review.
Artifact: The stated answer. Everything below attacks this specific text.
Not generic concerns. Attacks on THIS specific answer for THIS specific problem.
For each attack, use this template:
ATTACK [N]: [one-line summary]
Claim: [the specific thing that is wrong, incomplete, or dangerous]
Evidence: [why this attack is plausible — cite specific aspects
of the answer, the domain, or the context]
If true: [what happens — the specific consequence]Attack axis checklist — write at least one attack per axis:
Anti-fake rule: Each attack must reference specific content from the stated answer (Step 1) or specific facts about THIS problem's context. Generic attacks like "there may be edge cases" are not attacks — they are noise.
Artifact: Five numbered attacks. Step 3 must process each one.
For each of the five attacks, assign one rating with written justification:
ATTACK [N]: [FATAL / SIGNIFICANT / MINOR / DISMISSED]
Justification: [specific evidence-based reasoning — not "I don't think so"]Rating criteria:
Dismissal rules:
Artifact: Five rated attacks with written justifications. Step 4 depends on these ratings.
If any FATAL exists: Stop. The answer does not ship. Rebuild from the attack's insight.
If SIGNIFICANT exists (no FATAL): Revise the answer to address each SIGNIFICANT attack. OR explicitly flag the limitation: "This recommendation assumes X. If X is false, the alternative is Y."
If MINOR only: Ship the answer with limitations named. The user deserves to know them.
If all DISMISSED: Ship with the opposition record attached. Transparency proves the answer was tested, not just generated.
OPPOSITION RECORD
────────────────────────────────────────
Answer reviewed: [summary from Step 1]
Attacks mounted: [count]
Fatal: [count] — [list]
Significant: [count] — [list]
Minor: [count]
Dismissed: [count]
Answer status: [passed / revised / rebuilt]
Surviving risks: [what to watch for — from MINOR/SIGNIFICANT attacks]
Confidence: [high / medium / low — earned by this record]
────────────────────────────────────────The model's default is to confirm its own answer — each token after the initial conclusion is more likely to support than to challenge. This skill creates a structural break where the model generates content that actively opposes its own conclusion. The five mandatory attacks force the model to activate knowledge pathways that confirmation bias suppresses. The most useful insight in any opposition is often not the attack itself — it is the weakness it reveals that the answer needs to address to be genuinely correct.
~30 seconds. Free. No account. Every finding cites a rule and a line of evidence.