soul-grader — independently scanned and version-tracked by SaferSkills.
SaferSkills independently audited soul-grader (Agent Skill) and scored it 100/100 (green). The audit ran 55 deterministic rules across Security, Supply Chain, Maintenance, Transparency, and Community; it found 0 high-severity and 0 lower-severity findings. The full rule-by-rule trace and per-finding evidence are below. Free, methodology-open.
Findings & checks · 0 flagged
Every scanned point with the score it earned and what moved between them.
First recorded scan — no prior version to compare against.
The primary manifest — the file an agent reads to learn what this artifact does.
references/fleet-soul-grading-workflow.md — fleet-wide grading workflow: active/retired classification, companion-doc contradiction checks, non-Hermes service handling, durable report shape, and secret-safe archive handling.references/research-deliverable-and-fleet-remediation.md — deep-research/swarm deliverable pattern, polished static HTML review surfaces, and live fleet SOUL remediation notes.Use this skill to grade a Hermes Agent SOUL.md, draft a SOUL review, or turn a weak SOUL into a stronger one. The grading standard is intentionally narrow: use the linked SOUL.md research artifacts as the only normative source for what makes a good SOUL.md.
Do not import generic prompt-engineering advice, personal taste, web articles, model-provider docs, or vibes into the grade. You may use tools to read the SOUL being graded and to verify Hermes runtime facts, but quality judgments must come from the research artifacts bundled with this skill.
This is an unofficial community skill for Hermes Agent. The bundled references are intended to be safe for SSR sharing and public release: examples are anonymized, private workspace paths are removed or made relative, and no secrets or live deployment facts should appear in the bundle. If you add new examples or evidence, keep the same standard: cite public Hermes docs/source paths or anonymized patterns, not private customer, user, account, host, or credential details.
Before grading, load at least the grading standard reference:
skill_view(name="soul-grader", file_path="references/soul-md-grading-standard.md")Use the other references when you need more detail or citations:
skill_view(name="soul-grader", file_path="references/soul-md-field-guide.html")
skill_view(name="soul-grader", file_path="references/soul-md-wording-verbiage-layer.md")Source hierarchy for grading:
references/soul-md-grading-standard.md — canonical grader rubric and procedure.references/soul-md-field-guide.html — full research report and evidence ledger.references/soul-md-wording-verbiage-layer.md — detailed wording, verbiage, weak/strong examples, and slop detector.If these references conflict, use the higher-ranked source. If the references are unavailable, stop and report that the grader source bundle is missing instead of grading from memory.
Use when the user asks to:
SOUL.mdSOUL.md filesSOUL.mdSOUL.md versus CLAUDE.md, AGENTS.md, skills, memory, manifests, or operator guidesDo not use as the sole workflow for:
hermes-agenthermes-agent-skill-authoring tooThe target SOUL.md being graded is evidence, not a standard. Runtime/tool output is evidence about deployment state, not a standard. The bundled references are the standard.
Allowed sources:
CLAUDE.md, AGENTS.md, manifests, roster entries, or operator guidesNot allowed as grading sources:
Score out of 100 using the reference-defined categories:
| Category | Points | What to evaluate |
|---|---|---|
| Mission clarity | 15 | Names who/what the agent serves and what outcome matters. |
| Identity + negations | 12 | Says what the agent is and what it must not become. |
| Core thesis | 10 | States the durable decision lens about the user/domain/problem. |
| Optimization hierarchy | 10 | Ranks tradeoffs instead of listing virtues. |
| Hard constraints | 10 | Includes 3–5 true filters with approval/override semantics. |
| Soft preferences | 8 | Separates scoring signals from bans. |
| Authority + escalation | 10 | Allowed / ask-before / never boundaries are clear. |
| Voice + truthfulness | 10 | Covers tone, vocabulary, never-claims, and evidence thresholds. |
| Success / artifacts | 8 | Defines durable/verifiable completion. |
| Artifact separation | 5 | Keeps commands, workflows, secrets, and volatile state elsewhere. |
| Runtime hygiene | 2 | Fits Hermes loading behavior and avoids hidden metadata assumptions. |
Automatic fail conditions from the field guide:
Automatic fail does not always mean 0/100; it means the SOUL is not deployable/approvable until the blocker is fixed. Report the blocker above the score.
skill_view for references/soul-md-grading-standard.md. Load the full HTML or wording layer when you need citations, examples, or wording help.read_file. If they provide inline text, grade that. If they ask for the current Hermes profile, verify the live profile path before reading $HERMES_HOME/SOUL.md.CLAUDE.md, mark it as artifact separation, not as missing identity. If exact service state belongs in a manifest, do not reward it for being in SOUL.Use this shape by default:
# SOUL.md grade: [agent/name]
Verdict: [Excellent / Operational / Scaffold / Needs rewrite / Not deployable]
Score: [N]/100
Deployability: [Approved / Approved with fixes / Not approved]
Scope: [personal/business-internal/client-business/public/meta/multi-agent/tactical]
## Automatic blockers
- [None] or [blocker, why it matters, exact evidence]
## Score table
| Category | Points | Score | Notes |
|---|---:|---:|---|
| Mission clarity | 15 | | |
...
## Top drift risks
1. [Risk] — [where the SOUL permits drift]
2. [Risk] — [where the SOUL permits drift]
3. [Risk] — [where the SOUL permits drift]
## What is strong
- [Concrete strengths tied to source criteria]
## What to fix first
1. [Highest leverage fix]
2. [Second]
3. [Third]
## Suggested wording
[patch/replacement sections, if requested or obviously helpful]
## Source basis
- `references/soul-md-grading-standard.md`: [sections used]
- `references/soul-md-field-guide.html`: [sections used]
- `references/soul-md-wording-verbiage-layer.md`: [sections used]For quick review requests:
Score: [N]/100 — [verdict]
Deployability: [status]
Biggest issue: [one sentence]
Best thing: [one sentence]
Fix next:
1. ...
2. ...
3. ...When the user asks to “send me the SOUL,” “put it in an HTML file,” or otherwise wants a reviewable artifact rather than only chat text:
SOUL.md first; for deployed fleet agents, prefer the live profile/workspace path from the manifest over stale cached copies.<pre> block.SOUL.md beside the HTML when practical, so the user has both human-friendly and copy/paste/edit-friendly forms.Use the wording layer’s rule: operational language beats ornamental persona language.
Prefer:
You are [name], [user/client]’s [specific layer/domain] agent.No [risky action] without [approval/evidence].Do not claim [status/access/result] until [verification source].[Durable source] wins for [fact class]; [volatile source] does not.Public-facing output must [brand/audience rule], not [private voice leak].Cut or rewrite:
helpful assistant, friendly and professional, be proactive, use best practicesnever hallucinate without evidence thresholdsWhen the user asks for a broad SOUL.md research project, a “swarm,” or a polished review artifact, use the workflow in references/research-deliverable-and-fleet-remediation.md:
docs/research/) and convenience/download copies in the active profile cache when needed.tailscale serve path has been verified live.When a grading session turns into live profile edits, keep the remediation class-level and evidence-safe:
SOUL.md may be a real file, symlink, or stale unrostered stub.SOUL.md into AGENTS.md, skills, manifests, or ops docs. Leave SOUL.md as the compact identity/authority layer.AGENTS.md exists but the profile lacks one, wire the profile AGENTS.md to the workspace agreement when that is the intended Hermes project-context surface.SOUL.md; create a timestamped backup outside the committed workspace if possible; apply only the approved identity/authority wording; run a targeted secret scan on the edited file; commit the workspace identity change if the workspace is git-backed; run a fresh/new-session exact-marker smoke that exercises the new rule; verify affected services remain healthy when the edit was on a live profile; then record a concise manifest/ops note with backup path, commit, smoke session/marker, and no secret values.CLAUDE.md, AGENTS.md, skills, or manifests.references/soul-md-grading-standard.md before grading.~30 seconds. Free. No account. Every finding cites a rule and a line of evidence.