aoa-eval-32d006 — independently scanned and version-tracked by SaferSkills.
SaferSkills independently audited aoa-eval-32d006 (Agent Skill) and scored it 100/100 (green). The audit ran 55 deterministic rules across Security, Supply Chain, Maintenance, Transparency, and Community; it found 0 high-severity and 0 lower-severity findings. The full rule-by-rule trace and per-finding evidence are below. Free, methodology-open.
Findings & checks · 0 flagged
Every scanned point with the score it earned and what moved between them.
First recorded scan — no prior version to compare against.
The primary manifest — the file an agent reads to learn what this artifact does.
Use this skill as the front door for AoA evaluation work. It decides whether a task should inspect existing evals, apply an existing eval or validator, record repo-local eval pressure, design a local eval suite, or mine .aoa session evidence for missed eval triggers.
When the OS Abyss aoa-evals workspace is available, this skill also starts by raising the read-only eval session readiness packet so a fresh agent can see the current tools, local eval ports, stop lines, freshness blockers, and candidate-only packet contract before choosing a subskill.
When that packet exposes eval_forge_front_door, use it as the live Eval Forge orientation surface. It should point to the Forge operating path mechanics/proof-object/parts/eval-authoring/docs/EVAL_FORGE_OPERATING_PATH.md, the session-mining criteria/reject taxonomy mechanics/proof-object/parts/eval-authoring/docs/SESSION_MINING_CRITERIA.md, the local-port decision matrix mechanics/proof-object/parts/eval-authoring/docs/LOCAL_PORT_DECISION_MATRIX.md, the latest route-review report, the worksheet example, and exact route commands. These references are routing aids, not proof authority.
Use this skill when:
connect evals to a repository
local evals/ port appears during repository work
eval:repeated agent route failure, skipped validator/test/script evidence, unsafe proof/local/MCP/session mixing, or missed trigger behavior is enough pressure
aoa-evals, aoa-evals-mcp, local eval ports, eval intake,graders, traces, regressions, validators, tests, or scripts as evaluation surfaces
tools or touching local eval files
owner surfaces have been checked
raw session evidence
Do not use this skill when:
use the normal engineering workflow or aoa-contract-test
eval, test, landing, or done are notsufficient without route pressure
aoa-source-of-truth-checkaoa-decisionaoa-memo-writebackaoa-evals proof doctrine; route tothe aoa-evals repo and its validators instead of making this skill the owner
evals/PORT.yaml, evals/intake/, scripts, tests, validators, and reporoute cards
aoa-evals docs, schemas, validators, reports, and review surfacesAbyss this may be aoa-evals/scripts/build_local_eval_port_inventory.py
aoa-evals/scripts/aoa_eval_session_start.py
eval_forge_front_door packet fields from session-start/readiness:surface_refs, exact_commands, surface_status, stop-lines, and non-proof boundary flags
aoa-evals/scripts/validate_eval_candidate_packets.py
aoa-evals-mcp packets when available, treated as access-plane data.aoa session evidence through aoa-session-memory-evidence-routewhen the task asks how an eval, validator, test, MCP, failure, or trigger was used in prior sessions
.aoa search hits, segments, raw refs, and freshness only when session miningis the chosen route
aoa-eval-select, aoa-eval-apply,aoa-eval-local-need, aoa-eval-design, or aoa-eval-session-mining
.aoa evidence role
freshness blockers and stop lines that constrain the route
exposes them, with proof authority explicitly kept false
session-mining report
aoa-evals checkout is available, raise the read-onlysession readiness packet before choosing a subskill:
aoa-evals repo, runpython scripts/aoa_eval_session_start.py --json
freshness blockers, and stop lines as advisory routing evidence
eval_forge_front_door, inspect surface_refs,exact_commands, surface_status, and non-proof boundary flags before classifying the route
EVAL_FORGE_OPERATING_PATH.md for the first operating path,SESSION_MINING_CRITERIA.md before session mining, LOCAL_PORT_DECISION_MATRIX.md before local intake/design, and the worksheet example before owner-review worksheet work
and central source inspection instead of inventing a route
.aoa or trace-derived eval candidates, validate thecandidate-only contract when the validator is available: python scripts/validate_eval_candidate_packets.py --schema-only; validate any actual packet path before using it as candidate evidence
aoa-eval-selectaoa-eval-applyaoa-eval-local-need
aoa-eval-design.aoa evidence should be mined for missed trigger cases: useaoa-eval-session-mining
before local need/design/session mining to distinguish missing, skeleton, active, invalid, and stale local surfaces; treat inventory route keys as advisory routing evidence, not proof or mutation authority
AGENTS.md, evals/PORT.yaml,nearby validators, tests, scripts, and local route docs
aoa-evals only for proof doctrine, local-port contract,central bundles, scoring, verdicts, or review rules
aoa-evals-mcp only as a runtime access plane; do not let an MCP packetcreate proof truth or write central eval bundles
aoa-session-memory-evidence-routeor the equivalent aoa-session-memory-mcp read-only packet to find usage, consequence, failure, and raw refs; keep those refs candidate-only until the local or central eval owner accepts them
of context unless the route changes
narrow source to inspect
validation command, and remaining proof risk
aoa-evals owns proof doctrine and central eval verdictsevals/ ports and can hold intake pressurewithout becoming proof authorities
.aoa raw traces are candidate evidence, not reviewed truthaoa-session-memory evidence routes can locate usage and consequences, butthey do not own proof doctrine, verdicts, scoring, baselines, or eval promotion
queue summaries are read-only routing aids, not proof objects
owner route, but they do not score, promote, accept, or prove an eval
local or central surface
central proof
files and owner validators remain stronger
inspected
aoa-evals objectmemory
.aoa search hits override repo-local source filesverdict, score, baseline, or proof promotion
route into existing surfaces, local ports, worksheets, rejects, or review
evals/PORT.yaml was inspected or the absence was namedand the route recommendation it returned
name any freshness blocker or stop line that constrained the route
eval_forge_front_door is available, confirm its operating path,criteria, local-port matrix, route-review or worksheet refs, and exact commands were considered as routing evidence only
aoa-evals authority was kept separate from local intakedesign
.aoa evidence is marked candidate-only when usedrefs and did not replace local or central eval owner review
packet validator when the validator exists
Manifest-backed techniques:
8Dionysus/aoa-techniques at 1a7d146957108ecefc24219c7d56357c5a4a2c2c using path techniques/proof/evaluation-chain/contract-first-smoke-summary/TECHNIQUE.md and sections: Intent, Inputs, Outputs, Core procedure, Contracts, Risks, Validation8Dionysus/aoa-techniques at 1a7d146957108ecefc24219c7d56357c5a4a2c2c using path techniques/governance/decision-routing/owner-layer-triage/TECHNIQUE.md and sections: Intent, Inputs, Outputs, Core procedure, Contracts, Risks, Validation8Dionysus/aoa-techniques at 1a7d146957108ecefc24219c7d56357c5a4a2c2c using path techniques/proof/owner-truth-closeout/canonical-owner-with-validated-mirror/TECHNIQUE.md and sections: Intent, Inputs, Outputs, Core procedure, Contracts, Risks, Validationaoa-evals-mcp tool names~30 seconds. Free. No account. Every finding cites a rule and a line of evidence.