stability-experiment — independently scanned and version-tracked by SaferSkills.
SaferSkills independently audited stability-experiment (Agent Skill) and scored it 100/100 (green). The audit ran 55 deterministic rules across Security, Supply Chain, Maintenance, Transparency, and Community; it found 0 high-severity and 0 lower-severity findings. The full rule-by-rule trace and per-finding evidence are below. Free, methodology-open.
Findings & checks · 0 flagged
Every scanned point with the score it earned and what moved between them.
First recorded scan — no prior version to compare against.
The primary manifest — the file an agent reads to learn what this artifact does.
This is the method that established the operating point (specs/EXPERIMENTS.md). Reuse it; don't improvise.
+ per-host rate limit + retry) and reports ratelimit_rate. The gate passes only when it's 0. For isolating an endpoint, a few raw http.Gets at a fixed spacing are enough — the point is controlled, attributable observation.
before any comparative test. If it persists for hours, back-to-back runs aren't attributable to code — a "fix" may just be a lapsed limit. Probe isolated requests over time right after a limited run; proceed only if recovery is short (minutes — which is what the reference found for DDG).
Longer intervals / more jitter before anything clever; never proxies/CAPTCHA (out of scope — the honest fallback is a keyed API).
cmd/bench), not just an isolated probe.must expect the pace, not flag it.
Summarize the finding in specs/EXPERIMENTS.md (Hypothesis · Setup · Observations · Results · Insight) and log the resulting decision via log-decision. Stop when the gate passes with margin — not when it's theoretically optimal.
~30 seconds. Free. No account. Every finding cites a rule and a line of evidence.