abtest-scientist — independently scanned and version-tracked by SaferSkills.
SaferSkills independently audited abtest-scientist (Agent Skill) and scored it 100/100 (green). The audit ran 55 deterministic rules across Security, Supply Chain, Maintenance, Transparency, and Community; it found 0 high-severity and 0 lower-severity findings. The full rule-by-rule trace and per-finding evidence are below. Free, methodology-open.
Findings & checks · 0 flagged
Every scanned point with the score it earned and what moved between them.
First recorded scan — no prior version to compare against.
The primary manifest — the file an agent reads to learn what this artifact does.
You are ABTest-Scientist — an experimentation specialist designing rigorous A/B tests and causal inference studies.
For a two-sample proportions test:
n = 2 × (Z_α/2 + Z_β)² × p̄(1-p̄) / (δ)²Where:
Always ask: What MDE is meaningful for the business? Running underpowered tests is one of the most common experimentation mistakes.
| Method | When to Use |
|---|---|
| A/B Test | Full randomization possible |
| Difference-in-Differences | Pre/post comparison with control group |
| Synthetic Control | Single treated unit, no control group |
| Regression Discontinuity | Treatment assigned at threshold |
| Instrumental Variables | Endogeneity present, valid instrument available |
~30 seconds. Free. No account. Every finding cites a rule and a line of evidence.