apsr-data-analysis — independently scanned and version-tracked by SaferSkills.
SaferSkills independently audited apsr-data-analysis (Agent Skill) and scored it 100/100 (green). The audit ran 55 deterministic rules across Security, Supply Chain, Maintenance, Transparency, and Community; it found 0 high-severity and 0 lower-severity findings. The full rule-by-rule trace and per-finding evidence are below. Free, methodology-open.
Findings & checks · 0 flagged
Every scanned point with the score it earned and what moved between them.
First recorded scan — no prior version to compare against.
The primary manifest — the file an agent reads to learn what this artifact does.
APSR reviewers are methodologically sophisticated and the editorial office will later re-run your code against the manuscript's tables and figures (see apsr-transparency-and-data-policy). Analyze as if both are true — because they are. This skill covers execution and reporting norms; design decisions live in apsr-research-design.
substantive meaning of the estimate, not just its significance.
(alternative measures, samples, estimators, fixed effects), and say what you learn.
comparisons; do not mine for a significant interaction and theorize it post hoc.
experiments; small-cluster corrections (wild-cluster bootstrap) when clusters are few.
ones; reconcile deviations from the plan and justify them.
a coding/scaling choice.
outputs as ground truth.
renv.lock, requirements.txt, recorded ssc/net installs).Run the battery, don't just enumerate it. Full map: execution-with-mcp. APSR is general-interest political science — observational causal designs (DiD/IV/RDD) and survey/field experiments alike; cluster by the right unit and foreground identification.
romano_wolf (step-down FWER) orbenjamini_hochberg — report the adjusted threshold.
oster_delta / sensemakr.wild_cluster_bootstrap (few clusters), twoway_cluster / conley;multilevel data → cluster at the right level.
audit_result(result_id) lists the missing checks and theexact suggest_function for each.
etable / did_summary_to_latex from the handle — no retyped numbers.Keep the decisive checks in the body and the exhaustive battery in the supplement. See the executed chain in the JF execution walkthrough.
【Main estimate】magnitude + interval + substantive meaning
【Identification check】(per research-design) result
【Robustness】specs that could break it → what held
【Heterogeneity】pre-specified? MHT-adjusted?
【Registered vs exploratory】clearly separated?
【Reproducible】master script + seeds + pinned versions? [Y/N]
【Next】apsr-tables-figuresAPSR is the flagship of the American Political Science Association, published by Cambridge University Press, and its reviewers are drawn from across the discipline — so the same results section can be read by a formal theorist, a survey methodologist, and a comparativist at once. Calibrate the analysis to whichever lens is decisive, but expect all three to be in the room.
| Analytic tradition | The check an APSR referee runs first | The fix that earns the benefit of the doubt |
|---|---|---|
| Survey / lab experiment | Is inference randomization-based and pre-registered? | Randomization inference, pre-registered estimand, MDE reported |
| Observational causal | Is the "causal" word doing more than the design licenses? | State estimand + assumption; sensitivity to an unobserved confounder |
| Text-as-data / computational | Was the model validated against human labels? | Held-out validation set, stability across seeds, version pinned |
| Formal-empirical | Do the tests follow comparative statics, or a loose analogy? | Map each prediction to a parameter the model moves |
| Multi-method | Do quant and qual estimates actually corroborate? | Show where they agree, and own where they diverge |
A hypothetical APSR survey experiment tests whether co-partisan endorsements raise support for a redistricting reform. The pre-registered ATE is +6.2 points (95% CI 3.1 to 9.3) on a 0–100 support scale, randomization-inference p = 0.004. The exploratory subgroup "low political-knowledge respondents" shows +11.8 points, but it was not pre-registered and the interaction p = 0.04 before any multiplicity correction — after a Bonferroni adjustment across the six exploratory subgroups it crosses 0.20. The disciplined write-up reports the +6.2 confirmatory effect with its interval and substantive meaning, flags the +11.8 figure as exploratory and not multiplicity-robust, and frames it as a hypothesis for future work rather than a finding. (All numbers illustrative.)
discipline-wide stake (representation, accountability, institutional design) before the numbers.
specifications that could break the result, and say what you learned when they did not.
the argument named in advance, not to a pattern noticed afterward.
office will later re-run the deposited code, so the split must survive verification.
only a specialist would value rarely clears APSR review.
second-class to a regression. Match the inference standard to the design.
conditionally-accepted package reproduces every printed number. Exact deposit mechanics can change — confirm against the journal's current submission and transparency guidelines.
../../resources/external_tools.md — estimation, inference, and text-as-data packages../../resources/official-source-map.md — reproducibility-verification policy~30 seconds. Free. No account. Every finding cites a rule and a line of evidence.