cps-research-design — independently scanned and version-tracked by SaferSkills.
SaferSkills independently audited cps-research-design (Agent Skill) and scored it 100/100 (green). The audit ran 55 deterministic rules across Security, Supply Chain, Maintenance, Transparency, and Community; it found 0 high-severity and 0 lower-severity findings. The full rule-by-rule trace and per-finding evidence are below. Free, methodology-open.
Findings & checks · 0 flagged
Every scanned point with the score it earned and what moved between them.
First recorded scan — no prior version to compare against.
The primary manifest — the file an agent reads to learn what this artifact does.
CPS is methodologically pluralist but demanding about each tradition. The design must credibly connect the comparative argument (cps-theory-building) to evidence and rule out the leading rival (cps-literature-positioning). This skill is mode-aware: pick the section that matches your work and defend the comparative leverage — the variation across cases or time that identifies the claim.
(parallel trends, exclusion, continuity, ignorability). Defend them; don't assert them.
(use modern staggered-adoption estimators, not naive TWFE); RD around institutional thresholds; IV (first-stage strength, exclusion, weak-IV-robust inference); survey experiments fielded comparatively.
coding, harmonized surveys); address country-level confounding and cross-national measurement error.
corrections (wild bootstrap) when the number of countries is small; multiple-comparison adjustment.
(shared outcome despite different contexts) — justified by design, not convenience.
case of and avoid selecting on the outcome.
would have disconfirmed the mechanism.
cps-transparency-and-data).contexts; sampling frames; what the comparative contrast licenses about generalization.
mechanism; state how each method covers the other's blind spot, not as decoration.
For the single strongest rival, write one sentence: "If the rival were true rather than my argument, the cross-case/over-time pattern would look like ___; instead it looks like ___." If you cannot, the design does not yet identify the comparative contribution.
Estimate and audit the design, don't only describe it. Full map: execution-with-mcp. CPS is comparative politics — cross-national and sub-national designs; emphasize identification and clustered / multiway inference.
detect_design → recommend → fit with as_handle=true → audit_result.callaway_santanna / sun_abraham +bacon_decomposition + honest_did_from_result); IV (effective_f_test + anderson_rubin_ci); RDD (rdrobust + mccrary_test).
romano_wolf for many-outcomefamily-wise control, and mediate for mediation (not naive controlling-away).
oster_delta / sensemakr for observational claims.Report the effect size in interpretable units; route the full battery to the appendix/supplement. A run end-to-end (synthetic data, real returns) is in the JF execution walkthrough.
【Mode】comparative-causal / case-based / experiment / multi-method
【Comparative leverage】the across-case / over-time variation that identifies the claim
【Estimand or claim】what is being identified/shown
【Key assumption(s)】and how each is defended (incl. comparability)
【Rival ruled out】the adjudication sentence
【Robustness/sensitivity】planned checks
【Next】cps-data-analysis../../resources/external_tools.md — comparative datasets, identification packages, and CAQDAS for qualitative work../../resources/code/ — staggered-DiD / IV / RDD / DML command chain to adapt~30 seconds. Free. No account. Every finding cites a rule and a line of evidence.