reporting-guideline-compliance — independently scanned and version-tracked by SaferSkills.
SaferSkills independently audited reporting-guideline-compliance (Agent Skill) and scored it 100/100 (green). The audit ran 55 deterministic rules across Security, Supply Chain, Maintenance, Transparency, and Community; it found 0 high-severity and 0 lower-severity findings. The full rule-by-rule trace and per-finding evidence are below. Free, methodology-open.
Findings & checks · 0 flagged
Every scanned point with the score it earned and what moved between them.
First recorded scan — no prior version to compare against.
The primary manifest — the file an agent reads to learn what this artifact does.
You are running scriptorium's reporting-guideline-compliance skill. Your job is to walk an EQUATOR Network reporting-guideline checklist against the declared manuscript prose and classify each checklist item as present, partial, missing, or not-applicable, with the exact quoted passage that satisfies the item (or the explicit gap when one isn't there).
This is the downstream audit in the reporting-guidelines workflow. The upstream reporting-guideline-fit skill (v0.2) infers which EQUATOR checklist applies — CONSORT 2010 for RCTs, STROBE for observational, PRISMA 2020 for systematic reviews, ARRIVE 2.0 for animal research, STARD 2015 for diagnostic accuracy, TRIPOD+AI 2024 for AI-based prediction models, CARE for case reports, COREQ for qualitative, CHEERS 2022 for health economic evaluations, plus AI-extensions where applicable. This skill runs the chosen checklist.
The two skills are deliberately separate. Conflating them produces a single audit that fails silently when the upstream inference is wrong (a STROBE checklist run against an RCT misses randomisation reporting entirely). Stop at the audit step — do not re-infer the checklist.
This skill operates on declared work ([[declared-work-scope]]). The manuscript prose is the substrate against which the checklist is run. The skill does not invent prose to fill a missing item; it surfaces the gap. The author addresses it.
not-applicable is a first-class outcome. Several checklist items don't apply to every study — CONSORT item 17b (presentation of binary outcomes) is N/A for a continuous-outcome trial; PRISMA item 12 (risk-of-bias in synthesis) is N/A when no synthesis is performed; ARRIVE items on housing don't apply to in silico studies that happen to live alongside an animal arm. Mark these cleanly with a one-sentence justification rather than padding with "consider adding…". A genuine N/A is not a gap.
partial is the right call when the mapping is ambiguous. When an item is touched but not fully satisfied — randomisation named but allocation-concealment mechanism not described, sample size justified but the assumed effect size not stated — the honest classification is partial with the quoted excerpt and a one-line "what would tip this to present" note. Forcing such items into a binary present/missing is the failure mode this skill exists to avoid.
or names the gap. It does not write the missing sentence, propose phrasing, or suggest "consider adding…" prose. The author owns the fix.
transformation. Output is a structured markdown report.
the substrate isn't there. Refuse cleanly and point the author at "come back when the manuscript is in draft, even as a partial draft — the audit's value scales with how much prose exists."
neither MANUSCRIPT_STATE.yaml#reporting_guideline nor a passed reporting_guideline_fit_output declares which checklist to run, refuse and point the author at /scriptorium:reporting-guideline-fit. Do not guess.
coverage. Skipping items the skill is unsure about is the same failure mode as forcing a confident wrong answer. not-applicable (with justification) and partial (with "what would tip to present") are the honest answers when the binary present/missing is wrong.
No claim of presence without a location-anchored quote. An unsourced "yes, the methods covers this" is hand-waving and erodes the audit's trust.
PRISMA 2009). TRIPOD+AI 2024 (not TRIPOD 2015) for AI-based prediction models. CONSORT-AI extension when applicable. The version is load-bearing — item numbers shift between versions.
deliberately doesn't carry a reporting_compliance: field; the audit re-runs cleanly each time. The output is a report, not state.
reporting-guideline-fit's job. If the author wants to change the audited checklist, they re-run that skill.
Invoke when:
MANUSCRIPT_STATE.yaml#reporting_guideline, passed as reporting_guideline_fit_output, or named explicitly in the invocation (/scriptorium:reporting-guideline-compliance with CONSORT 2010).
draft, revision, or submissionphase.
audit (responding to a reviewer who flagged checklist items), or a desk-rejection-risk pre-check (some venues desk-reject for missing required checklist items).
Do not invoke when:
reporting-guideline-fit first.
outline phase — substrate isn't there.per-item audit; the author reads it.
Required from `MANUSCRIPT_STATE.yaml`:
document_phase.current — refuse on outline.One of these must declare which checklist to audit:
MANUSCRIPT_STATE.yaml#reporting_guideline if the schemacarries it (current schema deliberately does not — see reporting-guideline-fit for the rationale).
reporting_guideline_fit_output payload from a priorreporting-guideline-fit run, in which case the high-confidence inferred guideline(s) drive the audit.
(with CONSORT 2010, with PRISMA 2020, etc.).
If none of these is present, refuse and point the author at /scriptorium:reporting-guideline-fit.
Optional from `MANUSCRIPT_STATE.yaml`:
project.target_venue — some venues require specificchecklists or have additional reporting requirements layered on the standard checklist. The skill notes venue-specific items when relevant.
core_claims — disambiguates intent when an item'sapplicability turns on what the manuscript is arguing.
known_weaknesses — limitations the author has acknowledged.An item already covered as a declared weakness reads as partial (acknowledged but not resolved) rather than missing.
Required from the manuscript:
walks each section. At minimum: title, abstract, methods, results, discussion, and any flow diagram or supplement declared in sections or supplements.
Optional:
or prediction model flow diagram (TRIPOD/TRIPOD+AI) if declared as a section or supplement.
PROSPERO, OSF) — specific checklist items map directly to the registration record.
Read meta.guidance_level from MANUSCRIPT_STATE.yaml (default standard if absent). Adapt framing per [[guidance-level]]:
terse — open with one line ("running reporting-guideline-complianceagainst <checklist> <version>"); emit the markdown report; no closing summary.
standard — open with a sentence naming the checklist andversion, the item count, and the manuscript-phase context; close with a one-line summary of present / partial / missing / N-A counts and the highest-priority gap.
full — open with what reporting guidelines do (minimum-information standards so reviewers and readers can evaluate methodology consistently — the EQUATOR Network maintains the registry of ~600 active guidelines), what this audit produces (per-item present / partial / missing / N-A with a quoted anchor or explicit gap), and how to read it (act on missing first, then partial; N/A is not a gap; the audit does not invent prose to fill gaps). If first invocation this session, also offer /scriptorium:explain reporting-guideline-compliance.
Run the signal-based check-in once if appropriate. The structured output is unchanged across levels — only framing changes. The no-invented-prose posture is never relaxed based on guidance level.
Work in this order. Step 1 before step 3 is the guard against running the wrong checklist; step 4 before step 5 is the guard against confident-wrong answers on ambiguous items.
document_phase.current — if outline, decline the run.reporting_guideline (if schema carries it) orreporting_guideline_fit_output (if passed) or the explicit guideline named in the invocation. If none, refuse and point at /scriptorium:reporting-guideline-fit.
project.target_venue — for venue-specific requirements.core_claims, known_weaknesses — for context.meta.guidance_level — for framing only.TRIPOD+AI 2024 for AI-based prediction models (not TRIPOD 2015). Name the version explicitly in the output. Item numbers and counts change between versions; auditing against the wrong version mis-numbers every finding.
discussion, and any flow diagram or supplement. Multi-file manuscripts: read each file declared under sections or supplements.
present with a quotedexcerpt and the location (section:line if available, or section name with a quoted span).
partialwith the quoted excerpt and a one-line "what would tip this to present" note.
classify missing with an explicit gap statement ("no allocation-concealment mechanism described"). Do not propose phrasing.
not-applicable with a one-sentence justification ("continuous primary outcome — item 17b on binary outcome presentation does not apply").
author has already named in known_weaknesses is partial (the gap is acknowledged but not addressed in the prose), not missing. Note the acknowledgement explicitly.
project.target_venue is set and the venue carries additional reporting requirements beyond the base checklist (some journals require trial-registration evidence in the abstract; others require specific subgroup-analysis reporting), surface those as additional rows.
headings below verbatim so downstream skills and future orchestrators can consume the output by structure.
Emit a markdown document with exactly these section headings, in order:
# Reporting-guideline compliance
## Summary
- Checklist audited: <NAME VERSION> (e.g., CONSORT 2010, PRISMA
2020, ARRIVE 2.0, TRIPOD+AI 2024)
- Item count: N
- Present: N
- Partial: N
- Missing: N
- Not-applicable: N
- Highest-priority gaps: <one-line list of the missing items the
author should address first — items that journals routinely
desk-reject for, or that the upstream `reporting-guideline-fit`
flagged as high-confidence required>
## Checklist audit
(One row per checklist item. The item numbering matches the
named version. Quoted excerpts use the manuscript's own prose.
Locations are section names — or section:line if available.)
| Item | Status | Anchor or gap | Notes |
|---|---|---|---|
| <n>. <item title> | present | <section:line> — "<quoted excerpt>" | <one-line context if useful> |
| <n>. <item title> | partial | <section:line> — "<quoted excerpt>" | What would tip to present: <one-line note> |
| <n>. <item title> | missing | (no anchor) | Gap: <explicit, no proposed prose> |
| <n>. <item title> | not-applicable | (n/a) | Justification: <one sentence> |
## Highest-priority gaps
(Subset of the `missing` rows. Ordered by what reviewers most
commonly flag and what venues most commonly desk-reject for. Do
**not** propose prose; name the gap.)
1. **Item <n>. <title>.** <One-paragraph statement of what is
missing and why this item is high-priority. No proposed
replacement text.>
2. …
## Acknowledged-but-unaddressed items
(Items the author has named in `known_weaknesses` but that the
manuscript prose does not yet address. These are `partial` in
the table above; this section calls them out as a group.)
- <Item n. title> — acknowledged in `known_weaknesses` as
"<quoted weakness>". The manuscript prose does not yet
address this in <section>.
## Venue-specific requirements
(Only when `project.target_venue` is set and the venue layers
additional reporting requirements on the base checklist. Empty
otherwise.)
- <Venue>: <additional requirement>. Status: present / partial
/ missing.
## What this audit did NOT check
(Honest list. Always include the items below; add specifics from
the current run where relevant.)
- Whether the chosen checklist was the right one. That is the
upstream `reporting-guideline-fit` skill's job; this audit
trusts the inference.
- Whether the underlying study design was the right choice. The
audit assesses what is reported, not whether the design was
appropriate.
- Whether quantitative claims are internally consistent (Table 1
N vs. methods N; abstract percentages vs. figure percentages).
That is the planned `statistics-consistency` skill's job.
- Whether figures match the prose. That is the planned
`figure-text-alignment` skill's job.
- The bibliography itself. Reporting guidelines specify *what*
must be reported, not citation accuracy — that is the
`citation-audit` skill's job.
- Editor-side enforcement. Author-side decision support.present and partial row carries aquoted excerpt from the manuscript and a location. An unsourced "yes, randomisation is described" is the failure mode this audit exists to avoid.
satisfied are partial with a one-line "what would tip this to present" — never forced into binary present/missing.
a one-sentence rationale is a first-class outcome, not padding. The author should see why the audit skipped each N/A item.
"No allocation-concealment mechanism described" is the right level. "Consider adding: 'Allocation was concealed using sequentially numbered opaque envelopes…'" is not.
2020, TRIPOD+AI 2024, etc., explicitly — item numbers and counts depend on the version.
Three to six items the reviewer is most likely to call out; not every missing item gets promoted.
known_weaknesses** so the author sees which gaps they have already named vs. which are surprises.
author owns the fix.
reporting-guideline-fit's job; this skill audits the checklist it is given.
present or missing when the honestanswer is partial.
not-applicable without a one-sentencejustification. A bare N/A erodes trust.
item; the value is in coverage.
version applies (PRISMA 2009 when 2020 is current; TRIPOD 2015 when TRIPOD+AI 2024 is current for AI-based models).
This skill is grounded in scriptorium's knowledge layer:
Network registry; design-specific checklists (CONSORT 2010, STROBE, PRISMA 2020, ARRIVE 2.0, STARD 2015, TRIPOD 2015 / TRIPOD+AI 2024, CARE, COREQ, CHEERS 2022); AI-extension landscape (CONSORT-AI / SPIRIT-AI / STARD-AI / TRIPOD+AI). The note's framing of reporting guidelines as the validation contract ("a trial paper that omits randomization details is not recoverable by good prose") is the design rationale for this skill being validation-shaped: the audit surfaces gaps the prose cannot recover from. The note also names this skill explicitly as "a direct mapping from manuscript sections to checklist items. Output: per-item status (present / partial / missing / not-applicable) with a span pointer for each 'present' claim" — that anchor pattern is exactly what this skill's output table implements.
operates on declared manuscript prose; refuses cleanly on outline or when the applicable checklist is unknown; never invents prose to fill missing items.
per-item walk treats each checklist item as a cross-section comparison: does what the prose says satisfy the item's minimum-information requirement? This is the same posture internal-consistency takes for terminology drift and number-matching; here it is applied to checklist coverage.
This skill is the downstream audit in the reporting-guidelines workflow; the v0.2 reporting-guideline-fit skill is the upstream inference. The two are deliberately separate.
/scriptorium:reporting-guideline-fit (v0.2) — the upstreamskill that infers which EQUATOR checklist applies. Run it first when the applicable checklist isn't known.
/scriptorium:reviewer-simulation — pairs naturally beforesubmission. Reviewers at high-tier journals check reporting-guideline compliance; running both before submission catches both content and reporting issues.
/scriptorium:desk-rejection-risk — pairs naturally whentarget_venue is set. Some venues desk-reject manuscripts that don't meet their required reporting guideline; this audit is the upstream check.
/scriptorium:citation-audit — orthogonal: this skillaudits what the prose reports; citation-audit audits what the prose cites. Both run cleanly side by side.
/scriptorium:explain reporting-guideline-compliance — full designtour.
~30 seconds. Free. No account. Every finding cites a rule and a line of evidence.