media-processor — independently scanned and version-tracked by SaferSkills.
SaferSkills independently audited media-processor (Agent Skill) and scored it 100/100 (green). The audit ran 55 deterministic rules across Security, Supply Chain, Maintenance, Transparency, and Community; it found 0 high-severity and 0 lower-severity findings. The full rule-by-rule trace and per-finding evidence are below. Free, methodology-open.
Findings & checks · 0 flagged
Every scanned point with the score it earned and what moved between them.
First recorded scan — no prior version to compare against.
The primary manifest — the file an agent reads to learn what this artifact does.
Specialized tools for extracting precise visual details (exact colors, spacing, hierarchy), processing audio/video, and generating images.
All scripts live in scripts/ relative to this skill's directory. They auto-select the best model per task and handle retries, large file uploads, and error reporting.
| Script | Purpose |
|---|---|
gemini_batch_process.py | Analyze images, transcribe audio/video, extract data from PDFs |
image_gen.py | Generate and edit images (paid plan required) |
document_converter.py | Convert PDF, DOCX, XLSX, PPTX to Markdown; extract page ranges and images |
Requires GEMINI_API_KEY in environment or .env in this skill's directory. Run any script with --help for setup details and available parameters.
Quick start — image analysis:
python <skill-dir>/scripts/gemini_batch_process.py \
--files <image-path> \
--task analyze \
--prompt "<tailored prompt>" \
--output <output-path>.mdThe prompt sent to the processing model is the single biggest factor in output quality. Tailor prompts to what the task actually needs — generic prompts produce generic results.
What makes a good analysis prompt:
Example prompt patterns:
UI implementation: "Extract component hierarchy, layout type, exact hex colors, typography (sizes/weights), spacing in px, interactive states, icons and decorative elements"
Chart data: "Extract chart type, axes with units, every data point with exact values, legend entries with colors. Output as a markdown table"
Design review: "Compare this screenshot against the design. Flag differences in spacing, colors, alignment, missing elements, and visual inconsistencies. Note exact values for each discrepancy"
When a user pastes images in chat, they are auto-saved to:
$CLAUDE_DIR/image-cache/<current_session_id>/<image_number>.pngUse ls "$CLAUDE_DIR/image-cache/" to discover the session ID, then list its contents to find available images.
Scripts auto-select models per task (see model-routing.md). Override with --model <model-id> when the default isn't enough — for example, --model gemini-3.1-pro-preview for complex visual analysis where the pro model catches more detail than flash.
| Reference | When to read |
|---|---|
| api-gotchas.md | Before using image generation, video processing, or raw API calls — prevents common failures |
| model-routing.md | When choosing or overriding the default model for a task |
| media-optimization.md | When files are too large to upload — ffmpeg compression recipes |
~30 seconds. Free. No account. Every finding cites a rule and a line of evidence.