github Python re-analysis due

babywyrm/stoneburner

github

Local-first LLM eval: cost, quality, and security suites. Install stoneburner-atomics; CLI is atomics.

maintainer
babywyrm
licence
MIT
first seen
2026-08-16
last seen
2026-09-19
releases · 30d
9
short id
risk 32/100 · heuristic grade
B low inferred analyzer v33 (current)
  • capability exposure inferred + 35
  • trust mitigators mixed − 3

inferred mixed

The A–E grade is our heuristic synthesis — a "review this" prompt, not a verdict. Each factor is tagged by what backs it: attested (a verifiable record), reported (a third party's claim), or inferred (our own heuristic, e.g. permissions). See methodology.

grade last moved 1w ago · see ecosystem CVEs →

capability exposure grade factor +35
Inferred surface — each links to servers holding it:
vulnerabilities 0 CVEs

No known CVEs for this server.

tool safety all quiet

No tool-safety findings — heuristic detectors run on the compute-risk cadence; a finding appears when a tool trips a rule.

skills & danger signals github-tarball
prompt-surface shipped agent-instruction files + hidden-content / dangerous-code findings — quoted from the analyzed source

analyzed commit b5fc7d8 · analyzer v33 · 17h ago

skills & prompt files 4

danger signals4

embed badge readme-ready
live risk-grade badge preview [![MCP Observatory risk grade](https://mcpobservatory.com/servers/github:babywyrm/stoneburner/badge.svg)](https://mcpobservatory.com/servers/github:babywyrm/stoneburner/security)

Heuristic, inferred signals — false positives (legitimately powerful tools, forks, language ports) are expected. Treat each as "review this", not a verdict. See the ecosystem-wide picture on the security hub, or the fleet security of babywyrm.