A from-scratch Model Context Protocol server (stdio, stdlib-only, zero deps): safe calc (OWASP LLM06), BM25 search, a no-LLM-judge answer grader, and live model-drift / eval-run lookups. 39 tests.
- capability exposure inferred + 14
inferred
The A–E grade is our heuristic synthesis — a "review this" prompt, not a verdict. Each factor is tagged by what backs it: attested (a verifiable record), reported (a third party's claim), or inferred (our own heuristic, e.g. permissions). See methodology.
graded 4m ago · see ecosystem CVEs →
No known CVEs for this server.
No tool-safety findings — heuristic detectors run on the compute-risk cadence; a finding appears when a tool trips a rule.
Heuristic, inferred signals — false positives (legitimately powerful tools, forks, language ports) are expected. Treat each as "review this", not a verdict. See the ecosystem-wide picture on the security hub, or the fleet security of egnaro9.