The MCP ecosystem often has several independent packages implementing the same tool — and "Test" currently has 9. Every one of them was scanned with the same deterministic engine, so the numbers below are directly comparable: the Trust Score (A–F), the capability blast radius, and how much of the package the scan could actually inspect. They are ranked by score; ties break toward unscoped, longer-established packages.
| Implementation | Package | Trust | Capability | Coverage | Findings | Scanned |
|---|---|---|---|---|---|---|
| Testbest by paveldudka | mcp-server-test
npm |
A 100/100 | Minimal | Source | 0 | 2026-07-22 |
| Test by lorien | test-server
PyPI |
A 98/100 | Minimal | Source | 1 | 2026-07-22 |
| Test by PyPI | mcp-test-server
PyPI |
A 98/100 | Minimal | Source | 1 | 2026-07-22 |
| Test by npm | mcp-test-server
npm |
A 98/100 | Minimal | Source | 1 | 2026-07-22 |
| Test by npm | test-mcp-server
npm |
A 98/100 | Minimal | Source | 1 | 2026-07-22 |
| Test by PyPI | test-mcp-server
PyPI |
A 98/100 | Minimal | Source | 1 | 2026-07-22 |
| Test by npm | test-mcp
npm |
A 98/100 | Moderate | Source | 2 | 2026-07-22 |
| Test by npm | mcp-test
npm |
A 98/100 | High | Source | 4 | 2026-07-22 |
| Test by rdwj | mcp-test-mcp
npm |
A 94/100 | High | Source | 3 | 2026-07-22 |
A higher-ranked implementation is not "the official one" — ranking reflects only what the deterministic scan found in each published package. Open an implementation to read its individual findings with evidence, its scan history per version, and to grab an embeddable badge. Scores are automated opinions, recomputed on every rescan.