Nitpick MCP Server

@buyhatke-dev/nitpick-mcp npm v0.2.0

Published by @buyhatke-dev — no publish provenance, so origin is unverified, but the source is public: the repository link below is self-declared yet readable, so you can inspect the code before adopting it.

Nitpick Testing MCP — agent-agnostic exploratory UI reviewer (host-is-brain, driver abstraction, persistent .qa/ memory)

Trust grade
A
93/100
Last scanned get badge →
Trust
A · 93/100
Adoption risk for you: the threat score, then adjusted down for blast radius, publisher verification and how much the scan could see. Deterministic; every point is auditable.
Capability
High
Blast radius if it went rogue — what the server’s tools could reach. Independent of trust.
Coverage
Source
How much the scan could actually inspect. Shallow coverage is stated, never hidden.
Share this Trust Score
𝕏 Share LinkedIn Reddit
A Why this grade threat 100 − adoption risk = 93/100

The grade answers one question — how safe is this server for you to adopt — so it is computed in two auditable stages. Nothing below is an opinion or an LLM's guess; every line is a real term the deterministic engine applied, and the same input always yields the same number.

1. Threat score — 100 − 0 = 100. What the published surface and source actually contain:

The deterministic scan raised no scored threat in the surface it inspected — the threat score stayed at 100. Capability observations and advisory notes are recorded but never lower it.

2. Client adoption risk — 100 − 7 = 93. Three small, subtract-only factors that reflect your risk in adopting it — a clean scan proves less on a powerful, unverified or barely-inspectable package, so the grade says so plainly:

PointsAdoption-risk factor
−6 capability blast radius (high) — client exposure if the model is manipulated
−1 publisher verification (public source) — no provenance, but the source is public and inspectable

Capability observations and info notes are shown under Findings but never scored. Open any row's finding below for the file, line and evidence behind a deduction.

Findings 9

high Sensitive-source and external-sink co-existMTC-FLOW-004

Tools that read sensitive data ([export_data]) and tools that can send data out ([evaluate, upload_file]) are exposed together. An agent can move private data to the sink.

Evidence: sources [export_data] → sinks [evaluate, upload_file]

Fix: Keep secret-reading and egress capabilities on separate, separately-approved servers.

Location: flow export_data → evaluate

high Tool "evaluate" exposes command/code executionMTC-CAP-001

Tool "evaluate" appears to run shell commands or evaluate code (parameter "script"). Arbitrary execution driven by model input is one of the most dangerous MCP capabilities; combined with any untrusted input it becomes RCE.

Fix: Sandbox execution, allowlist commands/arguments, and never pass model output to a shell unescaped.

Location: tool evaluate

high Shell/command execution in server code (dist/report.js)MTC-SRC-002

In the server's implementation (`dist/report.js:2`): Spawning a shell/process is command-execution capability; with unsanitized tool input it is command injection / RCE. This is read from the code itself — not from the tool description — so a poisoned server cannot hide it behind honest-looking metadata.

Evidence: ecFileSync } from "node:child_process"; const SEV_COLOR = { high: "#b91c1c", medium: "#c2410c", low: "#15803d" }; const

Fix: Review this call path: confirm it never receives unsanitized tool input, constrain it, or remove it. Treat a server whose code reaches these sinks as high-capability regardless of what its tools claim.

Location: server dist/report.js

high Shell/command execution in server code (dist/util/sh.js)MTC-SRC-002

In the server's implementation (`dist/util/sh.js:1`): Spawning a shell/process is command-execution capability; with unsanitized tool input it is command injection / RCE. This is read from the code itself — not from the tool description — so a poisoned server cannot hide it behind honest-looking metadata.

Evidence: { execFile } from "node:child_process"; const MAX_BUFFER = 64 * 1024 * 1024; // 64MB — screenshots/hierarchy dumps are l

Fix: Review this call path: confirm it never receives unsanitized tool input, constrain it, or remove it. Treat a server whose code reaches these sinks as high-capability regardless of what its tools claim.

Location: server dist/util/sh.js

medium Unconstrained command parameter "script" on "evaluate"MTC-CAP-006

Tool "evaluate" takes a command-shaped parameter "script" with no enum/pattern constraint. Free-form, model- or attacker-controlled arguments reaching a shell is the command-injection precondition.

Fix: Constrain the parameter (enum/pattern), or build the command from a fixed template with escaped args.

Location: tool evaluate · inputSchema.properties.script

medium Dynamic module load from a non-literal (dist/doctor.js)MTC-SRC-005

In the server's implementation (`dist/doctor.js:28`): Loading a module chosen at runtime (from a variable) can pull in and run attacker-influenced code paths. This is read from the code itself — not from the tool description — so a poisoned server cannot hide it behind honest-looking metadata.

Evidence: st { chromium } = await import(spec); const p = chromium.executablePath(); return !!p && existsSync(p);

Fix: Review this call path: confirm it never receives unsanitized tool input, constrain it, or remove it. Treat a server whose code reaches these sinks as high-capability regardless of what its tools claim.

Location: server dist/doctor.js

medium Dynamic module load from a non-literal (dist/drivers/playwright.js)MTC-SRC-005

In the server's implementation (`dist/drivers/playwright.js:42`): Loading a module chosen at runtime (from a variable) can pull in and run attacker-influenced code paths. This is read from the code itself — not from the tool description — so a poisoned server cannot hide it behind honest-looking metadata.

Evidence: ({ chromium } = await import(spec)); } catch { throw new Error("Playwright is not installe

Fix: Review this call path: confirm it never receives unsanitized tool input, constrain it, or remove it. Treat a server whose code reaches these sinks as high-capability regardless of what its tools claim.

Location: server dist/drivers/playwright.js

low Mutating tool "evaluate" declares no destructiveHintMTC-CAP-005

Tool "evaluate" can mutate/egress but declares no destructiveHint. Clients that don't default to spec-safe behavior may not prompt before running it.

Fix: Declare accurate annotations, and gate destructive tools on user confirmation regardless.

Location: tool evaluate

low Unconstrained path parameter "filename" on "export_data"MTC-CAP-008

Tool "export_data" takes a path parameter "filename" with no constraint. Without a canonicalize-and-contain check (not visible statically), this permits ../ traversal outside the intended root.

Fix: Resolve and verify the path stays within an allowed root; reject traversal sequences.

Location: tool export_data · inputSchema.properties.filename

Tools 39

Each tool and what it can reach — statically extracted from the published source.

  • evaluateruns code / shell
  • export_datareads sensitive data
  • upload_filenetwork egress
  • audit_uino sensitive capability
  • captureno sensitive capability
  • close_tabno sensitive capability
  • compare_ui_stateno sensitive capability
  • current_pageno sensitive capability
  • doctorno sensitive capability
  • dragno sensitive capability
Show 29 more tools ↓
  • export_reportno sensitive capability
  • extract_uino sensitive capability
  • focusno sensitive capability
  • hoverno sensitive capability
  • inputno sensitive capability
  • inspect_uino sensitive capability
  • list_findingsno sensitive capability
  • list_tabsno sensitive capability
  • loginno sensitive capability
  • long_pressno sensitive capability
  • navigateno sensitive capability
  • observeno sensitive capability
  • open_tabno sensitive capability
  • pressno sensitive capability
  • record_findingno sensitive capability
  • run_flowno sensitive capability
  • save_authno sensitive capability
  • save_flowno sensitive capability
  • save_ui_stateno sensitive capability
  • scrollno sensitive capability
  • scroll_tono sensitive capability
  • select_optionno sensitive capability
  • set_checkedno sensitive capability
  • start_sessionno sensitive capability
  • switch_tabno sensitive capability
  • tapno sensitive capability
  • update_briefno sensitive capability
  • update_mapno sensitive capability
  • waitno sensitive capability

Toxic flows 1

Cross-tool combinations that form a data-exfiltration primitive (untrusted input → sensitive source → external sink).

What this scan could not see

Versions 1

Scan history per published version. The engine is deterministic — the same version always yields the same score, so a changed score means the package itself changed.

VersionScoreFindingsEngineScanned
v0.2.0 latest A 93/100 9 1.13.0 2026-08-25

Embed this score

Show this server's live Trust Score in your README, docs or website. The badge is served straight from the registry and updates automatically after every rescan — no API key needed. It links back to this page, so anyone who sees the grade can also read the findings behind it instead of taking a number on faith.

MCP Trust Score: A · 93/100
Markdown (GitHub README)
[![MCP Trust Score](https://mcptrustchecker.com/registry/buyhatke-dev-nitpick-mcp/badge.svg)](https://mcptrustchecker.com/registry/buyhatke-dev-nitpick-mcp)
HTML
<a href="https://mcptrustchecker.com/registry/buyhatke-dev-nitpick-mcp"><img src="https://mcptrustchecker.com/registry/buyhatke-dev-nitpick-mcp/badge.svg" alt="MCP Trust Score" height="20"></a>
Prefer shields.io styling? Point it at https://mcptrustchecker.com/registry/buyhatke-dev-nitpick-mcp/badge.json via https://img.shields.io/endpoint?url=…

Verify this score yourself

The score above is reproducible: the same package version always yields the same result. Run it locally or over the free API — no account, no LLM, fully deterministic.

npx mcptrustchecker scan @buyhatke-dev/nitpick-mcp --online

Use the free API → How scoring works

More in AI & Agents