# dev.safeprompt/mcp (npm · @safeprompt.dev/mcp)

Detect prompt injection, jailbreaks, and code injection in untrusted text before it reaches an LLM.

- Trust score: 70/100 (medium)
- Change this week: +24
- Registry status: active
- Liveness: live
- Owner verified: no
- Last scored: 2026-08-03

## Components

- npm · `@safeprompt.dev/mcp`: 70/100 (this document), [markdown](https://verifymcp.io/servers/dev-safeprompt-mcp/safeprompt-dev-mcp.md), [page](https://verifymcp.io/servers/dev-safeprompt-mcp/safeprompt-dev-mcp)

## Channel facts

- Registry: `npm`
- Package: `@safeprompt.dev/mcp`
- Version: `0.1.0`
- Transport: `stdio`

## Trust breakdown

How this component scores in each security and reliability category. Every signal is checked automatically from public evidence about the published package, including repeated runs of it in an isolated sandbox, and we only credit what we can confirm. Scores are 0–100 per category. Scoring method: https://verifymcp.io/docs/scoring (what has changed: https://verifymcp.io/docs/scoring/changelog)

Scored 2026-08-03.

- **Supply Chain Security**: 87/100
  - No malware found by supply-chain analysis.
  - Only part of the dependency tree could be resolved (95 of 99), so this covers what we could see, not the whole tree.
  - No install/post-install scripts declared.
  - Only part of the dependency tree could be resolved (95 of 99), so this covers what we could see, not the whole tree.
- **Provenance & Transparency**: 45/100
  - Source repository is publicly reachable at the declared URL.
  - Provenance check failed: no build-provenance attestation is published.
  - Clear OSI-approved license (MIT).
  - Actively maintained (last published 42 days ago).
  - Disclosure check failed: no security disclosure policy was found in the source repository.
- **Schema Quality & AI Usability**: 77/100
  - AI-judged instruction clarity (excellent).
  - Tool/resource definitions use about 276 tokens (~138/item across 2 items; 2 tools + 0 resources), lean.
  - Usage-examples check failed: none of the tools include examples.
- **Stability & Change Management**: 27/100
  - Stability observed for 8 of 30 days with no destabilising changes; credit accrues until the full window elapses.
- **Tool Coverage**: 100/100
  - 100% of tools have a non-trivial description (not blank, and not just the tool's name).
  - 100% of tool parameters carry a description.
- **Capabilities**: 100/100
  - Implements a supported MCP spec version (2025-11-25); the latest is 2026-07-28.

## Install

### Claude

```bash
claude mcp add dev-safeprompt-mcp -- npx -y @safeprompt.dev/mcp
```

### Codex

```bash
codex mcp add dev-safeprompt-mcp -- npx -y @safeprompt.dev/mcp
```

### opencode

```json
{
  "$schema": "https://opencode.ai/config.json",
  "mcp": {
    "dev-safeprompt-mcp": {
      "type": "local",
      "command": [
        "npx",
        "-y",
        "@safeprompt.dev/mcp"
      ],
      "enabled": true
    }
  }
}
```

### OpenClaw

```bash
openclaw mcp add dev-safeprompt-mcp --command npx --arg -y --arg @safeprompt.dev/mcp
```

### Hermes

```yaml
mcp_servers:
  dev-safeprompt-mcp:
    command: "npx"
    args: ["-y", "@safeprompt.dev/mcp"]
```

### Other

```json
{
  "mcpServers": {
    "dev-safeprompt-mcp": {
      "command": "npx",
      "args": [
        "-y",
        "@safeprompt.dev/mcp"
      ]
    }
  }
}
```

## Changelog

Every change recorded for this component, newest first. Days that predate change tracking, or that we cannot explain, say so: "we were watching and nothing happened" and "we were not watching" are different claims.

### 2026-08-03 (score 70, +1)

No change was recorded against any check on this day. Stability & Change Management went from 23 to 27. That category is still filling its 30-day observation window: 7 days of observed history at the previous scan, 8 at this one. The score rises as the window fills, whether or not the server changes.

### 2026-08-02 (score 69, +64)

- [security regression] Provenance: unverified → fail
- [security improvement] Install scripts: unverified → pass
- [security improvement] Known CVEs: unverified → partial
- [security improvement] Malware scan: unverified → pass
- [functional improvement] Tool coverage: unverified → 100
- [functional improvement] MCP protocol: unverified → pass
- [functional improvement] Stability: unverified → 0.23
- [functional improvement] Schema quality: unverified → excellent
- [functional improvement] License: unverified → pass
- [functional improvement] Dependency health: unverified → partial
- [functional improvement] Maintenance: unverified → pass
- [functional] Licence: MIT

### 2026-07-31 (score 5, −41)

- [functional] We updated how we score, so this day's move reflects our rubric, not a change to the server

### 2026-07-27 (score 46)

First indexed and scored.

## MCP tools (2)

### `validate_prompt` (~154 tokens)

Validate a prompt for injection attacks

Check a single piece of untrusted text (a user message, a retrieved RAG document, tool output, etc.) for prompt injection, jailbreaks, instruction-override, data-exfiltration, and code injection BEFORE passing it to an LLM. Returns a safe/unsafe verdict with confidence, threat category, and reasoning. Call this on any untrusted input you are about to feed to a model.

Input parameters:

- `prompt` (string, required): The untrusted text to validate.
- `sensitivity` (string): Block threshold. 'lenient' (0.95) blocks only high-confidence attacks, 'balanced' (default) is the recommended setting, 'strict' (0.75) blocks aggressively. Omit for balanced.

### `validate_prompts` (~122 tokens)

Batch-validate multiple prompts

Validate many pieces of untrusted text in one call (processed in parallel server-side). Use for bulk pre-screening, indexing pipelines, or checking a batch of retrieved documents. Returns one verdict per input, in order.

Input parameters:

- `prompts` (array, required): Array of untrusted texts to validate (max 100).
- `sensitivity` (string): Block threshold. 'lenient' (0.95) blocks only high-confidence attacks, 'balanced' (default) is the recommended setting, 'strict' (0.75) blocks aggressively. Omit for balanced.

## Diagnostics

Captured diagnostic sections: Provenance, Dependencies. The full working is on the page: https://verifymcp.io/servers/dev-safeprompt-mcp/safeprompt-dev-mcp#diagnostics

## Score history

- 2026-08-03: 70
- 2026-08-02: 69
- 2026-08-01: 5
- 2026-07-31: 5
- 2026-07-30: 46
- 2026-07-28: 46
- 2026-07-27: 46

## Links

- npm package: https://www.npmjs.com/package/@safeprompt.dev/mcp
- Socket report: https://socket.dev/npm/package/@safeprompt.dev/mcp
- Repository: https://github.com/ianreboot/safeprompt
- Website: https://safeprompt.dev/
- Changelog RSS feed: https://verifymcp.io/servers/dev-safeprompt-mcp/safeprompt-dev-mcp/changelog.xml
- Changelog JSON feed: https://verifymcp.io/servers/dev-safeprompt-mcp/safeprompt-dev-mcp/changelog.json
- HTML version of this page: https://verifymcp.io/servers/dev-safeprompt-mcp/safeprompt-dev-mcp
