# InferIndex (remote · mcp.inferindex.dev)

LLM API prices across 70+ providers: cheapest offer, comparisons, history and cost estimates.

- Trust score: 68/100 (medium)
- Registry status: active
- Liveness: live
- Owner verified: no
- Last scored: 2026-09-20

## Components

- remote · `mcp.inferindex.dev`: 68/100 (this document), [markdown](https://verifymcp.io/servers/inferindex-inferindex/mcp.md), [page](https://verifymcp.io/servers/inferindex-inferindex/mcp)

## Channel facts

- Endpoint: `https://mcp.inferindex.dev/mcp`
- Transports: `streamable-http`
- Auth: `none`
- Version: `1.0.0`

## Trust breakdown

How this component scores in each security and reliability category. Every signal is checked automatically against the live server, and we only credit what we can confirm. Scores are 0–100 per category. Scoring method: https://verifymcp.io/docs/scoring (what has changed: https://verifymcp.io/docs/scoring/changelog)

Scored 2026-09-20.

- **Endpoint Security**: 69/100
  - The endpoint's TLS certificate is valid, in date, and uses a strong key.
  - No authorisation is required to call this server. Every tool declares its destructiveHint and none is destructive, so open access doesn't expose one.
  - HTTPS enforcement could not be verified: the plaintext port answered with HTTP 405, which proves neither a plaintext path nor enforcement.
  - The HSTS (Strict-Transport-Security) header is present.
  - DNSSEC check failed: this domain isn't protected by DNSSEC.
- **Transport & Reachability**: 100/100
  - Verified streamable-http transport via a live MCP handshake.
- **Schema Quality & AI Usability**: 66/100
  - AI-judged instruction clarity (excellent).
  - Context-footprint check failed: tool/resource definitions use about 1351 tokens (~270/item across 5 items; 5 tools + 0 resources), over budget; trim descriptions and params.
  - Usage-examples check failed: none of the tools include examples.
- **Stability & Change Management**: 7/100
  - Stability observed for 2 of 30 days with no destabilising changes; credit accrues until the full window elapses.
- **Tool Coverage**: 100/100
  - 100% of tools have a non-trivial description (not blank, and not just the tool's name).
  - 100% of tool parameters carry a description.
- **Tool Safety**: 100/100
  - No prompt-injection markers were found in the server instructions, tool names or descriptions we captured.
  - We read all 5 captured tool definition(s), and no name or description among them implies an irreversible operation.
  - An AI judge read all 6 captured unit(s) of tool text and found none that tries to manipulate the model reading it.
- **Capabilities**: 100/100
  - Implements a supported MCP spec version (2025-11-25); the latest is 2026-07-28.

## Install

### How do I install the InferIndex MCP server?

InferIndex is a hosted endpoint at https://mcp.inferindex.dev/mcp, so there is nothing to install locally. Ready-made configuration for Claude, Cursor, VS Code, Codex and 5 more is on this page, copied from each client's own documentation.

### Claude

```bash
claude mcp add --transport http inferindex-inferindex 'https://mcp.inferindex.dev/mcp'
```

### Cursor

```json
{
  "mcpServers": {
    "inferindex-inferindex": {
      "url": "https://mcp.inferindex.dev/mcp"
    }
  }
}
```

### VS Code

```json
{
  "servers": {
    "inferindex-inferindex": {
      "type": "http",
      "url": "https://mcp.inferindex.dev/mcp"
    }
  }
}
```

### Codex

```toml
[mcp_servers.inferindex-inferindex]
url = "https://mcp.inferindex.dev/mcp"
```

### opencode

```json
{
  "$schema": "https://opencode.ai/config.json",
  "mcp": {
    "inferindex-inferindex": {
      "type": "remote",
      "url": "https://mcp.inferindex.dev/mcp",
      "enabled": true
    }
  }
}
```

### OpenClaw

```bash
openclaw mcp add inferindex-inferindex --url 'https://mcp.inferindex.dev/mcp' --transport streamable-http
```

### Hermes

```yaml
mcp_servers:
  inferindex-inferindex:
    url: "https://mcp.inferindex.dev/mcp"
```

### Netclaw

```json
{
  "McpServers": {
    "inferindex-inferindex": {
      "Transport": "http",
      "Url": "https://mcp.inferindex.dev/mcp"
    }
  }
}
```

### Vellum

```bash
assistant mcp add inferindex-inferindex -t streamable-http -u 'https://mcp.inferindex.dev/mcp'
```

### Other

```json
{
  "mcpServers": {
    "inferindex-inferindex": {
      "type": "http",
      "url": "https://mcp.inferindex.dev/mcp"
    }
  }
}
```

The mcpServers block is a cross-client convention. Remote transports vary, so check your client's docs.

## Changelog

Every change recorded for this component, newest first. Days that predate change tracking, or that we cannot explain, say so: "we were watching and nothing happened" and "we were not watching" are different claims.

### 2026-09-19 (score 68, +1)

- [functional improvement] Stability: unverified → 0.03

### 2026-09-18 (score 67)

First indexed and scored.

## MCP tools (5)

### `search_models` (~69 tokens)

Search models

Find the exact id of an LLM tracked by InferIndex from a name or partial name (e.g. 'deepseek', 'qwen3 max', 'claude opus'). Returns matching model ids and names, best match first.

Input parameters:

- `query` (string, required): Model name or part of it

### `cheapest` (~419 tokens)

Cheapest offers for a model

Cheapest current API offers for one model across direct providers and aggregators, in USD per 1M tokens (input, output, blended 3:1). Stale prices, and flex/batch tiers, are excluded by default. Optional filters (context, tools, JSON, vision, region, no training on prompts, open sign-up) and usage (tokens per request, requests per day) to get an estimated cost per request and per month. Returns the winner and the first offers.

Input parameters:

- `cached_ratio` (number): Share of input tokens served from the provider's prompt cache (0 to 1)
- `include_tiers` (string): Also include lower-priority service tiers, comma-separated: flex, batch (hidden by default)
- `json` (boolean): Only offers that support JSON output
- `limit` (integer): Number of offers to return (default 5, max 25)
- `min_context` (integer): Minimum context window in tokens
- `model` (string, required): Model id or name, e.g. 'deepseek-v3.2', 'deepseek/deepseek-v4-pro', 'gpt-5.6-luna'. Use search_models when unsure.
- `no_training` (boolean): Only providers whose published terms say they do not train on your prompts
- `no_waitlist` (boolean): Only providers with open sign-up (no waitlist or invitation)
- `output_tokens` (integer): Output tokens per request, for the estimated cost
- `prompt_tokens` (integer): Input tokens per request, for the estimated cost
- `region` (string): Only providers that process data in this region: eu, us, …
- `requests_per_day` (integer): Requests per day, to also get an estimated monthly cost
- `strict` (boolean): Exclude offers whose provider does not publish the filtered information (by default they are kept and flagged)
- `tools` (boolean): Only offers that support tool calling
- `vision` (boolean): Only offers that accept image input

### `compare_providers` (~345 tokens)

Compare providers for a model

Current offers for one model, one line per provider and source (direct or via an aggregator), cheapest first (10 by default), with price, context, quantization, published conditions (training on prompts, data regions, sign-up) and reliability from official status pages.

Input parameters:

- `cached_ratio` (number): Share of input tokens served from the provider's prompt cache (0 to 1)
- `include_tiers` (string): Also include lower-priority service tiers, comma-separated: flex, batch (hidden by default)
- `limit` (integer): Number of offers to return (default 10, max 50)
- `model` (string, required): Model id or name, e.g. 'deepseek-v3.2', 'deepseek/deepseek-v4-pro', 'gpt-5.6-luna'. Use search_models when unsure.
- `no_training` (boolean): Only providers whose published terms say they do not train on your prompts
- `no_waitlist` (boolean): Only providers with open sign-up (no waitlist or invitation)
- `output_tokens` (integer): Output tokens per request, for the estimated cost
- `prompt_tokens` (integer): Input tokens per request, for the estimated cost
- `region` (string): Only providers that process data in this region: eu, us, …
- `requests_per_day` (integer): Requests per day, to also get an estimated monthly cost
- `sort` (string): Sort order (default blended); estimated_cost needs prompt_tokens or output_tokens
- `strict` (boolean): Exclude offers whose provider does not publish the filtered information (by default they are kept and flagged)

### `price_history` (~240 tokens)

Price history of a model

Price history of one model: every offer tracked by InferIndex (daily or weekly min/max/last price in USD, or raw price changes), plus the official price of the model's lab over time. Give either days, or from/to (YYYY-MM-DD), or at (a date) for the prices in effect that day.

Input parameters:

- `at` (string): A single date, YYYY-MM-DD: prices in effect that day
- `days` (integer): Number of days back from today (default 7)
- `from` (string): Start date, YYYY-MM-DD (with to, instead of days)
- `granularity` (string): day (default), week, or raw price changes
- `limit` (integer): Maximum number of points (default 100, max 500)
- `model` (string, required): Model id or name, e.g. 'deepseek-v3.2', 'deepseek/deepseek-v4-pro', 'gpt-5.6-luna'. Use search_models when unsure.
- `provider` (string): Only this provider
- `to` (string): End date, YYYY-MM-DD

### `estimate_cost` (~203 tokens)

Estimate the cost of a workload

Estimated cost of a workload on one model at each provider: cost per request, and per month if requests_per_day is given, taking the provider's tiered pricing and prompt-cache price into account. Offers sorted by estimated cost, cheapest first.

Input parameters:

- `cached_ratio` (number): Share of input tokens served from the provider's prompt cache (0 to 1)
- `limit` (integer): Number of offers to return (default 5, max 25)
- `model` (string, required): Model id or name, e.g. 'deepseek-v3.2', 'deepseek/deepseek-v4-pro', 'gpt-5.6-luna'. Use search_models when unsure.
- `output_tokens` (integer): Output tokens per request, for the estimated cost
- `prompt_tokens` (integer): Input tokens per request, for the estimated cost
- `requests_per_day` (integer): Requests per day, to also get an estimated monthly cost

## Diagnostics

Captured diagnostic sections: TLS, DNSSEC, Authorisation, Transports. The full working is on the page: https://verifymcp.io/servers/inferindex-inferindex/mcp#diagnostics

## Score history

- 2026-09-20: 68
- 2026-09-19: 68
- 2026-09-18: 67

## Common questions

### What is the InferIndex MCP server?

InferIndex is an MCP server listed in the public MCP registry as io.github.InferIndex/inferindex. LLM API prices across 70+ providers: cheapest offer, comparisons, history and cost estimates. This page covers its hosted endpoint (https://mcp.inferindex.dev/mcp).

### Is the InferIndex MCP server safe to use?

InferIndex scores 68 out of 100 on VerifyMCP. That is a record of what we were able to check automatically, not an endorsement. The category breakdown on this page shows every signal behind the number, including the ones we could not confirm.

### What tools does the InferIndex MCP server expose?

InferIndex exposes 5 tools: search_models, cheapest, compare_providers, price_history, estimate_cost. Their descriptions and schemas cost roughly 1,276 tokens of context every time the server is loaded.

### Does the InferIndex MCP server require authentication?

No. We connected to InferIndex without credentials and it answered, so anything it exposes is reachable by anyone who knows the address.

### Is the InferIndex MCP server still maintained?

InferIndex is still listed as active in the MCP registry. We last reached this channel on 20 September 2026. Those dates come from our own scans of the registry and the channel itself, not from anything the publisher announced.

## Links

- Remote endpoint: https://mcp.inferindex.dev/mcp
- Repository: https://github.com/InferIndex/inferindex-docs
- Website: https://api.inferindex.dev/
- Changelog RSS feed: https://verifymcp.io/servers/inferindex-inferindex/mcp.xml
- Changelog JSON feed: https://verifymcp.io/servers/inferindex-inferindex/mcp.json
- HTML version of this page: https://verifymcp.io/servers/inferindex-inferindex/mcp
