# Agentic RL: Credit Assignment and CLI Agents (remote · hoyant-su-agentic-rl.hf.space)

Filter agent RL methods by supervision, critic and task setting; retrieve source links and BibTeX.

- Trust score: 70/100 (medium)
- Change this week: +3
- Registry status: active
- Liveness: live
- Owner verified: no
- Last scored: 2026-09-28

## Components

- remote · `hoyant-su-agentic-rl.hf.space`: 70/100 (this document), [markdown](https://verifymcp.io/servers/space-hf-hoyant-su-agentic-rl-agentic-rl/gradio-api-mcp.md), [page](https://verifymcp.io/servers/space-hf-hoyant-su-agentic-rl-agentic-rl/gradio-api-mcp)

## Channel facts

- Endpoint: `https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/`
- Transports: `streamable-http`
- Auth: `none`
- Version: `1.1.0`

## Trust breakdown

How this component scores in each security and reliability category. Every signal is checked automatically against the live server, and we only credit what we can confirm. Scores are 0–100 per category. Scoring method: https://verifymcp.io/docs/scoring (what has changed: https://verifymcp.io/docs/scoring/changelog)

Scored 2026-09-28.

- **Endpoint Security**: 57/100
  - The endpoint's TLS certificate is valid, in date, and uses a strong key.
  - Authorisation not fully verified: no authorisation is required to call this server, and 8 tool(s) never declared a destructiveHint. The MCP spec treats an absent hint as destructive by default, so we cannot call this surface safe.
  - HTTPS is enforced; there's no plaintext access path.
  - HSTS check failed: the Strict-Transport-Security header is absent.
  - DNSSEC check failed: this domain isn't protected by DNSSEC.
- **Transport & Reachability**: 100/100
  - Verified streamable-http transport via a live MCP handshake.
- **Schema Quality & AI Usability**: 70/100
  - AI-judged instruction clarity (good).
  - Tool/resource definitions use about 589 tokens (~73/item across 8 items; 8 tools + 0 resources), lean.
  - Usage-examples check failed: none of the tools include examples.
- **Stability & Change Management**: 67/100
  - Stability observed for 20 of 30 days with no destabilising changes; credit accrues until the full window elapses.
- **Tool Coverage**: 67/100
  - 100% of tools have a non-trivial description (not blank, and not just the tool's name).
  - 0% of tool parameters carry a description.
- **Tool Safety**: 100/100
  - No prompt-injection markers were found in the server instructions, tool names or descriptions we captured.
  - We read all 8 captured tool definition(s), and no name or description among them implies an irreversible operation.
  - An AI judge read all 8 captured unit(s) of tool text and found none that tries to manipulate the model reading it.
- **Capabilities**: 100/100
  - Implements a supported MCP spec version (2025-11-25); the latest is 2026-07-28.

## Install

### How do I install the Agentic RL: Credit Assignment and CLI Agents MCP server?

Agentic RL: Credit Assignment and CLI Agents is a hosted endpoint at https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/, so there is nothing to install locally. Ready-made configuration for Claude, Cursor, VS Code, Codex and 5 more is on this page, copied from each client's own documentation.

### Claude

```bash
claude mcp add --transport http space-hf-hoyant-su-agentic-rl-agentic-rl 'https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/'
```

### Cursor

```json
{
  "mcpServers": {
    "space-hf-hoyant-su-agentic-rl-agentic-rl": {
      "url": "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/"
    }
  }
}
```

### VS Code

```json
{
  "servers": {
    "space-hf-hoyant-su-agentic-rl-agentic-rl": {
      "type": "http",
      "url": "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/"
    }
  }
}
```

### Codex

```toml
[mcp_servers.space-hf-hoyant-su-agentic-rl-agentic-rl]
url = "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/"
```

### opencode

```json
{
  "$schema": "https://opencode.ai/config.json",
  "mcp": {
    "space-hf-hoyant-su-agentic-rl-agentic-rl": {
      "type": "remote",
      "url": "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/",
      "enabled": true
    }
  }
}
```

### OpenClaw

```bash
openclaw mcp add space-hf-hoyant-su-agentic-rl-agentic-rl --url 'https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/' --transport streamable-http
```

### Hermes

```yaml
mcp_servers:
  space-hf-hoyant-su-agentic-rl-agentic-rl:
    url: "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/"
```

### Netclaw

```json
{
  "McpServers": {
    "space-hf-hoyant-su-agentic-rl-agentic-rl": {
      "Transport": "http",
      "Url": "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/"
    }
  }
}
```

### Vellum

```bash
assistant mcp add space-hf-hoyant-su-agentic-rl-agentic-rl -t streamable-http -u 'https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/'
```

### Other

```json
{
  "mcpServers": {
    "space-hf-hoyant-su-agentic-rl-agentic-rl": {
      "type": "http",
      "url": "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/"
    }
  }
}
```

The mcpServers block is a cross-client convention. Remote transports vary, so check your client's docs.

## Changelog

Every change recorded for this component, newest first. Days that predate change tracking, or that we cannot explain, say so: "we were watching and nothing happened" and "we were not watching" are different claims.

### 2026-09-28 (score 70, 0)

- [functional] We updated how we score, so this day's move reflects our rubric, not a change to the server

### 2026-09-27 (score 70, +1)

No change was recorded against any check on this day. Stability & Change Management went from 60 to 63. That category is still filling its 30-day observation window: 18 days of observed history at the previous scan, 19 at this one. The score rises as the window fills, whether or not the server changes.

### 2026-09-25 (score 69, +1)

- [functional] We updated how we score, so this day's move reflects our rubric, not a change to the server

### 2026-09-23 (score 68, +1)

No change was recorded against any check on this day. Stability & Change Management went from 47 to 50. That category is still filling its 30-day observation window: 14 days of observed history at the previous scan, 15 at this one. The score rises as the window fills, whether or not the server changes.

### 2026-09-21 (score 67, +1)

No change was recorded against any check on this day. Stability & Change Management went from 40 to 43. That category is still filling its 30-day observation window: 12 days of observed history at the previous scan, 13 at this one. The score rises as the window fills, whether or not the server changes.

### 2026-09-19 (score 66, +1)

No change was recorded against any check on this day. Stability & Change Management went from 33 to 37. That category is still filling its 30-day observation window: 10 days of observed history at the previous scan, 11 at this one. The score rises as the window fills, whether or not the server changes.

### 2026-09-16 (score 65, +1)

No change was recorded against any check on this day. Stability & Change Management went from 23 to 27. That category is still filling its 30-day observation window: 7 days of observed history at the previous scan, 8 at this one. The score rises as the window fills, whether or not the server changes.

### 2026-09-14 (score 64, +1)

No change was recorded against any check on this day. Stability & Change Management went from 17 to 20. That category is still filling its 30-day observation window: 5 days of observed history at the previous scan, 6 at this one. The score rises as the window fills, whether or not the server changes.

## MCP tools (8)

### `Agentic_RL_list_sources` (~46 tokens)

List original papers and retrieval coverage. Discover source-linked comparisons of credit assignment, agent memory, selective observation and terminal benchmarks, with JSON, CSV and BibTeX links.

### `Agentic_RL_search_evidence` (~65 tokens)

Search original papers on agentic reinforcement learning, credit assignment and CLI agents. Use English keywords (AND), OR and quoted phrases. Return relevant passages, source citations, equations and table cells.

Input parameters:

- `limit` (integer)
- `query` (string, required)

### `Agentic_RL_fetch_evidence` (~53 tokens)

Fetch a complete original evidence block by the evidence_id returned from search_evidence, including section anchor, version, equations, table cells, links, and attribution.

Input parameters:

- `evidence_id` (string, required)

### `Agentic_RL_dataset_overview` (~40 tokens)

Inspect ShellOps and ShellOps-Pro task counts, train/test splits, task types, published schemas, source files, license and citation.

### `Agentic_RL_search_tasks` (~136 tokens)

Find real ShellOps CLI benchmark tasks by case-insensitive literal substring in the complete instruction, task ID or published task type. Empty query lists all tasks. Select partition 'all', 'shellops' or 'shellops_pro'; select published split 'all', 'train_src', 'train' or 'test'. Results are ordered by partition then task ID, with explicit pagination and no relevance scoring. The train subset is not double-counted.

Input parameters:

- `limit` (integer)
- `offset` (integer)
- `partition` (string)
- `query` (string, required)
- `split` (string)

### `Agentic_RL_get_task` (~99 tokens)

Inspect one published ShellOps or ShellOps-Pro task by its exact task_id and partition ('shellops' or 'shellops_pro'). Returns the complete instruction, actual reward specification, published reference answer/command, file-entry metadata, pinned parquet rows and workspace asset links. File content is available at the source links. No shell execution or solution verification is performed.

Input parameters:

- `partition` (string, required)
- `task_id` (string, required)

### `Agentic_RL_list_method_facets` (~42 tokens)

List exact filter values for agent RL credit granularity, supervision, value critics and evaluation settings. Each value reports its source-supported method count.

### `Agentic_RL_filter_methods` (~108 tokens)

Filter agent RL credit-assignment methods by research conditions and return original section evidence and BibTeX. Discover accepted values with list_method_facets. Filters combine with AND; empty strings leave a facet unrestricted. Unknown critic status never matches no. Results use publication order without a relevance or quality ranking.

Input parameters:

- `credit_granularity` (string)
- `evaluation_setting` (string)
- `learned_value_critic` (string)
- `required_supervision` (string)

## Diagnostics

Captured diagnostic sections: TLS, DNSSEC, Authorisation, Transports. The full working is on the page: https://verifymcp.io/servers/space-hf-hoyant-su-agentic-rl-agentic-rl/gradio-api-mcp#diagnostics

## Score history

- 2026-09-28: 70
- 2026-09-27: 70
- 2026-09-26: 69
- 2026-09-25: 69
- 2026-09-24: 68
- 2026-09-23: 68
- 2026-09-22: 67
- 2026-09-21: 67
- 2026-09-20: 66
- 2026-09-19: 66
- 2026-09-18: 65
- 2026-09-17: 65
- 2026-09-16: 65
- 2026-09-15: 64
- 2026-09-14: 64
- 2026-09-13: 63
- 2026-09-12: 63
- 2026-09-11: 62
- 2026-09-10: 62
- 2026-09-09: 61
- 2026-09-08: 61

## Common questions

### What is the Agentic RL: Credit Assignment and CLI Agents MCP server?

Agentic RL: Credit Assignment and CLI Agents is an MCP server listed in the public MCP registry as space.hf.hoyant-su-agentic-rl/agentic-rl. Filter agent RL methods by supervision, critic and task setting; retrieve source links and BibTeX. This page covers its hosted endpoint (https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/).

### Is the Agentic RL: Credit Assignment and CLI Agents MCP server safe to use?

Agentic RL: Credit Assignment and CLI Agents scores 70 out of 100 on VerifyMCP. That is a record of what we were able to check automatically, not an endorsement. The category breakdown on this page shows every signal behind the number, including the ones we could not confirm.

### What tools does the Agentic RL: Credit Assignment and CLI Agents MCP server expose?

Agentic RL: Credit Assignment and CLI Agents exposes 8 tools: Agentic_RL_list_sources, Agentic_RL_search_evidence, Agentic_RL_fetch_evidence, Agentic_RL_dataset_overview, Agentic_RL_search_tasks, and 3 more. Their descriptions and schemas cost roughly 589 tokens of context every time the server is loaded.

### Does the Agentic RL: Credit Assignment and CLI Agents MCP server require authentication?

No. We connected to Agentic RL: Credit Assignment and CLI Agents without credentials and it answered, so anything it exposes is reachable by anyone who knows the address.

### Is the Agentic RL: Credit Assignment and CLI Agents MCP server still maintained?

Agentic RL: Credit Assignment and CLI Agents is still listed as active in the MCP registry. We last reached this channel on 28 September 2026. Those dates come from our own scans of the registry and the channel itself, not from anything the publisher announced.

## Links

- Remote endpoint: https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/
- Website: https://huggingface.co/spaces/Hoyant-Su/Agentic-RL
- Changelog RSS feed: https://verifymcp.io/servers/space-hf-hoyant-su-agentic-rl-agentic-rl/gradio-api-mcp.xml
- Changelog JSON feed: https://verifymcp.io/servers/space-hf-hoyant-su-agentic-rl-agentic-rl/gradio-api-mcp.json
- HTML version of this page: https://verifymcp.io/servers/space-hf-hoyant-su-agentic-rl-agentic-rl/gradio-api-mcp
