Agentic RL: Credit Assignment and CLI Agents
REMOTE · HOYANT-SU-AGENTIC-RL.HF.SPACE · SCANNED SEP 28
Filter agent RL methods by supervision, critic and task setting; retrieve source links and BibTeX.
Available components
How this component scores in each security and reliability category. Every signal is checked automatically against the live server, and we only credit what we can confirm. How we score → Why this is hard to score →
Endpoint Security57
- The endpoint's TLS certificate is valid, in date, and uses a strong key. View diagnostics → Pass
- Authorisation not fully verified: no authorisation is required to call this server, and 8 tool(s) never declared a destructiveHint. The MCP spec treats an absent hint as destructive by default, so we cannot call this surface safe. See how to fix → View diagnostics → Unverified
- HTTPS is enforced; there's no plaintext access path. View diagnostics → Pass
- HSTS check failed: the Strict-Transport-Security header is absent. See how to fix → View diagnostics → Fail
- DNSSEC check failed: this domain isn't protected by DNSSEC. See how to fix → View diagnostics → Fail
Transport & Reachability100
- Verified streamable-http transport via a live MCP handshake. View diagnostics → Pass
Schema Quality & AI Usability70
- AI-judged instruction clarity (good).Pass
- Tool/resource definitions use about 589 tokens (~73/item across 8 items; 8 tools + 0 resources), lean.Pass
- Usage-examples check failed: none of the tools include examples. See how to fix → Fail
Stability & Change Management67
- Stability observed for 20 of 30 days with no destabilising changes; credit accrues until the full window elapses.Partial
Tool Coverage67
- 100% of tools have a non-trivial description (not blank, and not just the tool's name).Pass
- 0% of tool parameters carry a description.Fail
Tool Safety100
- No prompt-injection markers were found in the server instructions, tool names or descriptions we captured.Pass
- We read all 8 captured tool definition(s), and no name or description among them implies an irreversible operation.Pass
- An AI judge read all 8 captured unit(s) of tool text and found none that tries to manipulate the model reading it.Pass
Capabilities100
- Implements a supported MCP spec version (2025-11-25); the latest is 2026-07-28.Pass
How do I install the Agentic RL: Credit Assignment and CLI Agents MCP server?
Agentic RL: Credit Assignment and CLI Agents is a hosted endpoint at https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/, so there is nothing to install locally. Ready-made configuration for Claude, Cursor, VS Code, Codex and 5 more is on this page, copied from each client's own documentation.
remote · hoyant-su-agentic-rl.hf.space
claude mcp add --transport http space-hf-hoyant-su-agentic-rl-agentic-rl 'https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/'
{
"mcpServers": {
"space-hf-hoyant-su-agentic-rl-agentic-rl": {
"url": "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/"
}
}
} {
"servers": {
"space-hf-hoyant-su-agentic-rl-agentic-rl": {
"type": "http",
"url": "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/"
}
}
} [mcp_servers.space-hf-hoyant-su-agentic-rl-agentic-rl] url = "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/"
{
"$schema": "https://opencode.ai/config.json",
"mcp": {
"space-hf-hoyant-su-agentic-rl-agentic-rl": {
"type": "remote",
"url": "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/",
"enabled": true
}
}
} openclaw mcp add space-hf-hoyant-su-agentic-rl-agentic-rl --url 'https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/' --transport streamable-http
mcp_servers:
space-hf-hoyant-su-agentic-rl-agentic-rl:
url: "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/" {
"McpServers": {
"space-hf-hoyant-su-agentic-rl-agentic-rl": {
"Transport": "http",
"Url": "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/"
}
}
} assistant mcp add space-hf-hoyant-su-agentic-rl-agentic-rl -t streamable-http -u 'https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/'
{
"mcpServers": {
"space-hf-hoyant-su-agentic-rl-agentic-rl": {
"type": "http",
"url": "https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/"
}
}
} The mcpServers block is a cross-client convention. Remote transports vary, so check your client's docs.
Every change we have recorded for this component, newest first. Security-relevant changes are always shown. ▲ marks a change for the better, ▼ a change for the worse; unmarked changes are neutral.
- 28 Sept 26 0
- We updated how we score, so this day's move reflects our rubric, not a change to the server See what changed → functional
- 27 Sept 26 +1
No change was recorded against any check on this day. Stability & Change Management went from 60 to 63. That category is still filling its 30-day observation window: 18 days of observed history at the previous scan, 19 at this one. The score rises as the window fills, whether or not the server changes.
- 25 Sept 26 +1
- We updated how we score, so this day's move reflects our rubric, not a change to the server See what changed → functional
- 23 Sept 26 +1
No change was recorded against any check on this day. Stability & Change Management went from 47 to 50. That category is still filling its 30-day observation window: 14 days of observed history at the previous scan, 15 at this one. The score rises as the window fills, whether or not the server changes.
- 21 Sept 26 +1
No change was recorded against any check on this day. Stability & Change Management went from 40 to 43. That category is still filling its 30-day observation window: 12 days of observed history at the previous scan, 13 at this one. The score rises as the window fills, whether or not the server changes.
- 19 Sept 26 +1
No change was recorded against any check on this day. Stability & Change Management went from 33 to 37. That category is still filling its 30-day observation window: 10 days of observed history at the previous scan, 11 at this one. The score rises as the window fills, whether or not the server changes.
- 16 Sept 26 +1
No change was recorded against any check on this day. Stability & Change Management went from 23 to 27. That category is still filling its 30-day observation window: 7 days of observed history at the previous scan, 8 at this one. The score rises as the window fills, whether or not the server changes.
- 14 Sept 26 +1
No change was recorded against any check on this day. Stability & Change Management went from 17 to 20. That category is still filling its 30-day observation window: 5 days of observed history at the previous scan, 6 at this one. The score rises as the window fills, whether or not the server changes.
Diagnostic detail from the automated scan of this channel: what the scanner observed at each step, so you can see exactly where a check passed or failed. It is informational only and never changes the trust score.
Captured 28 Sept 2026 · Probed https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/
TLS valid
Negotiated TLS 1.3 with TLS_AES_128_GCM_SHA256 .
| Subject | Issuer | Valid from | Valid until | Key | Signature | Serial |
|---|---|---|---|---|---|---|
| CN=hf.space | CN=Amazon RSA 2048 M04,O=Amazon,C=US | 27 Sept 2026 | 12 Apr 2027 | RSA 2048 | SHA256-RSA | f51719dcdf7239b73ac39d3e38d8e9d |
| SANs: hf.space, *.hf.space | ||||||
| CN=Amazon RSA 2048 M04,O=Amazon,C=US (CA) | CN=Amazon Root CA 1,O=Amazon,C=US | 23 Aug 2022 | 23 Aug 2030 | RSA 2048 | SHA256-RSA | 773124f2a952e3ed18a58bdb85d1bc0ce5f27 |
| CN=Amazon Root CA 1,O=Amazon,C=US (CA) | CN=Starfield Services Root Certificate Authority - G2,O=Starfield Technologies\, Inc.,L=Scottsdale,ST=Arizona,C=US | 25 May 2015 | 31 Dec 2037 | RSA 2048 | SHA256-RSA | 67f944a2a27cdf3fac2ae2b01f908eeb9c4c6 |
Background: What to check on a remote MCP endpoint →
DNSSEC insecure
Validation of hoyant-su-agentic-rl.hf.space. — Not signed
| Zone | DS | Keys | Algorithms | Outcome |
|---|---|---|---|---|
| . | trust_anchor | 20326, 38696 | 8, 8 | Verified |
| space. | present | 51168 | 13 | Verified |
| hf.space. | absent | Unsigned (proven) parent-signed NSEC/NSEC3 proves an unsigned delegation |
Authentication No authorisation required
The endpoint answered without asking for a token. Anyone who knows the URL can reach it.
| Result | No authorisation required |
|---|---|
| HTTP status | 200 |
Background: How OAuth 2.1 works in the 2026 MCP spec →
Transports 2 probes
| Transport | URL | Outcome | Status | Location |
|---|---|---|---|---|
| streamable-http | https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/ | Verified | 200 | |
| http (plaintext) | http://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/ | HTTPS enforced | 301 | https://hoyant-su-agentic-rl.hf.space:443/gradio_api/mcp/ |
The tools this component advertises to a client, with an estimated token cost for each. Expand a tool to see its parameters and schema. The per-tool counts are indicative and are not scored directly; the schema's total context footprint is one signal in Schema Quality & AI Usability. A tool's description is untrusted text the model reads on every call, which is what makes this list a security surface and not just an inventory: how tool poisoning works →
Agentic_RL_dataset_overview ~40
Inspect ShellOps and ShellOps-Pro task counts, train/test splits, task types, published schemas, source files, license and citation.
Input schema present but exposes no named parameters.
No output schema declared.
No examples provided.
Agentic_RL_fetch_evidence ~53
Fetch a complete original evidence block by the evidence_id returned from search_evidence, including section anchor, version, equations, table cells, links, and attribution.
| Name | Type | Req | Description |
|---|---|---|---|
| evidence_id | string | yes | – |
No output schema declared.
No examples provided.
Agentic_RL_filter_methods ~108
Filter agent RL credit-assignment methods by research conditions and return original section evidence and BibTeX. Discover accepted values with list_method_facets. Filters combine with AND; empty strings leave a facet unrestricted. Unknown critic status never matches no. Results use publication order without a relevance or quality ranking.
| Name | Type | Req | Description |
|---|---|---|---|
| credit_granularity | string | – | – |
| evaluation_setting | string | – | – |
| learned_value_critic | string | – | – |
| required_supervision | string | – | – |
No output schema declared.
No examples provided.
Agentic_RL_get_task ~99
Inspect one published ShellOps or ShellOps-Pro task by its exact task_id and partition ('shellops' or 'shellops_pro'). Returns the complete instruction, actual reward specification, published reference answer/command, file-entry metadata, pinned parquet rows and workspace asset links. File content is available at the source links. No shell execution or solution verification is performed.
| Name | Type | Req | Description |
|---|---|---|---|
| partition | string | yes | – |
| task_id | string | yes | – |
No output schema declared.
No examples provided.
Agentic_RL_list_method_facets ~42
List exact filter values for agent RL credit granularity, supervision, value critics and evaluation settings. Each value reports its source-supported method count.
Input schema present but exposes no named parameters.
No output schema declared.
No examples provided.
Agentic_RL_list_sources ~46
List original papers and retrieval coverage. Discover source-linked comparisons of credit assignment, agent memory, selective observation and terminal benchmarks, with JSON, CSV and BibTeX links.
Input schema present but exposes no named parameters.
No output schema declared.
No examples provided.
Agentic_RL_search_evidence ~65
Search original papers on agentic reinforcement learning, credit assignment and CLI agents. Use English keywords (AND), OR and quoted phrases. Return relevant passages, source citations, equations and table cells.
| Name | Type | Req | Description |
|---|---|---|---|
| limit | integer | – | – |
| query | string | yes | – |
No output schema declared.
No examples provided.
Agentic_RL_search_tasks ~136
Find real ShellOps CLI benchmark tasks by case-insensitive literal substring in the complete instruction, task ID or published task type. Empty query lists all tasks. Select partition 'all', 'shellops' or 'shellops_pro'; select published split 'all', 'train_src', 'train' or 'test'. Results are ordered by partition then task ID, with explicit pagination and no relevance scoring. The train subset is not double-counted.
| Name | Type | Req | Description |
|---|---|---|---|
| limit | integer | – | – |
| offset | integer | – | – |
| partition | string | – | – |
| query | string | yes | – |
| split | string | – | – |
No output schema declared.
No examples provided.
What is the Agentic RL: Credit Assignment and CLI Agents MCP server?
Agentic RL: Credit Assignment and CLI Agents is an MCP server listed in the public MCP registry as space.hf.hoyant-su-agentic-rl/agentic-rl. Filter agent RL methods by supervision, critic and task setting; retrieve source links and BibTeX. This page covers its hosted endpoint (https://hoyant-su-agentic-rl.hf.space/gradio_api/mcp/).
Is the Agentic RL: Credit Assignment and CLI Agents MCP server safe to use?
Agentic RL: Credit Assignment and CLI Agents scores 70 out of 100 on VerifyMCP. That is a record of what we were able to check automatically, not an endorsement. The category breakdown on this page shows every signal behind the number, including the ones we could not confirm.
What tools does the Agentic RL: Credit Assignment and CLI Agents MCP server expose?
Agentic RL: Credit Assignment and CLI Agents exposes 8 tools: Agentic_RL_list_sources, Agentic_RL_search_evidence, Agentic_RL_fetch_evidence, Agentic_RL_dataset_overview, Agentic_RL_search_tasks, and 3 more. Their descriptions and schemas cost roughly 589 tokens of context every time the server is loaded.
Does the Agentic RL: Credit Assignment and CLI Agents MCP server require authentication?
No. We connected to Agentic RL: Credit Assignment and CLI Agents without credentials and it answered, so anything it exposes is reachable by anyone who knows the address.
Is the Agentic RL: Credit Assignment and CLI Agents MCP server still maintained?
Agentic RL: Credit Assignment and CLI Agents is still listed as active in the MCP registry. We last reached this channel on 28 September 2026. Those dates come from our own scans of the registry and the channel itself, not from anything the publisher announced.