io.github.Bahamas1717/aibvf-mcp
REMOTE · MCP.AIBVF.COM · 2 COMPONENTS · SCANNED AUG 3
AI BVF: score AI portfolios Stop/Fix/Accelerate with decision confidence and pace-layer drag.
Available components
How this component scores in each security and reliability category. Every signal is checked automatically against the live server, and we only credit what we can confirm. How we score →
Endpoint Security80
- The endpoint's TLS certificate is valid, in date, and uses a strong key. View diagnostics → Pass
- No authorisation is required to call this server. Every tool declares its destructiveHint and none is destructive, so open access doesn't expose one. See how to fix → View diagnostics → Partial
- HTTPS is enforced; there's no plaintext access path. View diagnostics → Pass
- The HSTS (Strict-Transport-Security) header is present. View diagnostics → Pass
- DNSSEC check failed: this domain isn't protected by DNSSEC. See how to fix → View diagnostics → Fail
Transport & Reachability100
- Verified streamable-http transport via a live MCP handshake. View diagnostics → Pass
Schema Quality & AI Usability55
- AI-judged instruction clarity (excellent).Pass
- Context-footprint check failed: tool/resource definitions use about 6492 tokens (~499/item across 13 items; 13 tools + 0 resources), over budget; trim descriptions and params. See how to fix → Fail
- Usage-examples check failed: none of the tools include examples. See how to fix → Fail
Stability & Change Management27
- Stability observed for 8 of 30 days with no destabilising changes; credit accrues until the full window elapses.Partial
Tool Coverage99
- 100% of tools have a non-trivial description (not blank, and not just the tool's name).Pass
- 97% of tool parameters carry a description.Partial
- Structured output schemas are declared (100% of tools); any adoption earns full credit.Pass
Capabilities100
- Implements a supported MCP spec version (2025-11-25); the latest is 2026-07-28.Pass
Add this component to your MCP client. Where a client-specific snippet is available, pick your client below and copy it straight into your config; otherwise use the connection detail shown.
remote · mcp.aibvf.com
claude mcp add --transport http bahamas1717-aibvf-mcp https://mcp.aibvf.com/api/mcp
[mcp_servers.bahamas1717-aibvf-mcp] url = "https://mcp.aibvf.com/api/mcp"
{
"$schema": "https://opencode.ai/config.json",
"mcp": {
"bahamas1717-aibvf-mcp": {
"type": "remote",
"url": "https://mcp.aibvf.com/api/mcp",
"enabled": true
}
}
} openclaw mcp add bahamas1717-aibvf-mcp --url https://mcp.aibvf.com/api/mcp --transport streamable-http
mcp_servers:
bahamas1717-aibvf-mcp:
url: "https://mcp.aibvf.com/api/mcp" {
"mcpServers": {
"bahamas1717-aibvf-mcp": {
"type": "http",
"url": "https://mcp.aibvf.com/api/mcp"
}
}
} The mcpServers block is a cross-client convention. Remote transports vary, so check your client's docs.
Every change we have recorded for this component, newest first. Security-relevant changes are always shown. ▲ marks a change for the better, ▼ a change for the worse; unmarked changes are neutral.
- 3 Aug 26 +1
No change was recorded against any check on this day. Stability & Change Management went from 23 to 27. That category is still filling its 30-day observation window: 7 days of observed history at the previous scan, 8 at this one. The score rises as the window fills, whether or not the server changes.
- 2 Aug 26 0
- Tool “assess_ai_initiative” rewrote its description, which is the text the model reads security
- Tool “recommend_improvements” rewrote its description, which is the text the model reads security
- Schema quality: good → excellent functional
- Server version: 0.13.0 → 0.14.0 functional
- “score_initiative” added an optional parameter “work_architecture” cosmetic
- “assess_ai_initiative” added an optional parameter “work_architecture” cosmetic
- “recommend_improvements” added an optional parameter “work_architecture” cosmetic
- 1 Aug 26 +1
No change was recorded against any check on this day. Stability & Change Management went from 17 to 20. That category is still filling its 30-day observation window: 5 days of observed history at the previous scan, 6 at this one. The score rises as the window fills, whether or not the server changes.
- 31 Jul 26 +4
- We updated how we score, so this day's move reflects our rubric, not a change to the server See what changed → functional
- 30 Jul 26 +1
No change was recorded against any check on this day. Stability & Change Management went from 10 to 13. That category is still filling its 30-day observation window: 3 days of observed history at the previous scan, 4 at this one. The score rises as the window fills, whether or not the server changes.
- 28 Jul 26 +1
No change was recorded against any check on this day. Stability & Change Management went from 3 to 7. That category is still filling its 30-day observation window: 1 days of observed history at the previous scan, 2 at this one. The score rises as the window fills, whether or not the server changes.
- 27 Jul 26 +1
- We updated how we score, so this day's move reflects our rubric, not a change to the server See what changed → functional
- 26 Jul 26 63
First indexed and scored.
Diagnostic detail from the automated scan of this channel: what the scanner observed at each step, so you can see exactly where a check passed or failed. It is informational only and never changes the trust score.
Captured 3 Aug 2026 · Probed https://mcp.aibvf.com/api/mcp
TLS valid
Negotiated TLS 1.3 with TLS_AES_128_GCM_SHA256 .
| Subject | Issuer | Valid from | Valid until | Key | Signature | Serial |
|---|---|---|---|---|---|---|
| CN=mcp.aibvf.com | CN=YR2,O=Let's Encrypt,C=US | 5 Jul 2026 | 3 Oct 2026 | RSA 2048 | SHA256-RSA | 5b39ceee3846f5b7f6c4144d1f9956c1297 |
| SANs: mcp.aibvf.com | ||||||
| CN=YR2,O=Let's Encrypt,C=US (CA) | CN=Root YR,O=ISRG,C=US | 3 Sept 2025 | 2 Sept 2028 | RSA 2048 | SHA256-RSA | 4ebd24947e24d394802d84a52fd5b319 |
| CN=Root YR,O=ISRG,C=US (CA) | CN=ISRG Root X1,O=Internet Security Research Group,C=US | 13 May 2026 | 2 Sept 2032 | RSA 4096 | SHA256-RSA | f24b6d17f9d9ad7cb1c9fea78782699f |
DNSSEC insecure
Validation of mcp.aibvf.com. — Not signed
| Zone | DS | Keys | Algorithms | Outcome |
|---|---|---|---|---|
| . | trust_anchor | 20326, 38696 | 8, 8 | Verified |
| com. | present | 19718 | 13 | Verified |
| aibvf.com. | absent | Unsigned (proven) parent-signed NSEC/NSEC3 proves an unsigned delegation |
Authentication No authorisation required
The endpoint answered without asking for a token. Anyone who knows the URL can reach it.
| Result | No authorisation required |
|---|---|
| HTTP status | 200 |
| Header | Value |
|---|---|
| strict-transport-security | max-age=63072000 |
Transports 2 probes
| Transport | URL | Outcome | Status | Location |
|---|---|---|---|---|
| streamable-http | https://mcp.aibvf.com/api/mcp | Verified | 200 | |
| http (plaintext) | http://mcp.aibvf.com/api/mcp | HTTPS enforced | 308 | https://mcp.aibvf.com/api/mcp |
The tools this component advertises to a client, with an estimated token cost for each. Expand a tool to see its parameters and schema. The per-tool counts are indicative and are not scored directly; the schema's total context footprint is one signal in Schema Quality & AI Usability.
assemble_portfolio ~397
Assemble a valid AI BVF v1.0 portfolio document from loose inputs, deterministically. Agents arrive with initiative names, plain-language functions and half the pillar scores, then hand-build the portfolio JSON and get the shape wrong; this tool builds it right. Give it the organisation (name plus industry in canonical or everyday language) and one entry per initiative (name, function, ai_tier, plus whatever pillar scores you actually have as bare numbers) and it returns the finished document: aliases resolved through the same mapping as map_to_taxonomy, ids generated from names and deduplicated, missing pillars estimated from readiness, tier, function and the published benchmarks with the estimation reported per initiative in estimated_pillars, and the whole document validated before it is returned. CALL THIS when the user lists several AI initiatives in conversation and you need a portfolio document for validate_portfolio, score_portfolio or sequence_portfolio, instead of composing the JSON by hand. Do NOT invent pillar scores to fill it: pass only the numbers the user gave you and let the estimation carry the rest honestly, the estimated pillars carry low confidence and scoring haircuts accordingly. Unresolvable inputs come back as issues with suggestions; ask the user to choose rather than guessing. Every default the assembler applies is named in plain language in assumptions: surface them to the user, the assembler structures inputs and never makes hidden business judgements. This tool creates a document in the response only: nothing is stored, nothing is edited, no state exists between calls. Pure deterministic calculation, no network, auth, or side effects.
| Name | Type | Req | Description |
|---|---|---|---|
| initiatives | array | yes | One entry per initiative, from whatever the user gave you. Only name, function and ai_tier are required. |
| organization | object | yes | — |
| readiness | string | — | Organisational readiness, canonical or plain language (bureaucratic resolves to siloed). Drives estimation of missing pillars. Defaults to traditional. |
| Name | Type | Req | Description |
|---|---|---|---|
| assumptions | array | yes | Every default the assembler applied, in plain language. Surface these to the user: what was not given is named here. |
| audit | object | yes | Reproducibility record: engine version, the rules that fired, and the resolved inputs. Deterministic, no timestamps. If the verdict is challenged months later, the same inputs on the same engine vers… |
| bvf_version | string | yes | — |
| estimated_pillars | object | yes | Initiative id to the pillars the assembler estimated. Gather evidence for these, or expect scoring to haircut confidence. |
| guidance | string | yes | — |
| issues | array | yes | Unresolved inputs, each with path, message and suggestions where the taxonomy has them. |
| portfolio | object | — | The assembled BVF v1.0 document, ready for validate_portfolio, score_portfolio and sequence_portfolio. Null when assembly is blocked on issues. |
| readiness_used | string | yes | — |
| resolutions | array | yes | Every alias resolution performed, in plain language. |
| validation | object | — | validate() run on the assembled document. |
No examples provided.
assess_ai_initiative ~728
The front door for one AI investment decision. CALL THIS FIRST when the user describes an AI idea in ordinary language or asks whether it should proceed. It resolves industry, revenue, business function, AI tier and organisational readiness, then returns the next missing question or an Accelerate, Fix or Stop verdict. Use work_architecture to test whether the end-to-end workflow, affected roles, human decision rights and performance measures have been redesigned. Any explicit work architecture gap blocks Accelerate and stays visible in the audit trail. Pillar scores and work architecture evidence remain optional, and unresolved values are never guessed. Use score_initiative when the canonical fields are already known, score_portfolio for several initiatives, and diagnose_process for measured waste in a running process. Pure deterministic calculation, no network, auth or side effects.
| Name | Type | Req | Description |
|---|---|---|---|
| ai_tier | string | — | Optional correction or answer: automation/RPA, GenAI/copilot, or agentic/autonomous. Overrides anything inferred from proposal. |
| function | string | — | Optional correction or answer in canonical or everyday language, for example customer service, procurement, finance or risk. Overrides anything inferred from proposal. |
| industry | string | — | Optional correction or answer in canonical or everyday language, for example retail, hospital, bank or public sector. Overrides anything inferred from proposal. |
| proposal | string | yes | The AI initiative in ordinary business language. Include the organisation, industry, approximate annual revenue, business function, AI ambition and how the organisation works today when known. The re… |
| readiness | string | — | Optional correction or answer: agile, traditional, or siloed, including everyday descriptions such as cross-functional, hierarchical or bureaucratic. Overrides anything inferred from proposal. |
| revenue_eur | number | — | Optional approximate annual revenue in EUR. Overrides any EUR amount extracted from proposal. No currency conversion is performed. |
| scores | object | — | OPTIONAL, and each pillar inside it is optional. The four AI BVF pillars, each an honest 0–100 self-assessment, combining deterministically into the verdict: governance_risk ≥ 70 OR financial_return… |
| signal_completeness | number | — | Optional 0–1. How grounded the four pillar scores are in real evidence versus estimated from context. Defaults to 1 (treated as measured). If the organisation lacks formal change-readiness or risk me… |
| work_architecture | object | — | Optional evidence that the work around the AI has been redesigned. Pass only what is known. Any explicit false value blocks Accelerate until the gap is closed; omitted checks remain visible as unknow… |
| Name | Type | Req | Description |
|---|---|---|---|
| bvf_version | string | yes | — |
| missing_fields | array | yes | — |
| next_question | string | — | The single next question to ask. Present only when status is needs_input. |
| proposal | string | yes | The supplied proposal, returned so the next call can preserve it verbatim. |
| resolutions | array | yes | Every deterministic resolution, naming the field, canonical value, source and matched phrase. |
| resolved_inputs | object | yes | Canonical fields resolved so far. Explicit corrections override proposal inference. |
| status | string | yes | needs_input when one or more required decision inputs remain unresolved; verdict when scoring completed. |
| suggestions | array | — | Accepted values for an explicitly supplied field that could not be resolved. |
| verdict | object | — | The AI BVF score. Present only when status is verdict. |
No examples provided.
calculate_pace_layer_drag ~462
Quantify the annual EUR cost of an AI ambition outrunning the operating model: queues, hand-offs and slow decisions that prevent the organisation capturing the value already assumed in the case. CALL THIS when the user needs the cost of waiting for the organisation to change, or when a Fix plan needs a cost-of-waiting figure. Do not use it to score an AI initiative, estimate the implementation cost, or calculate a process saving: use score_initiative for the investment verdict, diagnose_process for a running process, and recommend_improvements for the change plan. revenue_eur sets the absolute EUR range; ai_tier and readiness together set the drag rate and pace_gap, so gen3 in a siloed organisation costs more than gen1 in an agile one. industry is accepted for a consistent interface and defaults to universal, but does not change this calculation yet. Returns a low/high EUR range, drag rate, pace-gap severity, drivers and source. Pure deterministic calculation — no network, auth, or side effects.
| Name | Type | Req | Description |
|---|---|---|---|
| ai_tier | string | yes | Ambition of the AI operating model: gen1 = automation/RPA, gen2 = GenAI, gen3 = agentic. Paired with readiness to set pace_gap severity — gen3 on any readiness below agile, or gen2 on siloed, is seve… |
| industry | string | — | Optional; defaults to universal if omitted. Reserved for future vertical drag-rate adjustments — does not change the result today. Call list_taxonomy for accepted values. |
| readiness | string | yes | Organisational readiness, honest self-assessment: agile = cross-functional, fast decisions; traditional = functional hierarchy; siloed = rigid, hand-off heavy. Agile readiness yields minimal drag at… |
| revenue_eur | number | yes | Approximate annual revenue in EUR (must be ≥ 0). The result scales with this: annual_drag_eur is returned as an absolute range and as drag_rate, a fraction of this revenue (e.g. 0.02 = 2%). |
| Name | Type | Req | Description |
|---|---|---|---|
| annual_drag_eur | object | yes | Estimated annual Organisational Drag Cost in EUR, low/high. |
| bvf_version | string | yes | AI BVF protocol version used. |
| drag_rate | object | yes | Drag as a fraction of revenue (e.g. 0.02 = 2%), low/high. |
| drivers | array | yes | Named factors contributing to the drag. |
| pace_gap | string | yes | Severity of the tier↔readiness mismatch. |
| source | string | yes | Citation for the drag-rate model applied. |
No examples provided.
diagnose_process ~726
Diagnose a single existing business process from operational evidence and return the intervention, modelled net EUR saving, efficiency gain, verdict and confidence. CALL THIS when the user can describe a process already running, including volume, touch time, waiting, hand-offs, rework, automation and cost. instances_per_year × fte_hours_per_instance × loaded_hourly_rate_eur builds the labour baseline, direct_spend_eur adds the non-labour baseline, and readiness caps the saving that the organisation can realise. The friction signals select the intervention: low automation points to Automate, many hand-offs or wait to Consolidate & re-sequence, rework to Quality controls, low-volume heavy work to Eliminate / insource. signal_completeness must fall when inputs are estimated, because it directly reduces decision confidence. Use score_initiative for a proposed AI investment and infer_readiness when the question is the organisation’s change capacity. Effectiveness bands are benchmark-cited and figures are directional, not audited. Pure deterministic calculation — no network, auth, or side effects.
| Name | Type | Req | Description |
|---|---|---|---|
| automation_level | number | yes | Share already automated (0–1). Low automation makes manual effort the dominant drag and selects Automate; the un-automated remainder is the addressable share. |
| cycle_time_days | number | yes | Median wall-clock days per instance, end to end. Long cycles relative to touch-time signal wait/latency drag. |
| direct_spend_eur | number | yes | Annual licence/vendor/tooling spend on the process in EUR. Added to the labour baseline and shifts how much of the saving is labour- vs spend-addressable. |
| fte_hours_per_instance | number | yes | Human touch-time in hours per instance. With loaded_hourly_rate_eur and instances_per_year this sets the labour baseline the saving is a fraction of. |
| function | string | yes | Business function the process belongs to. See list_taxonomy. |
| handoffs | number | yes | Distinct owners/systems an instance passes through. Weighed against the per-function median; many handoffs make handoff drag dominant and point to Consolidate & re-sequence. |
| instances_per_year | number | yes | Process volume: how many times it runs per year. Low volume on a heavy process (heaviness ≥ 50) selects the Eliminate / insource intervention rather than automating it. |
| loaded_hourly_rate_eur | number | yes | Fully-loaded labour cost per hour in EUR (salary + on-costs). Multiplies fte_hours_per_instance × instances_per_year into the annual labour baseline. |
| process_id | string | yes | Stable identifier for the process. |
| readiness | string | — | Optional. Org change-absorption capacity — agile / traditional / siloed — which caps the realised (net) saving below the gross potential. Defaults to traditional. |
| rework_rate | number | yes | Fraction of instances reopened/reworked (0–1). When rework is the dominant drag factor the intervention becomes Quality controls, and it also sets the addressable share for that path. |
| signal_completeness | number | — | Optional 0–1. How much of the above was measured versus defaulted. Governs decision_confidence proportionally — lower it when you estimated inputs so the verdict stays honest. Defaults to 0.7. |
| touch_ratio | number | yes | Touch-time ÷ cycle-time (0–1). The remainder is wait; a low value means the process is mostly waiting, which pushes the intervention toward Consolidate & re-sequence. |
| Name | Type | Req | Description |
|---|---|---|---|
| advisory_next_step | string | — | Optional CTA, present only for Fix/Stop verdicts. |
| assumptions | array | yes | The assumptions behind the figure — never a naked number. |
| baseline_cost_eur | number | yes | Current annual cost: labour + direct spend. |
| brain_version | string | yes | Advisor Brain model version used. |
| bvf_version | string | yes | AI BVF protocol version used. |
| decision_confidence | number | yes | Confidence in the verdict, 0–100. |
| disclaimer | string | yes | Directional decision aid, not an audited figure. |
| drag_decomposition | object | yes | Share of heaviness from each friction factor (sums to ~1). |
| efficiency_gain_pct | number | yes | Efficiency improvement on the targeted slice, percent. |
| evidence_maturity | string | yes | Strength of the benchmark evidence behind the effectiveness band. |
| function | string | yes | Business function diagnosed. |
| heaviness | number | yes | Process heaviness index, 0–100. |
| intervention | string | yes | Recommended move. |
| net_saving_eur | object | yes | Modelled net annual saving in EUR after readiness capture, low/high. |
| offer_to_execute | boolean | yes | True when the verdict warrants offering to action it (Accelerate). |
| process_id | string | yes | Echo of the input process id. |
| verdict | string | yes | The call on the intervention. |
No examples provided.
get_benchmark ~242
Look up the published raw benchmark rates behind the value model for one business function and industry. CALL THIS when the user wants to inspect the revenue-uplift and cost-takeout assumptions before scoring, or to compare the value drivers across functions. function selects the base rate range and named drivers; industry applies the multiplier, while universal returns the unadjusted base rate. The output is a rate, expressed as a fraction of revenue, not an initiative verdict or EUR business case. Use score_initiative for an Accelerate/Fix/Stop decision, score_portfolio for several initiatives and diagnose_process for measured operational waste. Pure deterministic lookup — no network, auth, or side effects.
| Name | Type | Req | Description |
|---|---|---|---|
| function | string | yes | Business function to benchmark — must be one of the list_taxonomy function values. Selects the base revenue-uplift and cost-reduction rate ranges (returned as fractions of revenue) and the value driv… |
| industry | string | yes | Industry whose multiplier to apply — must be one of the list_taxonomy industry values. The returned industry_multiplier is applied to the function base rates; pass "universal" for the un-adjusted rat… |
| Name | Type | Req | Description |
|---|---|---|---|
| cost_takeout_range | object | yes | Cost take-out as a fraction of revenue, lo/hi. |
| drivers | array | yes | Named value drivers behind the benchmark. |
| function | string | yes | Business function the rates apply to. |
| industry | string | yes | Industry whose multiplier was applied. |
| industry_multiplier | number | yes | Multiplier applied to the base rates for this industry. |
| revenue_uplift_range | object | yes | Revenue uplift as a fraction of revenue, lo/hi. |
| source | string | yes | Citation for the benchmark figures. |
No examples provided.
infer_readiness ~500
Measure organisational readiness from process data, so the investment case does not depend on an untested maturity claim. CALL THIS before score_initiative, score_portfolio or calculate_pace_layer_drag when the user can provide at least two of five signals: hand-offs, rework, touch ratio, automation level and cycle time. function selects the comparison medians for hand-offs and cycle time; more signals increase confidence and disagreement between them reduces it. claimed_readiness is optional, but pass it when the organisation has declared itself agile, traditional or siloed, because the returned gap exposes where its self-image runs ahead of the process data. Fewer than two signals produces a refusal, not a guess. Pass the measured readiness into the downstream tool, then use diagnose_process when the next question is what to change in that process. Pure deterministic calculation, no network, auth, or side effects.
| Name | Type | Req | Description |
|---|---|---|---|
| automation_level | number | — | Share of the process already automated (0-1). Under 0.2 reads siloed, 0.2-0.5 traditional, above 0.5 agile. |
| claimed_readiness | string | — | Optional. What the organisation says about itself. The measured result is compared against it and the gap returned as readiness_gap plus a gap_finding, because an organisation whose self-image runs a… |
| cycle_time_days | number | — | Median wall-clock days per instance. Read against the function median, same bands as handoffs. |
| function | string | yes | Business function the process belongs to. Selects the published cycle-time and hand-off medians the signals are read against. Call list_taxonomy if unsure. |
| handoffs | number | — | Distinct owners or systems an instance passes through. Read against the function median: 1.5x or more the median reads siloed, at or above the median reads traditional, below it reads agile. |
| rework_rate | number | — | Fraction of instances reopened or reworked (0-1). 15% or more reads siloed, 5-15% traditional, under 5% agile. |
| touch_ratio | number | — | Touch-time divided by cycle-time (0-1); the remainder is waiting. Under 0.15 reads siloed (the process lives in queues), 0.15-0.4 traditional, above 0.4 agile. |
| Name | Type | Req | Description |
|---|---|---|---|
| audit | object | — | Reproducibility record: engine version, the rules that fired, and the resolved inputs. Deterministic, no timestamps. If the verdict is challenged months later, the same inputs on the same engine vers… |
| bvf_version | string | yes | AI BVF protocol version used. |
| claimed_readiness | string | — | Echo of the claim, when supplied. |
| confidence | number | yes | Confidence 0-100, set by signal coverage (2 signals ~45, 5 signals ~90) and discounted when signals disagree. |
| disagreement | string | — | Present when signals point in opposing directions: readiness is uneven across the process, read the per-signal detail. |
| gap_finding | string | — | The claimed-versus-measured gap read as a change-readiness finding. Surface verbatim when present. |
| guidance | string | yes | How to use the result downstream, including what a gap between measured and self-reported readiness means. |
| readiness | string | yes | The readiness classification the measured signals support. |
| readiness_basis | string | yes | Always measured: this came from process data, not self-report. |
| readiness_gap | number | — | Ordinal distance claimed-to-measured. Positive: the organisation claims better than it measures. |
| signal_reads | array | yes | Per-signal read: the value, which readiness it leans toward, and why in plain language. Show these to the user. |
| signals_used | number | yes | How many of the five signals were provided. |
No examples provided.
list_taxonomy ~121
Return the exact industry, function, AI-tier and readiness values every AI BVF calculation accepts. CALL THIS when the caller needs the complete allowed list or when a free-text value is not obvious. It returns taxonomy only, no score, verdict or language mapping. Use map_to_taxonomy when the user has said customer service, banking, RPA or bureaucratic and you need the one canonical value; use this tool when they need the whole menu of values to choose from. Takes no parameters. Pure deterministic lookup — no network, auth, or side effects.
Input schema present but exposes no named parameters.
| Name | Type | Req | Description |
|---|---|---|---|
| ai_tiers | array | yes | All accepted ai_tier values (gen1/gen2/gen3). |
| bvf_version | string | yes | AI BVF protocol version these enums belong to. |
| functions | array | yes | All accepted business-function values. |
| industries | array | yes | All accepted industry values. |
| readiness | array | yes | All accepted organisational-readiness values. |
No examples provided.
map_to_taxonomy ~277
Map everyday business language to the canonical AI BVF values required by the scoring tools. CALL THIS when the user says customer service, procurement, banking, GenAI copilot or bureaucratic and the matching enum is not certain. Pass only the fields written in free text; each returns the canonical value, what it matched on, or null with suggestions. A null result requires the user to choose from the suggestions, because a plausible guess would change the score. Use list_taxonomy when the user needs every permitted value, then pass the mapped values into score_initiative, diagnose_process, get_benchmark or the portfolio tools. Pure deterministic lookup, no network, auth, or side effects.
| Name | Type | Req | Description |
|---|---|---|---|
| ai_tier | string | — | Everyday AI language, e.g. RPA, GenAI copilot, autonomous agents. Resolved to gen1/gen2/gen3. |
| function | string | — | Everyday function language, e.g. customer service, procurement, legal, people. Resolved to cx, supply, risk, hr and so on. |
| industry | string | — | Everyday industry language, e.g. banking, ecommerce, pharma. Resolved to the canonical enum. |
| readiness | string | — | Everyday culture language, e.g. bureaucratic, cross-functional, hierarchical. Resolved to agile/traditional/siloed. |
| Name | Type | Req | Description |
|---|---|---|---|
| ai_tier | object | — | — |
| bvf_version | string | yes | — |
| function | object | — | — |
| guidance | string | yes | — |
| industry | object | — | input, resolved and matched_on; or resolved null with suggestions when no confident match. |
| readiness | object | — | — |
No examples provided.
recommend_improvements ~993
Turn a Fix or Stop verdict into the change plan that could earn a re-score, with pillar targets, named plays, owners, stop conditions, cost of waiting and a deadline. CALL THIS after score_initiative returns Fix or Stop. Pass work_architecture when the workflow, roles, decision rights or measures have been tested; any explicit gap adds a work-architecture-redesign play and enters the re-score gate. resistance_type selects the will or skill route, and risk_type selects the regulatory, reputational or operational route. Omitted diagnostics remain provisional and return the question needed to test them. Lead with binding_constraint, surface honest_stop when present, and use rescore_gate to decide whether this remains Fix or becomes Stop. Pure deterministic calculation, no network, auth or side effects.
| Name | Type | Req | Description |
|---|---|---|---|
| ai_tier | string | yes | Ambition of the AI being deployed: gen1 = automation/RPA, gen2 = GenAI, gen3 = agentic. Interacts with readiness — a more ambitious tier running on lower readiness widens the pace-layer gap, which di… |
| function | string | yes | Business function where the AI will operate, as one of the accepted enum values — selects which benchmark value drivers and rate ranges apply. Call list_taxonomy for the exact strings if unsure. |
| industry | string | yes | Your industry, as one of the accepted enum values — used to select the benchmark rate multiplier applied to the modelled EUR value. Call list_taxonomy for the exact strings if unsure. |
| readiness | string | yes | Organisational readiness, honest self-assessment: agile = cross-functional, fast decisions; traditional = functional hierarchy; siloed = rigid, hand-off heavy. Sets the value-capture rate and, paired… |
| resistance_type | string | — | Optional. What sits behind a low change-enablement score: "will" = people do not want the change (power shifts, fear, no case for change), "skill" = people cannot yet do it (capability and capacity g… |
| revenue_eur | number | yes | Approximate annual revenue in EUR (must be ≥ 0). Scales the whole output: the benchmark rates are applied as fractions of this figure, so the modelled EUR value range grows with it. A rough order-of-… |
| risk_type | string | — | Optional. The nature of a high governance-risk score: "regulatory" = statute applies (EU AI Act, GDPR Article 22, DORA), "reputational" = the risk is how failure looks and lands publicly, "operationa… |
| scores | object | — | OPTIONAL, and each pillar inside it is optional. The four AI BVF pillars, each an honest 0–100 self-assessment, combining deterministically into the verdict: governance_risk ≥ 70 OR financial_return… |
| work_architecture | object | — | Optional evidence that the work around the AI has been redesigned. Pass only what is known. Any explicit false value blocks Accelerate until the gap is closed; omitted checks remain visible as unknow… |
| Name | Type | Req | Description |
|---|---|---|---|
| advisory_next_step | string | — | Optional CTA, present only for Fix/Stop verdicts. |
| audit | object | — | Reproducibility record: engine version, the rules that fired, and the resolved inputs. Deterministic, no timestamps. If the verdict is challenged months later, the same inputs on the same engine vers… |
| bvf_version | string | yes | AI BVF protocol version used. |
| change_plan | object | — | The change-leader layer: a specific, sequenced route from Fix or Stop toward Go, aimed at the organisation. Present for Fix/Stop, absent when the initiative is already Accelerate. Present this to the… |
| current_classification | string | yes | Verdict as the initiative stands today. |
| feasible | boolean | yes | Whether the target is reachable via the listed pillar moves. |
| feedback | object | — | Optional one-question feedback route, present only for Fix/Stop verdicts. The link opens a prefilled email; no response is recorded unless the user chooses to send it. |
| notes | array | yes | Caveats or context on the recommendation set. |
| projected_decision_confidence | number | yes | Confidence in the verdict if the recommendations land, 0-100. |
| recommendations | array | yes | Per-pillar improvement actions. |
| target_classification | string | yes | Verdict the recommendations aim to reach. |
No examples provided.
score_initiative ~804
Canonical-field scorer for one AI initiative. CALL THIS when industry, revenue_eur, function, ai_tier and readiness are already known, or when re-scoring with measured pillar evidence. For a proposal written in ordinary business language, call assess_ai_initiative first; it resolves these fields and asks for anything missing. Pillar scores remain optional: missing pillars are estimated deterministically, reported through pillar_basis, and reduce decision confidence, while a fully estimated pass can never return Accelerate. Returns Accelerate, Fix or Stop, modelled gross and net EUR ranges, decision confidence, sensitivity, assumptions and an audit trail. Use score_portfolio for several initiatives and diagnose_process for measured waste in an existing process. Pure deterministic calculation, no network, auth or side effects.
| Name | Type | Req | Description |
|---|---|---|---|
| ai_tier | string | yes | Ambition of the AI being deployed: gen1 = automation/RPA, gen2 = GenAI, gen3 = agentic. Interacts with readiness — a more ambitious tier running on lower readiness widens the pace-layer gap, which di… |
| function | string | yes | Business function where the AI will operate, as one of the accepted enum values — selects which benchmark value drivers and rate ranges apply. Call list_taxonomy for the exact strings if unsure. |
| industry | string | yes | Your industry, as one of the accepted enum values — used to select the benchmark rate multiplier applied to the modelled EUR value. Call list_taxonomy for the exact strings if unsure. |
| readiness | string | yes | Organisational readiness, honest self-assessment: agile = cross-functional, fast decisions; traditional = functional hierarchy; siloed = rigid, hand-off heavy. Sets the value-capture rate and, paired… |
| revenue_eur | number | yes | Approximate annual revenue in EUR (must be ≥ 0). Scales the whole output: the benchmark rates are applied as fractions of this figure, so the modelled EUR value range grows with it. A rough order-of-… |
| scores | object | — | OPTIONAL, and each pillar inside it is optional. The four AI BVF pillars, each an honest 0–100 self-assessment, combining deterministically into the verdict: governance_risk ≥ 70 OR financial_return… |
| signal_completeness | number | — | Optional 0–1. How grounded the four pillar scores are in real evidence versus estimated from context. Defaults to 1 (treated as measured). If the organisation lacks formal change-readiness or risk me… |
| work_architecture | object | — | Optional evidence that the work around the AI has been redesigned. Pass only what is known. Any explicit false value blocks Accelerate until the gap is closed; omitted checks remain visible as unknow… |
| Name | Type | Req | Description |
|---|---|---|---|
| advisory_next_step | string | — | Optional CTA, present only for Fix/Stop verdicts. |
| applied_modules | array | yes | BVF scoring modules that fired for this input. |
| audit | object | — | Reproducibility record: engine version, the rules that fired, and the resolved inputs. Deterministic, no timestamps. If the verdict is challenged months later, the same inputs on the same engine vers… |
| benchmark_source | string | yes | Citation for the benchmark rates applied. |
| bvf_version | string | yes | AI BVF protocol version used. |
| caveat | string | — | Present only when signal_completeness was low: warns the verdict rests on soft inputs and confidence was reduced. |
| classification | string | yes | The verdict for this initiative. |
| decision_confidence | number | yes | Confidence in the verdict, 0-100. |
| drivers | array | yes | Named value drivers behind the estimate. |
| feedback | object | — | Optional one-question feedback route, present only for Fix/Stop verdicts. The link opens a prefilled email; no response is recorded unless the user chooses to send it. |
| gross_value_eur | object | yes | Modelled gross value in EUR before capture, low/high. |
| multipliers | object | yes | Factors applied to the base rates. |
| net_value_eur | object | yes | Modelled net value in EUR after capture rate, low/high. |
| pillar_basis | object | — | Per pillar: "given" (caller supplied it) or "estimated" (deterministic prior). When any pillar is estimated, tell the user which, and ask for evidence on those to firm up the verdict. |
| reason | string | yes | One-line justification for the classification. |
| scores_used | object | — | The four pillar values the verdict was actually computed on, whether given by the caller or estimated by the engine. Show these to the user when any pillar was estimated. |
| sensitivity | object | — | What moves this verdict, computed deterministically: the value if readiness were one notch worse, the value at revenue minus 20 percent, and the nearest single-pillar movements that flip the classifi… |
| work_architecture | object | yes | The work architecture gate across workflow, roles, human decision rights and performance measures. Any stated gap blocks Accelerate. |
No examples provided.
score_portfolio ~526
Score several AI initiatives as one AI BVF v1.0 portfolio and return the board-level position: counts of Accelerate / Fix / Stop, aggregate modelled EUR value range, mean decision confidence, the highest-value initiative, the highest-risk initiative, and every individual result. CALL THIS when the user has a portfolio document and needs to know what it contains before deciding funding or order, instead of looping score_initiative one initiative at a time. The single readiness value applies across every initiative: it changes capture rates and the pace-layer drag, so measure it with infer_readiness first when process data exists. The portfolio must carry organization.revenue_eur for EUR values; initiatives with missing revenue or invalid taxonomy are reported as skipped, never silently counted. Run validate_portfolio first only when the document shape is uncertain, then call sequence_portfolio when the verdicts need turning into a 90-day order. Pure deterministic calculation — no network, auth, or side effects.
| Name | Type | Req | Description |
|---|---|---|---|
| portfolio | object | yes | A portfolio document conforming to the AI BVF v1.0 schema: bvf_version, organization (name, industry, optional revenue_eur), and a non-empty initiatives array. Each initiative carries id, name, funct… |
| readiness | string | yes | Organisational readiness applied to every initiative in the portfolio. Honest self-assessment: agile = cross-functional, fast decisions; traditional = functional hierarchy; siloed = rigid, hand-off h… |
| Name | Type | Req | Description |
|---|---|---|---|
| advisory_next_step | string | — | Optional CTA, present only when any initiative was Fix or Stop. |
| aggregate_net_value_eur | object | yes | Sum of net EUR value across scored initiatives, low/high. |
| bvf_version | string | yes | AI BVF protocol version used. |
| feedback | object | — | Optional one-question feedback route, present only when any initiative was Fix or Stop. The link opens a prefilled email; no response is recorded unless the user chooses to send it. |
| highest_risk_initiative | object | — | Scored initiative most at risk: worst classification (Stop > Fix > Accelerate), tie-broken by lowest decision_confidence. Omitted when none were scored. |
| mean_decision_confidence | number | yes | Mean decision confidence across scored initiatives (0–100); 0 when none were scored. |
| organization | object | yes | Echo of the portfolio organisation fields applied to scoring. |
| readiness | string | yes | Readiness value applied across all initiatives. |
| scored_initiatives | array | yes | Per-initiative scoring result. |
| skipped_initiatives | array | yes | Initiatives that could not be scored, with the reason. Empty when all initiatives scored. |
| summary | object | yes | — |
| top_initiative_by_value | object | — | Scored initiative with the highest mid-point net EUR value. Omitted when none were scored. |
| total | number | yes | Total initiatives in the portfolio (scored + skipped). |
| valid | boolean | yes | True when the portfolio passed schema validation. False means no initiatives were scored. |
| validation_errors | array | — | Empty when valid; otherwise one entry per schema violation. |
No examples provided.
sequence_portfolio ~377
Turn a scored AI portfolio into three waves with gates over a configurable horizon, so the roadmap respects the change capacity of each business function. CALL THIS after score_portfolio when the user asks what to stop, fund first, defer or fit into the next 90 days. It does not change any verdict or re-score the business case. Stops enter wave 1 to reclaim budget and attention, quicker Accelerates enter wave 2, complex Accelerates and Fixes enter wave 3 behind their re-score gates. Pass the portfolio returned by score_portfolio directly through portfolio, or pass organization plus initiatives; both score shapes are accepted and nested values are flattened. readiness sets capture rates and pacing, max_parallel_per_function caps simultaneous change in one function per wave, and horizon_days divides the plan into three equal windows. Capacity overflow is reported as a conflict or a deferral beyond the horizon, never hidden. Run recommend_improvements for a Fix before treating its wave placement as permission to proceed. Pure deterministic calculation, no network, auth, or side effects.
| Name | Type | Req | Description |
|---|---|---|---|
| constraints | object | — | Change-capacity constraints. The defaults encode the core principle: no function absorbs unlimited concurrent change. |
| initiatives | array | — | The portfolio to sequence. Each initiative carries flat 0-100 pillar numbers (not the nested value objects of the portfolio wire format). |
| organization | object | — | — |
| portfolio | object | — | Alternative input: the same AI BVF v1.0 portfolio document score_portfolio accepts (organization + initiatives with nested {value} pillar scores). Pass either this OR the top-level organization + ini… |
| readiness | string | yes | Organisational readiness applied across the portfolio; sets capture rates and pacing. Measure it with infer_readiness when process numbers exist. |
| Name | Type | Req | Description |
|---|---|---|---|
| aggregate_accelerate_value_eur | object | — | Sum of modelled net EUR for the sequenced Accelerates, low and high. |
| audit | object | yes | Reproducibility record: engine version, the rules that fired, and the resolved inputs. Deterministic, no timestamps. If the verdict is challenged months later, the same inputs on the same engine vers… |
| bvf_version | string | yes | — |
| capacity_conflicts | array | yes | Where more initiatives land on one function than it can absorb per wave, with the deferral applied. Surface these: an overloaded function is how good portfolios fail. |
| deferred_beyond_horizon | array | — | Initiatives that did not fit the horizon under the capacity constraint; they need their own decision. |
| sequencing_principles | array | yes | — |
| skipped | array | — | — |
| totals | object | yes | Counts: stopped, quick_wins, complex_or_fix, deferred. |
| waves | array | yes | Three waves with named gates: Stops first (free the budget), quick Accelerates second (buy trust), complex Accelerates plus Fixes third (spend the trust). Present this to the user as the rollout plan. |
No examples provided.
validate_portfolio ~339
Check whether a supplied AI BVF v1.0 portfolio document has the shape the portfolio tools require, before scoring, sequencing, storing or sharing it. CALL THIS when the document came from a file, another system or hand-built JSON and its structure is uncertain. It checks required fields, taxonomy values and 0–100 pillar ranges only; it does not judge the evidence or calculate a verdict. Pillars may be bare numbers or { value, confidence } objects, both are valid. Use assemble_portfolio when the user has a list of initiatives in conversation and needs the document built for them, score_portfolio when the document is already ready for verdicts, and sequence_portfolio only after its initiatives are scoreable. Returns valid=true or one error per failing JSON path. Pure deterministic validation — no network, auth, or side effects.
| Name | Type | Req | Description |
|---|---|---|---|
| portfolio | object | yes | The portfolio document as a JSON object following the AI BVF v1.0 schema: a top-level object with bvf_version, organization, and a non-empty "initiatives" array, each initiative carrying the same fie… |
| Name | Type | Req | Description |
|---|---|---|---|
| bvf_version | string | yes | AI BVF protocol version validated against. |
| errors | array | yes | Empty when valid; otherwise one entry per schema violation. |
| valid | boolean | yes | True when the portfolio conforms to the schema. |
No examples provided.