io.github.aborruso/ckan-mcp-server
NPM · @ABORRUSO/CKAN-MCP-SERVER · SCANNED AUG 3
MCP server for interacting with CKAN open data portals
Available components
How this component scores in each security and reliability category. Every signal is checked automatically from public evidence about the published package, including repeated runs of it in an isolated sandbox, and we only credit what we can confirm. How we score →
Supply Chain Security92
- No malware found by supply-chain analysis.Pass
- Only part of the dependency tree could be resolved (135 of 137), so this covers what we could see, not the whole tree.Partial
- No install/post-install scripts declared.Pass
- Only part of the dependency tree could be resolved (135 of 137), so this covers what we could see, not the whole tree. View diagnostics → Partial
Provenance & Transparency97
- Source repository is publicly reachable at the declared URL. View diagnostics → Pass
- Cryptographically verified build provenance (signed, bound to ondata/ckan-mcp-server). View diagnostics → Pass
- Clear OSI-approved license (MIT).Pass
- Actively maintained (last published 0 days ago).Pass
- Disclosure check failed: no security disclosure policy was found in the source repository. See how to fix → Fail
Schema Quality & AI Usability71
- 100% of prompts and resources have a non-trivial description (not blank, and not just the item's name).Pass
- AI-judged instruction clarity (good).Pass
- Context-footprint check failed: tool/resource definitions use about 8264 tokens (~413/item across 20 items; 20 tools + 0 resources), over budget; trim descriptions and params. See how to fix → Fail
- Usage-examples check failed: none of the tools include examples. See how to fix → Fail
Stability & Change Management27
- Stability observed for 8 of 30 days with no destabilising changes; credit accrues until the full window elapses.Partial
Tool Coverage100
- 100% of tools have a non-trivial description (not blank, and not just the tool's name).Pass
- 100% of tool parameters carry a description.Pass
Capabilities100
- Implements a supported MCP spec version (2025-11-25); the latest is 2026-07-28.Pass
Add this component to your MCP client. Where a client-specific snippet is available, pick your client below and copy it straight into your config; otherwise use the connection detail shown.
npm · @aborruso/ckan-mcp-server
claude mcp add aborruso-ckan-mcp-server -- npx -y @aborruso/ckan-mcp-server
codex mcp add aborruso-ckan-mcp-server -- npx -y @aborruso/ckan-mcp-server
{
"$schema": "https://opencode.ai/config.json",
"mcp": {
"aborruso-ckan-mcp-server": {
"type": "local",
"command": [
"npx",
"-y",
"@aborruso/ckan-mcp-server"
],
"enabled": true
}
}
} openclaw mcp add aborruso-ckan-mcp-server --command npx --arg -y --arg @aborruso/ckan-mcp-server
mcp_servers:
aborruso-ckan-mcp-server:
command: "npx"
args: ["-y", "@aborruso/ckan-mcp-server"] {
"mcpServers": {
"aborruso-ckan-mcp-server": {
"command": "npx",
"args": [
"-y",
"@aborruso/ckan-mcp-server"
]
}
}
} Every change we have recorded for this component, newest first. Security-relevant changes are always shown. ▲ marks a change for the better, ▼ a change for the worse; unmarked changes are neutral.
- 3 Aug 26 +19
- CVE-2026-53509 no longer affects this package ▲ security
- CVE-2026-33060 no longer affects this package ▲ security
- Provenance: fail → pass ▲ security
- Known CVEs: fail → partial ▲ security
- Stability: Stability not yet verified: we do not have a sandbox capture of the MCP schema this version of the package serves yet. security
- The attested source repository moved: ondata/ckan-mcp-server security
- Schema quality: 374 → 413 ▼ functional
- Schema quality: 100 → unverified ▼ functional
- Tool coverage: 100 → unverified ▼ functional
- Capabilities: pass → unverified ▼ functional
- Tool coverage: 86% → 100% ▲ functional
- Stability: unverified → 0.27 ▲ functional
- Package version: 0.4.83 → 0.4.115 functional
- Package version: 0.4.83 → 0.4.114 functional
- 2 Aug 26 −3
No change was recorded against any check on this day. Supply Chain Security went from 88 to 78.
- 1 Aug 26 +59
- Known CVEs: unverified → fail ▼ security
- Provenance: unverified → fail ▼ security
- Malware scan: unverified → pass ▲ security
- Install scripts: unverified → pass ▲ security
- Stability: Stability not yet verified: not enough scan history yet (needs a 30-day window). security
- MCP protocol: unverified → pass ▲ functional
- Tool coverage: unverified → 100 ▲ functional
- License: unverified → pass ▲ functional
- Schema quality: unverified → 100 ▲ functional
- Dependency health: unverified → partial ▲ functional
- Maintenance: unverified → pass ▲ functional
- Licence: MIT functional
- 31 Jul 26 −33
- We updated how we score, so this day's move reflects our rubric, not a change to the server See what changed → functional
- 30 Jul 26 −4
- Known CVEs: fail → unverified ▼ security
- Malware scan: pass → unverified ▼ security
- Provenance: fail → unverified ▼ security
- Install scripts: pass → unverified ▼ security
- CVE-2026-33060 no longer affects this package ▲ security
- CVE-2026-53509 no longer affects this package ▲ security
- Security disclosure: unverified → fail ▼ functional
- Maintenance: pass → unverified ▼ functional
- Dependency health: partial → unverified ▼ functional
- License: pass → unverified ▼ functional
- Schema quality: unverified → 100 ▲ functional
- Tool coverage: unverified → 100 ▲ functional
- Licence: MIT functional
- 28 Jul 26 −32
- Schema quality: 100 → unverified ▼ functional
- Security disclosure: fail → unverified ▼ functional
- Tool coverage: 100 → unverified ▼ functional
- 27 Jul 26 +32
- We updated how we score, so this day's move reflects our rubric, not a change to the server See what changed → functional
- 26 Jul 26 42
First indexed and scored.
Diagnostic detail from the automated scan of this channel: what the scanner observed at each step, so you can see exactly where a check passed or failed. It is informational only and never changes the trust score.
Captured 3 Aug 2026 · Analysed npm/@aborruso/[email protected]
Provenance verified
Ecosystem: npm · Outcome: verified
Reason: verified
- Source repo:
- ondata/ckan-mcp-server
- Certificate issuer:
- https://token.actions.githubusercontent.com
- Certificate SAN:
- https://github.com/ondata/ckan-mcp-server/.github/workflows/release.yml@refs/tags/v0.4.115
- Rekor log index:
- 2336625021
- Predicate type:
- https://slsa.dev/provenance/v1
- Subject digest:
- sha512:af995c2418e3b88d6dc65a1a1eaa841011e079151178a84e2b448d6df8a06626ece37640d6286a6204c8cde09aadc9214130015750437d2de4a447ad8
- Discovery method:
- attestation_endpoint
Dependencies 135 packages
135 packages in the resolved dependency tree · 66 deprecated · 38 stale.
The dependency tree was only partially resolved, so these counts may be incomplete.
The tools this component advertises to a client, with an estimated token cost for each. Expand a tool to see its parameters and schema. The per-tool counts are indicative and are not scored directly; the schema's total context footprint is one signal in Schema Quality & AI Usability.
ckan_analyze_datasets Analyze CKAN Datasets and DataStore Schema ~318
Search datasets and inspect the DataStore schema of queryable resources. For each dataset found, lists all resources. For DataStore-enabled resources, fetches the full field schema (name, type, and label/notes when available) plus total record count — all in one call. Use this before ckan_datastore_search to understand what fields are available and what data to expect. Args: - server_url (string): Base URL of CKAN server - q (string): Solr search query (e.g. "incidenti", "title:ambiente") - rows (number): Max datasets to analyze (default 5, max 20) - response_format ('markdown' | 'json'): Output format Returns: For each dataset: title, ID, organization, and per DataStore resource: field schema with label/notes (when available from DataStore Dictionary) and record count. Typical workflow: ckan_analyze_datasets → ckan_datastore_search (with known field names)
| Name | Type | Req | Description |
|---|---|---|---|
| q | string | yes | Solr search query (e.g. 'incidenti', 'title:ambiente') |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| rows | integer | — | Max datasets to analyze (default 5, max 20) |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.comune.messina.it) |
No output schema declared.
No examples provided.
ckan_catalog_stats Get CKAN Portal Statistics ~215
Get a statistical overview of a CKAN portal: total dataset count and breakdown by category, format, and organization. Single CKAN call (package_search with rows=0 and facets). No query needed. Args: - server_url (string): Base URL of the CKAN server - facet_limit (number): Max entries per facet section (default 20) - response_format ('markdown' | 'json'): Output format Returns: Total dataset count, categories ranked by count, file formats ranked by count, organizations ranked by count. Typical workflow: ckan_catalog_stats (understand the portal) → ckan_package_search (query specific data)
| Name | Type | Req | Description |
|---|---|---|---|
| facet_limit | integer | — | Max entries per facet section (default 20) |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.comune.messina.it) |
No output schema declared.
No examples provided.
ckan_datastore_search Search CKAN DataStore ~622
Query data from a CKAN DataStore resource. The DataStore allows SQL-like queries on tabular data. Not all resources have DataStore enabled. The response always includes a Fields section listing all available column names and types. Use limit=0 to discover column names without fetching data — do this before using filters to avoid guessing column names and getting HTTP 400 errors. Args: - server_url (string): Base URL of CKAN server - resource_id (string): ID of the DataStore resource - q (string): Full-text search query (optional) - filters (object): Key-value filters (e.g., { "anno": 2023 }) - limit (number): Max rows to return (default: 100, max: 32000) - offset (number): Pagination offset (default: 0) - fields (array): Specific fields to return (optional) - sort (string): Sort field with direction (e.g., "anno desc") - distinct (boolean): Return distinct values (default: false) - response_format ('markdown' | 'json'): Output format Returns: DataStore records matching query, always including available column names and types Examples: - { server_url: "...", resource_id: "abc-123", limit: 0 } ← discover columns first - { server_url: "...", resource_id: "abc-123", limit: 50 } - { server_url: "...", resource_id: "...", filters: { "regione": "Sicilia" } } - { server_url: "...", resource_id: "...", sort: "anno desc", limit: 100 } Typical workflow: ckan_package_search → ckan_package_show (find resource_id with datastore_active=true) → ckan_datastore_search (limit=0 to get columns) → ckan_datastore_search (with filters)
| Name | Type | Req | Description |
|---|---|---|---|
| distinct | boolean | — | Return only distinct rows |
| fields | array | — | Specific field names to return; omit to return all fields |
| filters | object | — | Key-value filters for exact matches (e.g., { "regione": "Sicilia", "anno": 2023 }) |
| limit | integer | — | Max rows to return (default 100, max 32000); use 0 to get only column names without data |
| offset | integer | — | Pagination offset |
| q | string | — | Full-text search across all fields |
| resource_id | string | yes | UUID of the DataStore resource (from ckan_package_show resource.id where datastore_active is true) |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.gov.it/opendata) |
| sort | string | — | Sort expression (e.g., 'anno desc', 'nome asc') |
No output schema declared.
No examples provided.
ckan_datastore_search_sql Search CKAN DataStore with SQL ~318
Run SQL queries on a CKAN DataStore resource. This endpoint is only available on CKAN portals with DataStore enabled and SQL access exposed. Args: - server_url (string): Base URL of CKAN server - sql (string): SQL query (e.g., SELECT * FROM "resource_id" LIMIT 10) - response_format ('markdown' | 'json'): Output format Returns: SQL query results from DataStore Examples: - { server_url: "...", sql: "SELECT * FROM "abc-123" LIMIT 10" } - { server_url: "...", sql: "SELECT COUNT(*) AS total FROM "abc-123"" } Typical workflow: ckan_package_show (get resource_id) → ckan_datastore_search_sql (run SQL on it) Security note: SQL queries are forwarded directly to the CKAN DataStore API. The CKAN server enforces its own access controls and read-only permissions. No local database is exposed. Queries are limited to public DataStore resources on the target portal.
| Name | Type | Req | Description |
|---|---|---|---|
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.gov.it/opendata) |
| sql | string | yes | SQL SELECT query; resource_id is the table name, must be double-quoted (e.g., SELECT * FROM "abc-123" LIMIT 10) |
No output schema declared.
No examples provided.
ckan_find_portals Find CKAN Portals ~417
Search the live datashades.info registry of ~950 CKAN portals worldwide. Use this tool to discover which CKAN portals exist for a country, language, or topic before querying them with other CKAN tools. **IMPORTANT — country parameter**: always pass country name in English. If the user writes in another language (e.g. "Italia", "España", "Brasil"), translate to English ("Italy", "Spain", "Brazil") before calling this tool. Args: - country (string): Country name in English (e.g. "Italy", "Brazil", "France") - query (string): Keyword to match against portal title (e.g. "transport", "health") - min_datasets (number): Minimum number of datasets (e.g. 100) - language (string): Portal default locale code (e.g. "it", "en", "pt_BR", "fr") - has_datastore (boolean): If true, return only portals with DataStore enabled (supports SQL queries) - limit (number): Max results to return (default 10, max 50) Returns: Ranked list of matching portals with URL, country, CKAN version, dataset count, and DataStore status. Typical workflow: ckan_find_portals (discover portal URL) → ckan_status_show (verify) → ckan_package_search (search datasets)
| Name | Type | Req | Description |
|---|---|---|---|
| country | string | — | Country name in English (e.g. 'Italy', 'Brazil'). Translate from any language before passing. |
| has_datastore | boolean | — | If true, return only portals with DataStore plugin (required for SQL queries) |
| language | string | — | Portal default locale code (e.g. 'it', 'en', 'pt_BR') |
| limit | integer | — | Max results (default 10, max 50) |
| min_datasets | integer | — | Minimum number of datasets |
| query | string | — | Keyword matched against portal title (case-insensitive) |
No output schema declared.
No examples provided.
ckan_find_relevant_datasets Find Relevant CKAN Datasets ~627
Find and rank datasets by relevance to a query using weighted fields. Use this instead of ckan_package_search when you want relevance-ranked results with explicit scoring across title, notes, tags, and organization fields. Use ckan_package_search instead when you need Solr filter syntax, facets, or pagination. Uses package_search for discovery and applies a local scoring model. Args: - server_url (string): Base URL of CKAN server (e.g., "https://dati.gov.it/opendata") - query (string): Natural language or keyword query (e.g., "mobilità urbana", "air quality") - limit (number): Number of datasets to return (default: 10) - weights (object): Field weights for scoring — higher weight = more influence on rank Default: title=4, tags=3, notes=2, organization=1, holder=4, publisher=2 Note on holder vs organization: on federated catalogs (e.g. dati.gov.it), `organization` is the harvesting catalog (e.g. Regione Puglia), while `holder` (DCAT-AP_IT dct:rightsHolder) is the actual data owner (e.g. Comune di Lecce). Queries like "datasets from a specific Comune" match `holder` correctly; matching only `organization` misses datasets harvested via aggregators. `publisher` (dct:publisher) is scored separately at lower weight as it can contain technical roles ("Redazione OD") rather than the institutional owner. - query_parser ('default' | 'text'): Override search parser behavior - response_format ('markdown' | 'json'): Output format Returns: Ranked datasets with relevance scores and per-field score breakdowns Examples: - { server_url: "https://dati.gov.it/opendata", query: "mobilità" } - { server_url: "...", query: "trasporti", limit: 5, weights: { title: 5, notes: 2 } } - { server_url: "...", query: "defibrillatori Comune di Lecce", weights: { holder: 5 } } Typical workflow: ckan_find_relevant_datasets → ckan_package_show (inspect top results) → ckan_datastore_search (query data)
| Name | Type | Req | Description |
|---|---|---|---|
| limit | integer | — | Number of datasets to return |
| query | string | yes | Natural language or keyword query to match against dataset title, notes, tags, organization, holder and publisher |
| query_parser | string | — | Override search parser ('text' forces text:(...) on non-fielded queries) |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.gov.it/opendata) |
| weights | object | — | Per-field scoring weights; unspecified fields use defaults |
No output schema declared.
No examples provided.
ckan_get_mqa_quality Get MQA Quality Score ~163
Get MQA (Metadata Quality Assurance) quality metrics for a dataset on dati.gov.it. Returns quality score and detailed metrics (accessibility, reusability, interoperability, findability, contextuality) from data.europa.eu. Only works with dati.gov.it server. Typical workflow: ckan_package_show (get dataset ID) → ckan_get_mqa_quality → ckan_get_mqa_quality_details (for non-max dimensions)
| Name | Type | Req | Description |
|---|---|---|---|
| dataset_id | string | yes | Dataset ID or name |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of dati.gov.it (e.g., https://www.dati.gov.it/opendata) |
No output schema declared.
No examples provided.
ckan_get_mqa_quality_details Get MQA Quality Details ~150
Get detailed MQA (Metadata Quality Assurance) quality reasons for a dataset on dati.gov.it. Returns dimension scores, non-max reasons, and raw MQA flags from data.europa.eu. Only works with dati.gov.it server. Typical workflow: ckan_get_mqa_quality (get overview scores) → ckan_get_mqa_quality_details (inspect failing metrics)
| Name | Type | Req | Description |
|---|---|---|---|
| dataset_id | string | yes | Dataset ID or name |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of dati.gov.it (e.g., https://www.dati.gov.it/opendata) |
No output schema declared.
No examples provided.
ckan_group_list List CKAN Groups ~313
List all groups on a CKAN server. Groups are thematic collections of datasets. Args: - server_url (string): Base URL of CKAN server - all_fields (boolean): Return full objects vs just names (default: false) - sort (string): Sort field (default: "name asc") - limit (number): Maximum results (default: 100). Use 0 to get only the count via faceting - offset (number): Pagination offset (default: 0) - response_format ('markdown' | 'json'): Output format Returns: List of groups with metadata. When limit=0, returns only the count of groups with datasets. Typical workflow: ckan_group_list → ckan_group_show (inspect one) → ckan_package_search with fq="groups:name" (browse its datasets)
| Name | Type | Req | Description |
|---|---|---|---|
| all_fields | boolean | — | Return full group objects (true) or just name slugs (false) |
| limit | integer | — | Max groups to return. Use 0 to get only the count via faceting |
| offset | integer | — | Pagination offset |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.gov.it/opendata) |
| sort | string | — | Sort field and direction (e.g., 'name asc', 'package_count desc') |
No output schema declared.
No examples provided.
ckan_group_search Search CKAN Groups by Name ~206
Search for groups by name pattern. This tool provides a simpler interface than package_search for finding groups. Wildcards are automatically added around the search pattern. Args: - server_url (string): Base URL of CKAN server - pattern (string): Search pattern (e.g., "energia", "salute") - response_format ('markdown' | 'json'): Output format Returns: List of matching groups with dataset counts Typical workflow: ckan_group_search → ckan_group_show (get details) → ckan_package_search with fq="groups:name"
| Name | Type | Req | Description |
|---|---|---|---|
| pattern | string | yes | Name pattern to search for (wildcards added automatically, e.g., 'energia', 'salute') |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.gov.it/opendata) |
No output schema declared.
No examples provided.
ckan_group_show Show CKAN Group Details ~208
Get details of a specific group. Args: - server_url (string): Base URL of CKAN server - id (string): Group ID or name - include_datasets (boolean): Include list of datasets (default: true) - response_format ('markdown' | 'json'): Output format Returns: Group details with optional datasets Typical workflow: ckan_group_show → ckan_package_show (inspect a dataset) → ckan_datastore_search (query its data)
| Name | Type | Req | Description |
|---|---|---|---|
| id | string | yes | Group ID (UUID) or machine-readable name slug (e.g., 'transport', 'energia') |
| include_datasets | boolean | — | Include the list of datasets belonging to this group |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.gov.it/opendata) |
No output schema declared.
No examples provided.
ckan_list_resources List CKAN Dataset Resources ~448
List all resources in a dataset with a compact summary. Returns a focused table of resources showing format, size, DataStore availability, and download URL. Use this to quickly assess what files a dataset contains before deciding how to access the data. Args: - server_url (string): Base URL of CKAN server - id (string): Dataset ID or name - format_filter (string): Filter resources by format, case-insensitive (e.g., "CSV", "json", "XLSX") - response_format ('markdown' | 'json'): Output format Returns: Compact resource summary with name, ID, format, size, DataStore flag, and URL Examples: - { server_url: "https://dati.gov.it/opendata", id: "dataset-name" } - { server_url: "...", id: "dataset-name", format_filter: "CSV" } Typical workflow: ckan_package_search → ckan_list_resources (assess available files) → ckan_datastore_search (for resources with DataStore=true) When a resource has DataStore=false but its download URL belongs to a different (source) portal, the tool can probe the source portal for DataStore availability and report source_datastore_active and source_portal_url so you can query the data there instead. This probing is OFF by default (it issues extra HTTP requests to hosts taken from the dataset's own resource URLs); set check_source_portal=true to enable it.
| Name | Type | Req | Description |
|---|---|---|---|
| check_source_portal | boolean | — | Opt-in (default false): when true, probes the source portal for DataStore availability when a resource URL points to a different CKAN instance. Issues extra HTTP requests to hosts taken from the data… |
| format_filter | string | — | Filter resources by format, case-insensitive (e.g., 'CSV', 'json', 'XLSX') |
| id | string | yes | Dataset ID or name |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server |
No output schema declared.
No examples provided.
ckan_organization_list List CKAN Organizations ~318
List all organizations on a CKAN server. Organizations are entities that publish and manage datasets. Args: - server_url (string): Base URL of CKAN server - all_fields (boolean): Return full objects vs just names (default: false) - sort (string): Sort field (default: "name asc") - limit (number): Maximum results (default: 100). Use 0 to get only the count via faceting - offset (number): Pagination offset (default: 0) - response_format ('markdown' | 'json'): Output format Returns: List of organizations with metadata. When limit=0, returns only the count of organizations with datasets. Typical workflow: ckan_organization_list → ckan_organization_show (inspect one) → ckan_package_search with fq="organization:name" (browse its datasets)
| Name | Type | Req | Description |
|---|---|---|---|
| all_fields | boolean | — | Return full organization objects (true) or just name slugs (false) |
| limit | integer | — | Max organizations to return. Use 0 to get only the count via faceting |
| offset | integer | — | Pagination offset |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.gov.it/opendata) |
| sort | string | — | Sort field and direction (e.g., 'name asc', 'package_count desc') |
No output schema declared.
No examples provided.
ckan_organization_search Search CKAN Organizations by Name ~259
Search for organizations by name pattern. This tool provides a simpler interface than package_search for finding organizations. Wildcards are automatically added around the search pattern. Args: - server_url (string): Base URL of CKAN server - pattern (string): Search pattern (e.g., "toscana", "salute") - response_format ('markdown' | 'json'): Output format Returns: List of matching organizations with dataset counts Examples: - { server_url: "https://www.dati.gov.it/opendata", pattern: "toscana" } - { server_url: "https://catalog.data.gov", pattern: "health" } Typical workflow: ckan_organization_search → ckan_organization_show (get details) → ckan_package_search with fq="organization:name"
| Name | Type | Req | Description |
|---|---|---|---|
| pattern | string | yes | Name pattern to search for (wildcards added automatically, e.g., 'toscana', 'health') |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.gov.it/opendata) |
No output schema declared.
No examples provided.
ckan_organization_show Show CKAN Organization Details ~247
Get details of a specific organization. Args: - server_url (string): Base URL of CKAN server - id (string): Organization ID or name - include_datasets (boolean): Include list of datasets (default: true) - include_users (boolean): Include list of users (default: false) - response_format ('markdown' | 'json'): Output format Returns: Organization details with optional datasets and users Typical workflow: ckan_organization_show → ckan_package_show (inspect a dataset) → ckan_datastore_search (query its data)
| Name | Type | Req | Description |
|---|---|---|---|
| id | string | yes | Organization ID (UUID) or machine-readable name slug (e.g., 'regione-siciliana') |
| include_datasets | boolean | — | Include the list of datasets published by this organization |
| include_users | boolean | — | Include the list of users belonging to this organization |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.gov.it/opendata) |
No output schema declared.
No examples provided.
ckan_package_search Search CKAN Datasets ~2,182
Search for datasets (packages) on a CKAN server using Solr query syntax. Supports full Solr search capabilities including filters, facets, and sorting. Use this to discover datasets matching specific criteria. Note on parser behavior: Some CKAN portals use a restrictive default query parser that can break long OR queries. For those portals, this tool may force the query into 'text:(...)' based on per-portal config. You can override with 'query_parser' to force or disable this behavior per request. Important - Date field semantics: - issued: publisher's content publish date when available (best proxy for "created/published") - modified: publisher's content update date when available - metadata_created: CKAN record creation timestamp (publish time on source portals, harvest time on aggregators; fallback for "created" if issued missing) - metadata_modified: CKAN record update timestamp (publish time on source portals, harvest time on aggregators; use for "updated/modified in last X") Natural language mapping (important for tool callers): - "created"/"published" -> prefer issued; fallback to metadata_created - "updated"/"modified" -> prefer modified; fallback to metadata_modified - For "recent in last X", consider using content_recent (issued with metadata_created fallback) Content-recent helper: - content_recent: if true, rewrites the query to use issued with a fallback to metadata_created when issued is missing. - content_recent_days: window for content_recent (default 30 days). Args: - server_url (string): Base URL of CKAN server (e.g., "https://dati.gov.it/opendata") - q (string): Search query using Solr syntax (default: "*:*" for all) - fq (string): Filter query (e.g., "organization:comune-palermo") IMPORTANT — Solr fq syntax rules: 1. OR inside a single field: use field:(val1 OR val2), NOT field:val1 OR field:val2. Wrong: fq=type:"A" OR type:"B" → silently ignored, returns entire catalog. Right: fq…
| Name | Type | Req | Description |
|---|---|---|---|
| content_recent | boolean | — | Use issued date with fallback to metadata_created for recent content |
| content_recent_days | integer | — | Day window for content_recent (default 30) |
| facet_field | array | — | Fields to facet on |
| facet_limit | integer | — | Maximum facet values per field |
| fq | string | — | Filter query in Solr syntax; applied after scoring, does not affect relevance. CKAN extras fields use prefix 'extras_' (e.g. extras_hvd_category). For OR on same field use field:(val1 OR val2), never… |
| include_drafts | boolean | — | Include draft datasets |
| page | integer | — | Page number (1-based); alias for start. Overrides start if provided. |
| page_size | integer | — | Results per page when using page (default: 10) |
| q | string | — | Search query in Solr syntax |
| query_parser | string | — | Override search parser ('text' forces text:(...) on non-fielded queries) |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| rows | integer | — | Number of results to return |
| server_url | string | yes | Base URL of the CKAN server |
| sort | string | — | Sort field and direction (e.g., 'metadata_modified desc') |
| start | integer | — | Offset for pagination |
No output schema declared.
No examples provided.
ckan_package_show Show CKAN Dataset Details ~418
Get complete metadata for a specific dataset (package). Returns full details including resources, organization, tags, and all metadata fields. Notes: - metadata_modified is a CKAN record timestamp (publish time on source portals, harvest time on aggregators), not the content date. - issued/modified are content dates when provided by the publisher. - JSON output adds metadata_harvested_at (same as metadata_modified). Args: - server_url (string): Base URL of CKAN server - id (string): Dataset ID or name (machine-readable slug) - include_tracking (boolean): Include view/download statistics (default: false) - response_format ('markdown' | 'json'): Output format Returns (JSON format): id, name, title, notes, organization, tags, state, license_title, metadata_created, metadata_modified, issued, modified, author, maintainer, frequency, language, publisher_name, holder_name, hvd_category, applicable_legislation, resources (id, name, format, url, size, datastore_active, created, last_modified, api_json_url), view_url, api_json_url Examples: - { server_url: "https://dati.gov.it/opendata", id: "dataset-name" } - { server_url: "...", id: "abc-123-def", include_tracking: true } Typical workflow: ckan_package_show → pick a resource with datastore_active=true → ckan_datastore_search (query its data)
| Name | Type | Req | Description |
|---|---|---|---|
| id | string | yes | Dataset ID (UUID) or machine-readable name slug (e.g., 'raccolta-differenziata-comuni') |
| include_tracking | boolean | — | Include tracking statistics |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.gov.it/opendata) |
No output schema declared.
No examples provided.
ckan_status_show Check CKAN Server Status ~124
Check if a CKAN server is available and get version information. Useful to verify server accessibility before making other requests. Also shows the count of High-Value Datasets (HVD) when the portal supports it. Args: - server_url (string): Base URL of CKAN server Returns: Server status, version information, and HVD dataset count (if available) Typical workflow: ckan_status_show (verify server is up) → ckan_package_search (discover datasets)
| Name | Type | Req | Description |
|---|---|---|---|
| server_url | string | yes | Base URL of the CKAN server |
No output schema declared.
No examples provided.
ckan_tag_list List CKAN Tags ~348
List tags from a CKAN server using faceting. This returns tag names with counts, optionally filtered by dataset query or tag substring. Args: - server_url (string): Base URL of CKAN server - q (string): Dataset search query (default: "*:*") - fq (string): Filter query (optional) - tag_query (string): Filter tags by substring (optional) - limit (number): Max tags to return (default: 100, max: 1000) - response_format ('markdown' | 'json'): Output format Returns: List of tags with counts (from faceting) Typical workflow: ckan_tag_list → ckan_package_search with fq="tags:tag_name" (find datasets by tag) → ckan_package_show
| Name | Type | Req | Description |
|---|---|---|---|
| fq | string | — | Filter query in Solr syntax (e.g., 'organization:comune-palermo') to restrict which datasets contribute to tag counts |
| limit | integer | — | Max tags to return (default 100, max 1000); tags are sorted by count descending |
| q | string | — | Dataset search query in Solr syntax to scope the tag facet (default: '*:*' for all datasets) |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
| server_url | string | yes | Base URL of the CKAN server (e.g., https://dati.gov.it/opendata) |
| tag_query | string | — | Substring filter applied to tag names after faceting (e.g., 'acqua' to keep only tags containing 'acqua') |
No output schema declared.
No examples provided.
sparql_query SPARQL Query ~363
Execute a SPARQL SELECT query against any public HTTPS SPARQL endpoint. Useful for querying open data portals and knowledge graphs that expose SPARQL endpoints, including: - data.europa.eu (European open data portal) - publications.europa.eu (EU Publications Office) - DBpedia, Wikidata - Any DCAT-AP compliant data catalog Only HTTPS endpoints are allowed. Queries timeout after 15 seconds. Only SELECT queries are supported (read-only). If the query does not contain a LIMIT clause, one is injected automatically (default: 25, max: 1000). Args: - endpoint_url (string): HTTPS URL of the SPARQL endpoint - query (string): SPARQL SELECT query to execute - limit (number): Max rows to return (default: 25). Ignored if query already contains LIMIT. - response_format ('markdown' | 'json'): Output format Examples: - Count Italian HVD datasets by publisher on data.europa.eu - Query Wikidata for entities related to a dataset topic - Explore EU controlled vocabularies on publications.europa.eu Typical workflow: sparql_query (explore schema) → sparql_query (targeted query) → ckan_package_search (get dataset details)
| Name | Type | Req | Description |
|---|---|---|---|
| endpoint_url | string | yes | HTTPS URL of the SPARQL endpoint |
| limit | integer | — | Max rows to return (default: 25, max: 1000). Injected as SPARQL LIMIT if not already present in query. |
| query | string | yes | SPARQL SELECT query to execute |
| response_format | string | — | Output format: 'markdown' for human-readable or 'json' for machine-readable |
No output schema declared.
No examples provided.