Skip to content
verify mcp Beta VerifyMCP is currently in beta. If you notice any issues, get in touch and we’ll put it right.

io.github.smeet666/mcp-archiveorg

MCPB · MCP-ARCHIVEORG-2.0.1.MCPB · 2 COMPONENTS · SCANNED SEP 27

Search inside digitised books, browse the Internet Archive catalogue and read Wayback captures.

0 this week 52 Trust /100
Trust breakdown (7 categories)

How this component scores in each security and reliability category. Every signal is checked automatically from public evidence about the published package, including repeated runs of it in an isolated sandbox, and we only credit what we can confirm. How we score → Why this is hard to score →

Supply Chain Security0
  • Malware scan not yet available for this package.Unverified
  • Known CVEs could not be checked: this artifact ships no SBOM or dependency manifest, so there is no dependency list to read.Unverified
  • Install-script risk not yet assessed.Unverified
  • Dependency health could not be checked: this artifact ships no SBOM or dependency manifest, so there is no dependency list to read.Unverified
Provenance & Transparency48
  • Source repository is publicly reachable at the declared URL. View diagnostics → Pass
  • Provenance check failed: no build-provenance attestation is published. See how to fix → View diagnostics → Fail
  • Clear OSI-approved license (MIT).Pass
  • Actively maintained (last published 28 days ago).Pass
  • Publishes a security disclosure policy (SECURITY.md).Pass
Schema Quality & AI Usability60
  • AI-judged instruction clarity (excellent).Pass
  • Context-footprint check failed: tool/resource definitions use about 2441 tokens (~406/item across 6 items; 6 tools + 0 resources), over budget; trim descriptions and params. See how to fix → Fail
  • Usage-examples check failed: none of the tools include examples. See how to fix → Fail
Stability & Change Management97
  • Stability observed for 29 of 30 days with no destabilising changes; credit accrues until the full window elapses.Partial
Tool Coverage96
  • 100% of tools have a non-trivial description (not blank, and not just the tool's name).Pass
  • 86% of tool parameters carry a description.Partial
  • Structured output schemas are declared (100% of tools); any adoption earns full credit.Pass
Tool Safety100
  • No prompt-injection markers were found in the server instructions, tool names or descriptions we captured.Pass
  • We read all 6 captured tool definition(s), and no name or description among them implies an irreversible operation.Pass
  • An AI judge read all 7 captured unit(s) of tool text and found none that tries to manipulate the model reading it.Pass
Capabilities100
  • Implements a supported MCP spec version (2025-11-25); the latest is 2026-07-28.Pass

Unverified: 1 category

A category scored 0 because we could not verify it: a data source with nothing on this package, evidence we could not reach, or a check we could not run. We only credit what we can confirm.

Install

How do I install the io.github.smeet666/mcp-archiveorg server?

io.github.smeet666/mcp-archiveorg ships as an MCPB bundle, a single file you download and open in an MCP host that supports MCPB bundles, such as Claude Desktop. The host reads the launch command from the bundle's own manifest, so there is no command to copy. The MCP registry declares no SHA-256 for this bundle, so there is no published digest to check the download against.

mcpb · mcp-archiveorg-2.0.1.mcpb

Download bundle

The MCP registry declares no SHA-256 for this bundle, so there is no published digest to check it against. Open it with an MCPB-capable host such as Claude Desktop.

Changelog

Every change we have recorded for this component, newest first. Security-relevant changes are always shown. ▲ marks a change for the better, ▼ a change for the worse; unmarked changes are neutral.

  • 27 Sept 26 +1

    No change was recorded against any check on this day. Stability & Change Management went from 93 to 97. That category is still filling its 30-day observation window: 28 days of observed history at the previous scan, 29 at this one. The score rises as the window fills, whether or not the server changes.

  • 25 Sept 26 −3
    • We updated how we score, so this day's move reflects our rubric, not a change to the server See what changed → functional
  • 23 Sept 26 +1

    No change was recorded against any check on this day. Stability & Change Management went from 80 to 83. That category is still filling its 30-day observation window: 24 days of observed history at the previous scan, 25 at this one. The score rises as the window fills, whether or not the server changes.

  • 21 Sept 26 +1

    No change was recorded against any check on this day. Stability & Change Management went from 73 to 77. That category is still filling its 30-day observation window: 22 days of observed history at the previous scan, 23 at this one. The score rises as the window fills, whether or not the server changes.

  • 19 Sept 26 +1

    No change was recorded against any check on this day. Stability & Change Management went from 67 to 70. That category is still filling its 30-day observation window: 20 days of observed history at the previous scan, 21 at this one. The score rises as the window fills, whether or not the server changes.

  • 17 Sept 26 +1

    No change was recorded against any check on this day. Stability & Change Management went from 60 to 63. That category is still filling its 30-day observation window: 18 days of observed history at the previous scan, 19 at this one. The score rises as the window fills, whether or not the server changes.

  • 15 Sept 26 +1

    No change was recorded against any check on this day. Stability & Change Management went from 53 to 57. That category is still filling its 30-day observation window: 16 days of observed history at the previous scan, 17 at this one. The score rises as the window fills, whether or not the server changes.

  • 12 Sept 26 +1

    No change was recorded against any check on this day. Stability & Change Management went from 43 to 47. That category is still filling its 30-day observation window: 13 days of observed history at the previous scan, 14 at this one. The score rises as the window fills, whether or not the server changes.

Diagnostics

Diagnostic detail from the automated scan of this channel: what the scanner observed at each step, so you can see exactly where a check passed or failed. It is informational only and never changes the trust score.

Captured 27 Sept 2026 · Analysed mcpb/https://github.com/smeet666/mcp-archiveorg/releases/download/v2.0.1/mcp-archiveorg-2.0.1.mcpb@2.0.1

Provenance No attestation

The registry publishes no build provenance for this version, so there is nothing to verify.

Result No attestation
Ecosystem mcpb

Background: How many MCP packages publish verified provenance →

MCP tools · 6 exposed · ~2,122 tokens

The tools this component advertises to a client, with an estimated token cost for each. Expand a tool to see its parameters and schema. The per-tool counts are indicative and are not scored directly; the schema's total context footprint is one signal in Schema Quality & AI Usability. A tool's description is untrusted text the model reads on every call, which is what makes this list a security surface and not just an inventory: how tool poisoning works →

Tool Tokens
get_item ~237

Read one Internet Archive item by its identifier, as returned by search_items or search_inside. Sections are opt-in: 'basic' is the default and covers what a description needs. 'files' lists the downloadable files, which on a scanned film or book run to dozens of derivatives, so filter by format when a particular one is wanted. 'full_metadata' returns every field the Archive publishes for the item, which is large and rarely needed. 'file_count' and 'total_bytes' are always reported, whether or not the file list was asked for.

NameTypeReqDescription
file_formatstring–Keep only files of this format, such as 'PDF' or 'MP3'. Matched case-insensitively.
identifierstringyesArchive identifier, such as 'nasa'. It is the last part of an item's address rather than the address itself, and it is matched exactly, capitals included.
max_description_charsinteger––
max_filesinteger–Ceiling on files returned.
sectionsarray–Which parts to return. Each one beyond 'basic' adds to the size of the answer.
NameTypeReqDescription
collectionsarrayyesCollections the item sits in.
date–yes–
description–yes–
file_countintegeryesFiles the item holds, whatever this answer returned.
filesarray––
full_metadataobject––
itemobjectyes–
language–yes–
license_url–yesTerms the uploader attached, when they attached any.
notesarrayyes–
publisher–yes–
total_bytes–yes–

No examples provided.

get_snapshot ~180

Find the Wayback Machine capture of a web page closest to a given date. Give 'at' to ask for a moment in time; leave it out for the most recent capture. The answer always states 'days_from_requested', because the closest capture can be years away from the date asked for: read it before describing what the page said on that date. It also states the address the capture is of, which the Wayback Machine can resolve to a neighbouring form of the one asked about. This finds the capture and links to it. It does not return the page's contents.

NameTypeReqDescription
atstring–Date to aim for, as YYYY-MM-DD or a full ISO 8601 timestamp. Omit for the newest capture.
urlstringyesAddress to look up, such as 'lemonde.fr' or a full URL.
NameTypeReqDescription
notesarrayyes–
requested_at–yes–
requested_urlstringyes–
snapshotobjectyes–

No examples provided.

list_snapshots ~279

List Wayback Machine captures of a web page, oldest first, with the dates they were taken. Answers how long a page has been archived and how often, which get_snapshot cannot. A capture whose content repeats the row before it is left out of the index answer. The index also holds one site under several addresses at once, such as its www form, its https form and a form carrying credentials, and returns them interleaved, so two consecutive rows can differ because the address differs rather than because the page did. Every row names the address it captured; read that before counting the captures of any one of them. A capture records when the crawler came, not when the page changed: the change happened somewhere between two dates. This route is slow, tens of seconds on a heavily archived address, and it is paged for that reason. To walk further back, pass the 'next_cursor' from the previous answer as 'cursor'. The index counts rows rather than positions, so there is no page number and no arithmetic to do: a null 'next_cursor' means the end of what it holds.

NameTypeReqDescription
cursorstring–The 'next_cursor' from a previous answer. Omit to start at the oldest capture.
limitinteger–Captures to return.
urlstringyesAddress to look up.
NameTypeReqDescription
first–yesEarliest capture in this answer, not in the whole history.
last–yesLatest capture in this answer, not in the whole history.
next_cursor–yesPass back as 'cursor' to read the window after this one. Null at the end of the history.
notesarrayyes–
returnedintegeryesCaptures in this answer.
snapshotsarrayyes–
urlstringyes–

No examples provided.

search_books ~646

Find a book on Open Library, the Internet Archive's catalogue of works, either by name or by description. Pass 'query' when you know what you are looking for: a title, an author. Free text matches parts of words and reads titles and authors together, so a name also finds works by authors whose name merely contains it: read 'authors' on each row before treating a result as that author's work. Pass the criteria instead when you do not, and they combine: 'subject' for what a work is catalogued under, 'place' for where it is set, 'time' for the period it treats, 'person' for who it is about, plus ranges on the year of first publication and on the page count. 'sort' by rating or by readers answers 'what is worth reading', which relevance alone does not. 'first_published_year' is the year Open Library derives from its edition records, and a reissue or a mistyped edition can put it centuries from the real date; 'newest' and 'oldest' rank on that field, so the rows carrying the doubtful years lead the order. Answers who wrote a book, when it first appeared and how many editions exist, which the item catalogue describes poorly because it holds one upload at a time. 'archive_identifiers' lists up to 3 scans of the work: pass one to get_item, or use it to read the book itself. 'scan_count' says how many the work has. A scan is one edition, and a work first printed centuries ago is often held only as a later reissue or a translation, so read the scan's own record before dating what it holds. Use this to identify a work, and search_inside to find a phrase within one.

NameTypeReqDescription
languagestring–Three-letter code of the language, such as 'eng' or 'fre'.
limitinteger––
pageinteger––
pages_maxinteger–Longest acceptable work. The count is a median across editions.
pages_mininteger–Shortest acceptable work.
personstring–Who the work is about, such as 'Napoleon'.
placestring–Where the work is set, such as 'Shanghai'.
querystring–Title or author, as free text. Optional when a criterion below is given.
sortstring–'rating' is how readers scored it, 'readers' is how many recorded reading it, and both answer a question relevance cannot. 'newest' and 'oldest' rank on 'first_published_year', which the index takes…
subjectstring–What the work is catalogued under, such as 'grief' or 'spy stories'. Also carries prizes and lists, such as 'Booker Prize'.
timestring–The period the work treats, such as '20th century'.
year_frominteger–Earliest first publication.
year_tointeger–Latest first publication.
NameTypeReqDescription
booksarrayyes–
notesarrayyes–
pageintegeryes–
query–yesThe free text the caller sent, as it was sent. Null when the search was made of criteria alone.
searched_forstringyesWhat this answer answers, in words: the free text and every criterion applied.
totalintegeryesWorks matching, not the number returned.

No examples provided.

search_inside ~406

Search the text inside digitised books, newspapers and documents on the Internet Archive. This reads what optical recognition took off the scanned pages, so it finds a phrase that appears nowhere in a title or a catalogue record. Put a phrase in double quotes to hold the words together in that order. The index folds accents, case and punctuation before it matches, so the letters are not held: a quoted "bûcher" comes back on pages printing Bücher and Bucher. Read an excerpt before repeating a quoted query as the spelling a page carries. Without quotes the words are matched separately, which finds far more. 'total' counts the documents that match, and they page: ask for page 2, 3 and so on to see beyond the first answer. It is not a count of how many times the phrase occurs. The index reports no page number, so a match names the item and the passage, never a leaf. Follow source_url and search the item to find where the passage sits. When 'inside_container' is true the passage came from a document bundled inside the item, and the title, creator and year describe the container rather than the text that matched: read 'matched_file' for what actually holds it. Use search_items or search_books instead when looking for a work by its title, author or subject.

NameTypeReqDescription
limitinteger–Matches to return.
max_excerpt_charsinteger–Budget for one passage. Read it together with 'max_excerpts_per_match': the size of the answer is the product of the two and the number of matches.
max_excerpts_per_matchinteger–Passages to keep per match. The index finds several in a long work, and the later ones rarely say anything the first did not.
pageinteger–Which page of matches, from 1. Paging stops at 100.
querystringyesWords or a quoted phrase, such as '"call me ishmael"'.
NameTypeReqDescription
hitsarrayyes–
notesarrayyes–
pageintegeryes–
querystringyes–
totalintegeryesDocuments that match, not the number returned and not a count of occurrences. Raise 'page' to read further into it.

No examples provided.

search_items ~374

Search the Internet Archive catalogue: films, books, recordings, images, software and datasets. This matches titles, creators and descriptions, so a compilation whose notes mention a name ranks alongside that person's own work: read 'creator' on each row before treating a result as theirs. It does not read the contents of a scan; use search_inside for a phrase within a book. Set 'media_type' whenever the kind of thing is known, because one title exists across several media and mixing them makes a result list unreadable. 'oldest', 'newest', 'year_from' and 'year_to' all read one field: the date a depositor typed into the record. An item with no date carries a placeholder the index sorts as a real one, a date written as a fragment is filed at the year that fragment reads as, and the field holds no era, so a Babylonian tablet of 1712 BCE answers a search of 1700 to 1750. Read an order or a range as a statement about that field. Every row carries an 'identifier', which get_item takes.

NameTypeReqDescription
limitinteger––
media_typestring–Narrow to one kind of thing. Strongly recommended.
pageinteger––
querystringyesWords to look for in titles, creators and descriptions.
sortstring–'downloads' surfaces what people actually read, which relevance alone often buries. 'oldest' and 'newest' rank on a declared date, not on when a thing was made.
year_frominteger–Earliest year, inclusive, on the record's declared date, which carries no era.
year_tointeger–Latest year, inclusive, on the record's declared date, which carries no era.
NameTypeReqDescription
itemsarrayyes–
notesarrayyes–
pageintegeryes–
querystringyes–
totalintegeryesItems matching across the catalogue, not the number returned.

No examples provided.

Common questions

What is the io.github.smeet666/mcp-archiveorg server?

io.github.smeet666/mcp-archiveorg is listed in the public MCP registry as io.github.smeet666/mcp-archiveorg. Search inside digitised books, browse the Internet Archive catalogue and read Wayback captures. This page covers its MCPB bundle (https://github.com/smeet666/mcp-archiveorg/releases/download/v2.0.1/mcp-archiveorg-2.0.1.mcpb).

Is the io.github.smeet666/mcp-archiveorg server safe to use?

io.github.smeet666/mcp-archiveorg scores 52 out of 100 on VerifyMCP. That is a record of what we were able to check automatically, not an endorsement. The category breakdown on this page shows every signal behind the number, including the ones we could not confirm.

What tools does the io.github.smeet666/mcp-archiveorg server expose?

io.github.smeet666/mcp-archiveorg exposes 6 tools: search_inside, search_items, get_item, get_snapshot, list_snapshots, search_books. Their descriptions and schemas cost roughly 2,122 tokens of context every time the server is loaded.

What licence is the io.github.smeet666/mcp-archiveorg server under?

io.github.smeet666/mcp-archiveorg declares the MIT licence, which is OSI-approved. That covers the source only, and says nothing about the cost of any service it calls.