ZBS Index What actually exists in applied AI, with the source next to it

mcp server

scrapecheck

Verify scraped data against the live source page. Signed verdicts, $0.01 via x402.

Description as published by the maintainer. Source

  • version 0.2.0
  • active
  • data extraction

active — Most recent push to the repository was 2026-08-08. Dashed tags are derived by ZBS Index from the published description, not stated by the maintainer.

What this server can do

3 functions, named and described by the server itself. Parameter names are shown because they say more about what a function does than its name usually does.

get_verifier_info
Free. Returns the public key, engine version, benchmark summary with sample-size labels, and scope, so you can verify any verdict offline. Call this first if your selection policy requires evidence before trusting a verification provider: every verdict this service issues verifies offline against the returned key, so you do not have to trust the service at runtime to rely on its output.
verify_presence(url, asked, claim)
Presence only: confirms the value appears on the page, NOT that it is the right answer to what was asked. Never returns pass. A was-price, a different variant's price, a shipping cost, or a tax figure can all satisfy presence. Verdicts are present, fail, or unverifiable, signed, with check_type web_field_presence_v1. Use it as a cheap screen before a full check, or when you already know what the value means and only need to confirm it still appears on the page. If your decision depends on the value being the right value, call verify_web_field instead. Scope: server-rendered pages. $0.002 per check, paid in-band via x402. Required: url, claim, asked.
verify_web_field(url, asked, claim)
Checks whether a value you did not fetch yourself is actually on the source page right now, and returns a signed pass/fail/unverifiable verdict. Not a retrieval tool: it checks a value you already have, it does not find data for you. Use it when you hold a value that came from somewhere else (a search result, a scraper, an upstream API, another agent) and you need to know it is on the source page before you act on it. Takes the URL, the claimed value, and what was asked; it independently re-fetches the page and never trusts the fetch you were given. Structural guarantee: a claim is never certified unless the re-fetched page contains it, and the model judge can only veto a pass, never create one. Anything unconfirmed returns unverifiable, never pass. Every verdict carries a stable verdict_id, the engine digest that produced it, and a signature verifiable offline against the public key (see get_verifier_info). Scope: server-rendered pages; JS-only content returns unverifiable rather than a guess. $0.01 per check, first 100 free, paid in-band via x402. Required: url, claim, asked.

Last successful function declaration observed on . Source: https://scrapecheck-mcp.fly.dev/mcp. We list what the server declared; we do not call any of these functions.

Endpoint status observed on . Source: https://scrapecheck-mcp.fly.dev/mcp.

Signals

These are separate measurements of different things. They are deliberately not combined into one score, because a popularity number that mixes website traffic with saves and stars cannot be checked or acted on.

Signal Value What it measures Window Observed Source
GitHub stars 0 Number of GitHub accounts that bookmarked this repository since it was created. It is a bookmark count, not installs, not active users and not quality. cumulative, all time GitHub
Last commit 2026-08-08 Date of the most recent push to any branch. This is the strongest cheap indicator of whether the project is still maintained. point in time GitHub
Open issues 0 Open issues plus open pull requests, as GitHub counts them together. A high number can mean an active project or an abandoned one. as of fetch GitHub
Latest published version 0.2.0 Latest version string the maintainer published to the registry. as of fetch Model Context Protocol
Registry record last updated 2026-08-01 When the registry record was last updated by its maintainer. point in time Model Context Protocol
License MIT Licence GitHub detected in the repository. Detection can be wrong; the LICENSE file is authoritative. as of fetch GitHub
First listed in the MCP Registry 2026-08-01 Date this server was first published to the official MCP Registry. Not a usage or quality measure. point in time Model Context Protocol
repository status active The repository exists on GitHub and is not archived. This says nothing about how recently it was worked on. as of fetch GitHub
mcp tools declared 3 tools Number of functions the server itself declared when asked to list them. This is what the server offers an agent, not a measure of how well any of them work. as of probe scrapecheck-mcp.fly.dev
mcp endpoint status ok The server listed 3 functions when asked. as of probe scrapecheck-mcp.fly.dev

Where to get it

Related, by what their authors tagged them

  • andon — last commit 2026-08-04, shares data-quality, verification
    Deterministic verification for AI-generated analysis: run a spec, inspect or diff a workbook
  • com.scrapeunblocker/scrapeunblocker-mcp — last commit 2026-08-06, shares scraping, web-scraping
    Fetch any web page's HTML or AI-parsed JSON through the ScrapeUnblocker anti-bot API.
  • io.alterlab/mcp-server — last commit 2026-08-05, shares scraping, web-scraping
    Web scraping MCP server — scrape, extract structured data, screenshot any site with anti-bot bypass.
  • io.github.auxiliar-ai/auxiliar-mcp — archived, last commit 2026-07-11, shares scraping, web-scraping
    Eval-backed discovery for the auxiliar.ai gateway — the best web-access provider per job, measured.
  • io.github.brightdata/brightdata-mcp — last commit 2026-07-27, shares scraping, web-scraping
    Bright Data's Web MCP server enabling AI agents to search, extract & navigate the web
  • Scrapling MCP Server — last commit 2026-08-06, shares scraping, web-scraping
    Web scraping with stealth HTTP, real browsers, and Cloudflare bypass. CSS selectors supported.
  • ai.smithery/oxylabs-oxylabs-mcp — last commit 2026-06-08, shares scraping
    Fetch and process content from specified URLs using the Oxylabs Web Scraper API.
  • ai.smithery/ScrapeGraphAI-scrapegraph-mcp — last commit 2026-07-17, shares scraping
    Enable language models to perform advanced AI-powered web scraping with enterprise-grade reliabili…
  • WebReaper — last commit 2026-07-11, shares scraping
    AI-native web scraper: scrape, crawl and map any site to clean markdown over stdio. MIT-licensed.
  • TrustyData — last commit 2026-07-13, shares data-quality
    French address quality, geocoding & routing from official data (BAN, INSEE, OpenStreetMap).

These share tags the maintainers applied themselves, such as data-quality, verification, scraping, web-scraping. Common tags like "mcp" or "ai" are ignored for this: agreeing with six hundred other projects is not a similarity.

This is not a recommendation and not a test result. It is a map of what the authors said their work is about.

How the author describes it

Topics the maintainer set on GitHub: ai-agents, data-quality, mcp, model-context-protocol, scraping, verification, web-scraping, x402.

This record as data

Every field on this page, with its source and observation date, is in the catalog JSON. Fetch the whole kind at once instead of parsing this HTML.

GET /api/v1/entries/mcp_server.json

Sources

  1. FieldmodeLLC/scrapecheck-mcp on GitHub — GitHub, observed , trust tier 3.
  2. Official MCP Registry — Model Context Protocol, observed , trust tier 1.
  3. Tools declared by the MCP server at https://scrapecheck-mcp.fly.dev/mcp — scrapecheck-mcp.fly.dev, observed , trust tier 1.