mcp server
scrapecheck
Verify scraped data against the live source page. Signed verdicts, $0.01 via x402.
Description as published by the maintainer. Source
- version 0.2.0
- active
- data extraction
active — Most recent push to the repository was 2026-08-08. Dashed tags are derived by ZBS Index from the published description, not stated by the maintainer.
What this server can do
3 functions, named and described by the server itself. Parameter names are shown because they say more about what a function does than its name usually does.
get_verifier_info- Free. Returns the public key, engine version, benchmark summary with sample-size labels, and scope, so you can verify any verdict offline. Call this first if your selection policy requires evidence before trusting a verification provider: every verdict this service issues verifies offline against the returned key, so you do not have to trust the service at runtime to rely on its output.
verify_presence(url, asked, claim)- Presence only: confirms the value appears on the page, NOT that it is the right answer to what was asked. Never returns pass. A was-price, a different variant's price, a shipping cost, or a tax figure can all satisfy presence. Verdicts are present, fail, or unverifiable, signed, with check_type web_field_presence_v1. Use it as a cheap screen before a full check, or when you already know what the value means and only need to confirm it still appears on the page. If your decision depends on the value being the right value, call verify_web_field instead. Scope: server-rendered pages. $0.002 per check, paid in-band via x402. Required: url, claim, asked.
verify_web_field(url, asked, claim)- Checks whether a value you did not fetch yourself is actually on the source page right now, and returns a signed pass/fail/unverifiable verdict. Not a retrieval tool: it checks a value you already have, it does not find data for you. Use it when you hold a value that came from somewhere else (a search result, a scraper, an upstream API, another agent) and you need to know it is on the source page before you act on it. Takes the URL, the claimed value, and what was asked; it independently re-fetches the page and never trusts the fetch you were given. Structural guarantee: a claim is never certified unless the re-fetched page contains it, and the model judge can only veto a pass, never create one. Anything unconfirmed returns unverifiable, never pass. Every verdict carries a stable verdict_id, the engine digest that produced it, and a signature verifiable offline against the public key (see get_verifier_info). Scope: server-rendered pages; JS-only content returns unverifiable rather than a guess. $0.01 per check, first 100 free, paid in-band via x402. Required: url, claim, asked.
Last successful function declaration observed on . Source: https://scrapecheck-mcp.fly.dev/mcp. We list what the server declared; we do not call any of these functions.
Endpoint status observed on . Source: https://scrapecheck-mcp.fly.dev/mcp.
Signals
These are separate measurements of different things. They are deliberately not combined into one score, because a popularity number that mixes website traffic with saves and stars cannot be checked or acted on.
| Signal | Value | What it measures | Window | Observed | Source |
|---|---|---|---|---|---|
| GitHub stars | 0 | Number of GitHub accounts that bookmarked this repository since it was created. It is a bookmark count, not installs, not active users and not quality. | cumulative, all time | GitHub | |
| Last commit | 2026-08-08 | Date of the most recent push to any branch. This is the strongest cheap indicator of whether the project is still maintained. | point in time | GitHub | |
| Open issues | 0 | Open issues plus open pull requests, as GitHub counts them together. A high number can mean an active project or an abandoned one. | as of fetch | GitHub | |
| Latest published version | 0.2.0 | Latest version string the maintainer published to the registry. | as of fetch | Model Context Protocol | |
| Registry record last updated | 2026-08-01 | When the registry record was last updated by its maintainer. | point in time | Model Context Protocol | |
| License | MIT | Licence GitHub detected in the repository. Detection can be wrong; the LICENSE file is authoritative. | as of fetch | GitHub | |
| First listed in the MCP Registry | 2026-08-01 | Date this server was first published to the official MCP Registry. Not a usage or quality measure. | point in time | Model Context Protocol | |
| repository status | active | The repository exists on GitHub and is not archived. This says nothing about how recently it was worked on. | as of fetch | GitHub | |
| mcp tools declared | 3 tools | Number of functions the server itself declared when asked to list them. This is what the server offers an agent, not a measure of how well any of them work. | as of probe | scrapecheck-mcp.fly.dev | |
| mcp endpoint status | ok | The server listed 3 functions when asked. | as of probe | scrapecheck-mcp.fly.dev |
Where to get it
Related, by what their authors tagged them
-
andon
— last commit 2026-08-04, shares data-quality, verification
Deterministic verification for AI-generated analysis: run a spec, inspect or diff a workbook
-
com.scrapeunblocker/scrapeunblocker-mcp
— last commit 2026-08-06, shares scraping, web-scraping
Fetch any web page's HTML or AI-parsed JSON through the ScrapeUnblocker anti-bot API.
-
io.alterlab/mcp-server
— last commit 2026-08-05, shares scraping, web-scraping
Web scraping MCP server — scrape, extract structured data, screenshot any site with anti-bot bypass.
-
io.github.auxiliar-ai/auxiliar-mcp
— archived, last commit 2026-07-11, shares scraping, web-scraping
Eval-backed discovery for the auxiliar.ai gateway — the best web-access provider per job, measured.
-
io.github.brightdata/brightdata-mcp
— last commit 2026-07-27, shares scraping, web-scraping
Bright Data's Web MCP server enabling AI agents to search, extract & navigate the web
-
Scrapling MCP Server
— last commit 2026-08-06, shares scraping, web-scraping
Web scraping with stealth HTTP, real browsers, and Cloudflare bypass. CSS selectors supported.
-
ai.smithery/oxylabs-oxylabs-mcp
— last commit 2026-06-08, shares scraping
Fetch and process content from specified URLs using the Oxylabs Web Scraper API.
-
ai.smithery/ScrapeGraphAI-scrapegraph-mcp
— last commit 2026-07-17, shares scraping
Enable language models to perform advanced AI-powered web scraping with enterprise-grade reliabili…
-
WebReaper
— last commit 2026-07-11, shares scraping
AI-native web scraper: scrape, crawl and map any site to clean markdown over stdio. MIT-licensed.
-
TrustyData
— last commit 2026-07-13, shares data-quality
French address quality, geocoding & routing from official data (BAN, INSEE, OpenStreetMap).
These share tags the maintainers applied themselves, such as data-quality, verification, scraping, web-scraping. Common tags like "mcp" or "ai" are ignored for this: agreeing with six hundred other projects is not a similarity.
This is not a recommendation and not a test result. It is a map of what the authors said their work is about.
How the author describes it
Topics the maintainer set on GitHub: ai-agents, data-quality, mcp, model-context-protocol, scraping, verification, web-scraping, x402.
This record as data
Every field on this page, with its source and observation date, is in the catalog JSON. Fetch the whole kind at once instead of parsing this HTML.
GET /api/v1/entries/mcp_server.json