ZBS Index What actually exists in applied AI, with the source next to it

mcp server

Web Scraper to Markdown API

Extract clean markdown from any URL. Removes boilerplate. For RAG pipelines. x402.

Description as published by the maintainer. Source

  • version 1.2.0
  • active
  • data extraction
  • retrieval

active — Most recent push to the repository was 2026-07-19. Dashed tags are derived by ZBS Index from the published description, not stated by the maintainer.

What this server can do

2 functions, named and described by the server itself. Parameter names are shown because they say more about what a function does than its name usually does.

web_scrape_batch(urls)
Use this when you need to extract clean content from multiple web pages at once (up to 10 URLs). Returns the same structured markdown output as web_scrape_to_markdown for each URL. 1. results (array) -- each entry has title, description, author, content, wordCount, charCount, url 2. summary -- total pages scraped, total word count, failed URLs if any Example output: {"results":[{"url":"https://a.com","title":"Page A","wordCount":800},{"url":"https://b.com","title":"Page B","wordCount":1200}],"summary":{"total":2,"totalWords":2000,"failed":0}} Use this FOR building research corpora, comparing content across competitor pages, or bulk documentation extraction. Essential when you have 3+ URLs to process in one workflow. Do NOT use for single URLs -- use web_scrape_to_markdown instead. Do NOT use for SEO comparison -- use seo_audit_batch instead. Required: urls.
web_scrape_to_markdown(url)
Scrape and extract content from a URL with full JS rendering, returned as clean markdown. Alternative to Firecrawl scrape at 2.5x lower cost. Strips navigation, ads, scripts, and boilerplate — ideal for RAG pipelines and AI research agents. 1. title (string) -- page title from <title> tag 2. description (string) -- meta description 3. author (string) -- author from meta tags or schema 4. content (string) -- clean markdown body text, headings preserved 5. wordCount (number) -- total words in extracted content 6. charCount (number) -- total characters 7. url (string) -- final URL after redirects Example output: {"title":"How to Scale APIs","description":"A guide to...","content":"# How to Scale APIs\n\nScaling requires...","wordCount":1250,"charCount":7800,"url":"https://blog.example.com/scale-apis"} Use this BEFORE summarizing articles, building RAG corpora, researching topics from web sources, or extracting data from documentation pages. Essential for any workflow that needs to scrape and extract content from web pages as LLM input. Drop-in replacement for Firecrawl scrape. Do NOT use for screenshots -- use capture_screenshot instead. Do NOT use for SEO audit -- use seo_audit_page instead. Do NOT use for tech stack detection -- use website_detect_tech_stack instead. Do NOT use for web search -- use web_search_query instead. Required: url.

Last successful function declaration observed on . Source: https://web-scraper.api.klymax402.com/mcp. We list what the server declared; we do not call any of these functions.

Endpoint status observed on . Source: https://web-scraper.api.klymax402.com/mcp.

Signals

These are separate measurements of different things. They are deliberately not combined into one score, because a popularity number that mixes website traffic with saves and stars cannot be checked or acted on.

Signal Value What it measures Window Observed Source
GitHub stars 0 Number of GitHub accounts that bookmarked this repository since it was created. It is a bookmark count, not installs, not active users and not quality. cumulative, all time GitHub
Last commit 2026-07-19 Date of the most recent push to any branch. This is the strongest cheap indicator of whether the project is still maintained. point in time GitHub
Open issues 1 Open issues plus open pull requests, as GitHub counts them together. A high number can mean an active project or an abandoned one. as of fetch GitHub
Latest published version 1.2.0 Latest version string the maintainer published to the registry. as of fetch Model Context Protocol
Registry record last updated 2026-05-16 When the registry record was last updated by its maintainer. point in time Model Context Protocol
License MIT Licence GitHub detected in the repository. Detection can be wrong; the LICENSE file is authoritative. as of fetch GitHub
First listed in the MCP Registry 2026-05-16 Date this server was first published to the official MCP Registry. Not a usage or quality measure. point in time Model Context Protocol
repository status active The repository exists on GitHub and is not archived. This says nothing about how recently it was worked on. as of fetch GitHub
mcp tools declared 2 tools Number of functions the server itself declared when asked to list them. This is what the server offers an agent, not a measure of how well any of them work. as of probe web-scraper.api.klymax402.com
mcp endpoint status ok The server listed 2 functions when asked. as of probe web-scraper.api.klymax402.com

Where to get it

Related, by what their authors tagged them

  • browser-act-skills-1688-product-detail — last commit 2026-08-05, shares data-extraction, web-scraping
    Extracts comprehensive wholesale product data from 1688.com product detail pages: title, tiered pricing, SKU variants w…
  • browser-act-skills-airbnb-listing-detail — last commit 2026-08-05, shares data-extraction, web-scraping
    Fetches complete Airbnb listing details for a given numeric listing ID via the internal GraphQL API, returning title, r…
  • browser-act-skills-airbnb-search-listing — last commit 2026-08-05, shares data-extraction, web-scraping
    Extracts Airbnb accommodation search results from a destination query via SSR-embedded data, returning listing ID, URL,…
  • browser-act-skills-amazon-alexa-qa — last commit 2026-08-05, shares data-extraction, web-scraping
    Amazon Alexa for Shopping Q&A automation: submits questions to Amazon's Alexa/Rufus AI shopping assistant and collects…
  • browser-act-skills-amazon-asin-lookup-api-skill — last commit 2026-08-05, shares data-extraction, web-scraping
    This skill helps users extract structured product details from Amazon using a specific ASIN (Amazon Standard Identifica…
  • browser-act-skills-amazon-best-selling-products-finder-api-skill — last commit 2026-08-05, shares data-extraction, web-scraping
    This skill helps users extract structured best-selling product data from Amazon via the BrowserAct API. Agent should pr…
  • browser-act-skills-amazon-bestseller-listing — last commit 2026-08-05, shares data-extraction, web-scraping
    Amazon Best Sellers listing scraper: extract product cards from any Amazon Best Sellers (zgbs) or /gp/bestsellers/ cate…
  • browser-act-skills-amazon-buy-box-monitor-api-skill — last commit 2026-08-05, shares data-extraction, web-scraping
    This skill helps users extract basic product details other sellers prices and seller ratings from Amazon via ASIN autom…
  • browser-act-skills-amazon-competitor-analyzer — last commit 2026-08-05, shares data-extraction, web-scraping
    Scrapes Amazon product data from ASINs using browseract.com automation API and performs surgical competitive analysis.…
  • browser-act-skills-amazon-listing-competitor-analysis-skill — last commit 2026-08-05, shares data-extraction, web-scraping
    This skill helps users analyze Amazon competitor listings by ASIN and produce structured competitive intelligence plus…

These share tags the maintainers applied themselves, such as data-extraction, web-scraping. Common tags like "mcp" or "ai" are ignored for this: agreeing with six hundred other projects is not a similarity.

This is not a recommendation and not a test result. It is a map of what the authors said their work is about.

Also from br0ski777

  • Address Validator API — last commit 2026-07-19
    Parse, validate and normalize postal addresses. Detect country, verify format. x402.
  • AI Text Summarizer API — last commit 2026-07-19
    Summarize text or URLs into key points with compression ratio and reading time. x402 USDC.
  • Airdrop Checker API — last commit 2026-07-19
    Check wallet eligibility for active crypto airdrops — LayerZero, EigenLayer, Scroll. x402.
  • Barcode Generator API — last commit 2026-07-18
    Generate barcodes — EAN-13, UPC-A, Code128, Code39. Returns SVG. x402 micropayment.
  • Base DeFi Yield Optimizer API — last commit 2026-07-19
    Base chain DeFi yields — Aerodrome, Moonwell, APY, TVL, risk scores. x402.
  • Base64 Codec API — last commit 2026-07-18
    Encode/decode base64 and URL-safe base64. x402 micropayment.
  • Cross-Chain Bridge Route API — last commit 2026-07-19
    Best cross-chain bridge routes with fees, time, providers. LI.FI-powered. x402.
  • Code Sandbox API — last commit 2026-07-31
    Execute Python, JavaScript, SQL code in a sandbox. x402 micropayment.
  • Color Palette Generator API — last commit 2026-07-18
    Generate harmonious color palettes — complementary, analogous, triadic. x402 micropayment.
  • Company Enrichment API — last commit 2026-07-31
    Company firmographics from domain: name, socials, tech stack, emails, phone, address

How the author describes it

Topics the maintainer set on GitHub: ai-agents, data-extraction, mcp, mcp-server, web-scraping, x402.

This record as data

Every field on this page, with its source and observation date, is in the catalog JSON. Fetch the whole kind at once instead of parsing this HTML.

GET /api/v1/entries/mcp_server.json

Sources

  1. Br0ski777/web-scraper-x402 on GitHub — GitHub, observed , trust tier 3.
  2. Official MCP Registry — Model Context Protocol, observed , trust tier 1.
  3. Tools declared by the MCP server at https://web-scraper.api.klymax402.com/mcp — web-scraper.api.klymax402.com, observed , trust tier 1.