mcp server
Web Scraper to Markdown API
Extract clean markdown from any URL. Removes boilerplate. For RAG pipelines. x402.
Description as published by the maintainer. Source
- version 1.2.0
- active
- data extraction
- retrieval
active — Most recent push to the repository was 2026-07-19. Dashed tags are derived by ZBS Index from the published description, not stated by the maintainer.
What this server can do
2 functions, named and described by the server itself. Parameter names are shown because they say more about what a function does than its name usually does.
web_scrape_batch(urls)- Use this when you need to extract clean content from multiple web pages at once (up to 10 URLs). Returns the same structured markdown output as web_scrape_to_markdown for each URL. 1. results (array) -- each entry has title, description, author, content, wordCount, charCount, url 2. summary -- total pages scraped, total word count, failed URLs if any Example output: {"results":[{"url":"https://a.com","title":"Page A","wordCount":800},{"url":"https://b.com","title":"Page B","wordCount":1200}],"summary":{"total":2,"totalWords":2000,"failed":0}} Use this FOR building research corpora, comparing content across competitor pages, or bulk documentation extraction. Essential when you have 3+ URLs to process in one workflow. Do NOT use for single URLs -- use web_scrape_to_markdown instead. Do NOT use for SEO comparison -- use seo_audit_batch instead. Required: urls.
web_scrape_to_markdown(url)- Scrape and extract content from a URL with full JS rendering, returned as clean markdown. Alternative to Firecrawl scrape at 2.5x lower cost. Strips navigation, ads, scripts, and boilerplate — ideal for RAG pipelines and AI research agents. 1. title (string) -- page title from <title> tag 2. description (string) -- meta description 3. author (string) -- author from meta tags or schema 4. content (string) -- clean markdown body text, headings preserved 5. wordCount (number) -- total words in extracted content 6. charCount (number) -- total characters 7. url (string) -- final URL after redirects Example output: {"title":"How to Scale APIs","description":"A guide to...","content":"# How to Scale APIs\n\nScaling requires...","wordCount":1250,"charCount":7800,"url":"https://blog.example.com/scale-apis"} Use this BEFORE summarizing articles, building RAG corpora, researching topics from web sources, or extracting data from documentation pages. Essential for any workflow that needs to scrape and extract content from web pages as LLM input. Drop-in replacement for Firecrawl scrape. Do NOT use for screenshots -- use capture_screenshot instead. Do NOT use for SEO audit -- use seo_audit_page instead. Do NOT use for tech stack detection -- use website_detect_tech_stack instead. Do NOT use for web search -- use web_search_query instead. Required: url.
Last successful function declaration observed on . Source: https://web-scraper.api.klymax402.com/mcp. We list what the server declared; we do not call any of these functions.
Endpoint status observed on . Source: https://web-scraper.api.klymax402.com/mcp.
Signals
These are separate measurements of different things. They are deliberately not combined into one score, because a popularity number that mixes website traffic with saves and stars cannot be checked or acted on.
| Signal | Value | What it measures | Window | Observed | Source |
|---|---|---|---|---|---|
| GitHub stars | 0 | Number of GitHub accounts that bookmarked this repository since it was created. It is a bookmark count, not installs, not active users and not quality. | cumulative, all time | GitHub | |
| Last commit | 2026-07-19 | Date of the most recent push to any branch. This is the strongest cheap indicator of whether the project is still maintained. | point in time | GitHub | |
| Open issues | 1 | Open issues plus open pull requests, as GitHub counts them together. A high number can mean an active project or an abandoned one. | as of fetch | GitHub | |
| Latest published version | 1.2.0 | Latest version string the maintainer published to the registry. | as of fetch | Model Context Protocol | |
| Registry record last updated | 2026-05-16 | When the registry record was last updated by its maintainer. | point in time | Model Context Protocol | |
| License | MIT | Licence GitHub detected in the repository. Detection can be wrong; the LICENSE file is authoritative. | as of fetch | GitHub | |
| First listed in the MCP Registry | 2026-05-16 | Date this server was first published to the official MCP Registry. Not a usage or quality measure. | point in time | Model Context Protocol | |
| repository status | active | The repository exists on GitHub and is not archived. This says nothing about how recently it was worked on. | as of fetch | GitHub | |
| mcp tools declared | 2 tools | Number of functions the server itself declared when asked to list them. This is what the server offers an agent, not a measure of how well any of them work. | as of probe | web-scraper.api.klymax402.com | |
| mcp endpoint status | ok | The server listed 2 functions when asked. | as of probe | web-scraper.api.klymax402.com |
Where to get it
Related, by what their authors tagged them
-
browser-act-skills-1688-product-detail
— last commit 2026-08-05, shares data-extraction, web-scraping
Extracts comprehensive wholesale product data from 1688.com product detail pages: title, tiered pricing, SKU variants w…
-
browser-act-skills-airbnb-listing-detail
— last commit 2026-08-05, shares data-extraction, web-scraping
Fetches complete Airbnb listing details for a given numeric listing ID via the internal GraphQL API, returning title, r…
-
browser-act-skills-airbnb-search-listing
— last commit 2026-08-05, shares data-extraction, web-scraping
Extracts Airbnb accommodation search results from a destination query via SSR-embedded data, returning listing ID, URL,…
-
browser-act-skills-amazon-alexa-qa
— last commit 2026-08-05, shares data-extraction, web-scraping
Amazon Alexa for Shopping Q&A automation: submits questions to Amazon's Alexa/Rufus AI shopping assistant and collects…
-
browser-act-skills-amazon-asin-lookup-api-skill
— last commit 2026-08-05, shares data-extraction, web-scraping
This skill helps users extract structured product details from Amazon using a specific ASIN (Amazon Standard Identifica…
-
browser-act-skills-amazon-best-selling-products-finder-api-skill
— last commit 2026-08-05, shares data-extraction, web-scraping
This skill helps users extract structured best-selling product data from Amazon via the BrowserAct API. Agent should pr…
-
browser-act-skills-amazon-bestseller-listing
— last commit 2026-08-05, shares data-extraction, web-scraping
Amazon Best Sellers listing scraper: extract product cards from any Amazon Best Sellers (zgbs) or /gp/bestsellers/ cate…
-
browser-act-skills-amazon-buy-box-monitor-api-skill
— last commit 2026-08-05, shares data-extraction, web-scraping
This skill helps users extract basic product details other sellers prices and seller ratings from Amazon via ASIN autom…
-
browser-act-skills-amazon-competitor-analyzer
— last commit 2026-08-05, shares data-extraction, web-scraping
Scrapes Amazon product data from ASINs using browseract.com automation API and performs surgical competitive analysis.…
-
browser-act-skills-amazon-listing-competitor-analysis-skill
— last commit 2026-08-05, shares data-extraction, web-scraping
This skill helps users analyze Amazon competitor listings by ASIN and produce structured competitive intelligence plus…
These share tags the maintainers applied themselves, such as data-extraction, web-scraping. Common tags like "mcp" or "ai" are ignored for this: agreeing with six hundred other projects is not a similarity.
This is not a recommendation and not a test result. It is a map of what the authors said their work is about.
Also from br0ski777
-
Address Validator API
— last commit 2026-07-19
Parse, validate and normalize postal addresses. Detect country, verify format. x402.
-
AI Text Summarizer API
— last commit 2026-07-19
Summarize text or URLs into key points with compression ratio and reading time. x402 USDC.
-
Airdrop Checker API
— last commit 2026-07-19
Check wallet eligibility for active crypto airdrops — LayerZero, EigenLayer, Scroll. x402.
-
Barcode Generator API
— last commit 2026-07-18
Generate barcodes — EAN-13, UPC-A, Code128, Code39. Returns SVG. x402 micropayment.
-
Base DeFi Yield Optimizer API
— last commit 2026-07-19
Base chain DeFi yields — Aerodrome, Moonwell, APY, TVL, risk scores. x402.
-
Base64 Codec API
— last commit 2026-07-18
Encode/decode base64 and URL-safe base64. x402 micropayment.
-
Cross-Chain Bridge Route API
— last commit 2026-07-19
Best cross-chain bridge routes with fees, time, providers. LI.FI-powered. x402.
-
Code Sandbox API
— last commit 2026-07-31
Execute Python, JavaScript, SQL code in a sandbox. x402 micropayment.
-
Color Palette Generator API
— last commit 2026-07-18
Generate harmonious color palettes — complementary, analogous, triadic. x402 micropayment.
-
Company Enrichment API
— last commit 2026-07-31
Company firmographics from domain: name, socials, tech stack, emails, phone, address
How the author describes it
Topics the maintainer set on GitHub: ai-agents, data-extraction, mcp, mcp-server, web-scraping, x402.
This record as data
Every field on this page, with its source and observation date, is in the catalog JSON. Fetch the whole kind at once instead of parsing this HTML.
GET /api/v1/entries/mcp_server.json