ZBS Index What actually exists in applied AI, with the source next to it

mcp server

evalview-mcp

Regression testing for AI agents. Golden baselines, CI/CD, LangGraph, CrewAI, OpenAI, Claude.

Description as published by the maintainer. Source

  • version 0.6.0
  • slowing

slowing — Registry entry last updated 2026-03-27.

Signals

These are separate measurements of different things. They are deliberately not combined into one score, because a popularity number that mixes website traffic with saves and stars cannot be checked or acted on.

Signal Value What it measures Window Observed Source
Latest published version 0.6.0 Latest version string the maintainer published to the registry. as of fetch Model Context Protocol
Registry record last updated 2026-03-27 When the registry record was last updated by its maintainer. point in time Model Context Protocol
First listed in the MCP Registry 2026-03-27 Date this server was first published to the official MCP Registry. Not a usage or quality measure. point in time Model Context Protocol

Where to get it

Related, by what their authors tagged them

  • Hlido Agent Reviews — last commit 2026-07-15, shares agent-evaluation
    Independent AI-agent reviews: trust checks, evidence scorecards, incident registry, recommendations.
  • io.github.agustincf/arcade1v1 — last commit 2026-07-19, shares agent-benchmark
    Play 1v1 arcade games vs AI agents & humans, ranked by ELO. Replay-verified, on-chain escrow (Base).
  • Synap Memory — last commit 2026-08-06, shares autogen, crewai
    Persistent memory for AI agents — log and recall conversation context over MCP.
  • Agent Coherence — Stale Write Guard (FS) — last commit 2026-07-21, shares autogen, crewai
    Coherence guard for shared files: denies stale writes so agents don't silently overwrite each other.
  • io.github.enrichgateagent-png/beacon-mcp — last commit 2026-07-13, shares autogen, crewai
    Search 17,900+ open-source AI agents by capability, ranked by GitHub traction. Free, no key.
  • io.github.attenlabs/hotato — last commit 2026-08-02, shares regression-testing, testing
    Open-source, self-hosted conversation QA for voice agents. MIT.
  • ai.testiv/mcp — last commit 2026-08-04, shares regression-testing
    Local-first visual regression for AI agents: verdicts, diff images, explain_snapshot. No API key.
  • redline — last commit 2026-08-05, shares regression-testing
    Run local prompt regression checks from prompt-response logs in AI coding assistants.
  • agenthold — last commit 2026-07-07, shares crewai, langgraph
    Shared versioned state for multi-agent AI workflows
  • io.github.figuard/figuard-mcp — last commit 2026-08-04, shares crewai, langgraph
    Pre-flight spend authorization for AI agents. Set a budget, enforce limits, audit every decision.

These share tags the maintainers applied themselves, such as agent-evaluation, agent-benchmark, autogen, crewai. Common tags like "mcp" or "ai" are ignored for this: agreeing with six hundred other projects is not a similarity.

This is not a recommendation and not a test result. It is a map of what the authors said their work is about.

How the author describes it

Topics the maintainer set on GitHub: agent-benchmark, agent-evaluation, agentic-ai, ai-agents, anthropic, autogen, cli, crewai, evaluation, langchain-agent, langgraph, llm, mcp, openai-assistants, pytest, python, regression-testing, testing.

This record as data

Every field on this page, with its source and observation date, is in the catalog JSON. Fetch the whole kind at once instead of parsing this HTML.

GET /api/v1/entries/mcp_server.json

Sources

  1. Official MCP Registry — Model Context Protocol, observed , trust tier 1.