mcp server
evalview-mcp
Regression testing for AI agents. Golden baselines, CI/CD, LangGraph, CrewAI, OpenAI, Claude.
Description as published by the maintainer. Source
- version 0.6.0
- slowing
slowing — Registry entry last updated 2026-03-27.
Signals
These are separate measurements of different things. They are deliberately not combined into one score, because a popularity number that mixes website traffic with saves and stars cannot be checked or acted on.
| Signal | Value | What it measures | Window | Observed | Source |
|---|---|---|---|---|---|
| Latest published version | 0.6.0 | Latest version string the maintainer published to the registry. | as of fetch | Model Context Protocol | |
| Registry record last updated | 2026-03-27 | When the registry record was last updated by its maintainer. | point in time | Model Context Protocol | |
| First listed in the MCP Registry | 2026-03-27 | Date this server was first published to the official MCP Registry. Not a usage or quality measure. | point in time | Model Context Protocol |
Where to get it
Related, by what their authors tagged them
-
Hlido Agent Reviews
— last commit 2026-07-15, shares agent-evaluation
Independent AI-agent reviews: trust checks, evidence scorecards, incident registry, recommendations.
-
io.github.agustincf/arcade1v1
— last commit 2026-07-19, shares agent-benchmark
Play 1v1 arcade games vs AI agents & humans, ranked by ELO. Replay-verified, on-chain escrow (Base).
-
Synap Memory
— last commit 2026-08-06, shares autogen, crewai
Persistent memory for AI agents — log and recall conversation context over MCP.
-
Agent Coherence — Stale Write Guard (FS)
— last commit 2026-07-21, shares autogen, crewai
Coherence guard for shared files: denies stale writes so agents don't silently overwrite each other.
-
io.github.enrichgateagent-png/beacon-mcp
— last commit 2026-07-13, shares autogen, crewai
Search 17,900+ open-source AI agents by capability, ranked by GitHub traction. Free, no key.
-
io.github.attenlabs/hotato
— last commit 2026-08-02, shares regression-testing, testing
Open-source, self-hosted conversation QA for voice agents. MIT.
-
ai.testiv/mcp
— last commit 2026-08-04, shares regression-testing
Local-first visual regression for AI agents: verdicts, diff images, explain_snapshot. No API key.
-
redline
— last commit 2026-08-05, shares regression-testing
Run local prompt regression checks from prompt-response logs in AI coding assistants.
-
agenthold
— last commit 2026-07-07, shares crewai, langgraph
Shared versioned state for multi-agent AI workflows
-
io.github.figuard/figuard-mcp
— last commit 2026-08-04, shares crewai, langgraph
Pre-flight spend authorization for AI agents. Set a budget, enforce limits, audit every decision.
These share tags the maintainers applied themselves, such as agent-evaluation, agent-benchmark, autogen, crewai. Common tags like "mcp" or "ai" are ignored for this: agreeing with six hundred other projects is not a similarity.
This is not a recommendation and not a test result. It is a map of what the authors said their work is about.
How the author describes it
Topics the maintainer set on GitHub: agent-benchmark, agent-evaluation, agentic-ai, ai-agents, anthropic, autogen, cli, crewai, evaluation, langchain-agent, langgraph, llm, mcp, openai-assistants, pytest, python, regression-testing, testing.
This record as data
Every field on this page, with its source and observation date, is in the catalog JSON. Fetch the whole kind at once instead of parsing this HTML.
GET /api/v1/entries/mcp_server.json