mcp server
CompletionKit
Prompt evals over MCP: run a prompt on your dataset, score each output 1-5 with an LLM judge.
Description as published by the maintainer. Source
- version 1.0.0
- active
- evaluation
active — Most recent push to the repository was 2026-08-05. Dashed tags are derived by ZBS Index from the published description, not stated by the maintainer.
Signals
These are separate measurements of different things. They are deliberately not combined into one score, because a popularity number that mixes website traffic with saves and stars cannot be checked or acted on.
| Signal | Value | What it measures | Window | Observed | Source |
|---|---|---|---|---|---|
| GitHub stars | 1 | Number of GitHub accounts that bookmarked this repository since it was created. It is a bookmark count, not installs, not active users and not quality. | cumulative, all time | GitHub | |
| Last commit | 2026-08-05 | Date of the most recent push to any branch. This is the strongest cheap indicator of whether the project is still maintained. | point in time | GitHub | |
| Open issues | 11 | Open issues plus open pull requests, as GitHub counts them together. A high number can mean an active project or an abandoned one. | as of fetch | GitHub | |
| Latest published version | 1.0.0 | Latest version string the maintainer published to the registry. | as of fetch | Model Context Protocol | |
| Registry record last updated | 2026-07-18 | When the registry record was last updated by its maintainer. | point in time | Model Context Protocol | |
| First listed in the MCP Registry | 2026-07-18 | Date this server was first published to the official MCP Registry. Not a usage or quality measure. | point in time | Model Context Protocol | |
| repository status | active | The repository exists on GitHub and is not archived. This says nothing about how recently it was worked on. | as of fetch | GitHub |
Where to get it
Related, by what their authors tagged them
-
PasteHTML
— last commit 2026-08-05, shares rails, ruby
Publish, update, and organize HTML and Markdown pastes on PasteHTML as the authorized user.
-
Rails AI Context
— last commit 2026-08-01, shares rails, ruby
38 MCP tools give AI agents live Rails schema, routes, models, and conventions.
-
Lumen
— last commit 2026-06-07, shares llm-as-judge
Self-hostable agentic-AI LMS: catalog, RAG tutor, FSRS reviews, AI authoring, ingest.
-
Citadel
— last commit 2026-08-06, shares llm-evaluation
Encrypted-first embedded database with vector search and agent memory, exposed as MCP tools
-
io.github.ahmedEid1/forgejudge
— last commit 2026-07-25, shares llm-evaluation
Open eval leaderboard + CI gate for autonomous coding agents (solve, score, trace).
-
io.github.CodesWhat/portkey-admin-mcp
— last commit 2026-08-04, shares llmops
Full Portkey Admin API MCP server — configs, prompts, keys, analytics, and more.
-
io.github.dcondrey/misterdev
— last commit 2026-07-27, shares llmops
Autonomous LLM build orchestrator that plans, edits, and verifies code across languages.
-
ai.testiv/mcp
— last commit 2026-08-04, shares ruby
Local-first visual regression for AI agents: verdicts, diff images, explain_snapshot. No API key.
-
io.github.4DA-Systems/4da-mcp-server
— last commit 2026-08-05, shares ollama
Dependency intelligence for AI agents. CVE scanning, health checks, upgrade planning.
-
AI Guardian
— last commit 2026-08-03, shares ollama
Governed local-LLM (Ollama) observability: model policy, prompt scanner, 20 tools.
These share tags the maintainers applied themselves, such as rails, ruby, llm-as-judge, llm-evaluation. Common tags like "mcp" or "ai" are ignored for this: agreeing with six hundred other projects is not a similarity.
This is not a recommendation and not a test result. It is a map of what the authors said their work is about.
How the author describes it
Topics the maintainer set on GitHub: anthropic, evaluation-framework, evaluation-metrics, llm, llm-as-judge, llm-eval, llm-evaluation, llm-evaluation-framework, llm-evaluation-metrics, llmops, mcp, ollama, openai, prompt-engineering, prompt-testing, rails, rails-engine, ruby, ruby-on-rails.
Bring your own setup
We take apart real AI setups every week and show what broke, what cost too much, and what the trace actually said. If you run agents on real work, that is where the useful conversation is.
Join ZBS AI Practice Lab