mcp-llm-eval
A local MCP server that packages LLM evaluation gates as reusable CI/CD primitives, enabling AI agents to run datasets against models, score responses, and enforce quality thresholds.
- Score
- Unranked
- Evidence
- No signals yet
- Last commit
- as last read from GitHub; most reads are from 2 Sep 2026 or later
- Listed
Install
No one-command install. Set it up from its source.
Alternatives · MCPs
- safedep-vet1,107 stars · 111 forks97.723
- jpicklyk-task-orchestrator205 stars · 21 forks93.653
- raye-deng-open-code-review37 stars · 3 forks79.061
What it is
A local MCP server that packages LLM evaluation gates as reusable CI/CD primitives, enabling AI agents to run datasets against models, score responses, and enforce quality thresholds.
When to use it
A local MCP server that packages LLM evaluation gates as reusable CI/CD primitives, enabling AI agents to run datasets against models, score responses, and enforce quality thresholds.
How to install / invoke
See Glama for the install config.
Notes
Listed from the Glama MCP registry.