mcp-atlas
A large-scale benchmark that evaluates AI agents' tool-use competency across 36 real MCP servers using a reproducible Docker sandbox and LLM-as-judge scoring.
- Score
- Unranked
- Evidence
- No signals yet
- Last commit
- as last read from GitHub; most reads are from 2 Sep 2026 or later
- Listed
Install
No one-command install. Set it up from its source.
Alternatives · MCPs
- scrapegraphai-scrapegraph113 stars · 34 forks92.669
- macrocosmos-mcp29 stars · 7 forks81.008
- mrrobotke-django-migrations6 stars · 5 forks64.619
What it is
A large-scale benchmark that evaluates AI agents' tool-use competency across 36 real MCP servers using a reproducible Docker sandbox and LLM-as-judge scoring.
When to use it
A large-scale benchmark that evaluates AI agents' tool-use competency across 36 real MCP servers using a reproducible Docker sandbox and LLM-as-judge scoring.
How to install / invoke
See Glama for the install config.
Notes
Listed from the Glama MCP registry.