jwrede-llmprobe
Synthetic monitoring for LLM inference endpoints. Measure TTFT, latency, throughput, and errors across OpenAI, Anthropic, Google, Azure, Bedrock, and local servers (vLLM, SGLang, Ollama). CLI + MCP server with Prometheus and OpenTelemetry export.
- Score
- 46.6541 signal
- Evidence
- 5 stars
- Last commit
- as last read from GitHub; most reads are from 2 Sep 2026 or later
- Listed
Install
No one-command install. Set it up from its source.
Alternatives · MCPs
- pinkpixel-mindbridge37 stars · 9 forks83.721
- vic563-memgpt27 stars · 4 forks78.079
- feiskyer-ai-hub10 stars · 3 forks66.893
What it is
Synthetic monitoring for LLM inference endpoints. Measure TTFT, latency, throughput, and errors across OpenAI, Anthropic, Google, Azure, Bedrock, and local servers (vLLM, SGLang, Ollama). CLI + MCP server with Prometheus and OpenTelemetry export.
When to use it
When an agent needs the "Monitoring" capability this MCP server exposes.
Source
Migrated from the awesome-mcp-servers navigation directory (category: Monitoring). See https://github.com/Jwrede/llmprobe.