token-optimization-mcp
A fully offline MCP server for token estimation, prompt compression, model routing, and semantic caching to optimize LLM usage costs and efficiency.
- Score
- Unranked
- Evidence
- No signals yet
- Last commit
- as last read from GitHub; most reads are from 2 Sep 2026 or later
- Listed
Install
No one-command install. Set it up from its source.
Alternatives · MCPs
- token-optimizer536 stars · 65 forks96.719
- ai-elara-code-context-engine408 stars · 62 forks96.340
- ultra-multi-ai-provider275 stars · 19 forks94.100
What it is
A fully offline MCP server for token estimation, prompt compression, model routing, and semantic caching to optimize LLM usage costs and efficiency.
When to use it
A fully offline MCP server for token estimation, prompt compression, model routing, and semantic caching to optimize LLM usage costs and efficiency.
How to install / invoke
See Glama for the install config.
Notes
Listed from the Glama MCP registry.