screencye
Eyes for text-only LLMs: decodes screenshots into exact structured text (words, coordinates, sizes, colors) using pure-code CV and OCR. Enables text-only models to reason about UI layouts without vision models or VRAM usage.
- Score
- 43.4982 signals
- Evidence
- 2 stars · 2 forks
- Last commit
- as last read from GitHub; most reads are from 2 Sep 2026 or later
- Listed
Install
No one-command install. Set it up from its source.
Alternatives · MCPs
- textweb104 stars · 14 forks90.700
- atlas-vision-mcp10 stars · 3 forks66.893
- super-mcp-ocr-deepseek4 stars · 3 forks53.985
What it is
Eyes for text-only LLMs: decodes screenshots into exact structured text (words, coordinates, sizes, colors) using pure-code CV and OCR. Enables text-only models to reason about UI layouts without vision models or VRAM usage.
When to use it
Eyes for text-only LLMs: decodes screenshots into exact structured text (words, coordinates, sizes, colors) using pure-code CV and OCR. Enables text-only models to reason about UI layouts without vision models or VRAM usage.
How to install / invoke
See Glama for the install config.
Notes
Listed from the Glama MCP registry.