deepseek-vision
MCP server that gives text-only LLMs vision capabilities by using a free multimodal model to perceive images, audio, and video, returning text for the main model to reason with.
- Score
- 32.2231 signal
- Evidence
- 2 stars
- Last commit
- as last read from GitHub; most reads are from 2 Sep 2026 or later
- Listed
Install
No one-command install. Set it up from its source.
Alternatives · MCPs
- aidc-ai-pixelle1,108 stars · 150 forks97.855
- tianshu-mcp-server774 stars · 119 forks97.455
- comfyui-mcp-server409 stars · 83 forks96.548
What it is
MCP server that gives text-only LLMs vision capabilities by using a free multimodal model to perceive images, audio, and video, returning text for the main model to reason with.
When to use it
MCP server that gives text-only LLMs vision capabilities by using a free multimodal model to perceive images, audio, and video, returning text for the main model to reason with.
How to install / invoke
See Glama for the install config.
Notes
Listed from the Glama MCP registry.