multimodal-mcp-server
Enables MCP-compatible clients to leverage OpenAI's multimodal capabilities (vision, image generation, speech-to-text, text-to-speech) through file-oriented tools with a security-first architecture.
- Score
- 29.0132 signals
- Evidence
- 1 star · 1 fork
- Last commit
- as last read from GitHub; most reads are from 2 Sep 2026 or later
- Listed
Install
No one-command install. Set it up from its source.
Alternatives · MCPs
- minimax-tts-video1,573 stars · 279 forks98.301
- aidc-ai-pixelle1,108 stars · 150 forks97.855
- rlabs-inc-gemini217 stars · 45 forks94.885
What it is
Enables MCP-compatible clients to leverage OpenAI's multimodal capabilities (vision, image generation, speech-to-text, text-to-speech) through file-oriented tools with a security-first architecture.
When to use it
Enables MCP-compatible clients to leverage OpenAI's multimodal capabilities (vision, image generation, speech-to-text, text-to-speech) through file-oriented tools with a security-first architecture.
How to install / invoke
See Glama for the install config.
Notes
Listed from the Glama MCP registry.