luma-mcp
Multi-Model Visual Understanding MCP Server, GLM-4.6V, DeepSeek-OCR (free), and Qwen3-VL-Flash. Provide visual processing capabilities for AI coding models that do not support image understanding.多模型视觉理解MCP服务器,GLM-4.6V、DeepSeek-OCR(免费)和Qwen3-VL-Flash等。为不支持图片理解的 AI 编码模型提供视觉处理能力。
How to read this: tool names here are observed from a live tools/list handshake. The Risk label is a heuristic inferred from the tool name (write/destructive verbs), not from executing the tool — a conservative guess, not a verified capability. We never escalate risk from a description. Found one that's wrong? Tell us — we fix on report.
| Tool | Risk | Side effects | Approval |
|---|---|---|---|
| image_understand 图像理解工具(单一入口):
- 何时调用:用户提到看图/截图/界面/报错/OCR/布局,或对话中出现图片附件并询问图片相关问题时,优先调用本工具。
- 图片来源:粘贴图路径、本地路径、HTTP(S) URL、Data URI。
- prompt:直接传入用户原始问题即可;服务端会拼接基础视觉协议与可选 task 指引。
- task_type(可选):auto|general|ocr|ui|debug|describe。省略或 auto 时与旧版行为兼容(按 prompt 启发式)。 | read | false | unknown |
- repohttps://github.com/JochenYang/luma-mcp
- licenseMIT
- adoption87 stars · 11 forks
Add the “as seen on MCPExplorer” badge to your README.
This is one server. A loadout combines the right servers, governance, and proven plays for a whole job — assembled deliberately, not tool-dumped.
Explore loadouts →