
shadow-vision
Open-source MCP vision server that gives text-only LLMs and AI agents image understanding, OCR, visual analysis, UI inspection, and multimodal capabilities.
这是一个带 dsh-plugin 话题的 GitHub 仓库——仓库 README 才是权威安装来源。使用下面的模板,按 README 调整包名/路径:
dsh plugin --profile web add $shadow-vision本地开发? 用 dsh --profile web --patch ./cordis.yml 加绝对路径插件——见插件开发。GitHub 源码? dsh plugin --profile web add github:$WardLu/$shadow-vision——需要仓库里有 prepare 脚本(见打包发布)。
| 仓库 | WardLu/shadow-vision |
| 许可证 | MIT |
| 分类 | 工具 |
| 状态 | 社区 |
| 声明兼容 | not stated |
| 最近提交 | 2026-08-12 |
| 收录时间 | 2026-08-13 |
Declared 的兼容性是作者声明,dsh.so 尚未独立测试。版本兼容见更新日志。Built by DeepSeek, for DeepSeek — a Swift-native macOS coding agent
查看The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics).
查看为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode
查看问题或建议?官方 Discussions 是项目规范的支持渠道;Discord 里能找到活跃的社区成员。dsh.so 本身欢迎通过提交插件或反馈改进。