Vision plugins
$26 DeepSeek Harness plugins tagged Vision — tags are auto-extracted from repository descriptions; sorted by GitHub stars.
The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics).
为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.
是兄弟就来蹬我!DSH Web UI 广告:2005 年中文站点风格的侧栏广告 / 对话内信息流 / 角落弹窗 + 一个真实热区比视觉小得多的关闭叉。素材全虚构,域名打码。
dsh plugin: Chrome sidebar extension that lets DSH operate your browser directly—no vision capabilities required.
DSH 内测收官合影墙:GitHub OAuth 零权限登录 + 冻结白名单校验的拍立得合影站(含 DSH Skill 包装)
dsh 插件:给纯文本 DeepSeek 加视觉——view_image 工具桥接任意 OpenAI 兼容 VLM(默认智谱免费档,实测 4 厂商 10 模型)
No description provided.
OpenMAIC for DeepSeek Harness: classrooms, slides, interactive widgets, and Socratic teaching
Community Docker and Kubernetes packaging for DeepSeek Harness (@deepseek-ai/dsh), with a hardened image, Compose stack, Helm chart, Web UI, and headless CLI.
DSH 原生鸿蒙设备桥:hdc 工具让 Agent 完成截图-看图-装包-验证的闭环调试 / DSH-native HarmonyOS device bridge
DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.
DeepSeek Harness 插件:DeepSeek 大脑 + 自动识图。附加图片自动经 Qwen VLM 转译成文字后交给 DeepSeek 作答
Open-source MCP vision server that gives text-only LLMs and AI agents image understanding, OCR, visual analysis, UI inspection, and multimodal capabilities.
No description provided.
dsh browser-run: CF Browser Run web tools (markdown/screenshot/pdf) for DeepSeek Harness
No description provided.
No description provided.
DSH Pet Corner: a floating pet, keyless pet-image proxy, favorites, and plugin-owned settings API
Vision toolkit for DeepSeek Harness -- give text-only agents eyes
DSH plugin: pixel-diff two screenshots into diff.png + triptych (pixelmatch) — 像素对比工具
External vision model tool for DeepSeek Harness — inspect_image sends images to any OpenAI-compatible endpoint (GPT-4o, Qwen-VL, GLM-4V, Ollama...).
Eight DeepSeek Harness plugins: persona, language guard, per-request vision fallback, python/windows write guards, cross-agent memory, image generation, and skill shell injection.
DeepSeek Harness plugin: turn UI screenshots into structured, implementation-grade web frontend specs. Deterministic geometry (sharp) + optional vision-model semantics, merged into one JSON + Markdown spec.
让你能通过deepseek harness调用LM studio加载的本地视觉模型
CLI-first vision sidecar for text-only coding agents. Analyze screenshots, diagrams, charts, UI diffs, and videos with OpenAI-compatible multimodal models.