agent-vision-toolkit

为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode

Plugin视觉AI 模型开发自动化视觉 / OCR
验证
L2 · Structured
安全
Pending
健康
Active
信任
Bronze

功能介绍AI

A vision toolkit for text-only LLMs enabling image Q&A, OCR, UI restoration, and GUI automation with agent integrations.

  • Multi-image understanding and image Q&A
  • Long-screenshot OCR and frontend UI restoration
  • GUI automation with optional agent integrations

由 AI 基于 README 自动生成,仅供参考。

安装

dsh plugin --profile web add agent-vision-toolkit

Install method: npm · 尚未在容器中测试 (L3+)

兼容性

DSH VersionStatus
not statedDeclared — not tested

要求

  • • Node.js: not stated
  • • DSH: declared "not stated"
  • • External credentials: none detected

Security Report

自动扫描,非人工审核。

• Dependency vulnerabilities: — (pending)

• Suspicious permissions: — (pending)

• Hardcoded secrets: — (pending)

• Supply chain risks: — (pending)

⚠ 安全审计已排期 — 下次扫描后结果将显示在这里。

Activity

Last commit 2026-08-14 · activity: active

6 mo

Source

GitHub: github.com/Anionex/agent-vision-toolkit