
multimodal-bridge
multimodal-bridge 是一个多模态能力桥:把 Qwen 的视觉理解(Qwen-VL)与图像生成(Qwen-Image)带给没有原生多模态能力的纯文本模型(如 DeepSeek)。它有两种形态、同一套后端: MCP Server(qwen_vision / qwen_generate 工具):任何支持 MCP 的宿主(Claude Code、Kimi Code 等)直接挂载; DSH 插件(npm 包 dsh-multimodal-bridge,DeepSeek Harness bundle):dsh plugin add 一行安装,含模型自动 fallback、尺寸自适应与图片结果卡片。
Plugin视觉开发网络AI 模型视觉 / OCR
验证
L2 · Structured
安全
Pending
健康
Active
信任
Bronze
功能介绍AI
Bridges Qwen-VL vision understanding and Qwen-Image generation to text-only models like DeepSeek.
- Adds Qwen-VL visual understanding to text-only LLMs
- Adds Qwen-Image generation for multimodal output
- Installable as MCP server or DSH plugin
由 AI 基于 README 自动生成,仅供参考。
安装
dsh plugin --profile web add multimodal-bridgeInstall method: npm · 尚未在容器中测试 (L3+)
兼容性
| DSH Version | Status |
|---|---|
| not stated | Declared — not tested |
要求
- • Node.js: not stated
- • DSH: declared "not stated"
- • External credentials: none detected
Security Report
自动扫描,非人工审核。
• Dependency vulnerabilities: — (pending)
• Suspicious permissions: — (pending)
• Hardcoded secrets: — (pending)
• Supply chain risks: — (pending)
⚠ 安全审计已排期 — 下次扫描后结果将显示在这里。
Activity
Last commit 2026-08-13 · activity: active
6 mo