dsh-visionary

zhuiyueya

Give text-only DeepSeek models eyes — a DeepSeek Harness plugin that transparently converts chat images into OCR text + vision-model descriptions before they reach the LLM. Configure vision backends (GLM-4V, Qwen-VL, Gemini, Ollama…) right in the Models settings page; multi-backend fallback chain, double-layer caching, no config files.

deepseekdeepseek-harnessdsh-pluginimage-understandingllm-pluginmultimodalocrtext-only-llmvision-bridgevlm

安装

dsh plugin add github:zhuiyueya/dsh-visionary

首次安装 GitHub 来源的包时,可能需要在 profile 的 pnpm-workspace.yamlallowBuilds 中允许该包的构建脚本;也可用 github:zhuiyueya/dsh-visionary#<commit-sha> 锁定版本。详见官方文档

还没有 DSH?两步开始 →

第一步:准备 Node.js 环境

DSH 依赖 Node.js(建议 20 LTS 或更新版本)。终端里运行 node -v 能显示版本号即已就绪。

已有 Node.js:直接进入第二步。

没有 Node.js:去 nodejs.org 下载 LTS 安装包(Windows/macOS 双击安装);或用包管理器:

brew install node
winget install OpenJS.NodeJS.LTS

第二步:启动 DSH

无需安装,直接运行:

npx @deepseek-ai/dsh web

浏览器打开 http://127.0.0.1:3080,在 Web UI 的插件市场里搜索 dsh-visionary,或粘贴上面的安装命令。

在 GitHub 查看 分享到 X