deepseek-vl-support

by limccn

1 模型与账号接入github 检测到 manifest package.json#dsh收录于 08-17

让纯文本 DeepSeek 模型在 Claude Code 与 Codex 中获得视觉能力:把图片路由到任意 OpenAI 兼容视觉端点(OpenRouter、SiliconFlow、DashScope、Ollama、llama.cpp、vLLM、LM Studio 等);零运行时依赖,MIT 许可

Give DeepSeek (text-only) models **vision** in Claude Code and Codex by routing image files to any OpenAI-compatible vision endpoint (OpenRouter, SiliconFlow, DashScope, Ollama, llama.cpp, vLLM, LM Studio, …). Zero runtime dependencies, MIT licensed.

安装

dsh plugin --profile web add github:limccn/deepseek-vl-support

GitHub 源码安装:首次需按提示配置 allowBuilds 构建授权后重试

安装与环境配置指引、插件开发教程见 DSH 中文社区文档 ↗

安装即在你的机器上以你的权限运行第三方代码——它可读写文件、使用凭据、访问网络,DSH 的工具审批不会为插件代码加沙箱。「检测到 manifest」仅代表发现 dsh.bundle / dsh.plugin 清单,不构成兼容性或安全审查;安装前请审阅源码,不熟悉的插件先在不含密钥的环境试用。

README

目录

deepseek-vl-support

中文说明 → docs/README.zh-CN.md

What this does

Some AI models (like DeepSeek) can read your files, but they cannot see pictures. Screenshots of errors, UI mockups, charts — invisible to them.

This small tool gives them "eyes". Once installed, whenever the model tries to read a picture, the tool sends it to a vision service of your choice (Moonshot, OpenRouter, SiliconFlow, Ollama …), receives a detailed text description, and hands it to the model — as if the model could see the picture.

Model reads screenshot.png
  → the tool intercepts the read
  → picture → vision service → detailed text description comes back
  → the model receives: "[Vision of screenshot.png]: <description>"
  → the model answers from the description

No model settings to change, no config files to write — it works automatically after a one-time setup. One command to install, one command to remove. MIT licensed.

Who this is for

You use a text-only model (such as DeepSeek) in any AI coding agent or IDE and want it to understand pictures: error screenshots, UI mockups, charts, photos of notes. Pick your tool in the install wizard below — there is a one-command install for every supported agent, including Claude Code, Codex, Cursor, GitHub Copilot, VS Code, OpenCode, Trae, Qwen Code, and 14 more.

Before you start (what you need)

  1. Node.js 18 or newer — check with node -v. Not installed? Get it at https://nodejs.org.
  2. An account at a vision service, plus its API key — a vision service is the "eyes provider": a website that looks at pictures for you. Cloud options: Moonshot, OpenRouter, MiniMax, Zhipu GLM, StepFun, OpenCode Zen, SiliconFlow, DashScope. Free local options (run on your own computer): Ollama, llama.cpp, vLLM, LM Studio. The API key is a secret code from that service (usually under "API keys"); the installer asks for it once and stores it only on your computer.
  3. Your AI agent installed — any of the supported ones below.

Quick install wizard

Open a terminal in your project folder and run:

cd path/to/your/project
npx deepseek-vl-support@latest install

That's the whole install — the wizard auto-detects the agents on your machine and asks 7 short questions. Almost every question has a sensible default: just press Enter. The two that matter: which agents should get vision (pre-selected) and which vision service + API key to use (choose Decide later, the last option, if you want to sort that out afterwards).

When it finishes, restart your session — the installer prints this reminder, and it is required for the effect to kick in. Optional check:

npx deepseek-vl-support@latest doctor    # look for [OK]

Re-running on the same project? It asks whether to keep your current settings — Enter keeps them.

No terminal? Ask your agent to install it. If you use a tool that supports the Agent Plugins standard (GitHub Copilot, Cursor, Kiro, OpenClaw, Hermes Agent, VS Code, ChatGPT & Codex, Grok Bot, NanoClaw, and other spec-compliant agents), just say in the conversation:

Install the plugin from https://github.com/limccn/deepseek-vl-support and enable it

After a GitHub install, configure the vision endpoint once with npx deepseek-vl-support@latest install --target <your agent> (or environment variables — see Changing settings).

One-command install per agent

Everything below is equivalent to the wizard above — just narrowed to one agent. Pick yours:

1. Install

npx deepseek-vl-support@latest install --target claude

2. After install — restart your session, then read any picture: the description arrives automatically (/vision path.png for manual use).

1. Ask Codex to install it

Install the plugin from https://github.com/limccn/deepseek-vl-support and enable it

2. Or install via npx

npx deepseek-vl-support@latest install --target codex

3. After install — restart Codex, then ask it to describe a picture.

1. Install

npx deepseek-vl-support@latest install --target opencode

2. After install — restart OpenCode.

1. Install

npx deepseek-vl-support@latest install --target trae

2. After install — import the skill once: Settings → Rules & Skills → Create/Import.

1. Native install (recommended) — skill + extension in one command

pi install npm:deepseek-vl-support

2. Or install via npx

npx deepseek-vl-support@latest install --target pi

3. After install — restart Pi.

1. Native install (recommended)

omp install npm:deepseek-vl-support

2. Or install via npx

npx deepseek-vl-support@latest install --target omp

3. After install — run /reload-plugins (no restart needed).

1. Native install (recommended) — in-process tools, no subprocess

dsh plugin --profile web add deepseek-vl-support@latest

2. Or install via npx

npx deepseek-vl-support@latest install --target dsh

3. After install — restart the dsh web session.

1. Install

npx deepseek-vl-support@latest install --target qwen

2. After install — restart Qwen Code.

1. Install

npx deepseek-vl-support@latest install --target reasonix

2. After install — restart Reasonix.

1. Install

npx deepseek-vl-support@latest install --target kilo

2. After install — restart Kilo Code.

1. Install

npx deepseek-vl-support@latest install --target workbuddy

2. After install — restart WorkBuddy.

1. Install

npx deepseek-vl-support@latest install --target devin

2. After install — restart Devin. (Devin's CLI has no official npm package — download it from https://devin.ai/download.)

1. Ask Copilot to install it

Install the plugin from https://github.com/limccn/deepseek-vl-support and enable it

2. Or install via npx

npx deepseek-vl-support@latest install --target copilot

3. After install — check copilot plugin list.

1. Ask Cursor to install it

Install the plugin from https://github.com/limccn/deepseek-vl-support and enable it

2. Or install via npx

npx deepseek-vl-support@latest install --target cursor

3. After install — reload the window (Developer → Reload Window).

1. Ask Kiro to install it

Install the plugin from https://github.com/limccn/deepseek-vl-support and enable it

2. Or install via npx

npx deepseek-vl-support@latest install --target kiro

3. After install — import once: Kiro → Powers → Add Custom Power → Import from folder → ~/.deepseek-vl/plugin.

1. Ask OpenClaw to install it

Install the plugin from https://github.com/limccn/deepseek-vl-support and enable it

2. Or install via npx

npx deepseek-vl-support@latest install --target openclaw

3. After install — restart the gateway, verify with openclaw plugins list.

1. Ask Hermes to install it

Install the plugin from https://github.com/limccn/deepseek-vl-support and enable it

2. Or install via npx

npx deepseek-vl-support@latest install --target hermes

3. After install — verify with hermes plugins list.

1. Ask in a VS Code chat

Install the plugin from https://github.com/limccn/deepseek-vl-support and enable it

2. Or install via npx

npx deepseek-vl-support@latest install --target vscode

3. After install — reload the window.

1. Ask ChatGPT or Codex to install it

Install the plugin from https://github.com/limccn/deepseek-vl-support and enable it

2. Or install via npx

npx deepseek-vl-support@latest install --target chatgpt-codex

3. After install — start a new Codex thread or ChatGPT session.

1. Ask Grok to install it

Install the plugin from https://github.com/limccn/deepseek-vl-support and enable it

2. Or install via npx

npx deepseek-vl-support@latest install --target grok

3. After install — press r in the Plugins tab or start a new session.

1. Ask NanoClaw to install it

Install the plugin from https://github.com/limccn/deepseek-vl-support and enable it

2. Or install via npx

npx deepseek-vl-support@latest install --target nanoclaw

3. After install — run ncl wirings create per the printed guidance.

1. Ask Agent to install it

Install the plugin from https://github.com/limccn/deepseek-vl-support and enable it

2. Or install via npx

npx deepseek-vl-support@latest install --target other

Any combination works, comma-separated:

npx deepseek-vl-support@latest install --target claude,copilot

Or all 10 plugin clients in one run:

npx deepseek-vl-support@latest install --target copilot,cursor,kiro,openclaw,hermes,vscode,chatgpt-codex,grok,nanoclaw,other

All supported agents at a glance:

Agent --target
Claude Code claude
Codex codex
OpenCode opencode
Trae trae
Pi Coding Agent pi
Oh My Pi omp
DeepSeek Harness dsh
Qwen Code qwen
Reasonix reasonix
Kilo Code kilo
WorkBuddy (CodeBuddy Code) workbuddy
Devin devin
GitHub Copilot copilot
Cursor cursor
Kiro kiro
OpenClaw openclaw
Hermes Agent hermes
VS Code vscode
ChatGPT & Codex chatgpt-codex
Grok Bot grok
NanoClaw nanoclaw
Other agents other

Try it out

Fastest check — describe a picture directly in the terminal:

npx deepseek-vl-support@latest describe path/to/a/picture.png

A good text description comes back → everything is wired up. From then on, just read pictures in your agent as usual — the description arrives automatically.

Choosing a vision service

The installer offers the same services as presets — no need to remember these URLs unless you configure manually:

Service base URL Example model
Moonshot https://api.moonshot.cn/v1 moonshot-v1-32k-vision-preview
OpenRouter https://openrouter.ai/api/v1 qwen/qwen2.5-vl-72b-instruct
MiniMax https://api.minimaxi.com/v1 MiniMax-VL-01
Zhipu GLM https://open.bigmodel.cn/api/paas/v4 glm-4v-flash
StepFun https://api.stepfun.com/v1 step-1o-turbo-vision
OpenCode Zen https://opencode.ai/zen/v1 mimo-v2.5-free
SiliconFlow https://api.siliconflow.cn/v1 Qwen/Qwen2.5-VL-72B-Instruct
DashScope https://dashscope.aliyuncs.com/compatible-mode/v1 qwen-vl-max
Ollama (local) http://localhost:11434/v1 qwen2.5vl:7b (run ollama pull qwen2.5vl:7b first)
llama.cpp (local) http://localhost:8080/v1 llava (llama-server -m llava.gguf)
vLLM (local) http://localhost:8000/v1 deepseek-ai/deepseek-vl2
LM Studio (local) http://localhost:1234/v1 qwen2.5-vl-7b-instruct

Everyday commands

What you want Command
Install npx deepseek-vl-support@latest install
Health check npx deepseek-vl-support@latest doctor
Describe a picture now npx deepseek-vl-support@latest describe picture.png
See current settings npx deepseek-vl-support@latest config get
Change a setting npx deepseek-vl-support@latest config set maxBytes 5242880
Remove the tool npx deepseek-vl-support@latest uninstall

Changing settings

Your answers are saved in .deepseek-vl/config.json inside the project folder — usually you never need to touch it. The two settings worth knowing:

Setting Meaning Default
maxBytes Pictures bigger than this are skipped (saves time and money) 10485760 (10 MB)
timeoutMs How long to wait for one description 120000 (2 minutes)

Example — skip pictures over 5 MB:

npx deepseek-vl-support@latest config set maxBytes 5242880

Describing the same picture twice is free: results are cached on your machine (64 MB limit). Change the picture and it gets described again. Everything can also be set with environment variables (VISION_MODEL, VISION_BASE_URL, …) — see CLAUDE.md for the full reference.

Troubleshooting

Symptom What to do
The model still doesn't describe pictures Restart the session (required after install), then run … doctor and look for [OK].
doctor says no model configured You chose Decide later during install. Configure a model now: config set model <id> (plus config set baseUrl <url> if not using the default).
doctor shows "unreachable" / no [OK] The service address or key is wrong — check the base URL ends with /v1 and the API key is correct.
"image too large" hint Compress or crop the picture (e.g. under 5 MB, long side ~2000 px), or raise the limit with config set maxBytes ….
Descriptions are slow Lower the limit or switch to a faster service (see the table above).
Pasted (Ctrl+V) pictures are not described Pasted images bypass the read path — save the picture as a file first, then read it (or use /vision / describe_image).

More edge cases (Windows encoding, Codex-specific quirks, reasoning-model notes) live in CLAUDE.md and docs/README.zh-CN.md.

Acknowledgements

This project was inspired by pi-deepseek-vision — thanks to psychobarge for the open-source work.

Contributing

Contributions are welcome — see CONTRIBUTING.md for how to report issues and set up a development environment.

License

MIT

原始 README: https://github.com/limccn/deepseek-vl-support/blob/main/README.md ↗

同类插件

查看全部 →
模型与账号接入Anionex

agent-vision-toolkit

为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode

查看详情
928github+08-16
模型与账号接入toby-bridges

api-relay-audit

从 DeepSeek Harness 对 AI API 中转站和 LLM 代理运行本地安全审计,生成 Markdown 报告,覆盖提示词注入、模型替换信号、工具调用改写、错误泄漏、流完整性和按 profile 启用的 Web3 风险。

查看详情
791github+08-21
模型与账号接入ZJU-LLMs

OpenStory

✨ OpenStory 现已支持 DeepSeek Harness 插件! 现在可以通过 dsh-openstory 将 OpenStory 多智能体推演接入 DeepSeek Harness,让 agent 直接启动模拟、查看角色、下达指令并逐回合推进故事。查看 DSH 插件配置与使用指南。

查看详情
377github+08-17
模型与账号接入pulseaiclub

phi

pi的编码代理 ∞ 提供者、子代理、hashline编辑和权限门

查看详情
89github+08-16
模型与账号接入anysearch-team

anysearch-dsh

DeepSeek Harness(DSH)的 AnySearch 网络搜索提供方与高级搜索工具。

查看详情
79github+08-17
模型与账号接入kuangre123

codex-switch

Codex Switch 是一个 macOS 工具,一键配置 Codex 的自定义 API,同时保留官方 OpenAI 登录。保存后 Codex 的模型选择器里只会出现你选的那个 provider 的模型。也支持 Claude Code 的官方 / 自定义 API 切换。Codex Switch is a lightweight helper for configuring multiple coding-agent API routes. For Codex, it keeps Official OpenAI and a custom API provider configured in parallel, registers the custom model in Codex's mod

查看详情
67github+08-16