DSH Plugin Store

Category

Vision

OCR, screenshots, multimodal bridges

78 plugins

modlens

1.1k

liustack

The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | 全网第一个 DeepSeek Harness 视觉插件,为 DeepSeek、GLM 等纯文本模型外挂视觉能力,粘贴图片即得结构化 JSON 证据(OCR、版面、语义)。

agent-vision-toolkit

778

Anionex

为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode

dsh-vision-toolkit

285

Anionex

让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.

sealos-skills

70

labring

AI agent skills for Sealos — deploy any project, provision databases, object storage & more with one command. Works with Claude Code, Gemini CLI, Codex.

dsh-browser

69

Lum1104

dsh plugin: Chrome sidebar extension that lets DSH operate your browser directly—no vision capabilities required.

dsh-vision-router

17

ysr666

Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool-calling turns.

dsh-vision

15

william-jin-cmu

dsh 插件:给纯文本 DeepSeek 加视觉——view_image 工具桥接任意 OpenAI 兼容 VLM(默认智谱免费档,实测 4 厂商 10 模型)

dsh-vision

7

oil-oil

Near-native image understanding for DeepSeek Harness

image-vision

7

wangyang10

No description

dsh-vision-proxy

5

Flyvhidbwo

DeepSeek Harness 插件:DeepSeek 大脑 + 自动识图。附加图片自动经 VLM 转译成文字后交给 DeepSeek 作答

dsh-plugin-deepeye

4

Favio8

DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.

dsh-drop-to-path

3

loudMore

DSH 插件:图片与文件直达纯文本模型——图片保留原生附件体验,PDF/Office/压缩包/视频/音频显示为附件栏方块,点击发送时自动转为工作区路径,配合 dsh-vision-toolkit 粘贴即看图。A DSH plugin that delivers images AND files to text-only models as workspace paths: images keep the native attachment UI, other files show as square chips in the rail, paths append on send — pairs with dsh-vision-toolkit.

GrassVison

3

moduqishi

给 DeepSeek 等纯文本大模型外挂图像理解能力的实现无感添加视觉能力。提供 OpenAI 兼容的 API,自动将图片请求交给视觉模型分析,再将结构化结果注入文本模型,使增强后的模型体验接近原生多模态。

deepseek-harness-vision-plugin

2

sjscy05

No description

dsh-vision-provider

2

libinyam

Config-only DeepSeek Harness bundle for OpenAI-compatible vision models.

shadow-vision

2

WardLu

Open-source MCP vision server that gives text-only LLMs and AI agents image understanding, OCR, visual analysis, UI inspection, and multimodal capabilities.

free-vision-skill

2

niyongsheng

Local‑only vision skill for macOS 本地化识图技能

dsh-tool-vision

2

Scorp1o117

Vision model for DeepSeek Harness | DeepSeek Harness 外置视觉模型插件

dsh-vision-sidecar

2

121103qwq

Hosted free vision sidecar for DeepSeek Harness with durable session evidence

dsh-plugin-describe-image

2

whitelonng

DeepSeek Harness plugin: describe_image — give a text-only model vision through an OpenAI-compatible VLM endpoint

pi-mm-vision

2

Elohia

Synesthesia Encoder (通感编码器) — give any text-only LLM (DeepSeek, etc.) the ability to see images via structured spatial text encoding. A Pi agent extension.

dsh-PaddleOCR-Skills

1

Aidenwu0209

PaddleOCR skills for DeepSeek Harness with native tools and GUI configuration

dsh-vision-LMstudio

1

TiankunDai

让你能通过deepseek harness调用LM studio加载的本地视觉模型

dsh-vision-bridge

1

Xieweikang123

Give a text-only dsh model eyes: pasted images recognized into text via an OpenAI-compatible vision endpoint.

dsh-plugin-vision-toolkit

1

YYTbit

Vision toolkit for DeepSeek Harness -- give text-only agents eyes

multimodal-bridge

1

Spirit4471

multimodal-bridge 是一个多模态能力桥:把 Qwen 的视觉理解(Qwen-VL)与图像生成(Qwen-Image)带给没有原生多模态能力的纯文本模型(如 DeepSeek)。它有两种形态、同一套后端: MCP Server(qwen_vision / qwen_generate 工具):任何支持 MCP 的宿主(Claude Code、Kimi Code 等)直接挂载; DSH 插件(npm 包 dsh-multimodal-bridge,DeepSeek Harness bundle):dsh plugin add 一行安装,含模型自动 fallback、尺寸自适应与图片结果卡片。

better-model-provider

1

sanshanya

Per-model capability declaration for DeepSeek Harness: reasoning-effort levels (wire spellings) + request modalities (vision) for OpenAI-compatible providers. Settings section, zero runtime harness deps, no YAML.

deepseek-eyes

1

fryghost

Community plugin for DeepSeek Harness: give text-only models eyes - paste images natively, described via an OpenAI-compatible vision API

deepseek-hsrness-devkit

1

2472786266-spec

DSH DevKit: multimodal gallery + multi-agent supervision console (DeepSeek Harness dynamic Cordis plugin)

dsh-ui-spec

1

yumimanji

DeepSeek Harness plugin: turn UI screenshots into structured, implementation-grade web frontend specs. Deterministic geometry (sharp) + optional vision-model semantics, merged into one JSON + Markdown spec.

deepseek-visionary

1

xlight

使用 DeepSeek 官方多模态视觉模型让你的 Agent 不再眼瞎(支持 DSH、Zed、OpenCode、Codex、Claude Code、Cursor、Claude Desktop)

dsh-mimo-vision-hint

1

Isekai-Mfu

DSH plugin: dispatch image-recognition tasks to an opencode-go mimo-v2.5 subagent via system-prompt injection

dsh-toolbelt

1

cking000bigdemon

Eight DeepSeek Harness plugins: persona, language guard, per-request vision fallback, python/windows write guards, cross-agent memory, image generation, and skill shell injection.

slcatwujian-dsh-vision-plugin

1

yan5236

让不支持图片输入的主模型通过已配置的视觉模型理解图片的 DSH 插件:自动桥接、像素坐标描述、vision_ask 追问工具与设置页

dsh-media-skills

1

akqwpeter-prog

Free image reading & generation for DeepSeek Harness — paste an image into any chat, even text-only sessions. 免费读图·生图 · 9 种语言 · 无 Key 入库

dsh-mimo-agent-tools

1

ch1bug

Xiaomi MiMo search + multimodal tools for DeepSeek Harness agents: mimo_search/vision/audio/video/asr/tts

dsh-plugin-mm-vision

1

Elohia

No description

dsh-multimodal

1

MC5lan

给 DeepSeek 安装一双眼睛和一支画笔:会话里直接贴截图/图片,GLM 视觉模型先精确转写图片内容(报错信息、代码、界面逐字保留),然后 DeepSeek 继续处理你的问题——同一轮完成,全程无感;需要配图时,DeepSeek 自动调用文生图后端出图并显示在会话中。

dsh-legal-dashboard

1

ByronLeeeee

Matter-aware legal workspace dashboard and document agent tools for DeepSeek Harness

sidesight

1

ZhuXinAI

CLI-first vision sidecar for text-only coding agents. Analyze screenshots, diagrams, charts, UI diffs, and videos with OpenAI-compatible multimodal models.

ds-vision-plugin

1

Sorwcyra

Paste images into DeepSeek Harness with a four-model vision race, OCR, and an automatic text bridge.

dsh-vision-helper

1

Yuuz12

DeepSeek Harness Vision Helper/DeepSeek Harness 视觉辅助方案

dsh-computer-use

1

xiaoheizi1212

Model-agnostic Computer Use for DeepSeek Harness: isolated browser, Windows native helper, third-party vision perception, and a Chrome Cookie Bridge.

dsh-see-image

1

tiefeiyu

A see_image vision tool plugin for DeepSeek Harness — describe images through any OpenAI-compatible vision model (GitHub Copilot, OpenAI, Ollama, vLLM, LM Studio).

dsh-Unlimited-OCR-Skill

1

Aidenwu0209

Unlimited-OCR for DeepSeek Harness with a native tool and GUI configuration

doubao-vision-dsh

1

hawkongz

让纯文本模型通过桌面豆包看见聊天图片的 DeepSeek Harness 宿主插件(CDP 桥接,全预设生效,识别可取消)

dsh-vision

1

xiaoshihou514

DeepSeek Harness: vision

dsh-docling

1

Sqhao-O

Native Docling document intelligence for DeepSeek Harness.

dsh-qwen-multimodal

0

wuwangmao

DSH bundle: Qwen multimodal bridge — vision (qwen3-vl), speech-to-text (qwen3-asr), text-to-image (qwen-image), for DeepSeek Harness

vision_kit

0

Seom-ingit

Make your AI agent a math tutor. Structured extraction of vectors, matrices & geometry from math figures, with dimension-consistency + geometric self-check. Vision plugins for DeepSeek Harness, opencode (MCP) & CLI. Verify, don't believe.

dsh-mmx-multimodal

0

welsione

MiniMax multimodal capability hub for DeepSeek Harness (DSH): image understanding (VLM), text/image-to-video, speech, music, audio cover, web search, quota — one mmx_multimodal model tool wrapping the mmx-cli.

DSH_plugins_4U

0

honghudavy-star

DSH 自建插件集合:微信桥接器 + GUI 微信入口补丁,一键安装

dsh-vision-tools

0

moon09300731

DeepSeek Harness 视觉能力全家桶:vision_understand 工具 + 粘贴/拖拽/按钮三入口识图

analyze_image_tool

0

CaseyTso

No description

dsh-eye

0

wenliang9527

No description

dsh-vision-plugin

0

Xin-Zhang-IceMan

DeepSeek Harness 视觉插件:让纯文本模型拥有视觉能力 / Vision plugin for DSH: vision_analyze tool + automatic image transcription for text-only models.

dsh-mac-vision

0

Kevoyuan

On-device macOS OCR and Apple Vision for DeepSeek Harness — one native plugin with a bundled Skill.

dsh-plugins

0

Bernardxu123

DeepSeek Harness (dsh) 插件集合: dsh-sensenova-image 生图 + dsh-vision 看图, 克隆即装

dsh-ccswitch-import

0

chenhaolove89

DeepSeek Harness 插件:从 CCSWITCH 批量导入模型供应商 + visual_describe 视觉描述工具

prismrelay-mcp

0

Arnoldkevin

Vision-first local MCP that gives text-only Agents image understanding through Agnes AI (BYOK).

dsh-vision

0

237229953-create

DSH plugin: text-only models (e.g. DeepSeek-V4) automatically see images via a vision model. Official surface-replace, cache-friendly, human transcript untouched. 纯文本模型自动识图桥

dsh-llm-vision-router

0

dmsobtl

DSH 插件:消息含图片时自动路由到多模态模型,无图片时继续用便宜的 DeepSeek。

dsh-plugin-vision

0

qizhen2021

No description

dsh-vision-relay

0

junhongchashui

零修改、零切换的 DeepSeek Harness 视觉能力插件:纯文本模型粘贴即读图片,云端 + 本地 Ollama 双后端自动切换,ModLens v2 风格结构化证据输出。

dsh-vlm-bridge

0

me9rez

DeepSeek Harness (dsh) bundle plugin: vision_analyze tool lets text-only LLM agents read images via SenseNova VLM, with Schemastery config and single-source credentials

dsh-qwen-mm

0

RRRosmontis

Qwen-MM-Plugins integration bundle for DeepSeek Harness (dsh) — multimodal MCP tools (vision, OCR, ASR, search, video, Blender, FreeCAD) + image attachment bridge. 让 DeepSeek Harness 原生支持多模态。

dsh-voice

0

zhuiyueya

Voice for DeepSeek Harness(dsh) — speech-to-text input + read-aloud TTS for text-only DeepSeek, zero API key.

dsh-multimodal-skill

0

v587d

给纯文本 LLM 一双慧眼。 一个 DeepSeek Harness(DSH)原生 skill + 零依赖 Python CLI, 为 DeepSeek 等纯文本模型补上图像理解与文档解析(OCR、表格、公式、PDF → Markdown), 使用免费额度优先的三方多模态 API,国内网络直连、无需代理。

dsh-vision-android

0

superclaude1

DeepSeek Harness plugin: multimodal vision (OpenAI-compatible) + Android adb UI automation for real-tap mobile app testing

dsh-vision

0

lakeofsky347

Vision Bridge for DeepSeek Harness: 识图路由插件,图片交给视觉模型描述后回填给纯文本 DeepSeek

dsh-vision-tool

0

visail

Paste an image into the chat box and text-only DSH models can "see" it — auto-rewrite of pasted images + analyze_image tool routed to a Kimi vision model.

dsh-luna-vision-bridge

0

ycp424c

DSH adapter that transcribes native image attachments with Codex Luna before delegating to DeepSeek

dsh-auxiliary

0

Gu-ZT

Auxiliary models for DeepSeek Harness: vision understanding and context compression through dedicated model routes.

dsh-image-reader

0

zcXie777

Give DeepSeek Harness agents native image reading: a read_image tool backed by any OpenAI-compatible vision endpoint.

vision-mcp

0

weekitmo

MCP server for image understanding through OpenAI-compatible vision APIs. To provide image recognition capabilities for those large models that do not support Multimodal.

dsh-plugin-vision

0

tdf1995

Vision for text-only LLMs in DeepSeek Harness (DSH): describe images / OCR / VQA via free Gemini & GLM vision APIs

dsh-vision

0

sjakdhasdh

Vision tool plugin for DeepSeek Harness (DSH): give text-only models like deepseek-v4-flash image recognition via Alibaba Bailian / any OpenAI-compatible vision API. 给 DeepSeek Harness 无识图能力模型加识图工具。

dsh-visionary

0

zhuiyueya

Give text-only DeepSeek models eyes — a DeepSeek Harness plugin that transparently converts chat images into OCR text + vision-model descriptions before they reach the LLM. Configure vision backends (GLM-4V, Qwen-VL, Gemini, Ollama…) right in the Models settings page; multi-backend fallback chain, double-layer caching, no config files.