modlens
1.1k ★liustack
The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | 全网第一个 DeepSeek Harness 视觉插件,为 DeepSeek、GLM 等纯文本模型外挂视觉能力,粘贴图片即得结构化 JSON 证据(OCR、版面、语义)。
agent-vision-toolkit
778 ★Anionex
为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode
dsh-vision-toolkit
285 ★Anionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.
sealos-skills
70 ★labring
AI agent skills for Sealos — deploy any project, provision databases, object storage & more with one command. Works with Claude Code, Gemini CLI, Codex.
dsh-browser
69 ★Lum1104
dsh plugin: Chrome sidebar extension that lets DSH operate your browser directly—no vision capabilities required.
dsh-vision-router
17 ★ysr666
Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool-calling turns.
dsh-vision
15 ★william-jin-cmu
dsh 插件:给纯文本 DeepSeek 加视觉——view_image 工具桥接任意 OpenAI 兼容 VLM(默认智谱免费档,实测 4 厂商 10 模型)
dsh-vision
7 ★oil-oil
Near-native image understanding for DeepSeek Harness
image-vision
7 ★wangyang10
No description
dsh-vision-proxy
5 ★Flyvhidbwo
DeepSeek Harness 插件:DeepSeek 大脑 + 自动识图。附加图片自动经 VLM 转译成文字后交给 DeepSeek 作答
dsh-plugin-deepeye
4 ★Favio8
DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.
dsh-drop-to-path
3 ★loudMore
DSH 插件:图片与文件直达纯文本模型——图片保留原生附件体验,PDF/Office/压缩包/视频/音频显示为附件栏方块,点击发送时自动转为工作区路径,配合 dsh-vision-toolkit 粘贴即看图。A DSH plugin that delivers images AND files to text-only models as workspace paths: images keep the native attachment UI, other files show as square chips in the rail, paths append on send — pairs with dsh-vision-toolkit.
GrassVison
3 ★moduqishi
给 DeepSeek 等纯文本大模型外挂图像理解能力的实现无感添加视觉能力。提供 OpenAI 兼容的 API,自动将图片请求交给视觉模型分析,再将结构化结果注入文本模型,使增强后的模型体验接近原生多模态。
deepseek-harness-vision-plugin
2 ★sjscy05
No description
dsh-vision-provider
2 ★libinyam
Config-only DeepSeek Harness bundle for OpenAI-compatible vision models.
shadow-vision
2 ★WardLu
Open-source MCP vision server that gives text-only LLMs and AI agents image understanding, OCR, visual analysis, UI inspection, and multimodal capabilities.
free-vision-skill
2 ★niyongsheng
Local‑only vision skill for macOS 本地化识图技能
dsh-tool-vision
2 ★Scorp1o117
Vision model for DeepSeek Harness | DeepSeek Harness 外置视觉模型插件
dsh-vision-sidecar
2 ★121103qwq
Hosted free vision sidecar for DeepSeek Harness with durable session evidence
dsh-plugin-describe-image
2 ★whitelonng
DeepSeek Harness plugin: describe_image — give a text-only model vision through an OpenAI-compatible VLM endpoint
pi-mm-vision
2 ★Elohia
Synesthesia Encoder (通感编码器) — give any text-only LLM (DeepSeek, etc.) the ability to see images via structured spatial text encoding. A Pi agent extension.
dsh-PaddleOCR-Skills
1 ★Aidenwu0209
PaddleOCR skills for DeepSeek Harness with native tools and GUI configuration
dsh-vision-LMstudio
1 ★TiankunDai
让你能通过deepseek harness调用LM studio加载的本地视觉模型
dsh-vision-bridge
1 ★Xieweikang123
Give a text-only dsh model eyes: pasted images recognized into text via an OpenAI-compatible vision endpoint.
dsh-plugin-vision-toolkit
1 ★YYTbit
Vision toolkit for DeepSeek Harness -- give text-only agents eyes
multimodal-bridge
1 ★Spirit4471
multimodal-bridge 是一个多模态能力桥:把 Qwen 的视觉理解(Qwen-VL)与图像生成(Qwen-Image)带给没有原生多模态能力的纯文本模型(如 DeepSeek)。它有两种形态、同一套后端: MCP Server(qwen_vision / qwen_generate 工具):任何支持 MCP 的宿主(Claude Code、Kimi Code 等)直接挂载; DSH 插件(npm 包 dsh-multimodal-bridge,DeepSeek Harness bundle):dsh plugin add 一行安装,含模型自动 fallback、尺寸自适应与图片结果卡片。
better-model-provider
1 ★sanshanya
Per-model capability declaration for DeepSeek Harness: reasoning-effort levels (wire spellings) + request modalities (vision) for OpenAI-compatible providers. Settings section, zero runtime harness deps, no YAML.
deepseek-eyes
1 ★fryghost
Community plugin for DeepSeek Harness: give text-only models eyes - paste images natively, described via an OpenAI-compatible vision API
deepseek-hsrness-devkit
1 ★2472786266-spec
DSH DevKit: multimodal gallery + multi-agent supervision console (DeepSeek Harness dynamic Cordis plugin)
dsh-ui-spec
1 ★yumimanji
DeepSeek Harness plugin: turn UI screenshots into structured, implementation-grade web frontend specs. Deterministic geometry (sharp) + optional vision-model semantics, merged into one JSON + Markdown spec.
deepseek-visionary
1 ★xlight
使用 DeepSeek 官方多模态视觉模型让你的 Agent 不再眼瞎(支持 DSH、Zed、OpenCode、Codex、Claude Code、Cursor、Claude Desktop)
dsh-mimo-vision-hint
1 ★Isekai-Mfu
DSH plugin: dispatch image-recognition tasks to an opencode-go mimo-v2.5 subagent via system-prompt injection
dsh-toolbelt
1 ★cking000bigdemon
Eight DeepSeek Harness plugins: persona, language guard, per-request vision fallback, python/windows write guards, cross-agent memory, image generation, and skill shell injection.
slcatwujian-dsh-vision-plugin
1 ★yan5236
让不支持图片输入的主模型通过已配置的视觉模型理解图片的 DSH 插件:自动桥接、像素坐标描述、vision_ask 追问工具与设置页
dsh-media-skills
1 ★akqwpeter-prog
Free image reading & generation for DeepSeek Harness — paste an image into any chat, even text-only sessions. 免费读图·生图 · 9 种语言 · 无 Key 入库
dsh-mimo-agent-tools
1 ★ch1bug
Xiaomi MiMo search + multimodal tools for DeepSeek Harness agents: mimo_search/vision/audio/video/asr/tts
dsh-plugin-mm-vision
1 ★Elohia
No description
dsh-multimodal
1 ★MC5lan
给 DeepSeek 安装一双眼睛和一支画笔:会话里直接贴截图/图片,GLM 视觉模型先精确转写图片内容(报错信息、代码、界面逐字保留),然后 DeepSeek 继续处理你的问题——同一轮完成,全程无感;需要配图时,DeepSeek 自动调用文生图后端出图并显示在会话中。
dsh-legal-dashboard
1 ★ByronLeeeee
Matter-aware legal workspace dashboard and document agent tools for DeepSeek Harness
sidesight
1 ★ZhuXinAI
CLI-first vision sidecar for text-only coding agents. Analyze screenshots, diagrams, charts, UI diffs, and videos with OpenAI-compatible multimodal models.
ds-vision-plugin
1 ★Sorwcyra
Paste images into DeepSeek Harness with a four-model vision race, OCR, and an automatic text bridge.
dsh-vision-helper
1 ★Yuuz12
DeepSeek Harness Vision Helper/DeepSeek Harness 视觉辅助方案
dsh-computer-use
1 ★xiaoheizi1212
Model-agnostic Computer Use for DeepSeek Harness: isolated browser, Windows native helper, third-party vision perception, and a Chrome Cookie Bridge.
dsh-see-image
1 ★tiefeiyu
A see_image vision tool plugin for DeepSeek Harness — describe images through any OpenAI-compatible vision model (GitHub Copilot, OpenAI, Ollama, vLLM, LM Studio).
dsh-Unlimited-OCR-Skill
1 ★Aidenwu0209
Unlimited-OCR for DeepSeek Harness with a native tool and GUI configuration
doubao-vision-dsh
1 ★hawkongz
让纯文本模型通过桌面豆包看见聊天图片的 DeepSeek Harness 宿主插件(CDP 桥接,全预设生效,识别可取消)
dsh-vision
1 ★xiaoshihou514
DeepSeek Harness: vision
dsh-docling
1 ★Sqhao-O
Native Docling document intelligence for DeepSeek Harness.
dsh-qwen-multimodal
0 ★wuwangmao
DSH bundle: Qwen multimodal bridge — vision (qwen3-vl), speech-to-text (qwen3-asr), text-to-image (qwen-image), for DeepSeek Harness
vision_kit
0 ★Seom-ingit
Make your AI agent a math tutor. Structured extraction of vectors, matrices & geometry from math figures, with dimension-consistency + geometric self-check. Vision plugins for DeepSeek Harness, opencode (MCP) & CLI. Verify, don't believe.
dsh-mmx-multimodal
0 ★welsione
MiniMax multimodal capability hub for DeepSeek Harness (DSH): image understanding (VLM), text/image-to-video, speech, music, audio cover, web search, quota — one mmx_multimodal model tool wrapping the mmx-cli.
DSH_plugins_4U
0 ★honghudavy-star
DSH 自建插件集合:微信桥接器 + GUI 微信入口补丁,一键安装
dsh-vision-tools
0 ★moon09300731
DeepSeek Harness 视觉能力全家桶:vision_understand 工具 + 粘贴/拖拽/按钮三入口识图
analyze_image_tool
0 ★CaseyTso
No description
dsh-eye
0 ★wenliang9527
No description
dsh-vision-plugin
0 ★Xin-Zhang-IceMan
DeepSeek Harness 视觉插件:让纯文本模型拥有视觉能力 / Vision plugin for DSH: vision_analyze tool + automatic image transcription for text-only models.
dsh-mac-vision
0 ★Kevoyuan
On-device macOS OCR and Apple Vision for DeepSeek Harness — one native plugin with a bundled Skill.
dsh-plugins
0 ★Bernardxu123
DeepSeek Harness (dsh) 插件集合: dsh-sensenova-image 生图 + dsh-vision 看图, 克隆即装
dsh-ccswitch-import
0 ★chenhaolove89
DeepSeek Harness 插件:从 CCSWITCH 批量导入模型供应商 + visual_describe 视觉描述工具
prismrelay-mcp
0 ★Arnoldkevin
Vision-first local MCP that gives text-only Agents image understanding through Agnes AI (BYOK).
dsh-vision
0 ★237229953-create
DSH plugin: text-only models (e.g. DeepSeek-V4) automatically see images via a vision model. Official surface-replace, cache-friendly, human transcript untouched. 纯文本模型自动识图桥
dsh-llm-vision-router
0 ★dmsobtl
DSH 插件:消息含图片时自动路由到多模态模型,无图片时继续用便宜的 DeepSeek。
dsh-plugin-vision
0 ★qizhen2021
No description
dsh-vision-relay
0 ★junhongchashui
零修改、零切换的 DeepSeek Harness 视觉能力插件:纯文本模型粘贴即读图片,云端 + 本地 Ollama 双后端自动切换,ModLens v2 风格结构化证据输出。
dsh-vlm-bridge
0 ★me9rez
DeepSeek Harness (dsh) bundle plugin: vision_analyze tool lets text-only LLM agents read images via SenseNova VLM, with Schemastery config and single-source credentials
dsh-qwen-mm
0 ★RRRosmontis
Qwen-MM-Plugins integration bundle for DeepSeek Harness (dsh) — multimodal MCP tools (vision, OCR, ASR, search, video, Blender, FreeCAD) + image attachment bridge. 让 DeepSeek Harness 原生支持多模态。
dsh-voice
0 ★zhuiyueya
Voice for DeepSeek Harness(dsh) — speech-to-text input + read-aloud TTS for text-only DeepSeek, zero API key.
dsh-multimodal-skill
0 ★v587d
给纯文本 LLM 一双慧眼。 一个 DeepSeek Harness(DSH)原生 skill + 零依赖 Python CLI, 为 DeepSeek 等纯文本模型补上图像理解与文档解析(OCR、表格、公式、PDF → Markdown), 使用免费额度优先的三方多模态 API,国内网络直连、无需代理。
dsh-vision-android
0 ★superclaude1
DeepSeek Harness plugin: multimodal vision (OpenAI-compatible) + Android adb UI automation for real-tap mobile app testing
dsh-vision
0 ★lakeofsky347
Vision Bridge for DeepSeek Harness: 识图路由插件,图片交给视觉模型描述后回填给纯文本 DeepSeek
dsh-vision-tool
0 ★visail
Paste an image into the chat box and text-only DSH models can "see" it — auto-rewrite of pasted images + analyze_image tool routed to a Kimi vision model.
dsh-luna-vision-bridge
0 ★ycp424c
DSH adapter that transcribes native image attachments with Codex Luna before delegating to DeepSeek
dsh-auxiliary
0 ★Gu-ZT
Auxiliary models for DeepSeek Harness: vision understanding and context compression through dedicated model routes.
dsh-image-reader
0 ★zcXie777
Give DeepSeek Harness agents native image reading: a read_image tool backed by any OpenAI-compatible vision endpoint.
vision-mcp
0 ★weekitmo
MCP server for image understanding through OpenAI-compatible vision APIs. To provide image recognition capabilities for those large models that do not support Multimodal.
dsh-plugin-vision
0 ★tdf1995
Vision for text-only LLMs in DeepSeek Harness (DSH): describe images / OCR / VQA via free Gemini & GLM vision APIs
dsh-vision
0 ★sjakdhasdh
Vision tool plugin for DeepSeek Harness (DSH): give text-only models like deepseek-v4-flash image recognition via Alibaba Bailian / any OpenAI-compatible vision API. 给 DeepSeek Harness 无识图能力模型加识图工具。
dsh-visionary
0 ★zhuiyueya
Give text-only DeepSeek models eyes — a DeepSeek Harness plugin that transparently converts chat images into OCR text + vision-model descriptions before they reach the LLM. Configure vision backends (GLM-4V, Qwen-VL, Gemini, Ollama…) right in the Models settings page; multi-backend fallback chain, double-layer caching, no config files.