DeepSeek Harness Plugin Hub

CATEGORY

Vision & Media DSH plugins

Verified manifests and exact versions in this category.

All plugins

376 plugins
v0.4.0
dsh-image-gen

Image Gen

Bring ChatGPT-like image generation to DeepSeek Harness — Gemini, OpenAI, Seedream, DashScope, local ComfyUI & more.

vision-media02.7kv0.4.0 scan passed
Updated 9/4/2026
v0.1.13
@eric.wen/dsh-sight

Sight

DeepSeek Harness plugin: direct multimodal image transfer declarations + per-session image clearing, reasoning-effort auto-fill, and progressive Figma MCP bridging (design-to-code + AI-driven design).

vision-mediaintegrations-communication11.3kv0.1.13 scan passed
Updated 9/4/2026
v1.0.3
dsh-codex-tools

Codex Tools

Codex-backed web search, image generation, and image understanding tools for the DeepSeek Harness.

search-researchvision-media5441v1.0.3 scan passed
Updated 9/4/2026
v0.1.1-rc.2
dsh-open-file

Open File

Workspace-bound arbitrary file upload, reading, OCR, and rendering for DeepSeek Harness.

vision-mediaproductivity-workflow4519
Updated 9/4/2026
v0.2.8
dsh-mindseye

Mindseye

MindsEye: model-driven vision tools, structured evidence, and exact cache for DeepSeek Harness

vision-media2794v0.2.8 scan passed
Updated 9/4/2026
v0.3.0
@maxwell-feng/dsh-tesseract-ocr

Tesseract Ocr

dsh plugin: recognize attached images locally with Tesseract OCR and send only the recognized text to the model. Text models never receive image bytes; vision passthrough is opt-in.

vision-media5370v0.3.0 scan passed
Updated 9/4/2026
v0.1.9
@omdp/dsh-vision-bridge

Vision Bridge

DSH Vision Bridge plugin: Automatically distinguishes multimodal and text models. Multimodal models view images directly; text models use a configurable multimodal endpoint (baseUrl + apiKey + model) to view them on their behalf. Supports pasted images, the read_image tool, and converting images int

vision-mediaintegrations-communication2734v0.1.9 scan passed
Updated 9/4/2026
v0.1.0
dsh-vision-web

Vision Web

Eyes for text-only DeepSeek on DeepSeek Harness: image understanding via Doubao Web by default (zero cost, no API key), Antigravity IDE quota (flash/pro), Gemini API, or Cockpit proxy — model-invokable vision tool, wrapper adapters, evidence memory, and a

vision-media8220v0.1.0 scan passed
Updated 9/4/2026
v0.2.1
dsh-screenshot-feedback-hook-mcp

Screenshot Feedback Hook Mcp

Screenshot feedback for DeepSeek Harness — let your coding agent SEE what it builds

vision-mediadeveloper-tools3394v0.2.1 scan passed
Updated 9/4/2026
v0.5.2
dsh-video-understand

Video Understand

Low-cost video understanding tool: Bilibili links/BV/local videos → information layer (ASR + scenes + object trajectories + YOLO) → summaries + Q&A. Question-driven dynamic routing across levels (L0/L1/L2), semantic-layer reuse, and budget caps. Self-contained engine with no external dependencies.

vision-media5241v0.5.2 scan passed
Updated 9/4/2026
v0.4.0-alpha.9
dsh-dictate

Dictate

Context-aware voice input for DeepSeek Harness with Web Speech, local SenseVoice transcription, model polish, editable Composer drafts, and user-controlled sending

vision-mediaproductivity-workflow2460
Updated 9/4/2026
v0.4.0
dsh-auto-vision

Auto Vision

DeepSeek Harness Vision Bridge: automatically discovers your configured multimodal models, equips text-only primary models with a vision tool, and returns recognition results as plain text. Zero configuration, one-command installation.

vision-media2452v0.4.0 scan passed
Updated 9/4/2026
v0.7.2
@xlight-oss/visionary-dsh

Visionary Dsh

DeepSeek Visionary native plugin for DeepSeek Harness: deepseek_vision / status / login / logout native tools plus the text-model image bridge, all backed by the visionary-server CLI (DeepSeek web vision model, no API key).

vision-mediaintegrations-communication1870
Updated 9/4/2026
v0.2.0
dsh-plugin-deepeye

Plugin Deepeye

DeepSeek Harness native plugin: vision capabilities for text-only LLMs (describe, OCR, VQA, layout analysis, clipboard)

vision-media4233v0.2.0 scan passed
Updated 9/4/2026
v0.2.0
dsh-better-input

Better Input

Better input experience for DeepSeek Harness: voice input, AI polishing, PDF and image input (voice input first)

vision-media01.1k
Updated 9/4/2026
v0.7.2
@goodandready/dsh-fal-image-gen

Fal Image Gen

Image generation for DeepSeek Harness: a generate_image tool with pluggable providers — the FAL queue API or any OpenAI-compatible images API. The picture is shown inline in the conversation; the model receives either a link (works with any chat model) or

vision-media1541v0.7.2 scan passed
Updated 9/4/2026
v1.0.4
@alain-prot0s5/dsh-screenshot

Screenshot

Screenshot-to-input for DeepSeek Harness: composer camera button + global hotkey (Alt+A) + listener bound to the app lifecycle, configurable in settings. 截图自动粘贴到 DSH 输入框:相机按钮 + 全局快捷键 + 生命周期绑定 + 设置页配置。

vision-mediaui-customization2338v1.0.4 scan passed
Updated 9/4/2026
v0.3.6
computer-user

Computer User

Codex-style computer use for DeepSeek Harness (DSH): read the screen and drive the mouse & keyboard. Pairs with picturereader (image_scan/image_ocr) to close the look-act-verify loop. Windows.

vision-media1490v0.3.6 scan passed
Updated 9/4/2026
v3.3.1
picturereader

Picturereader

Unified image understanding plugin for DeepSeek Harness (DSH). Visual twin adapter for native thumbnails + auto-analysis on any text-only model (incl. pi-ai providers); privacy/smart/strict routing; local tools (scan/OCR×4 engines (windows/macos/paddle/ra

vision-media0976v3.3.1 scan passed
Updated 9/4/2026
v0.3.3
dsh-ocr-local

Ocr Local

Local OCR for DeepSeek Harness: paste/attach an image, get its text via PP-OCRv5 + ONNX Runtime, fully offline. TUI (cc-tui) and Web. / DeepSeek Harness 本地 OCR 插件:图片转文字,PP-OCRv5 + ONNX Runtime,完全离线,支持 TUI 与 Web。

vision-media6133v0.3.3 scan passed
Updated 9/4/2026
v0.2.0
dsh-dseyes

Dseyes

Give DeepSeek Harness free eyes — natively. Paste or drop an image into the Web GUI composer: it appears as a thumbnail attachment (like a normal AI chat), and before the text-only DeepSeek model is called, the host automatically reads the image with the

vision-media3217v0.2.0 scan passed
Updated 9/4/2026
v0.2.1
dsh-image-plugins

Image Plugins

Multimodal plugin for DeepSeek Harness: understand images and generate images through configurable OpenAI-compatible or DashScope endpoints.

vision-media1422v0.2.1 scan passed
Updated 9/4/2026
v0.2.2
dsh-vision-mix

Vision Mix

A DeepSeek Harness Mix plugin for vision routing plus GPT Image generation and editing.

vision-media3209v0.2.2 scan passed
Updated 9/4/2026
v0.1.3
dsh-vision-free-eyes

Vision Free Eyes

DeepSeek Harness plugin: a model-facing `vision` tool that describes and OCRs image files by calling the free Zhipu GLM vision API directly (no external CLI required).

vision-media2278v0.1.3 scan passed
Updated 9/4/2026
v0.2.0
dsh-vision-recognizer

Vision Recognizer

Adaptive image routing for DeepSeek Harness: pass images directly to native multimodal models and transcribe them only for text-only models.

vision-mediamodels-usage2268v0.2.0 scan passed
Updated 9/4/2026
v0.1.0
dsh-youreyes

Youreyes

Eyes for text-only DeepSeek on DeepSeek Harness: image understanding via Antigravity IDE quota (default, flash/pro), OpenAI-compatible VLM endpoints, Gemini API, or local Ollama — model-invokable vision tool, wrapper adapters for deepseek/opencode-go, evi

vision-media3199v0.1.0 scan passed
Updated 9/4/2026
v0.4.1
dsh-vision-proxy

Vision Proxy

DeepSeek brain + automatic image transcription for DeepSeek Harness: a deepseek-vision provider route that transcribes attached images to text via any OpenAI-compatible VLM before delegating to the text-only DeepSeek adapter. Paid fast path with a key (Da

vision-mediamodels-usage0791
Updated 9/4/2026
v0.1.3
dsh-vision-proxy-route

Vision Proxy Route

DeepSeek Harness plugin: a configurable provider route that transcribes pasted images via free Zhipu GLM vision models before delegating to a text-only adapter.

vision-mediamodels-usage2258v0.1.3 scan passed
Updated 9/4/2026
v0.1.4
dsh-design-qa

Design Qa

Let any text-only model in DeepSeek Harness read images. Image recognition is an on-demand tool—images do not enter the main model context, and you pay nothing if they are not viewed; includes an evaluation set with 23 defects for self-testing when switching models.

vision-media6109v0.1.4 scan passed
Updated 9/4/2026
v0.1.1
dsh-voice-webspeech

Voice Webspeech

DSH Web voice input plugin: no server, no keys, and no model downloads; directly uses the browser's built-in Web Speech API (Edge = Microsoft Azure Speech, Chrome = Google Speech).

vision-media2251
Updated 9/4/2026