DeepSeek Harness Plugin Hub

CATEGORY

Vision & Media DSH plugins

Verified manifests and exact versions in this category.

All plugins

376 plugins
v0.1.1
dsh-ui-spec

Ui Spec

DeepSeek Harness plugin that turns UI screenshots into implementation-grade web specs using OCR, deterministic geometry, scene graphs, assets, and render comparison.

vision-mediaui-customization3188v0.1.1 scan passed
Updated 9/4/2026
v0.1.0
dsh-voice-call

Voice Call

dsh-voice-call — give the agent a voice it owns: the agent decides when to speak (offer_call), the human holds the answer key (接听/拒接/稍后). Local-first TTS via CrispASR + Qwen3-TTS CustomVoice, plain audio files under ~/.dsh/voice/. Fork of dsh-voice with a

vision-mediaintegrations-communication3183v0.1.0 scan passed
Updated 9/4/2026
v1.0.1
dsh-model-capability

Model Capability

Explicit model input-modality selection for the DeepSeek Harness Web UI

vision-mediaui-customization3178v1.0.1 scan passed
Updated 9/4/2026
v1.3.2
dsh-sound-lab

Sound Lab

DSH 声音工坊(Sound Lab):Dialogue event sound effects, AI character voice generation, and sound library upload management, all through visual point-and-click controls; includes the 明日方舟安洁莉娜 desktop pet(60fps sprite animation, edge docking, and floating window configuration). Hot-pluggable dsh plugin.

vision-mediaui-customization2221
Updated 9/4/2026
v0.4.1
dsh-plugin-deepseek-vision

Plugin Deepseek Vision

DeepSeek Harness native vision Bundle: paste or drag in images and call OpenAI-compatible vision models through the hosted deepseek-vision-mcp.

vision-mediaintegrations-communication5103v0.4.1 scan passed
Updated 9/4/2026
v0.6.4
dsh-plugin-image-tools

Plugin Image Tools

DSH image plugin: ask_user_choice image/image-text mixed options (Web GUI renders image selection cards, with zoom support) + show_images embeds images in replies (mixed image and text) + click to enlarge all images in the chat area, with scroll wheel/button zoom and drag panning in the enlarged vie

vision-mediaui-customization0622v0.6.4 scan passed
Updated 9/4/2026
v1.2.2
dsh-vision-link

Vision Link

Lightweight, route-preserving vision link for DSH: a configured vision model sees while the selected text model stays in control.

vision-mediamodels-usage2204v1.2.2 scan passed
Updated 9/4/2026
v0.4.0
dsh-voice-kit

Voice Kit

Voice input (Web Speech API) and read-aloud (speechSynthesis) for the DeepSeek Harness web GUI — DSH 语音输入 + 回复朗读套件

vision-media1304v0.4.0 scan passed
Updated 9/4/2026
v0.1.7
dsh-deepseek-vision

Deepseek Vision

Out-of-tree dsh provider plugin: a DeepSeek gateway route that claims image input and transparently describes pasted images through a configured vision-language model (e.g. Qwen-VL) before the text-only DeepSeek wire sees them.

vision-media0572
Updated 9/4/2026
v0.3.1
dsh-sight

Sight

Plug-in vision for text-only DeepSeek Harness (dsh) models: a `vision` tool with built-in cheap/free VLM presets, multi-image batch analysis, paste-to-hint image admission, and a web settings page with hot-reload.

vision-mediaui-customization3127v0.3.1 scan passed
Updated 9/4/2026
v0.1.1-rc.2
dsh-open-eyes

Open Eyes

Open Eyes for DeepSeek Harness: delegate images to a configurable multimodal model through OpenAI Responses, Chat Completions, or Anthropic Messages.

vision-mediamodels-usage0459v0.1.1-rc.2 scan passed
Updated 9/4/2026
v0.6.7
@svenyu/dsvu

Dsvu

dsvu (DeepSeek-Powered Video Understanding), a low-cost video understanding tool: Bilibili links/BV/local videos → information layer (ASR + scenes + object trajectories + YOLO) → summaries + Q&A. Uses same-source instant viewing and answering with deepseek-v4-flash-vision-exp + dense grid frame samp

vision-media0441v0.6.7 scan passed
Updated 9/4/2026
v1.0.0
@dsh-extension/dsh-vision-bridge

Vision Bridge

On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

vision-mediaintegrations-communication657
Updated 9/4/2026
v0.3.0
@maxwell-feng/dsh-windows-ocr

Windows Ocr

dsh plugin: recognize attached images with the built-in Windows OCR engine (Windows.Media.Ocr) and send only the recognized text to the model. Text models never receive image bytes; vision passthrough is opt-in.

vision-media0403v0.3.0 scan passed
Updated 9/4/2026
v0.1.2
dsh-gemini-multimodal

Gemini Multimodal

DeepSeek Harness plugin: multimodal tools (image/audio/video/document understanding, transcription, image generation) via Gemini API or the local Antigravity CLI.

vision-mediamodels-usage1188v0.1.2 scan passed
Updated 9/4/2026
v0.1.0
dsh-mac-vision

Mac Vision

Native macOS OCR and Vision tools for DeepSeek Harness

vision-media1158v0.1.0 scan passed
Updated 9/4/2026
v0.10.0
dsh-vision-fallback

Vision Fallback

DSH Silent Vision Enhancement: Select the main model as usual; images are automatically sent to a fixed vision model and returned to the main model as hidden context.

vision-media2103v0.10.0 scan passed
Updated 9/4/2026
v1.3.0
dsh-tool-visual-primitives

Tool Visual Primitives

DSH visual primitives tool: Based on DeepSeek’s paper “Thinking with Visual Primitives,” routes images to an external vision model and returns text analysis with visual primitives. In a text-only loop, conversational models can “see” images without native vision capabilities.

vision-media2100
Updated 9/4/2026
v0.2.1
dsh-local-vision

Local Vision

DeepSeek Harness plugin: a model-callable local_vision tool that describes images with a local vision model through Ollama. No cloud, the image never leaves the machine. Two tiers: fast (text/colors/summary) and detailed (full description).

vision-media1148v0.2.1 scan passed
Updated 9/4/2026
v0.1.0
@zoytown/dsh-avatar

Avatar

DeepSeek Harness (dsh) plugin for wallpaper theming — upload images from the Settings page, pick one, and the whole dsh web UI renders over it with an adjustable readability mask and blur.

ui-customizationvision-media1143v0.1.0 scan passed
Updated 9/4/2026
v0.2.1
dsh-tool-generate-image

Tool Generate Image

Model-facing generate_image tool for DeepSeek Harness (dsh): generates brand-new images with the Google Gemini image-generation service via the Antigravity CLI, saves them, and returns the paths. Hot-pluggable — install with `dsh plugin --profile web add

vision-mediamodels-usage292
Updated 9/4/2026
v1.0.4
@woyeshishen/dsh-vision-plugin

Vision Plugin

Provides external vision model capabilities for DeepSeek Harness: the text-only primary model uses the describe_image tool to call an external vision model to analyze images and receive a text-only description (multimodal completion). A static Cordis plugin that loads automatically when DSH starts.

vision-media291
Updated 9/4/2026
v1.4.0
dsh-vision-plugin

Vision Pluginalt · Xin-Zhang-IceMan

dsh-vision-plugin: give DeepSeek Harness text-only models a pair of eyes — pasted images are transcribed by a vision model before they reach a text-only main model, plus the vision_analyze tool and a bilingual settings page.

vision-media1132v1.4.0 scan passed
Updated 9/4/2026
v0.1.5
dsh-subagent-vision

Subagent Vision

dsh bundle: subagent_vision — delegate image reading to a vision-capable model from a text-only session, plus paste-to-path so pasted images reach the subagent as file paths.

vision-mediaagents-orchestration1112v0.1.5 scan passed
Updated 9/4/2026
v0.1.3
dsh-vision-guard

Vision Guard

Transparent image guard + vision analysis for DeepSeek Harness: text-only models read pasted images without the 400 session deadlock.

vision-media1105v0.1.3 scan passed
Updated 9/4/2026
v0.1.0
mimo-vision

Mimo Vision

DeepSeek Harness (DSH) native plugin: the describe_image tool, a vision bridge (image -> mimo-v2.5 -> text description) over the ctx.fs / ctx.credentials seams

vision-media269
Updated 9/4/2026
v0.3.1
dsh-plugin-qwen-image

Plugin Qwen Image

DeepSeek Harness out-of-tree plugin: give a text-only coding model eyes by routing images to a Qwen-VL (DashScope) route through ctx.llm and returning text.

vision-mediadeveloper-tools267
Updated 9/4/2026
v0.4.3
@motong/dsh-voice

Voice

A community plugin that adds voice capabilities to DeepSeek Harness (DSH / DeepSeek Hermes): voice input in the input field (with configurable shortcut) and spoken responses (Microsoft Edge neural voices, with voice switching and preview), with no API key required.

vision-media1101v0.4.3 scan passed
Updated 9/4/2026
v0.1.1
dsh-speech-plugin

Speech Plugin

Speech plugin for DeepSeek Harness: per-message speak button, composer voice input, and auto-announce toggle, over cloud TTS/ASR with Web Speech API fallback

vision-media198
Updated 9/4/2026
v0.1.2
dsh-multimodal-bridge

Multimodal Bridge

DeepSeek Harness plugin bundle: qwen_vision (Qwen-VL image understanding) and qwen_generate (Qwen-Image text-to-image and image editing) tools for text-only models

vision-media348
Updated 9/4/2026