DeepSeek Harness Plugin Hub

CATEGORY

Vision & Media DSH plugins

Verified manifests and exact versions in this category.

All plugins

377 plugins
v0.1.1
@fzhiyu/dsh-image-view

Image View

Show the image a read_image call returned, inline in the tool row it belongs to

vision-media43
Updated 9/4/2026
v0.1.0
@evan-williams/vision-tool

Vision Tool

Conversation-style UI/UX visual review plugin for DeepSeek Harness: vision_review / vision_ask tools backed by any OpenAI-compatible multimodal endpoint (e.g. agnes-2.5-flash)

vision-mediaui-customization039
Updated 9/4/2026
v0.1.3
@pzqian123/dsh-tool-vision

Tool Vision

DSH plugin: read images through a user-configured vision-capable model when the main model route cannot accept image input

vision-media039
Updated 9/4/2026
v0.4.0
dsh-eyes

Eyes

A vision bridge profile bundle for DeepSeek Harness (dsh): gives non-vision models image-reading capability by delegating transcription to a vision-capable model, with GUI paste admission and a web settings page for vision model selection.

vision-media036
Updated 9/4/2026
v0.1.1
dsh-qr-toolkit

Qr Toolkit

QR code & barcode toolkit for DeepSeek Harness: decode images (PNG/JPEG, local or URL) and generate QR codes as PNG. 二维码/条码工具箱:解码与生成。

vision-media030
Updated 9/4/2026
v0.2.0
dsh-computer-use

Computer Use

Computer Use plugin: virtual mouse human-like operation + native vision model integration (deepseek-v4-flash-vision-exp directly reads screenshots), 12 model-friendly tools including screen_observe

vision-media230
Updated 8/26/2026
v0.15.0
dsh-ros2

Ros2

ROS2 debugging tools and diagnostics skills for DeepSeek Harness (DSH): package/workspace/dependency inspection, node/topic/service/action/param/interface enumeration, topic sampling, TF tree, graph topology JSON, approval-gated builds and message scaffolding, GUI screenshot and multimodal vision ob

developer-toolsvision-media180
Updated 8/24/2026
v0.3.2
dsh-visual-plugin

Visual Plugin

Visual media plugin for DeepSeek Harness: copy native image descriptions and securely normalize, inspect, and play scene-aware videos.

vision-media150
Updated 9/2/2026
v0.4.3
dsh-vision-opencode

Vision Opencode

DeepSeek Harness plugin: configurable vision model with vision_read_image tool, composer-bar vision-model selector, and automatic image-to-text conversion for text-only main models.

vision-mediamodels-usage130
Updated 8/30/2026
v0.4.1
dsh-chat-imagine

Chat Imagine

Generate and display images inline in the DSH chat via API channels or local CLIs (mmx / codex / agy), and read images into structured JSON evidence (OCR / layout / semantics) on any model — with backend probing and a bundled recovery skill.

vision-mediadeveloper-tools110
Updated 8/26/2026
v0.4.0
dsh-galgame-generator

Galgame Generator

Galgame generator for DeepSeek Harness (DSH): give it a script document plus character/background/BGM/animation assets in the workspace (img_human/, img_bg/, audio/, img_cg/) and it builds a playable visual-novel web page with configurable save slots, speaker-following multi-pose sprites, [表情]/[位置]/

ui-customizationvision-media100
Updated 9/4/2026
v0.3.8
dsh-windows-ocr

Windows Ocr

dsh plugin: recognize attached images with the built-in Windows OCR engine (Windows.Media.Ocr) and send only the recognized text to the model. Text models never receive image bytes; vision passthrough is opt-in.

vision-media90
Updated 9/4/2026
v0.2.0
dsh-highres-vision

Highres Vision

DeepSeek Harness vision enhancer: raises image limits to official values and provides high-resolution tiled image reading via a dedicated tool.

vision-media80
Updated 8/25/2026
v0.3.1
dsh-read-image-view

Read Image View

Display images read by the read_image tool inside the DeepSeek Harness Web GUI: a dedicated Read image row with a default-expanded message-style image card (PS-style transparency checkerboard), merged read-N-images rows with side-by-side frames for multi-image requests, and an in-page full-resolutio

vision-mediaui-customization60
Updated 9/2/2026
v0.2.6
@nn12138/dsh-voice

Voice

Voice input plugin for DeepSeek Harness: microphone → local/browser speech recognition → text submitted as a normal chat message (input-only, preset-agnostic)

vision-media60
Updated 8/21/2026
v0.3.8
dsh-tesseract-ocr

Tesseract Ocr

dsh plugin: recognize attached images locally with Tesseract OCR and send only the recognized text to the model. Text models never receive image bytes; vision passthrough is opt-in.

vision-media50
Updated 9/4/2026
v0.1.2
dsh-tool-see-image

Tool See Image

see_image tool for DSH: route image files to a configurable vision model (OpenAI-compatible API) and relay its description back to a text-only model

vision-mediamodels-usage50
Updated 8/27/2026
v0.5.1
dsh-tool-describe-image

Tool Describe Image

DSH plugin: image understanding via any OpenAI-compatible vision API, paste-to-describe, and an animated whale-buddy desktop pet with status bubbles and a floating settings panel

vision-mediaui-customization50
Updated 8/25/2026
v0.8.0
@haoku123/dsh-voice

Voicealt · @haoku123

Full-duplex voice mode for DeepSeek Harness: streamed ASR -> LLM -> TTS with barge-in

vision-media50
Updated 8/23/2026
v0.2.0
dsh-image-subagent

Image Subagent

Allow text-only primary models (DeepSeek V4, etc.) to receive image attachments: remove the modality declaration from the text-only route to pass the 0.1.1 admission gate, project images into placeholder text carrying the complete attachmentId, and have the primary model delegate reading to a visual

vision-mediaagents-orchestration50
Updated 8/22/2026
v0.1.2
dsh-paddleocr-skills

Paddleocr Skills

PaddleOCR text-recognition and document-parsing skills with native DeepSeek Harness tools and GUI configuration.

vision-media40
Updated 9/2/2026
v0.3.0
@limccn/deepseek-vl-support

Deepseek Vl Support

Give DeepSeek (text-only) models vision in Claude Code, Codex, and Agent Plugins clients: describe images via any OpenAI-compatible vision endpoint.

vision-media40
Updated 9/1/2026
v1.1.2
@paicat1/dsh-screenshot

Screenshot

Standalone screen capture for DeepSeek Harness (dsh): browser hotkeys plus an agent-facing capture+read tool. Forked out of @liustack/modlens#48 (upstream declined the feature).

vision-media40
Updated 8/30/2026
v0.2.9
deepseek-vl-support

Deepseek Vl Supportalt · limccn

Give DeepSeek (text-only) models vision in Claude Code, Codex, and Agent Plugins clients: describe images via any OpenAI-compatible vision endpoint.

vision-media40
Updated 8/20/2026
v0.1.0
dsh-stt-input

Stt Input

Speech-to-text voice input for the DeepSeek Harness (DSH) web GUI: click the mic button in the composer to turn speech into text in the input box. Supports the browser's built-in Web Speech API (zero-config, Chrome/Edge) and OpenAI-compatible Whisper endpoints (OpenAI / Groq) with the STT model sele

vision-media40
Updated 8/20/2026
v0.1.0
dsh-approval-voice

Approval Voice

DSH Web GUI approval voice notifications: when a popup requiring approval or a response appears, play a notification sound and announce it by voice to avoid missing it.

vision-mediaui-customization30
Updated 9/4/2026
v0.6.1
dsh-pseudo-vision

Pseudo Vision

Adds opt-in image-capable sibling routes for text-only providers and converts image blocks into budgeted, locally enhanced OCR, color-statistics, pixel-scan, and metadata evidence before delegation.

vision-mediamodels-usage30
Updated 9/4/2026
v2.9.2
@xiaoyuink/dsh-image-vision

Image Vision

Image recognition plugin: Automatically detects whether the current model has vision capabilities. If so, it uses the current model with the plugin’s preset prompts for analysis; otherwise, it calls the vision model configured in the plugin. Text models can paste, upload, or drag images directly int

vision-mediaui-customization30
Updated 9/4/2026
v0.1.0
teach-math-with-manim

Teach Math With Manim

Open-source companion repository for the book 《Manim,让数学看得见》: source code for 60+ educational animation examples + the manim-teaching AI Skill (including the DeepSeek Harness plugin)

vision-mediadeveloper-tools30
Updated 8/31/2026
v0.1.0
dsh-image-inline

Image Inline

Let the model render an image into the DSH web chat flow (QQ/WeChat style) via a show_image tool: the message content keeps only the path text, the image never enters model context, and the browser shows the picture inline, scrollable with history.

ui-customizationvision-media30
Updated 8/30/2026