DeepSeek Harness Plugin Hub

CATEGORY

Vision & Media DSH plugins

Verified manifests and exact versions in this category.

All plugins

377 plugins
v0.8.0
dsh-plugin-grok2api-media-tool

Plugin Grok2api Media Tool

dsh bundle plugin: grok2api image & video generation tools plus image recognition (generate_image / generate_video / recognize_image) for the DeepSeek Harness

vision-media30
Updated 8/24/2026
v0.2.3
dsh-vision-tiler

Vision Tiler

Loss-aware, full-coverage image tiling for DeepSeek Harness vision models

vision-media30
Updated 8/23/2026
v0.1.0
dsh-vision

Vision

Local vision bridge for DeepSeek Harness

vision-media30
Updated 8/22/2026
v0.1.1
dsh-antv-infographic

Antv Infographic

AntV Infographic renderer for DeepSeek Harness: render AI-generated infographic fences inline with streaming, editing, and SVG/PNG export.

vision-mediaui-customization30
Updated 8/21/2026
v0.1.0
@dsh-external/dsh-ocr1-memory

Ocr1 Memory

Based on the DeepSeek-OCR1 optical compressed memory system: renders memories as images for storage, with support for SoM segmentation, age-based decay/blur, activation-based recall, and DSH Agent retrieval

memory-contextvision-media30
Updated 8/21/2026
v0.1.0
meow-vision

Meow Vision

dsh plugin: gives text models a pair of eyes + gives multimodal models a pair of eyes to observe UI rendering — settings popup for selecting a vision model + meow_vision tool + meow_preview component screenshot tool

vision-mediaui-customization30
Updated 8/21/2026
v1.5.0
@xiaoyuink/dsh-image-create

Image Create

AI 生图 (image generation) plugin for the dsh web GUI: text-to-image and image-to-image through a configurable OpenAI-compatible endpoint, with multi-provider support, model discovery, and agent tool registration. 与 @xiaoyuink/dsh-image-vision 同系列的图像插件。

vision-media20
Updated 9/4/2026
v0.1.6
@allmodels/dsh-speech

Speech

Speech-to-text and spoken answer summaries for DeepSeek Harness using AllModels.io.

vision-media20
Updated 9/4/2026
v0.6.0
dsh-mmroute

Mmroute

多模态路由 for DeepSeek Harness:纯文本模型的每一步请求里的每张图片(用户上传 / 工具返回 / MCP 渲染)都先经多模态理解模型转述为文字再作答,图片类报错自动转述重试,两类模型全程交叉协作。Multimodal router for DSH: every image in every step of a text-only model's stream is transcribed to text by a multimodal understander first; image-related failures auto-recover with a transcr

vision-mediamodels-usage20
Updated 9/4/2026
v0.4.4
dsh-file-convert

File Convert

Local-first file conversion for DeepSeek Harness - images, PDF, data, audio/video and office documents. No API keys, no uploads, no token cost.

productivity-workflowvision-media20
Updated 9/3/2026
v0.1.2-rc.1
dsh-vision-analysis

Vision Analysis

Model-facing analyze_image tool for the DeepSeek Harness: multi-modal image understanding via any OpenAI- or Anthropic-compatible vision API, with 8 analysis modes, local path / http(s) URL / data URL input, and a Web UI hint that guides image-incapable models to the reliable local-path route.

vision-media20
Updated 9/3/2026
v0.1.3
dsh-auto-image

Auto Image

Bridges images into text for DeepSeek Harness: when the session's selected model cannot see images, a configured vision model describes them and the descriptions enter the durable session history as folded context rows — your message stays untouched.

vision-mediamemory-context20
Updated 8/31/2026
v0.2.0
dsh-photo-skins

Photo Skins

Photo skins for the dsh web GUI: import your own photos as UI skins (backdrop layer + auto accent extraction)

ui-customizationvision-media20
Updated 8/31/2026
v0.1.0
@crosery/dsh-viewer

Viewer

Renders images, video, audio, PDF and Office documents inline in the dsh web UI through a display_file tool, streaming bytes over a signed HTTP route with range support.

vision-media20
Updated 8/31/2026
v1.0.2
dsh-glm-vision

Glm Vision

GLM vision model plugin: registers the glm-vision provider route (glm-4.6v series, with image+text input declaration) and provides the glm_vision tool, allowing text-based primary models such as DeepSeek to directly call Zhipu vision models to analyze images.

vision-mediamodels-usage20
Updated 8/30/2026
v0.2.1
@max-null/dsh-capture

Capture

SSiD (思灵) quick screenshot capture: tray/hotkey → fullscreen box-select overlay → cropped image into the current conversation composer

vision-media20
Updated 8/26/2026
v0.1.0
dsh-show-media

Show Media

Show a local image or short video inside the current DeepSeek Harness conversation card, with click-to-preview. Does not replace the official read_image model path.

vision-mediaui-customization20
Updated 8/25/2026
v0.2.0
taxue-dsh-artisan

Taxue Dsh Artisan

taxue 画师: an integrated DSH toolkit for reverse-engineering image/video prompts, deterministic color palette extraction, audit optimization, and multi-provider image generation (with optional image-to-image control).

vision-media20
Updated 8/24/2026
v1.2.0
dsh-visibridge

Visibridge

Host-level vision bridge for text-only models: analyze_image tool (Ollama local / Xiaomi MiMo cloud / any OpenAI-compatible endpoint) returning structured evidence.

vision-mediaintegrations-communication20
Updated 8/23/2026
v0.2.1
dsh-medomni

Medomni

Medical imaging specialist tools for deepseek-harness: MAIRA-2 and MedGemma report generation for chest X-ray/CT/MRI/retinal (fundus) images, TotalSegmentator organ segmentation, BiomedCLIP zero-shot ultrasound classification, and BiomedParse text-prompted segmentation across X-ray, CT, MRI, ultraso

vision-media20
Updated 8/22/2026
v1.0.0
@kobenfang/dsh-eyes

Eyes

Agent Skill for DeepSeek Harness (dsh): eyes.

vision-media20
Updated 8/22/2026
v0.2.0
dsh-browser-vision

Browser Vision

Browser tool for DeepSeek Harness that can see the page: drives a real Chrome over CDP with browser-use and reads it with deepseek-v4-flash-vision-exp, so canvas text, text inside images and rendered charts are readable. Schema-driven JSON extraction, Markdown output, and per-run token/cost accounti

search-researchvision-media20
Updated 8/22/2026
v0.1.2
dsh-image-vision-bridge

Image Vision Bridge

DSH host plugin: automatically sends images in user messages to a vision model (default mimo-v2.5 on opencode-go) and feeds the returned text description to the main text model (e.g. deepseek-v4-pro), without touching the visible chat transcript.

vision-mediamodels-usage20
Updated 8/22/2026
v0.4.0
@chenjie1129/dsh-remotion-video-plugin

Remotion Video Plugin

Remotion video creation and verified rendering plugin for DeepSeek Harness

vision-media20
Updated 8/22/2026
v0.1.0-rc.6
@deepseek-ai/dsh-tool-vision-read

Tool Vision Read

DSH plugin: vision_read — route image reading to a dedicated vision model (e.g. Kimi K3) so text-only agents can see images

vision-media20
Updated 8/21/2026
v0.1.0
wavespeed-dsh-skill

Wavespeed Dsh Skill

WaveSpeed skill for DeepSeek Harness (dsh): generate and edit AI media (image, video, audio, 3D) via the open-source wavespeed CLI.

vision-media20
Updated 8/21/2026
v0.4.4
dsh-vision-skill

Vision Skill

DSH standard vision skill: Qwen dynamic-resolution preprocessing + OpenAI-compatible VLM chain with failover/429 backoff, structured evidence mode, local tesseract-first long-screenshot OCR, paste-to-path (no framework patch). 8 tools + runtime skill.

vision-media20
Updated 8/21/2026
v1.3.0
dsh-vision-imagen

Vision Imagen

DeepSeek Harness 全能插件:识别/生图/改图一体化,无需切换模型——用常规 DeepSeek 模型即可自动调用视觉与生图模型(gemini_vision / gemini_generate_image / gemini_optimize_image)。多后端:Gemini 原生 + 任意 OpenAI 兼容服务(GPT-4o、Qwen-VL、GLM-4V、Moonshot、gpt-image、DALL-E、Flux、Stable Diffusion、OpenRouter、硅基流动、各类中转等),生成后自动视觉自检反馈,优于 modlens。

vision-media20
Updated 8/21/2026
v1.0.3
@lp181818/dsh-vision-plugin

Vision Plugin

DSH vision/image-recognition plugin: enables AI to understand images with a configurable vision model

vision-media10
Updated 9/4/2026
v1.0.9
@wisdoverse/dsh-inline-media-viewer

Inline Media Viewer

Persistent inline image, video, and audio previews for DeepSeek Harness Web conversations, with workspace-confined local reads and an optional ComfyUI proxy.

vision-mediaui-customizationintegrations-communication10
Updated 9/4/2026