DeepSeek Harness Plugin Hub

CATEGORY

Vision & Media DSH plugins

Verified manifests and exact versions in this category.

All plugins

377 plugins
v0.1.0
@hackerfish/dsh-voice-input

Voice Input

Voice input: microphone button in the dialog, browser-built-in Web Speech API (zh-CN) recognition, results inserted into the input field draft

vision-media10
Updated 8/24/2026
v0.3.1
@zenk/vision

Vision

DSH local vision capabilities: macOS Vision OCR + ollama qwen3-vl semantic description + uploaded image bridge (converts image blocks to text, available on text-only channels)

vision-media10
Updated 8/24/2026
v0.1.0
dsh-deepseek-model-router

Deepseek Model Router

Auto-switch DeepSeek models for DeepSeek Harness: routes every model request to the vision model when images are present, to the pro model for complex tasks, and to the fast model otherwise; includes a switch_model tool for manual per-session overrides.

models-usagevision-media10
Updated 8/24/2026
v0.1.1
dsh-voice-plugin

Voice Plugin

Voice input plugin for DeepSeek Harness

vision-media10
Updated 8/24/2026
v0.3.0
dsh-funasr-voice

Funasr Voice

DSH Web local offline voice input plugin: the browser captures microphone audio → host launches local FunASR (SenseVoiceSmall) for recognition → text is entered into the input field.

vision-media10
Updated 8/23/2026
v1.0.0
dsh-imggen

Imggen

DeepSeek Harness plugin: text-to-image output with in-chat image cards, download button, history gallery, and provider selection tabs

vision-media10
Updated 8/22/2026
v0.1.0
dsh-attachment-downscale

Attachment Downscale

Automatic attachment fallback: DSH image limits (one side ≤2000px / ≤3.5MB / ≤40 million pixels) no longer reject oversized images; sharp automatically resizes or re-encodes them before storage, so original phone photos can be uploaded without exceeding limits.

vision-media10
Updated 8/22/2026
v0.1.2
dsh-plugin-88api-image

Plugin 88api Image

88API Image Studio for DSH: four Image2 and Nano Banana models for text-to-image, multi-reference editing, 2K/4K output, and sequential batches.

vision-media10
Updated 8/22/2026
v0.1.1
dock-media

Dock Media

Media player for the dock file explorer: plays audio (MP3, WAV, OGG, FLAC, M4A, ...) and video (MP4, WebM, MOV, ...) files, with a music player for pure audio and fullscreen playback for video.

vision-media10
Updated 8/22/2026
v0.1.3
dsh-text2img-compress

Text2img Compress

把长文本渲染成图片发送来压缩 LLM 输入 token(DeepSeek 视觉模型、每图 384 token 封顶)— text-as-image token compression for DeepSeek Harness

vision-mediamodels-usage10
Updated 8/22/2026
v0.1.1
dsh-computer-use-vision

Computer Use Vision

Windows computer-use capability for DeepSeek Harness: screenshot → vision model → simulated mouse/keyboard input, with self-evolving knowledge base.

vision-media10
Updated 8/22/2026
v0.1.0
dsh-ds-vision-auto-route

Ds Vision Auto Route

Route image-bearing turns to a configurable image-capable model for DeepSeek Harness

models-usagevision-media10
Updated 8/22/2026
v0.1.0
dsh-omnifile

Omnifile

DSH file adapter plugin: drag, paste, or click to select multiple local files; anydoc parses documents, multimodal models recognize images, and the content is integrated for the main model.

vision-mediaproductivity-workflow10
Updated 8/22/2026
v1.0.0
@dsh-extension/dsh-generation-image

Generation Image

On-demand image generation for DeepSeek Harness (DSH): a generate_image tool that calls your own OpenAI-compatible image URL + API key and delivers the image into the session

vision-media10
Updated 8/22/2026
v0.1.0
@dsh-external/dsh-image-tiler

Image Tiler

DSH agent tool: slice a large image into labeled ~800x800 tiles plus an overview thumbnail, preserving detail for vision models.

vision-media10
Updated 8/21/2026
v0.1.0
@local/dsh-image-describe

Image Describe

DeepSeek Harness host plugin: enables text-only models that do not support image input to "see" images (describe_image tool + image marker replacement)

vision-media10
Updated 8/21/2026
v0.1.0
dsh-pro-vision

Pro Vision

Let DeepSeek-V4-Pro (text-only) use V4-Flash-Vision-Exp for attached images. Mac/Windows/Linux.

vision-media10
Updated 8/21/2026
v0.1.0
dsh-official-vision

Official Vision

Direct DeepSeek official vision API bridge for DeepSeek Harness: registers the deepseek-v4-flash-vision-exp multimodal model as a provider route, with base64 inline and Files API image input. 直连 DeepSeek 官方视觉 API 的 DSH 插件

vision-mediaintegrations-communication10
Updated 8/21/2026
v0.1.0
dsh-vision-api-localorweb

Vision Api Localorweb

DSH plugin: Connects to image recognition model APIs (local large model image recognition tool + settings interface). Configure an OpenAI-compatible image recognition endpoint (LM Studio / vLLM / Ollama, etc.); leave the endpoint blank to disable the image recognition model.

vision-mediamodels-usage10
Updated 8/21/2026
v1.0.0
deepseek-visual-plugin

Deepseek Visual Plugin

DeepSeek Harness visual understanding plugin: Converts images in user messages and tool results into textual descriptions for text-only task models (such as DeepSeek).

vision-media10
Updated 8/21/2026
v0.1.0
dsh-culture-movies

Culture Movies

Movie genre

vision-media10
Updated 8/21/2026
v0.0.20
dsh-voice-input-cn

Voice Input Cn

Voice input plugin for DeepSeek Harness web (China-ready): Alibaba Cloud DashScope ASR via a local bridge. Mic button in the composer, streaming recognition, cursor-aware insertion, silence auto-stop.

vision-media10
Updated 8/20/2026
v0.1.0
dsh-imagedit

Imagedit

Local image editing toolkit for DeepSeek Harness: one image_edit tool for deterministic edits — cutout (quick flood-fill or rembg AI), trim, flip, rotate, brightness/contrast/saturation, blur, sharpen, rounded corners, border, canvas resize+center, sprite sheets, and PNG/JPEG/WebP export. Also ships

vision-media10
Updated 8/20/2026
v1.0.2
aura-vision

Aura Vision

Aura Vision — free vision OCR plugin for DeepSeek Harness web profile: Zhipu GLM-4V-Flash (free tier), adaptive tile recognition for long documents, history with favorites and Markdown/Excel/Word/PNG export.

vision-media10
Updated 8/20/2026
v0.6.4
dsh-attachment-formats

Attachment Formats

Codex-style attachment format expansion for the DeepSeek Harness Web GUI: PDF text-layer extraction (pymupdf4llm / pdfjs), Office text extraction, long-document spill + index cards, scanned-PDF OCR (tesseract.js), and browser-decodable images to PNG.

vision-mediaproductivity-workflow10
Updated 8/20/2026
v0.1.5
@mengruo/dsh-vision-toolkit

Vision Toolkit

DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, OCR, grounding, UI restoration, pixel diff, Artifacts, and Web UI.

vision-media00
Updated 9/4/2026
v0.1.0
dsh-reelsmaker

Reelsmaker

DeepSeek Harness plugin: turn a topic into a finished vertical reel. Free neural voice-over, burned-in captions, no API keys.

vision-media00
Updated 9/4/2026
v0.1.8
dsh-voice-announcer

Voice Announcer

End-of-conversation voice announcement: session name + turn count + result (edge-tts streaming / SAPI)

vision-media00
Updated 9/4/2026
v0.1.0
@mokuyoaxis/dsh-iris

Iris

Give DeepSeek Harness eyes and hands: multi-provider media generation, vision routing, and an integrated Iris workbench.

vision-mediaui-customization00
Updated 9/4/2026
v2.2.1
taishan-vision

Taishan Vision

Taishan Vision - Enables DeepSeek Harness text-only models to understand images: GLM visual recognition + inference by the selected model (static plugin package, permanently effective after restart; modified by JingQing v1.1.8)

vision-media00
Updated 9/4/2026