DeepSeek Harness Plugin Hub

CATEGORY

Vision & Media DSH plugins

Verified manifests and exact versions in this category.

All plugins

377 plugins
v0.1.0
@lamplitisles/kepos-speech

Kepos Speech

Kepos Speech plugin for Alibaba and ByteDance Chinese speech synthesis plus short-audio recognition in DeepSeek Harness Web

vision-media00
Updated 9/4/2026
v0.4.7
dsh-voice-scribe

Voice Scribe

Voice input plugin for DSH: tap or hold Alt to talk, then release or tap again to convert speech to text. Supports a hotword replacement list (hot.txt), custom polishing prompts, and a recording level indicator. Uses local offline recognition by default (SenseVoice; zero configuration, zero key, and

vision-media00
Updated 9/4/2026
v0.1.0-beta.9
dsh-image-viewer

Image Viewer

Zoom, pan, download, gallery, and region-note image viewer for DeepSeek Harness

vision-media00
Updated 9/4/2026
v0.1.0
dsh-live2d-avatar

Live2d Avatar

A switchable Live2D avatar stage and desktop companion for DeepSeek Harness.

ui-customizationvision-media00
Updated 9/3/2026
v0.3.1
dsh-grok-adaptation

Grok Adaptation

Normalize undersized Grok Responses images with Sharp while leaving supported large images unchanged.

vision-media00
Updated 9/3/2026
v1.0.0
dsh-vision-pro-bridge

Vision Pro Bridge

Give text-only DeepSeek-V4-Pro real vision with zero new dependencies and DeepSeek-only routing: images are described by deepseek-v4-flash-vision-exp (your existing DEEPSEEK_API_KEY), then the text is handed to V4-Pro.

vision-mediaintegrations-communication00
Updated 9/3/2026
v1.0.2
dsh-vision-toggle

Vision Toggle

DSH model vision toggle: adds a “Model Vision” row to the settings page for switching the input vision modality of manually declared models under llm-pi-ai custom routes via the official settings.mutate channel, with changes taking effect immediately.

vision-mediamodels-usage00
Updated 9/3/2026
v0.2.10
dsh-asr-voice

Asr Voice

Speak and it becomes text · Speak-to-prompt for DeepSeek Harness: cloud ASR speech recognition + prompt optimization + fill into draft/auto-send, cross-platform macOS / Windows.

vision-media00
Updated 9/3/2026
v1.3.0
@dshp-inx/vision-bridge

Vision Bridge

DeepSeek Harness (DSH) vision bridge: let text-only models delegate image understanding to multimodal vision models via vision_describe tool, with persistent settings UI and primary/fallback auto-retry. 纯文本模型通过 vision_describe 工具把图片转交给多模态视觉模型识别的桥接插件。

vision-mediaintegrations-communicationmodels-usage00
Updated 9/3/2026
v1.1.0
dsh-pdf-to-word

Pdf To Word

DeepSeek Harness plugin: PDF→Word (.docx) conversion with layout fidelity (fonts/tables/images/borders), OCR scan mode, and optional multimodal LLM verification. Registers the pdf_to_word model tool.

productivity-workflowvision-media00
Updated 9/3/2026
v0.1.0
dsh-vision-patch

Vision Patch

Per-model and per-route image-input (vision) checkboxes on the Models page's custom-provider cards, writing through the llm-pi-ai settings namespace.

vision-mediamodels-usage00
Updated 9/3/2026
v0.1.0
dsh-plugin-vision

Plugin Vision

Give the DSH agent a pair of eyes: call online VLMs (multi-provider, OpenAI-compatible) to analyze local images, URLs, and attachments uploaded in conversations; the selected models’ specialized capabilities are injected into the system prompt in real time

vision-media00
Updated 9/3/2026
v0.1.0
dsh-imgdraw

Imgdraw

Text-to-image for DeepSeek Harness: a `draw_image` model tool, an input-bar 生图 button with a prompt popup (async generation, 4-grid results, download / keep / delete), an /imgdraw image route, and persisted history. Backends: DashScope wan2.7-image (free default) and SiliconFlow Qwen-Image.

vision-media00
Updated 9/3/2026
v0.1.2
dsh-vision-autoswitch

Vision Autoswitch

DeepSeek Harness plugin: auto-route image-bearing requests to deepseek-v4-flash-vision-exp, then fall back to the original model.

vision-mediaagents-orchestration00
Updated 9/2/2026
v1.0.0
dsh-vision-assist

Vision Assist

DeepSeek Harness vision assistant plugin: equips models without vision capabilities with a switchable multimodal recognition model. Images in the input field are automatically saved to disk and rewritten as text prompts; the main model can view images by calling the vision_recognize tool. The recogn

vision-mediamodels-usage00
Updated 9/2/2026
v0.3.9
dsh-vision-bridge

Vision Bridgealt · alaxrpg

DSH Vision Bridge Plugin: Reuse DSH Provider or connect directly via OpenAI compatibility, with visual configuration and image pasting support

vision-mediaintegrations-communication00
Updated 9/2/2026
v0.1.0
dsh-plugin-show-image

Plugin Show Image

Render local image files inline in the DSH conversation via a global show_image tool.

vision-media00
Updated 9/1/2026
v0.1.2
media-preview

Media Preview

DSH 插件:在聊天记录中自动将本地音视频/图片路径渲染为可播放的预览组件。When an assistant message or tool result contains a local media path, the path is replaced inline with a playable <audio>/<video>/<img> element backed by same-origin /api/media-preview/* route.

vision-mediaui-customization00
Updated 9/1/2026
v0.2.5
dsh-plugin-voice

Plugin Voice

DeepSeek Harness plugin: voice + notification outputs—agents proactively contact users through cloud TTS (Volcano seed-tts / Xiaomi MiMo V2.5, with automatic fallback to SAPI on failure), desktop notifications, and alert sounds. Combines the native DSH integration of dsh-plugin-notify with the cloud

vision-mediaintegrations-communication00
Updated 9/1/2026
v0.1.3
@maiziman/dsh-model-capabilities

Model Capabilities

Automatic reasoning and image capability detection for custom DeepSeek Harness models

vision-media00
Updated 9/1/2026
v0.1.0
dsh-llm-multimodal

Llm Multimodal

DSH plugin: Provides image/video generation tools in DSH, based on OpenAI-compatible APIs. Models are automatically discovered from existing llm-pi-ai settings.

vision-media00
Updated 8/31/2026
v0.1.1
@xiaokaizhou/dsh-media-preview

Media Previewalt · @xiaokaizhou

DSH 插件:在聊天记录中自动将本地音视频/图片路径渲染为可播放的预览组件。When an assistant message or tool result contains a local media path, the path is replaced inline with a playable <audio>/<video>/<img> element backed by same-origin /api/media-preview/* route.

vision-mediaui-customization00
Updated 8/31/2026
v1.0.3
dsh-image-amnesia

Image Amnesia

Keep native vision, drop historical images before they hit relay providers. Global DeepSeek Harness bundle for every agent.

memory-contextvision-media00
Updated 8/31/2026
v1.0.0
dsh-voice-control

Voice Control

Voice control for DSH web: speech-to-text into the composer (with auto-send) and spoken playback of assistant replies via the Web Speech API

vision-media00
Updated 8/31/2026
v0.2.0
dsh-photos

Photos

Photo upload for the DeepSeek Harness composer — native picker feeding the stock paste pipeline

vision-media00
Updated 8/31/2026
v0.1.2
@jetecho/dsh-csv-and-image-preview

Csv And Image Preview

Preview images / SVG and CSV tables in the DeepSeek Harness chat, rendered as real browser <img> / <table> elements. Preview-first workflow: show the user the asset, wait for approval, then apply the real change.

vision-mediaui-customization00
Updated 8/31/2026
v0.2.1
@goodandready/dsh-im-hub-media

Im Hub Media

Multi-platform IM gateway for DeepSeek Harness (fork of dsh-im-hub with media): Telegram voice/photo/document/video + reply handling, STT (Deepgram primary, HF Whisper fallback), outbound MEDIA: markers. Feishu/WeCom kept as-is.

integrations-communicationvision-media00
Updated 8/30/2026
v0.1.2
dsh-svw-waveform

Svw Waveform

Native SVW waveform rendering plugin for DeepSeek Harness

vision-media00
Updated 8/30/2026
v0.1.0
dsh-multimodal-runtime

Multimodal Runtime

Mogu Multimodal Runtime - DeepSeek Harness 统一多模态能力运行时 (Comfy Local Provider V1)

vision-media00
Updated 8/29/2026
v0.2.3
dsh-speech-input

Speech Input

A microphone button for DeepSeek Harness that writes browser speech recognition into the composer draft.

vision-media00
Updated 8/29/2026