DeepSeek Harness Plugin Hub

Publish and manage complete Harness Profiles. Discover Plugins for your next setup.

Explore

PluginsPresetsDocsNews

Community

Publish a pluginContactReport an issue

Resources

Plugin Hub on GitHubDeepSeek HarnessSystem statusPrivacy notice
© 2026 DeepSeek Harness Plugin HubPowered byPaxTech

Independent and unofficial. Not affiliated with, authorized by, or endorsed by DeepSeek.

Vision & Media DSH plugins — Page 12

CATEGORY

Vision & Media DSH plugins

Verified manifests and exact versions in this category.

All categories

All pluginsAgents & Orchestration713Memory & Context852Coding & Developer Tools1.9kUI & Customization2.5kIntegrations & Communication1.3kVision & Media464Search & Research617Security & Access857Models & Usage1.1kProductivity & Workflow1.5k

All plugins

464 plugins
Sv0.4.0
dsh-speech

Speech

Speech capability plugin for the DeepSeek Harness (dsh) web host: a token-gated /s/api route family serving audio transcription (ASR) and synthesis (TTS) over configurable providers

↓ 0
Updated 9/8/2026
Mv1.0.0
dsh-model-modality

Model Modality

Declare whether a configured third-party model accepts image (multimodal) input; writes the modality into the owning provider settings and verifies it through runtime model resolution.

↓ 0★ 0
Updated 9/7/2026
Vv0.2.1
dsh-videogen

Videogen

AI video plugin for the dsh web GUI: multi-vendor video generation channels (OpenAI-compatible /v1/videos, Kling, MiniMax Hailuo, Volcengine ARK, custom generic async), FFmpeg video processing (probe / frame extraction / GIF / image compress), a lightweight storyboard studio (templates + shots + FFm

↓ 0★ 0
Updated 9/7/2026
Iv0.2.1
dsh-image2-draw

Image2 Draw

Image2 (gpt-image-2) generation plugin for DeepSeek Harness with simple relay settings, in-conversation previews, a chat-dock studio (text-to-image / image-to-image / multi-view character consistency) and a reusable prompt/flow preset library. Community fork of JuneLearn/dsh-image2-draw with gateway

↓ 0★ 0
Updated 9/6/2026
Ov0.5.0
dsh-ollama-vision-bridge

Ollama Vision Bridge

DSH vision bridge (DSH >= 0.1.2-rc.1): when the selected chat model is text-only, attached images are described by a local Ollama VL model (qwen3-vl:8b) with keep_alive VRAM cooling. Install-time patch of dsh-api-session-controller prompt admission + runtime status companion.

↓ 0★ 0
Updated 9/6/2026
Vv0.1.1
@dsh-external/dsh-vision-bridge

Vision Bridge

DeepSeek Harness vision bridge: session screenshots route by model capability. Image-capable models receive attachments inline; text-only models see the standard placeholder, and the bridge reads the image through a vision-language model — on demand via the vision_bridge_read tool or automatically v

↓ 0★ 0
Updated 9/6/2026
Vv0.2.2
@lovstudio/dsh-video-studio

Video Studio

A local-first video editing workbench for DeepSeek Harness with Remotion, GSAP and pluggable ASR

↓ 0★ 0
Updated 9/6/2026
Iv0.2.3
@goodandready/dsh-im-hub-media

Im Hub Media

Multi-platform IM gateway for DeepSeek Harness (fork of dsh-im-hub with media): Telegram voice/photo/document/video + reply handling, STT (Deepgram primary, HF Whisper fallback), outbound MEDIA: markers. Feishu/WeCom kept as-is.

↓ 0★ 0
Updated 9/5/2026
v0.1.1
dsh-voice-input-qwen-asr

Voice Input Qwen Asr

Voice input plugin (dual-face): mic button beside the composer send action, live recording bubble streaming PCM to a local Qwen3-ASR python service managed by the host, plus an ASR environment settings page (clone runtime/model repos, create venv, run server). Bilingual UI (zh/en)

↓ 0★ 0
Updated 9/5/2026
Fv0.1.0
dsh-file-upload-local

File Upload Local

Local file-upload plugin for DeepSeek Harness: a paperclip button and drag-and-drop that store files per-session under .dsh-uploads/<sessionId>/, plus a read_document tool that pages text and OCRs images (tesseract.js) so a text-only model can still read screenshots.

↓ 0★ 0
Updated 9/5/2026
Iv1.2.3
dsh-image-picker

Image Picker

DeepSeek Harness Web GUI input box 📎 image-picker button: Add reference images through the system file picker, bypassing drag-and-drop environment issues and reusing the official attachment pipeline (thumbnail generation, limit validation, and upload with the message).

↓ 0★ 0
Updated 9/5/2026
Bv0.1.0
@dsh-voice/bundle

Bundle

dsh-voice — voice notes in, spoken answers out: dictate audio that becomes user messages (transcribe), have the agent read replies aloud (speak), and leave walk-away narration on long headless runs. Local-first, plain audio files under ~/.dsh/voice/

↓ 0★ 0
Updated 9/5/2026
Vv0.1.2
dsh-video-director

Video Director

A project-scoped multimodal video director plugin for DeepSeek Harness.

↓ 0★ 0
Updated 9/4/2026
Kv0.1.0
@lamplitisles/kepos-speech

Kepos Speech

Kepos Speech plugin for Alibaba and ByteDance Chinese speech synthesis plus short-audio recognition in DeepSeek Harness Web

↓ 0★ 0
Updated 9/4/2026
Gv0.3.1
dsh-grok-adaptation

Grok Adaptation

Normalize undersized Grok Responses images with Sharp while leaving supported large images unchanged.

↓ 0★ 0
Updated 9/3/2026
v1.0.0
dsh-vision-pro-bridge

Vision Pro Bridge

Give text-only DeepSeek-V4-Pro real vision with zero new dependencies and DeepSeek-only routing: images are described by deepseek-v4-flash-vision-exp (your existing DEEPSEEK_API_KEY), then the text is handed to V4-Pro.

↓ 0★ 0
Updated 9/3/2026
Vv1.0.2
dsh-vision-toggle

Vision Toggle

DSH model vision toggle: adds a “Model Vision” row to the settings page for switching the input vision modality of manually declared models under llm-pi-ai custom routes via the official settings.mutate channel, with changes taking effect immediately.

↓ 0★ 0
Updated 9/3/2026
Vv1.3.0
@dshp-inx/vision-bridge

Vision Bridgealt · @dshp-inx

DeepSeek Harness (DSH) vision bridge: let text-only models delegate image understanding to multimodal vision models via vision_describe tool, with persistent settings UI and primary/fallback auto-retry. 纯文本模型通过 vision_describe 工具把图片转交给多模态视觉模型识别的桥接插件。

↓ 0★ 0
Updated 9/3/2026
Pv1.1.0
dsh-pdf-to-word

Pdf To Word

DeepSeek Harness plugin: PDF→Word (.docx) conversion with layout fidelity (fonts/tables/images/borders), OCR scan mode, and optional multimodal LLM verification. Registers the pdf_to_word model tool.

↓ 0★ 0
Updated 9/3/2026
Vv0.1.0
dsh-vision-patch

Vision Patch

Per-model and per-route image-input (vision) checkboxes on the Models page's custom-provider cards, writing through the llm-pi-ai settings namespace.

↓ 0★ 0
Updated 9/3/2026
Pv0.1.0
dsh-plugin-vision

Plugin Vision

Give the DSH agent a pair of eyes: call online VLMs (multi-provider, OpenAI-compatible) to analyze local images, URLs, and attachments uploaded in conversations; the selected models’ specialized capabilities are injected into the system prompt in real time

↓ 0★ 0
Updated 9/3/2026
Vv1.0.0
dsh-vision-assist

Vision Assist

DeepSeek Harness vision assistant plugin: equips models without vision capabilities with a switchable multimodal recognition model. Images in the input field are automatically saved to disk and rewritten as text prompts; the main model can view images by calling the vision_recognize tool. The recogn

↓ 0★ 0
Updated 9/2/2026
Vv0.3.9
dsh-vision-bridge

Vision Bridgealt · alaxrpg

DSH Vision Bridge Plugin: Reuse DSH Provider or connect directly via OpenAI compatibility, with visual configuration and image pasting support

↓ 0★ 0
Updated 9/2/2026
Pv0.1.0
dsh-plugin-show-image

Plugin Show Image

Render local image files inline in the DSH conversation via a global show_image tool.

↓ 0★ 0
Updated 9/1/2026
Mv0.1.2
media-preview

Media Preview

DSH 插件:在聊天记录中自动将本地音视频/图片路径渲染为可播放的预览组件。When an assistant message or tool result contains a local media path, the path is replaced inline with a playable <audio>/<video>/<img> element backed by same-origin /api/media-preview/* route.

↓ 0
Updated 9/1/2026
v0.2.5
dsh-plugin-voice

Plugin Voice

DeepSeek Harness plugin: voice + notification outputs—agents proactively contact users through cloud TTS (Volcano seed-tts / Xiaomi MiMo V2.5, with automatic fallback to SAPI on failure), desktop notifications, and alert sounds. Combines the native DSH integration of dsh-plugin-notify with the cloud

↓ 0★ 0
Updated 9/1/2026
Mv0.1.3
@maiziman/dsh-model-capabilities

Model Capabilities

Automatic reasoning and image capability detection for custom DeepSeek Harness models

↓ 0★ 0
Updated 9/1/2026
Lv0.1.0
dsh-llm-multimodal

Llm Multimodal

DSH plugin: Provides image/video generation tools in DSH, based on OpenAI-compatible APIs. Models are automatically discovered from existing llm-pi-ai settings.

↓ 0★ 0
Updated 8/31/2026
Mv0.1.1
@xiaokaizhou/dsh-media-preview

Media Previewalt · @xiaokaizhou

DSH 插件:在聊天记录中自动将本地音视频/图片路径渲染为可播放的预览组件。When an assistant message or tool result contains a local media path, the path is replaced inline with a playable <audio>/<video>/<img> element backed by same-origin /api/media-preview/* route.

↓ 0
Updated 8/31/2026
Iv1.0.3
dsh-image-amnesia

Image Amnesia

Keep native vision, drop historical images before they hit relay providers. Global DeepSeek Harness bundle for every agent.

↓ 0★ 0
Updated 8/31/2026
← Previous1…1011121314…16Next →
DeepSeek Harness Plugin Hub
ProfilesPluginsCategoriesNewsDocsSign inManage Profiles
ProfilesPluginsCategoriesNewsDocsSign in