CATEGORY
Vision & Media DSH plugins
Verified manifests and exact versions in this category.
All plugins
377 pluginsdsh-plugin-vision-toolkitPlugin Vision Toolkit
Vision toolkit for DeepSeek Harness -- glance, ground, detect, crop CLI tools for text-only agents to understand images
@chang416/deepseeDeepsee
DeepSee: DeepSeek Harness vision, multi-model routing, and Gemini visual self-checks before delivery
dsh-quicksightQuicksight
Two-tier image reading for text-only models in DeepSeek Harness: fast local OCR (RapidOCR, offline) first, then a vision model via modlens as fallback.
dsh-media-guardMedia Guard
Request image optimization, intelligent retention, and observability for DeepSeek Harness (DSH)
dsh-web-voice-inputWeb Voice Input
Voice input for the DeepSeek Harness (DSH) Web GUI: a microphone button in the chat composer that records audio, transcribes it through an OpenAI-compatible Whisper API (Groq / OpenAI / SiliconFlow), and inserts the recognized text into the input box.
dsh-mingmuMingmu
明眸 VisionBridge - In-house visual bridge: when a visionless model receives an image, it automatically calls a vision model for recognition and feeds the recognized text back to the main model—seamless, configurable, and upgrade-safe
dsh-ocr-bridgeOcr Bridge
Paste images into DeepSeek Harness chat and have them read by a free local backend (macOS Vision / Tesseract) before the text-only DeepSeek model answers
@muxiva/dsh-voiceVoice
Local-first, full-duplex voice for DeepSeek Harness, orchestrated by Muxiva
@zzdream67/dsh-vision-bridgeVision Bridge
Let text-only models see images in DeepSeek Harness: intercepts the llm/stream waterfall and transparently substitutes each image with a vision model's description.
@gelomen/glm-vision-pluginGlm Vision Plugin
DSH Web plugin: Call the 智谱 GLM vision model to analyze images (analyze_image tool + plugin settings card).
dsh-tool-accurate-visionTool Accurate Vision
Model-facing accurate_vision tool: precise image spatial reasoning via a vision model
@nextnowlabs/dsh-ark-toolkitArk Toolkit
DSH Ark Toolkit: A plugin that adds visual capabilities to DeepSeek Harness text-only models. Powered by 火山方舟 vision models, it supports image-based Q&A, OCR text recognition, and multi-image comparison, with built-in 豆包 Seedream text-to-image generation and 字节 TTS speech synthesis.
dsh-niulai-soundNiulai Sound
牛来确认音 — DSH approval sound effects from the movie 《牛来》: plays 「妈妈」 when user confirmation is requested, interrupts it and plays 「牛来!」 the moment the user approves. 审批音效:请求确认时播放「妈妈」,批准瞬间掐断并播放「牛来!」。
auto-visionAuto Vision
DSH (DeepSeek Harness) image integration plugin: pasted images are added only to the conversation, not the model request (llm/stream cleanup), and the see_image tool is registered so text models can view images on demand through any OpenAI-compatible vision API
dsh-vision-subagentVision Subagent
Vision for text-only DeepSeek Harness agents: delegate image reading to a one-shot subagent on a configurable vision route (MiniMax / Kimi / any OpenAI-compatible provider), keeping image bytes and the vision model's context out of the main session.
vision-translation-dshVision Translation Dsh
Native dsh (DeepSeek Harness) Cordis plugin adapter for vision-translation: grounds images into <vision-context> via the Python CLI (PROTOCOL v1). Spawns cli.py, never re-implements core logic.
local-ocr-cliLocal Ocr Cli
Fully-local OCR CLI for text-only LLMs: PaddleOCR-VL first-tier engine with tesseract fallback, dsh plugin included
dsh-seedance2Seedance2
Generate images and Seedance videos in DeepSeek Harness through the Seedance 2 AI API
@freespace8/dsh-free-visionFree Vision
Local image understanding for the DeepSeek Harness: OCR, table-layout detection, and semantic description of textless images via macOS Vision, fully on-device (images never leave your Mac). 本地化识图,图片不出本机。
dsh-nanobananaproNanobananapro
Generate images and videos in DeepSeek Harness through the NanoBananaPro API
@aalongaa/dsh-tool-visionTool Vision
Model-facing vision tool for DeepSeek Harness: analyze images via any OpenAI-compatible vision API.
@opensquad/dsh-voice-inputVoice Input
SenseVoice voice input plugin for DeepSeek Harness: adds a microphone button next to the conversation input box, records audio, and uses the local SenseVoice service to convert it to text and fill the input box. On first use, it automatically downloads the model and displays progress; the backend is
texteyeTexteye
DeepSeek Harness plugin: convert images to ASCII art + per-cell color map so text-only models can see them.
dsh-yali-image-generatorYali Image Generator
DeepSeek-Harness 图像生成插件。申请 Yali AI API Key:https://api.yaliai.com/
soyo-dsh-pluginSoyo Dsh Plugin
DSH-native video understanding with configurable multimodal providers
@acc1143/dsh-vision-bridgeVision Bridgealt · @acc1143
DSH dual-model routing plugin: keeps the main conversation in DeepSeek, automatically routes image tasks to an explicitly configured vision model, and returns the analysis results to DeepSeek to continue.
dsh-siliconflow-visionSiliconflow Vision
DSH plugin: Recognizes/analyzes images through the SiliconFlow vision model, supporting local file paths, http(s) image URLs, and base64 data URLs. Includes a persistent paste recognition panel (web).
@try-works/dsh-vision-workerVision Worker
DeepSeek Harness plugin: a vision worker over Cloudflare Workers AI (@cf/moonshotai/kimi-k2.6) that routes image requests from text-only callers, returns a versioned righthand.vision.v1 envelope, and supports follow-up questions.
dsh-bundle-visionBundle Vision
Vision bundle + plugin for DeepSeek Harness: the describe_image tool reads local images and asks any configured multimodal route, with zero core changes
dsh-tool-image-genTool Image Gen
DSH tool plugin: generate images through ToAPIs async GPT-Image-2 API (submit task, poll, download).