CATEGORY
Vision & Media DSH plugins
Verified manifests and exact versions in this category.
All plugins
377 plugins@fzhiyu/dsh-image-viewImage View
Show the image a read_image call returned, inline in the tool row it belongs to
@evan-williams/vision-toolVision Tool
Conversation-style UI/UX visual review plugin for DeepSeek Harness: vision_review / vision_ask tools backed by any OpenAI-compatible multimodal endpoint (e.g. agnes-2.5-flash)
@pzqian123/dsh-tool-visionTool Vision
DSH plugin: read images through a user-configured vision-capable model when the main model route cannot accept image input
dsh-eyesEyes
A vision bridge profile bundle for DeepSeek Harness (dsh): gives non-vision models image-reading capability by delegating transcription to a vision-capable model, with GUI paste admission and a web settings page for vision model selection.
dsh-qr-toolkitQr Toolkit
QR code & barcode toolkit for DeepSeek Harness: decode images (PNG/JPEG, local or URL) and generate QR codes as PNG. 二维码/条码工具箱:解码与生成。
dsh-computer-useComputer Use
Computer Use plugin: virtual mouse human-like operation + native vision model integration (deepseek-v4-flash-vision-exp directly reads screenshots), 12 model-friendly tools including screen_observe
dsh-ros2Ros2
ROS2 debugging tools and diagnostics skills for DeepSeek Harness (DSH): package/workspace/dependency inspection, node/topic/service/action/param/interface enumeration, topic sampling, TF tree, graph topology JSON, approval-gated builds and message scaffolding, GUI screenshot and multimodal vision ob
dsh-visual-pluginVisual Plugin
Visual media plugin for DeepSeek Harness: copy native image descriptions and securely normalize, inspect, and play scene-aware videos.
dsh-vision-opencodeVision Opencode
DeepSeek Harness plugin: configurable vision model with vision_read_image tool, composer-bar vision-model selector, and automatic image-to-text conversion for text-only main models.
dsh-chat-imagineChat Imagine
Generate and display images inline in the DSH chat via API channels or local CLIs (mmx / codex / agy), and read images into structured JSON evidence (OCR / layout / semantics) on any model — with backend probing and a bundled recovery skill.
dsh-galgame-generatorGalgame Generator
Galgame generator for DeepSeek Harness (DSH): give it a script document plus character/background/BGM/animation assets in the workspace (img_human/, img_bg/, audio/, img_cg/) and it builds a playable visual-novel web page with configurable save slots, speaker-following multi-pose sprites, [表情]/[位置]/
dsh-windows-ocrWindows Ocr
dsh plugin: recognize attached images with the built-in Windows OCR engine (Windows.Media.Ocr) and send only the recognized text to the model. Text models never receive image bytes; vision passthrough is opt-in.
dsh-highres-visionHighres Vision
DeepSeek Harness vision enhancer: raises image limits to official values and provides high-resolution tiled image reading via a dedicated tool.
dsh-read-image-viewRead Image View
Display images read by the read_image tool inside the DeepSeek Harness Web GUI: a dedicated Read image row with a default-expanded message-style image card (PS-style transparency checkerboard), merged read-N-images rows with side-by-side frames for multi-image requests, and an in-page full-resolutio
@nn12138/dsh-voiceVoice
Voice input plugin for DeepSeek Harness: microphone → local/browser speech recognition → text submitted as a normal chat message (input-only, preset-agnostic)
dsh-tesseract-ocrTesseract Ocr
dsh plugin: recognize attached images locally with Tesseract OCR and send only the recognized text to the model. Text models never receive image bytes; vision passthrough is opt-in.
dsh-tool-see-imageTool See Image
see_image tool for DSH: route image files to a configurable vision model (OpenAI-compatible API) and relay its description back to a text-only model
dsh-tool-describe-imageTool Describe Image
DSH plugin: image understanding via any OpenAI-compatible vision API, paste-to-describe, and an animated whale-buddy desktop pet with status bubbles and a floating settings panel
@haoku123/dsh-voiceVoicealt · @haoku123
Full-duplex voice mode for DeepSeek Harness: streamed ASR -> LLM -> TTS with barge-in
dsh-image-subagentImage Subagent
Allow text-only primary models (DeepSeek V4, etc.) to receive image attachments: remove the modality declaration from the text-only route to pass the 0.1.1 admission gate, project images into placeholder text carrying the complete attachmentId, and have the primary model delegate reading to a visual
dsh-paddleocr-skillsPaddleocr Skills
PaddleOCR text-recognition and document-parsing skills with native DeepSeek Harness tools and GUI configuration.
@limccn/deepseek-vl-supportDeepseek Vl Support
Give DeepSeek (text-only) models vision in Claude Code, Codex, and Agent Plugins clients: describe images via any OpenAI-compatible vision endpoint.
@paicat1/dsh-screenshotScreenshot
Standalone screen capture for DeepSeek Harness (dsh): browser hotkeys plus an agent-facing capture+read tool. Forked out of @liustack/modlens#48 (upstream declined the feature).
deepseek-vl-supportDeepseek Vl Supportalt · limccn
Give DeepSeek (text-only) models vision in Claude Code, Codex, and Agent Plugins clients: describe images via any OpenAI-compatible vision endpoint.
dsh-stt-inputStt Input
Speech-to-text voice input for the DeepSeek Harness (DSH) web GUI: click the mic button in the composer to turn speech into text in the input box. Supports the browser's built-in Web Speech API (zero-config, Chrome/Edge) and OpenAI-compatible Whisper endpoints (OpenAI / Groq) with the STT model sele
dsh-approval-voiceApproval Voice
DSH Web GUI approval voice notifications: when a popup requiring approval or a response appears, play a notification sound and announce it by voice to avoid missing it.
dsh-pseudo-visionPseudo Vision
Adds opt-in image-capable sibling routes for text-only providers and converts image blocks into budgeted, locally enhanced OCR, color-statistics, pixel-scan, and metadata evidence before delegation.
@xiaoyuink/dsh-image-visionImage Vision
Image recognition plugin: Automatically detects whether the current model has vision capabilities. If so, it uses the current model with the plugin’s preset prompts for analysis; otherwise, it calls the vision model configured in the plugin. Text models can paste, upload, or drag images directly int
teach-math-with-manimTeach Math With Manim
Open-source companion repository for the book 《Manim,让数学看得见》: source code for 60+ educational animation examples + the manim-teaching AI Skill (including the DeepSeek Harness plugin)
dsh-image-inlineImage Inline
Let the model render an image into the DSH web chat flow (QQ/WeChat style) via a show_image tool: the message content keeps only the path text, the image never enters model context, and the browser shows the picture inline, scrollable with history.