DeepSeek Harness Plugin Hub

CATEGORY

Vision & Media DSH plugins

Verified manifests and exact versions in this category.

All plugins

377 plugins
v1.0.0
xby-asr-1

Xby Asr 1

General-purpose speech recognition supporting multiple languages and lesser-spoken languages.

vision-media00
Updated 8/28/2026
v1.0.0
xby-asr-5

Xby Asr 5

Five commonly used speech recognition languages: Mandarin Chinese, English, Japanese, Korean, and Cantonese, with automatic language detection.

vision-media00
Updated 8/28/2026
v1.0.0
xby-asr-f

Xby Asr F

Speech recognition supporting Mandarin Chinese and over 20 dialects and accents.

vision-media00
Updated 8/28/2026
v1.0.0
xby-asr-zh

Xby Asr Zh

Chinese speech recognition

vision-media00
Updated 8/28/2026
v1.0.0
xby-bird

Xby Bird

Detect and identify birds in images.

vision-media00
Updated 8/28/2026
v1.0.0
xby-captcha

Xby Captcha

CAPTCHA recognition toolkit supporting the recognition of text, slider, rotation, selection, and other verification methods. Note: Always comply with the terms of use and applicable laws and regulations of the target website or system, and use it only where permitted.

vision-mediasecurity-access00
Updated 8/28/2026
v1.0.0
xby-cellphone-detection

Xby Cellphone Detection

Input an image to detect mobile phones and output the bounding boxes, confidence scores, and labels for all targets in the image.

vision-media00
Updated 8/28/2026
v1.0.0
xby-classify

Xby Classify

Classify images into 1000 ImageNet classes and return the Top-5 classes and confidence scores.

vision-media00
Updated 8/28/2026
v1.0.0
xby-daily-object-detection

Xby Daily Object Detection

Input an image, detect people, pets, cars, fire, and cardboard boxes, and output the bounding boxes, confidence scores, and labels for all detected objects in the image.

vision-media00
Updated 8/28/2026
v1.0.0
xby-detect-vehicle

Xby Detect Vehicle

Input an image to detect vehicle types (car/truck/bus/motorbike/tricycle/carplate) and output the bounding boxes, confidence scores, and labels for all detected objects.

vision-media00
Updated 8/28/2026
v1.0.0
xby-detect

Xby Detect

Includes everyday object detection, insect recognition, plant recognition, animal recognition, electric bicycle detection, mobile phone detection, gesture detection, flame detection, cigarette detection, human head and body detection, wildlife detection, bird recognition, pet emotion recognition, di

vision-media00
Updated 8/28/2026
v1.0.0
xby-dish

Xby Dish

Dish recognition, output possible dish names and probabilities.

vision-media00
Updated 8/28/2026
v1.0.0
xby-ebike-detection

Xby Ebike Detection

Input an image, detect the electric bicycles in it, and output the bounding boxes, confidence scores, and labels for all targets in the image.

vision-media00
Updated 8/28/2026
v1.0.0
xby-extract-image

Xby Extract Image

The MCP server provides functionality to extract images from local files and URLs and convert them to base64 format, suitable for LLM analysis.

vision-mediaintegrations-communication00
Updated 8/28/2026
v1.0.0
xby-fire-detection

Xby Fire Detection

Detect flames in various general scenarios; best suited for security camera and traffic camera views.

vision-media00
Updated 8/28/2026
v1.0.0
xby-general-recognition

Xby General Recognition

Recognizes labels in images containing a primary object and outputs the object's category label. Currently covers more than 50,000 object categories.

vision-media00
Updated 8/28/2026
v1.0.0
xby-gesture-detection

Xby Gesture Detection

Input an image, detect the gestures in it, and output the bounding boxes, confidence scores, and labels for all detected targets.

vision-media00
Updated 8/28/2026
v1.0.0
xby-head-person-detection

Xby Head Person Detection

Input an image, detect human heads and bodies, and output the bounding boxes, confidence scores, and labels for all detected targets.

vision-media00
Updated 8/28/2026
v1.0.0
xby-helmet-head

Xby Helmet Head

Input an image to detect human bodies, heads, and safety helmets, and output the bounding boxes, confidence scores, and labels for all detected objects.

vision-media00
Updated 8/28/2026
v1.0.0
xby-image-detect

Xby Image Detect

Detects 80 classes of COCO objects (people, vehicles, animals, everyday items, etc.) in images, outputting bounding boxes, confidence scores, and class labels.

vision-media00
Updated 8/28/2026
v1.0.0
xby-insect-recognition

Xby Insect Recognition

Identify the name of an insect or other arthropod (or its order, family, genus, or species).

vision-media00
Updated 8/28/2026
v1.0.0
xby-logo-analyze

Xby Logo Analyze

An intelligent logo extraction and processing MCP server that supports automatically identifying and extracting logo icons from website URLs, and provides image processing and vector conversion features.

vision-mediasearch-research00
Updated 8/28/2026
v0.1.3
dsh-omi-voice

Omi Voice

Read DeepSeek Harness conversations aloud: Doubao audio quality · click to read/pause/resume · Doubao Key stays only in the Omi engine (BYOK)

vision-media00
Updated 8/28/2026
v0.1.1
@lijian-ui/dsh-vision-toggle

Vision Toggle

Per-model vision (image input) toggle for DeepSeek Harness (dsh): list every configured model and flip a switch to enable/disable image support without hand-editing settings.yaml. 为 DeepSeek Harness 提供按模型的「支持图片」开关:无需手改 settings.yaml。

vision-mediamodels-usage00
Updated 8/28/2026
v0.1.0
@dopilot/dsh-plugin-image-gen

Plugin Image Gen

OpenAI-compatible image generation tool for DeepSeek Harness

vision-media00
Updated 8/27/2026
v0.1.1
dsh-pro-vision-std

Pro Vision Std

dsh-std Community v0.15 ModelProvider: DeepSeek-V4-Pro with images captioned by V4-Flash-Vision-Exp. Requires @dsh-std/adapter-dsh on DeepSeek Harness.

models-usagevision-media00
Updated 8/27/2026
v0.3.5
dsh-periscope

Periscope

DSH Desktop plugin: keep text-only models (deepseek-v4-flash / deepseek-v4-pro) as the session default and automatically route image-bearing requests to a configured vision model; also ships the generic file-attachment feature (drag/paste any file as a durable attachment the agent reads by path) as

vision-media00
Updated 8/27/2026
v0.1.0
dsh-handwritten-ocr

Handwritten Ocr

Local OCR for DSH: handwritten Chinese and math formulas to Markdown with LaTeX. GPU (DirectML) / CPU / NPU backends, settings-driven, one-click install.

vision-mediaproductivity-workflow00
Updated 8/27/2026
v1.2.0
dsh-image-picker

Image Picker

DeepSeek Harness Web GUI input box 📎 image-picker button: Add reference images through the system file picker, bypassing drag-and-drop environment issues and reusing the official attachment pipeline (thumbnail generation, limit validation, and upload with the message).

vision-mediaui-customization00
Updated 8/26/2026
v1.0.0
dsh-ocr-plugin

Ocr Plugin

DeepSeek Harness OCR plugin — based on 小笨羊 OCR capabilities, providing image text recognition for text-only models

vision-media00
Updated 8/26/2026