DSH Plugins Marketplace

DSH Plugins

DSH Plugins Marketplace

7,148 plugins indexed · 5,699 installable · 17,937 versions tracked

AllDevelopment & Infrastructure (485)Tools & Capabilities (1496)Memory & Context (199)Workflow & Automation (234)Models & Providers (201)Terminal & Clients (148)Security & Audit (138)Vision & Multimodal (155)UI & Experience (592)Notifications & Integrations (142)Themes & Skins (128)Sessions & Messages (224)Just for Fun (109)

Sort:

Most starsRecently updatedNewestName A–ZInstallable only

141 plugins

dsh-tesseract-ocr

Local OCR for attached images via Tesseract: only the recognized text is sent to the model, never the image bytes; vision passthrough is opt-in.

Vision & MultimodalManifest valid

0

778/wk

dsh plugin --profile web add dsh-tesseract-ocr

Give DSH text-only models vision: an image/OCR/document recognition skill (race pool → custom channels → local) plus an idempotent host patch so image messages reach the model.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add @dsh-user/dsh-vision-solution

Labnana image generation for DeepSeek Harness: text-to-image / image-to-image / precise editing with credits estimation, subscription balance and web settings UI.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-labnana

Voice conversation mode for the DSH Web GUI: tap the mic to start a full-screen call, talk hands-free, and hear each reply read aloud as it streams. Bilingual (Chinese/English) UI, light and dark them

Vision & MultimodalManifest valid

0

414/wk

dsh plugin --profile web add dsh-voice-talk

`describe_image` tool: a vision bridge that sends images to mimo-v2.5 through the opencode Zen API (credential `OPENCODE_GO_API_KEY`, free route first with paid fallback) and returns text descriptions

Vision & MultimodalManifest valid

0

dsh plugin --profile web add mimo-vision

Speaks agent replies in the DeepSeek Harness web UI through a provider fallback chain (OpenAI, ElevenLabs, Google, Azure, Groq, Deepgram, OpenRouter, Edge, Piper, eSpeak), so a failing or rate-limited

Vision & MultimodalManifest valid

0

dsh plugin --profile web add @goodandready/dsh-tts

Vision for any DSH route: paste images in the Web composer with intent-aware auto-analysis, delegate workspace image reads to a Kimi/MiniMax vision subagent, and materialize pasted originals for editi

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-vision-subagent

Sound reminders with one-click jump for concurrent sessions: the current session gets crisp dang/dang-dang tones, other sessions a soft ding/ding-ding plus a top-right card that jumps straight to the

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-dingo

DeepSeek Harness vision plugin: 8 analysis modes (describe, OCR, chart data, UI review, object detection, compare, code-gen, debug), any OpenAI- or Anthropic-compatible vision API, with a built-in fre

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-vision-analysis

Zero-dependency screen capture for DSH: Lightweight — zero deps, zero binaries; Stage & shoot — one-click full screen, window layout, hover-snap capture of occluded windows; Agent self-service — path-

Vision & MultimodalManifest valid

0

177/wk

dsh plugin --profile web add @paicat1/dsh-screenshot

Configurable image recognition for text-only DSH models: image messages are first transcribed by an OpenAI-compatible vision model (Base URL, model ID and API key set in a Settings section) and then p

Vision & MultimodalManifest valid

0

dsh plugin --profile web add @lp181818/dsh-vision-plugin

by 2021Heei

给 DeepSeek Harness 的语音朗读插件:AI 回复流式 TTS 朗读,思考等待期还有趣味语音短语反馈。Edge TTS 开箱即用,支持任意 OpenAI 兼容云端引擎。

Vision & MultimodalManifest valid

0

MIT

TypeScript

Sep 10, 2026

dsh plugin --profile web add dsh-tts-flash

Notification outbox: agent proactively notifies via toast / Chinese TTS voice / sound effects (explosion, victory, alarm), 60s confirmation window voice-calls you back, volume boost, settings panel.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-plugin-notify

Let your AI agent see and operate a real Android phone: phone_look (vision + UI-tree fusion), tap/swipe/type, screenshot — over adb, for any MCP client.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add phone-eye

Adds image input and recognition through configured DSH providers or an OpenAI-compatible endpoint.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-vision-bridge

Transparent image guard for text-only routes: paste images without the 400 session deadlock, plus a vision_analyze tool for OCR/PDF/docx/pptx/video.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-vision-guard

Turn every idea into an image or video with Kling AI. In DeepSeek Harness, use natural language for text-to-image, image-to-image, text-to-video, image-to-video, reference-image creation, task trackin

Vision & MultimodalManifest valid

0

dsh plugin --profile web add kling-ai-deepseek-harness

Hands repetitive text and vision labor (OCR, image analysis, comparison) to a local KoboldCpp (llama.cpp) server through koboldcpp_run and koboldcpp_vision tools, with on-demand server lifecycle manag

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-koboldcpp-hands

Continuous voice conversations for DSH with hands-free listening, push-to-talk, speech recognition, TTS replies, and background Agent delegation.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add @flowingspring/dsh-voco

Configure multiple image providers in Settings and call image_generate with the one selected model; images save under generate/image and show inline in the conversation.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-image-generation

Local WeChat OCR tool for DSH: `wechat_ocr_recognize` returns recognized text and the engine structured result for a local image path.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-wechat-ocr

Vision bridge for text-only DeepSeek: send images (alone or mixed with text) and a fast tiny vision model describes them behind the scenes — the chat keeps the picture, DeepSeek sees only text. Pure p

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dseyesopen

Listen to Ximalaya podcasts and audiobooks inside the DSH web UI: search albums, browse paginated tracks, and play through a now-playing bar (prev/play/next, seek, volume, rate). After QR login the "M

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-ximalaya

Browser Web Speech API voice input: zero server, zero keys, zero model downloads (Edge=Azure, Chrome=Google speech).

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-voice-webspeech

Page 4 of 6