DSH Plugins Marketplace

DSH Plugins

DSH Plugins Marketplace

7,148 plugins indexed · 5,699 installable · 18,270 versions tracked

AllDevelopment & Infrastructure (485)Tools & Capabilities (1496)Memory & Context (199)Workflow & Automation (234)Models & Providers (201)Terminal & Clients (148)Security & Audit (138)Vision & Multimodal (155)UI & Experience (592)Notifications & Integrations (142)Themes & Skins (128)Sessions & Messages (224)Just for Fun (109)

Sort:

Most starsRecently updatedNewestName A–ZInstallable only

141 plugins

dsh-soundscape

by berserk0501

DSH 本机思考与工具音效插件,支持 MediaPlayer、WAV/MP3、自定义映射和设置面板

Vision & MultimodalManifest valid

0

MIT

JavaScript

Aug 26, 2026

dsh plugin --profile web add @dsh-external/dsh-soundscape

Native image attachments for text-only DeepSeek in the Web GUI: pasted or dropped images appear as thumbnails in the session, and before dispatch the host reads them with the free Zhipu GLM-4V-Flash v

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-dseyes

Auto-discovery vision bridge for text-only DeepSeek Harness agents: automatically finds an image-capable model from your configured providers and returns picture descriptions as plain text via a visio

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-auto-vision

Native vision capability extension, using either Zhipu (free) or Qwen-VL (local).

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-vision

Free lip-sync video generation plugin for DSH with 3000+ voices and 500+ languages.

Vision & MultimodalManifest valid

0

53/wk

dsh plugin --profile web add dsh-tool-lipsync

Vision-augmented DeepSeek adapter: a vision-capable model describes image input, then a text-only DeepSeek model reasons over the description.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add @deepseek-ai/dsh-llm-deepseek-vision

Synesthesia Encoder for DSH: a vision model translates images into compact structured spatial text (canvas/elements/percentage coordinates), giving text-only LLMs pixel-level image understanding via t

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-plugin-mm-vision

Hands repetitive text and vision labor (OCR, image analysis, comparison) to a locally running Unsloth Desktop (Unsloth Studio) server through unsloth_run and unsloth_vision tools; pure HTTP client, ne

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-unsloth-hands

Consent-bound one-turn voice preview for DSH Web with a credential-free local synthetic demo, optional Qwen Audio, exact Session isolation, and explicit transcript-to-draft handoff without automatic s

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-live-voice

by SuCriss

Voice control for DeepSeek Harness web: speech-to-text into the composer and spoken playback of assistant replies, zero dependencies

Vision & MultimodalManifest valid

0

MIT

JavaScript

Sep 1, 2026

dsh plugin --profile web add dsh-voice-control

Free vision OCR with adaptive tile recognition for long documents and Markdown/Word/PNG/Excel export.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add aura-vision

Local-first structured vision for text-only agents: images go to a local OpenAI-compatible VLM and come back as JSON evidence (summary, verbatim OCR, layout regions, entities/relations, colors, explic

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-vision-local

Piano performance plugin: ask the agent to play a piece and it renders on a Canvas2D grand piano with real Salamander Grand samples, an immersive stage, and an interactive 88-key keyboard.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-pianist

Plug-in vision for text-only models on DSH, with native interaction for image understanding and generation, GUI automation, through layered evidence memory and cache.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-mindseye

Voice assistant for dsh web: say the wake phrase (e.g. "小鲸") to activate hands-free dictation — what you say is transcribed and typed into the chat box automatically. Supports spoken edit commands (se

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-voice-assistant

Gives dsh the ability to generate images and videos through the grok2api API.

Vision & MultimodalManifest valid

0

55/wk

dsh plugin --profile web add dsh-plugin-grok2api-media-tool

China-ready voice input for the composer. Requires a local Python bridge (pip install dashscope websockets, run bridge/voice-bridge.py) — the plugin alone does not work. Browser mic streams 16 kHz PCM

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-voice-input-cn

Give your text-only model eyes - chat image attachments are auto-described via a vision model (default prompt), with iterative re-parsing through model-generated prompts when details are missing; syst

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-vision-plugin

Mic button in the composer tool row: Web Speech API speech-to-text (Chrome/Edge), language switching, and optional auto-send, zero dependencies.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-voice-input-web

Lets a text-only DeepSeek agent read images in the same session by delegating to a vision-capable subagent, with send-time image-to-path conversion.

Vision & MultimodalManifest valid

0

105/wk

dsh plugin --profile web add dsh-subagent-vision

by Harzva

Bridge Apple on-device Vision framework (macOS) into DeepSeek Harness: OCR, image classification, face detection, document layout as local dsh tools. No network, no API key.

Vision & MultimodalManifest valid

0

MIT

Swift

Aug 29, 2026

dsh plugin --profile web add dsh-maclens

Reuses DeepSeek web's built-in vision mode for text-only models: the deepseek_vision tool drives the local deepseek-vision-cli browser automation (manual login helper, deep-think enabled, auto-closes

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-deepseek-vision

AI audio generation for the DeepSeek Harness web GUI — multi-vendor TTS, music, sound effects and voice design with a sidebar panel, model comparison, resource library and Agent tools.

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-audiogen

Local eyes for text-only models: eyes_render draws text/shapes/Mermaid onto a canvas in the Web GUI, eyes_paste captures pasted images, eyes_ocr reads text via the built-in Windows OCR (offline), and

Vision & MultimodalManifest valid

0

dsh plugin --profile web add dsh-omni-vision

Page 5 of 6