DSH Plugins Marketplace
7,148 plugins indexed · 5,699 installable · 18,270 versions tracked
155 plugins
Vision-augmented DeepSeek adapter: a vision-capable model describes image input, then a text-only DeepSeek model reasons over the description.
★ 0
dsh plugin --profile web add @deepseek-ai/dsh-llm-deepseek-visionTurn a phone camera into a live viewfinder and photo input for your dsh session, over LAN or USB.
★ 0
Synesthesia Encoder for DSH: a vision model translates images into compact structured spatial text (canvas/elements/percentage coordinates), giving text-only LLMs pixel-level image understanding via t
★ 0
dsh plugin --profile web add dsh-plugin-mm-visionHands repetitive text and vision labor (OCR, image analysis, comparison) to a locally running Unsloth Desktop (Unsloth Studio) server through unsloth_run and unsloth_vision tools; pure HTTP client, ne
★ 0
dsh plugin --profile web add dsh-unsloth-handsConsent-bound one-turn voice preview for DSH Web with a credential-free local synthetic demo, optional Qwen Audio, exact Session isolation, and explicit transcript-to-draft handoff without automatic s
★ 0
dsh plugin --profile web add dsh-live-voiceby SuCriss
Voice control for DeepSeek Harness web: speech-to-text into the composer and spoken playback of assistant replies, zero dependencies
★ 0
MIT
JavaScript
Sep 1, 2026
dsh plugin --profile web add dsh-voice-controlFree vision OCR with adaptive tile recognition for long documents and Markdown/Word/PNG/Excel export.
★ 0
dsh plugin --profile web add aura-visionLocal-first structured vision for text-only agents: images go to a local OpenAI-compatible VLM and come back as JSON evidence (summary, verbatim OCR, layout regions, entities/relations, colors, explic
★ 0
dsh plugin --profile web add dsh-vision-localPiano performance plugin: ask the agent to play a piece and it renders on a Canvas2D grand piano with real Salamander Grand samples, an immersive stage, and an interactive 88-key keyboard.
★ 0
dsh plugin --profile web add dsh-pianistPlug-in vision for text-only models on DSH, with native interaction for image understanding and generation, GUI automation, through layered evidence memory and cache.
★ 0
dsh plugin --profile web add dsh-mindseyeVoice assistant for dsh web: say the wake phrase (e.g. "小鲸") to activate hands-free dictation — what you say is transcribed and typed into the chat box automatically. Supports spoken edit commands (se
★ 0
dsh plugin --profile web add dsh-voice-assistantGives dsh the ability to generate images and videos through the grok2api API.
★ 0
↓ 55/wk
dsh plugin --profile web add dsh-plugin-grok2api-media-toolChina-ready voice input for the composer. Requires a local Python bridge (pip install dashscope websockets, run bridge/voice-bridge.py) — the plugin alone does not work. Browser mic streams 16 kHz PCM
★ 0
dsh plugin --profile web add dsh-voice-input-cnGive your text-only model eyes - chat image attachments are auto-described via a vision model (default prompt), with iterative re-parsing through model-generated prompts when details are missing; syst
★ 0
dsh plugin --profile web add dsh-vision-pluginMic button in the composer tool row: Web Speech API speech-to-text (Chrome/Edge), language switching, and optional auto-send, zero dependencies.
★ 0
dsh plugin --profile web add dsh-voice-input-webLets a text-only DeepSeek agent read images in the same session by delegating to a vision-capable subagent, with send-time image-to-path conversion.
★ 0
↓ 105/wk
dsh plugin --profile web add dsh-subagent-visionby Harzva
Bridge Apple on-device Vision framework (macOS) into DeepSeek Harness: OCR, image classification, face detection, document layout as local dsh tools. No network, no API key.
★ 0
MIT
Swift
Aug 29, 2026
dsh plugin --profile web add dsh-maclensAgent-callable vision tool that describes local images via any OpenAI-compatible vision endpoint you configure, with an optional multi-model cross-check and no built-in keys.
★ 0
Reuses DeepSeek web's built-in vision mode for text-only models: the deepseek_vision tool drives the local deepseek-vision-cli browser automation (manual login helper, deep-think enabled, auto-closes
★ 0
dsh plugin --profile web add dsh-deepseek-visionAI audio generation for the DeepSeek Harness web GUI — multi-vendor TTS, music, sound effects and voice design with a sidebar panel, model comparison, resource library and Agent tools.
★ 0
dsh plugin --profile web add dsh-audiogenLocal eyes for text-only models: eyes_render draws text/shapes/Mermaid onto a canvas in the Web GUI, eyes_paste captures pasted images, eyes_ocr reads text via the built-in Windows OCR (offline), and
★ 0
dsh plugin --profile web add dsh-omni-visionA conversational voice frontend Agent for dsh: speak naturally over ByteDance Duplex, delegate requests to background tasks, and hear their asynchronous results reported by voice.
★ 0
A `describe_image` vision tool for text-only models: images (local path, URL, attachment) go to a configurable OpenAI-compatible vision endpoint and only the returned text enters the session.
★ 0
by Elohia
Image-to-text input for the Web UI: paste or drag an image and it is transcribed into structured text and sent, giving text-only LLMs image-input takeover (OpenAI-compatible vision API).
★ 0
MIT
JavaScript
Aug 15, 2026
dsh plugin --profile web add dsh-plugin-image-input