DSH Plugins Marketplace

DSH Plugins

Plugins

/

dsh-bundle-vision

s

dsh-bundle-vision

Manifest valid

Zero-core-change vision capability for DeepSeek Harness: the describe_image tool + profile bundle, installable via 'dsh plugin add'

hasBundlePatch

dsh-bundle-vision

A zero-core-change vision capability for DeepSeek Harness, shipped as one installable npm package that is both a profile bundle and a plugin:

  • the plugin registers the model-facing describe_image tool;
  • the bundle patch mounts that plugin on any profile.

The tool reads a local PNG/JPEG/WebP/GIF file, commits the bytes through the shipped attachment service, and asks the named multimodal route about it in one direct LLM request (provider / model are tool arguments). The result is text only — no image block ever enters the calling session, so a text-only main model gains vision without any change to dsh itself.

How it works against the shipped seams

Everything the tool uses already ships with every dsh profile:

  • ctx.fs (bounded byte read, session-workspace resolution) — filesystem capability;
  • ctx.attachments (saveImage, image limits, magic-byte validation) — durable image storage;
  • ctx.llm (resolveModelInfo + stream) with the pi-ai multi-provider adapter — the multimodal request itself.

The one per-deployment prerequisite is the same as for any vision use of dsh: the multimodal model must declare image input in the llm-pi-ai settings section, e.g.:

llm-pi-ai:
  providers:
    my-vision:
      apiKeyEnv: MY_VISION_API_KEY
      api: openai-completions
      baseURL: https://example.invalid/v1
      models:
        - id: my-vision-model
          input: [text, image]

(On releases whose Models page has the input-modality control, the same declaration is one dropdown.)

Install (installed dsh, no source checkout)

From the npm registry:

dsh plugin --profile <name> add dsh-bundle-vision

Or from a packed tarball (e.g. before the first publish, or for a pinned version):

dsh plugin --profile <name> add ./dsh-bundle-vision-0.1.0.tgz

dsh plugin forwards to pnpm inside the profile directory and reconciles the profile's bundle layers automatically — a dependency declaring dsh.bundle joins the layer stack. Restart dsh <name>; the tool registers for every agent (profile-root registrations are visible to all preset scopes).

Use

Ask the main model, naming the route:

Use describe_image with file_path /path/to/photo.jpg, provider my-vision, model my-vision-model, and prompt "OCR the text in this image".

The main model supplies provider/model per call — configure several multimodal routes and switch per call, no profile edits. Refusals name the failing gate (unknown extension, deployment media types, a route whose model does not declare image input, a missing file, a type-mismatched file, an errored stream).

Version floor

The package declares >=0.1.0-rc.6 peer dependencies on the dsh seam packages. Bump the floor if a later release is required.

Development

npm install          # dev deps resolve the seam packages from the npm registry
npm run typecheck    # tsc --noEmit over src + tests
npm test             # vitest
npm run build        # tsdown (lib/index.js) + tsc declarations (lib/types)
npm pack             # the installable tarball

Comments

Loading…

Similar plugins

dsh-vlm-bridge

by me9rez

DeepSeek Harness (dsh) bundle plugin: vision_analyze tool lets text-only LLM agents read images via SenseNova VLM, with Schemastery config and single-source credentials

Manifest valid

★ 0

JavaScript

Aug 14, 2026

dsh plugin --profile web add dsh-vlm-bridge

by 123twtd

DeepSeek Harness (dsh) 模型视觉开关插件:为自定义模型声明图像输入能力(写入 llm-pi-ai 配置),适配 0.1.x

Tools & CapabilitiesManifest valid

★ 0

JavaScript

Sep 3, 2026

dsh plugin --profile web add dsh-vision-toggle

by Zh-U-hB

DeepSeek Harness plugin: route image-bearing messages to a user-configured OpenAI-compatible vision endpoint when the active text model cannot see images

Manifest valid

★ 0

MIT

TypeScript

Aug 16, 2026

dsh plugin --profile web add @deepseek-ai/dsh-vision-bridge

by NagasakiSoyo-ui

Vision-augmented DeepSeek adapter: a vision-capable model describes image input, then a text-only DeepSeek model reasons over the description.

Vision & MultimodalTools & CapabilitiesManifest valid

★ 1

MIT

TypeScript

Aug 14, 2026

dsh plugin --profile web add @deepseek-ai/dsh-llm-deepseek-vision

by Favio8

DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.

Tools & CapabilitiesManifest valid

★ 3

↓ 95/wk

MIT

TypeScript

Aug 17, 2026

dsh plugin --profile web add dsh-plugin-deepeye

by Harvey-Will

DeepSeek Harness vision plugin: 8 analysis modes (describe, OCR, chart data, UI review, object detection, compare, code-gen, debug), any OpenAI- or Anthropic-compatible vision API, with a built-in fre

Vision & MultimodalTools & CapabilitiesManifest valid

★ 1

MIT

TypeScript

Sep 30, 2026

dsh plugin --profile web add dsh-vision-analysis