DeepSeek Harness plugin

dsh-bundle-vision

Vision bundle + plugin for DeepSeek Harness: the describe_image tool reads local images and asks any configured multimodal route, with zero core changes

Jump to install

Source facts

Repository
skillre/dsh-bundle-vision
Latest update
Aug 15, 2026
Category
Tools & Capabilities
GitHub stars
0
Format
plugin
Catalog evidence
Upstream dsh.bundle evidence
Evidence path
package.json#dsh.bundle
Checked against
0.1.0-rc.8
Upstream check date
2026-08-20

This evidence comes from the upstream catalog. This site has not installed, run, or security-reviewed the plugin.

Install

Start with a prompt that asks an agent to review the GitHub repository and source. Switch to the command if you want to install it yourself.

Copy this prompt into DSH, Codex, or another agent and ask it to review the GitHub repository and source first.

Do not install or run any commands yet. Read this plugin's GitHub repository, README, and relevant source code. Then answer the questions below clearly and directly so I can decide whether it fits my needs:

1. What is this plugin, and what problem does it solve?
2. Who is it for, and what are its typical use cases?
3. How is it used after installation? Include one minimal example.
4. What known limitations or privacy, security, compatibility, or maintenance risks does it have?
5. Give a clear recommendation: recommend, conditionally recommend, or do not recommend, with reasons.

Distinguish statements documented by the repository, inferences from source code, and unknowns. If evidence is insufficient, say so explicitly. Do not guess or simply repeat the README.

GitHub: https://github.com/skillre/dsh-bundle-vision
Plugin: dsh-bundle-vision
Author: skillre

Check the source files

Read the README and other files from this plugin directory before installing.

File explorer3 files
README.mdSource · read only

dsh-bundle-vision

A zero-core-change vision capability for DeepSeek Harness, shipped as one installable npm package that is both a profile bundle and a plugin:

  • the plugin registers the model-facing describe_image tool;
  • the bundle patch mounts that plugin on any profile.

The tool reads a local PNG/JPEG/WebP/GIF file, commits the bytes through the shipped attachment service, and asks the named multimodal route about it in one direct LLM request (provider / model are tool arguments). The result is text only — no image block ever enters the calling session, so a text-only main model gains vision without any change to dsh itself.

How it works against the shipped seams

Everything the tool uses already ships with every dsh profile:

  • ctx.fs (bounded byte read, session-workspace resolution) — filesystem capability;
  • ctx.attachments (saveImage, image limits, magic-byte validation) — durable image storage;
  • ctx.llm (resolveModelInfo + stream) with the pi-ai multi-provider adapter — the multimodal request itself.

The one per-deployment prerequisite is the same as for any vision use of dsh: the multimodal model must declare image input in the llm-pi-ai settings section, e.g.:

llm-pi-ai:
  providers:
    my-vision:
      apiKeyEnv: MY_VISION_API_KEY
      api: openai-completions
      baseURL: https://example.invalid/v1
      models:
        - id: my-vision-model
          input: [text, image]

(On releases whose Models page has the input-modality control, the same declaration is one dropdown.)

Install (installed dsh, no source checkout)

From the npm registry:

dsh plugin --profile <name> add dsh-bundle-vision

Or from a packed tarball (e.g. before the first publish, or for a pinned version):

dsh plugin --profile <name> add ./dsh-bundle-vision-0.1.0.tgz

dsh plugin forwards to pnpm inside the profile directory and reconciles the profile's bundle layers automatically — a dependency declaring dsh.bundle joins the layer stack. Restart dsh <name>; the tool registers for every agent (profile-root registrations are visible to all preset scopes).

Use

Ask the main model, naming the route:

> Use describe_image with file_path /path/to/photo.jpg, provider my-vision, model my-vision-model, and prompt "OCR the text in this image".

The main model supplies provider/model per call — configure several multimodal routes and switch per call, no profile edits. Refusals name the failing gate (unknown extension, deployment media types, a route whose model does not declare image input, a missing file, a type-mismatched file, an errored stream).

Version floor

The package declares >=0.1.0-rc.6 peer dependencies on the dsh seam packages. Bump the floor if a later release is required.

Development

npm install          # dev deps resolve the seam packages from the npm registry
npm run typecheck    # tsc --noEmit over src + tests
npm test             # vitest
npm run build        # tsdown (lib/index.js) + tsc declarations (lib/types)
npm pack             # the installable tarball