Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

8 plugins found

A

dsh-vision-toolkit

anionex/dsh-vision-toolkit

Vision for text-only models: paste an image and the model switches to a Vision Toolkit variant for image Q&A, multi-image comparison, long-screenshot OCR, screenshot-to-UI reproduction, element grounding, and pixel diff. No API key by default — images are processed by the author-hosted free service, 100 per machine per day; configurable to your own provider.

8878 hours agoVision, Voice & MultimodalMIT
D

dsh-web-ui (dsh-tool-describe-image)

damonkoy/dsh-web-ui

Gives a text-only model image understanding via a vision-language model, exposed as a `describe_image` tool.

202 months agoVision, Voice & MultimodalApache-2.0
S

dsh-deepseek-vision

siegfly/dsh-deepseek-vision

A vision-language gateway provider route: pasted images are described by a configurable VL model (Qwen-VL by default) before the DeepSeek wire.

920 days agoVision, Voice & MultimodalMIT
S

dsh-vision-bridge

sfyyy/dsh-vision-bridge

On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

5last monthVision, Voice & MultimodalMIT
1

dsh-vision-sidecar

121103qwq/dsh-vision-sidecar

Hosted free vision sidecar for DeepSeek Harness with durable session evidence

42 months agoVision, Voice & MultimodalMIT
G

dsh-vision-bridge

gxx182/dsh-vision-bridge

Bridges session images to configurable vision providers and returns text-only analysis to eligible DeepSeek Harness model routes.

4last monthVision, Voice & MultimodalMIT
D

deepseekeyes

dttxorg/deepseekeyes

Auditable vision and cross-platform Computer Use runtime for DeepSeek Harness with source-preserving evidence.

3last monthVision, Voice & MultimodalMIT
W

visual-review

wang-bool/visual-review

Renders pasted/uploaded images inline in the DSH Web chat and gives text-only models vision: the model-invokable visual_review tool calls any OpenAI-compatible multimodal API first, falling back to a local Qwen3-VL worker.

3last monthVision, Voice & MultimodalMIT