Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

47 plugins found

Y

dsh-vision-router

ysr666/dsh-vision-router

Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.

1.1k15 hours agoVision, Voice & MultimodalMIT
A

dsh-vision-toolkit

anionex/dsh-vision-toolkit

Vision for text-only models: paste an image and the model switches to a Vision Toolkit variant for image Q&A, multi-image comparison, long-screenshot OCR, screenshot-to-UI reproduction, element grounding, and pixel diff. No API key by default — images are processed by the author-hosted free service, 100 per machine per day; configurable to your own provider.

8853 days agoVision, Voice & MultimodalMIT
O

dsh-vision

oil-oil/dsh-vision

Near-native image understanding for DeepSeek Harness

242 months agoVision, Voice & MultimodalMIT
F

dsh-vision-proxy

flyvhidbwo/dsh-vision-proxy

DeepSeek brain + automatic image transcription: attach images in the GUI and each one is transcribed via the official deepseek-v4-flash-vision-exp by default (a pure-text V4-Pro brain can see images), with any OpenAI-compatible VLM or local Ollama as alternatives.

15last monthVision, Voice & MultimodalMIT
P

dsh-vision-opencode

poiuyjie/dsh-vision-opencode

Adds a configurable vision model to text-only main models: a vision_read_image tool, a composer-bar vision-model selector, and automatic image-to-text conversion for text-only routes.

1325 days agoVision, Voice & MultimodalMIT
L

dsh-vision

linenxi-ctrl/dsh-vision

External vision plugin for DeepSeek Harness: whale-button config panel, image recognition with auto-reply, and agent screenshot/recognize tools.

92 months agoVision, Voice & MultimodalMIT
5

dsh-vision

54xkeee/dsh-vision

Vision for text-only DeepSeek via Doubao Web by default (zero-cost, no API key — drives your logged-in Chrome through a Windows CDP bridge), with Antigravity IDE quota (flash/pro) or Gemini fallback; auto detail escalation, vision evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.

72 months agoVision, Voice & MultimodalMIT
L

dsh-drop-to-path

loudmore/dsh-drop-to-path

DSH 插件:图片与文件直达纯文本模型——图片保留原生附件体验,PDF/Office/压缩包/视频/音频显示为附件栏方块,点击发送时自动转为工作区路径,配合 dsh-vision-toolkit 粘贴即看图。A DSH plugin that delivers images AND files to text-only models as workspace paths: images keep the native attachment UI, other files show as square chips in the rail, paths append on send — pairs with dsh-vision-toolkit.

62 months agoUI EnhancementsMIT
S

dsh-vision-bridge

sfyyy/dsh-vision-bridge

On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

52 months agoVision, Voice & MultimodalMIT
X

dsh-vision-hub (tool-vision)

xing666173/dsh-vision-hub

Enhanced vision toolbox: 14 pixel-level vision tools (describe, ground, detect, crop, pixel-diff, OCR, long-screenshot OCR, vectorize, colors, cutout, screenshot, present, materialize, html-screenshot) driven by one OpenAI-compatible endpoint, with clean \[图片: path] bridge markers, content-safety classification and rate-limit auto-retry.

4last monthVision, Voice & Multimodal
X

dsh-vision-hub (file-drop)

xing666173/dsh-vision-hub

Drag-and-drop file upload for PDF, Word, Excel and images: dropped files are saved to a local directory and referenced by path, no base64 bloat in the chat.

4last monthTools & Capabilities
X

dsh-vision-hub (bridge-preview)

xing666173/dsh-vision-hub

Inline preview for \[图片: path] image bridge markers: after the marker renders, the image shows in chat automatically, keeping long instructions out of the conversation.

4last monthSessions & Messages
1

dsh-vision-sidecar

121103qwq/dsh-vision-sidecar

Hosted free vision sidecar for DeepSeek Harness with durable session evidence

42 months agoVision, Voice & MultimodalMIT
G

dsh-vision-bridge

gxx182/dsh-vision-bridge

Bridges session images to configurable vision providers and returns text-only analysis to eligible DeepSeek Harness model routes.

42 months agoVision, Voice & MultimodalMIT
P

dsh-agent-router

peterwangze/dsh-agent-router

DeepSeek Harness 多模型路由插件:让专业的事情交给专业的 agent——自定义视觉/翻译/语音/子代理等专业 agent 并绑定独立模型,多模态账号一键登录、账号池健康路由与实时用量统计

32 months agoModels & ProvidersMIT
T

dsh-plugins (dsh-vision)

tzhr-invest/dsh-plugins

Agent-callable vision tool that describes local images via any OpenAI-compatible vision endpoint you configure, with an optional multi-model cross-check and no built-in keys.

37 days agoVision, Voice & MultimodalMIT
M

dsh-vision-tools

moon09300731/dsh-vision-tools

Full vision-capability bundle for DeepSeek Harness: a vision_understand tool (OpenAI-compatible vision APIs, free Zhipu GLM-4V-Flash by default) plus paste/drag-and-drop/button entry points for image recognition.

32 months agoVision, Voice & MultimodalMIT
H

dsh-vision-mix

haiziyao/dsh-vision-mix

Combine text, vision, and image-generation APIs into one Mix model with automatic routing: text-only requests go to the chat model, user images and agent screenshots go to the vision model, follow-ups keep using the same session image, and agents can generate or edit images with session-scoped call history.

326 days agoVision, Voice & MultimodalMIT
N

dsh-vision-plugin

nexsjournal/dsh-vision-plugin

Installs an image-input checkbox into the Models catalog so a custom model can declare image input and receive images directly; includes an optional BYO vision relay.

32 months agoUI EnhancementsMIT
B

dsh-vision-plugin

bug-huntter/dsh-vision-plugin

Configurable image recognition for text-only DSH models: image messages are first transcribed by an OpenAI-compatible vision model (Base URL, model ID and API key set in a Settings section) and then passed to the main model as text, while image-input support is advertised. The API key auth scheme is selectable — OpenAI, Anthropic, Gemini or Azure style request headers — and a missing key is reported before any request is sent.

26 days agoVision, Voice & MultimodalMIT
G

dsh-vision-bridge

goodandready/dsh-vision-bridge

Routes images to a vision model of your choice - auto-rewrite, explicit tools, or hybrid - so a text-only chat model does not fail a turn that contains a picture.

28 days agoVision, Voice & MultimodalMIT
X

dsh-vision

xiaoshihou514/dsh-vision

Native vision capability extension, using either Zhipu (free) or Qwen-VL (local).

2last monthVision, Voice & MultimodalMIT
K

dsh-vision-recognizer

kaixinbaba/dsh-vision-recognizer

Vision provider route that transcribes attached images to text through a configurable model (15+ OpenAI-compatible and Anthropic vendors) while DeepSeek keeps answering.

2last monthVision, Voice & MultimodalMIT
3

dsh-vision (vision-tool)

314857493/dsh-vision

Model-facing `vision` tool for DeepSeek Harness: describe and OCR image files by calling the free Zhipu GLM vision API directly (glm-4v-flash fallback chain), no external CLI required.

229 days agoVision, Voice & MultimodalMIT