Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

224 plugins found

R

dsh-longtask-orchestrator

rocker2018-droid/dsh-longtask-orchestrator

Long-task orchestration loop: Codex plans/scores/reviews, DeepSeek executes, Kimi supplements (vision acceptance, summarization, cross-check).

2last monthWorkflow & AutomationMIT
X

dsh-vision

xiaoshihou514/dsh-vision

Native vision capability extension, using either Zhipu (free) or Qwen-VL (local).

2last monthVision, Voice & MultimodalMIT
K

dsh-vision-recognizer

kaixinbaba/dsh-vision-recognizer

Vision provider route that transcribes attached images to text through a configurable model (15+ OpenAI-compatible and Anthropic vendors) while DeepSeek keeps answering.

2last monthVision, Voice & MultimodalMIT
H

dsh-open-eyes

hyp6666/dsh-open-eyes

Vision bridge for text-only DeepSeek routes that analyzes attached and local images through configurable OpenAI Responses, Chat Completions, or Anthropic Messages endpoints while leaving image-capable routes native.

224 days agoVision, Voice & MultimodalMIT
5

dsh-youreyes

54xkeee/dsh-youreyes

Vision toolkit for text-only DeepSeek: model-invokable `vision` tool, wrapper adapters for deepseek/opencode-go (v4 flash/pro), Antigravity IDE quota (default, flash/pro) / any OpenAI-compatible VLM / Gemini / local Ollama channels, evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.

2last monthVision, Voice & MultimodalMIT
3

dsh-vision (vision-tool)

314857493/dsh-vision

Model-facing `vision` tool for DeepSeek Harness: describe and OCR image files by calling the free Zhipu GLM vision API directly (glm-4v-flash fallback chain), no external CLI required.

225 days agoVision, Voice & MultimodalMIT
3

dsh-vision (vision-route)

314857493/dsh-vision

Registers a `deepseek-vision` provider route: the Web GUI accepts pasted images and transcribes them to text via the free Zhipu GLM vision API before delegating to the DeepSeek adapter.

225 days agoVision, Voice & MultimodalMIT
G

dsh-sub2api

godd6366/dsh-sub2api

Connect a sub2api gateway to DeepSeek Harness: OpenAI-compatible multi-provider routes (OpenAI / Claude / Grok / Gemini) behind one base URL, with per-key model discovery, usage lookup, and global vision/image tools.

222 days agoModels & ProvidersMIT
X

dsh-vision-bridge

ximengxiaolan/dsh-vision-bridge

Composer-attached images are transcribed to text by an OpenAI-compatible vision model before reaching text-only DeepSeek models.

22 months agoVision, Voice & MultimodalMIT
J

dsh-snapcompact

jessenreinhart/dsh-snapcompact

Bitmap-frame context compression plugin for DeepSeek Harness: renders discarded conversation history into pixel-font PNG frames for vision-capable models.

17 days agoSessions & Messages
P

deep-blend (bundle)

pearjelly/deep-blend

Render and iterate on Blender scenes from DSH: 16 tools for scene specs, previews, visual review and delivery renders, with immutable revisions and an approval gate.

12 days agoTools & CapabilitiesMIT
D

dsh-context-compression-improved

drscrewdriver/dsh-context-compression-improved

Same-origin in mechanism with the two loudest lines in context compression. The code-skeleton gate follows the skeletonization approach of Headroom (Apache-2.0), whose published headline is 20% fewer tokens for coding agents and 60–95% fewer tokens for JSON, same answers. The estimator channel follows TokenPilot (arXiv:2606.17016), which reports up to 60% lower cost for long-session agents. Both figures are theirs, quoted as-is; this plugin ships no benchmark of its own and claims no reduction of its own. What it adds for DeepSeek Harness: choose a compression profile, set the Auto Compact trigger level and toggle code-skeleton compression from one settings section, with exact DeepSeek V4 tokenizer measurement, same-revision count verification, and fail-open behaviour that keeps the original tool results on unsupported models.

114 hours agoSessions & MessagesMIT
F

dsh-drawai

fourzkw/dsh-drawai

An editable diagram canvas in the DSH sidebar plus two agent tools (diagram_read, diagram_apply) that read and edit the workspace in place, using native .drawio files: the canvas opens and saves them losslessly — cells it does not understand are preserved byte-for-byte — and a file-fingerprint revision powers optimistic locking against concurrent edits.

110 days agoTools & CapabilitiesMIT
C

dsh-vision-autoswitch

cultofluna/dsh-vision-autoswitch

DeepSeek Harness plugin: auto-route image-bearing requests to deepseek-v4-flash-vision-exp, then fall back to the original model.

128 days agoVision, Voice & MultimodalMIT
Z

vision-use

zzy6-a/vision-use

Desktop computer-use for DSH: captures the Windows screen into the agent vision channel and drives mouse/keyboard through a Codex-style overlay with Esc abort; auto-detects Windows native or WSL hosts.

119 days agoTools & CapabilitiesMIT
X

computer-user-vision

xie129716/computer-user-vision

Windows computer use forked from computer-user: 13 computer_* tools that read the screen and drive the mouse and keyboard. Image-capable routes get the screenshot as a real image with an exact image-to-screen mapping, so no external OCR; elements return as UI Automation refs so a click lands on the exact control rectangle; Ctrl+Alt+Esc stops every call.

115 days agoTools & CapabilitiesMIT
T

dsh-model-picker

ttmouse/dsh-model-picker

Model picker for the DSH web composer that replaces the model seat: a search-first popup whose left column indexes the suppliers and scrolls the grouped list, with a favorites group pinned at the top, image icons separating native from bridged vision input, and a reasoning-effort panel opened from the right zone of the trigger.

116 days agoUI Enhancements
C

kimi-webbridge-dsh

cfanmaoli/kimi-webbridge-dsh

Provisions the kimi-webbridge bridge daemon (auto-install with SHA-256 verification, auto-start) and registers a 25-action skill that drives the user's real browser through the local bridge.

12 days agoTools & CapabilitiesMIT
D

dsh-llm-openai-completions

drscrewdriver/dsh-llm-openai-completions

OpenAI-completions-compatible adapter for custom gateways (vLLM / LM Studio / self-hosted proxies): always role:"system", thinking driven by the model config, Qwen-style response split, and vision-model image input (single / multiple).

114 days agoModels & ProvidersMIT
J

dsh-mmroute

jmxsxwyzjdwl/dsh-mmroute

Transparent multimodal routing for text-only models: every image in every model call is fully transcribed (verbatim OCR, data, uncertainty zones, injection-hardened) by your own multimodal understander, with focused re-look via vision_relook and automatic retry on image-related failures. No bundled endpoints, no borrowed logins.

123 days agoVision, Voice & MultimodalMIT
C

aura-vision

ck-epsilon/aura-vision

Free vision OCR with adaptive tile recognition for long documents and Markdown/Word/PNG/Excel export.

1last monthVision, Voice & MultimodalMIT
X

dsh-profile-settings

xmoon/dsh-profile-settings

Per-profile settings overlay for DeepSeek Harness: global settings.yaml stays the baseline while each profile overrides any namespace through its own profiles/<name>/settings.patch.yml — object sections merge recursively, scalars and arrays replace wholesale, and !unset masks an inherited value. The overlay is transparent to existing plugins (they keep reading ctx.settings unchanged), writes land in the profile overlay only, and official schema validation, revision semantics, expectedRevision conflicts, watchers and events are untouched. Ships a settings command family (get/set/unset/mask/unmask/promote/demote/migrate/diff/layers) plus a Profile Settings section in the web Settings panel over a loopback RPC channel.

1last monthTools & CapabilitiesMIT
H

dsh-vision-analysis

harvey-will/dsh-vision-analysis

DeepSeek Harness vision plugin: 8 analysis modes (describe, OCR, chart data, UI review, object detection, compare, code-gen, debug), any OpenAI- or Anthropic-compatible vision API, with a built-in free vision model and automatic rate-limit failover.

14 days agoVision, Voice & MultimodalMIT
A

dsh-client-vision (tool-vision)

ankye/dsh-client-vision

Screen capture and external vision recognition: take_screenshot, list_windows, analyze_image and view_image tools with a configurable GPT vision channel (gpt-5.5 / gpt-5.6-sol / gpt-5.6-terra), API key via the credentials service, and a settings card; view_image shows the screenshot in the Web UI while the model context keeps text only.

120 days agoVision, Voice & MultimodalMIT