Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

78 plugins found

L

modlens

liustack/modlens

Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).

4.1k2 days agoVision, Voice & MultimodalMIT
Y

dsh-vision-router

ysr666/dsh-vision-router

Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.

1.1k5 hours agoVision, Voice & MultimodalMIT
T

tongflow (dsh-tongflow)

tong-io/tongflow

TongFlow film-crew studio for image, voice, music and video production: the agent writes per-asset TongFlow workflow files (.tongflow.json) that run through TongFlow plugins, with an embedded workflow canvas, a shot/character/take project layout and a manga-drama template; sessions starting with @tongflow open the Studio view.

1k3 days agoWorkflow & AutomationAGPL-3.0
T

tongflow

tong-io/tongflow

TongFlow studio plugin for DeepSeek Harness (dsh): film-crew style project model, agent-authored TongFlow workflows, deterministic media generation, embedded canvas.

870last monthVision, Voice & MultimodalAGPL-3.0
O

watch-skill

oxbshw/watch-skill

DeepWatch's capabilities, installable into an existing DeepSeek Harness profile

37815 days agoVision, Voice & MultimodalMIT
Z

dsh-crew

zseven-w/dsh-crew

Dispatch work to DSH agents from Claude Code or Codex: native subagent progress, in-host worker sessions with per-tier presets, and a multimodal bridge for vision and image generation.

1535 days agoIntegrations & RemoteMIT
V

ark-cli

volcengine/ark-cli

Volcengine Ark provider plugin for DeepSeek Harness (DSH): registers Agent Plan, Coding Plan, and postpaid routes on the native pi-ai adapter so Ark models appear in the model picker.

112last monthModels & ProvidersApache-2.0
S

dsh-AuthInOne

stormycry-cryp/dsh-authinone

Adds account login, API and custom Provider setup, model switching, image fallback for text-only models, and token/cost attribution to DeepSeek Harness 47f.

104last monthModels & ProvidersMIT
S

dsh-design-qa

sunxin-ai/dsh-design-qa

Design-fidelity QA for text-only models: a `deepseek_vision` tool borrows an eye from any OpenAI-compatible vision route, so the model can judge whether an implementation matches its mock — shipped with the benchmark behind that judgement (four fixtures, 23 injected defects, raw transcripts) and the questioning discipline it depends on.

4423 days agoVision, Voice & MultimodalMIT
J

picturereader

jing-hy/picturereader

Image "reading" for text-only models: downscale + reduce color depth + structure/color fingerprints into text grids fed back to the conversation, letting the model zoom, sample and OCR autonomously like a multimodal model; fully local with zero external model dependency, ships an image-reading methodology skill and optional PaddleOCR.

3711 days agoVision, Voice & MultimodalMIT
O

dsh-vision

oil-oil/dsh-vision

Near-native image understanding for DeepSeek Harness

242 months agoVision, Voice & MultimodalMIT
J

dsh-visual-plugin

jyh20030112/dsh-visual-plugin

Gives text-only models vision: forwards user images to an OpenAI-compatible vision model and shows the descriptions in a Web UI right panel.

1628 days agoVision, Voice & MultimodalMIT
F

dsh-vision-proxy

flyvhidbwo/dsh-vision-proxy

DeepSeek brain + automatic image transcription: attach images in the GUI and each one is transcribed via the official deepseek-v4-flash-vision-exp by default (a pure-text V4-Pro brain can see images), with any OpenAI-compatible VLM or local Ollama as alternatives.

14last monthVision, Voice & MultimodalMIT
W

dsh-mmx-bridge

welsione/dsh-mmx-bridge

MiniMax multimodal bridge: one `mmx_bridge` tool covers image understanding/generation, video, TTS, music, cover, web search and quota; optional `web_search`/`read_image` takeover; inline players/image previews right in the Web GUI (npm: `dsh-mmx-bridge`).

1012 days agoTools & CapabilitiesMIT
M

dsh-provider-qoder

mo-n/dsh-provider-qoder

Connects Qoder subscriptions to DeepSeek Harness, supporting Global and China regions, multimodal input, and tool calls.

910 hours agoModels & ProvidersMIT
S

dsh-deepseek-vision

siegfly/dsh-deepseek-vision

A vision-language gateway provider route: pasted images are described by a configurable VL model (Qwen-VL by default) before the DeepSeek wire.

919 days agoVision, Voice & MultimodalMIT
M

dsh-iris

mokuyoaxis/dsh-iris

Media and vision workspace for DeepSeek Harness: image, video and speech generation, image Q&A and element locating, long-image OCR, pixel diff, HTML-screenshot verification and video summarization, with DashScope and OpenAI-compatible providers, model pools and a workbench client.

79 days agoTools & CapabilitiesMIT
5

dsh-vision

54xkeee/dsh-vision

Vision for text-only DeepSeek via Doubao Web by default (zero-cost, no API key — drives your logged-in Chrome through a Windows CDP bridge), with Antigravity IDE quota (flash/pro) or Gemini fallback; auto detail escalation, vision evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.

7last monthVision, Voice & MultimodalMIT
Y

deepseek-harness-plugins (vision-bridge)

yinxe/deepseek-harness-plugins

Let text-only models see images: when a picture arrives with a placeholder like \[image omitted because this model accepts text only], the model calls the vision_describe tool and the plugin forwards the image reference plus the question to a multimodal model, retrying with a fallback model on failure. Configured in the settings page and persisted via the official settings API.

56 days agoVision, Voice & MultimodalMIT
S

dsh-vision-bridge

sfyyy/dsh-vision-bridge

On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

5last monthVision, Voice & MultimodalMIT
A

dsh-guide-dog

atropinoltt/dsh-guide-dog

MiniMax-powered multimodal plugin: real-time voice call mode (streaming conversation, floating dock UI), voice mode and mic voice input, plus image/video/music/speech generation and vision inspection tools.

5last monthVision, Voice & MultimodalMIT
G

deepseek-vision

gou-gee/deepseek-vision

DeepSeek Harness 原生视觉 Bundle:粘贴或拖入图片,通过托管的 deepseek-vision-mcp 调用 OpenAI 兼容视觉模型。

4last monthMCP & ConnectorsMIT
G

deepseek-vision (dsh-plugin-deepseek-vision)

gou-gee/deepseek-vision

Vision MCP and DSH bundle for text-only DeepSeek: analyze_image, analyze_clipboard, compare_images and vision_status tools, a visual settings page, free GLM-4.6V-Flash by default, result caching and rate-limit tolerance; keys stay out of logs.

423 days agoVision, Voice & MultimodalMIT
F

dsh-plugin-deepeye

favio8/dsh-plugin-deepeye

DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.

42 months agoVision, Voice & Multimodal