Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

86 plugins found

L

modlens

liustack/modlens

Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).

4k2 days agoVision, Voice & MultimodalMIT
Y

dsh-vision-router

ysr666/dsh-vision-router

Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.

1.1kyesterdayVision, Voice & MultimodalMIT
T

tongflow (dsh-tongflow)

tong-io/tongflow

TongFlow film-crew studio for image, voice, music and video production: the agent writes per-asset TongFlow workflow files (.tongflow.json) that run through TongFlow plugins, with an embedded workflow canvas, a shot/character/take project layout and a manga-drama template; sessions starting with @tongflow open the Studio view.

1k9 hours agoWorkflow & AutomationAGPL-3.0
T

tongflow

tong-io/tongflow

TongFlow studio plugin for DeepSeek Harness (dsh): film-crew style project model, agent-authored TongFlow workflows, deterministic media generation, embedded canvas.

870last monthVision, Voice & MultimodalAGPL-3.0
O

watch-skill

oxbshw/watch-skill

DeepWatch's capabilities, installable into an existing DeepSeek Harness profile

3786 days agoVision, Voice & MultimodalMIT
Z

dsh-crew

zseven-w/dsh-crew

Dispatch work to DSH agents from Claude Code or Codex: native subagent progress, in-host worker sessions with per-tier presets, and a multimodal bridge for vision and image generation.

1497 days agoIntegrations & RemoteMIT
V

ark-cli (ark-managed-agents)

volcengine/ark-cli

Adds a Managed Agents settings tab and MCP tools to dispatch long-running agent tasks to Ark cloud Managed Agents.

1332 days agoTools & CapabilitiesApache-2.0
V

ark-cli (ark-plan-api)

volcengine/ark-cli

Registers Ark Agent Plan, Coding Plan and postpaid model routes in the native DSH model picker.

1332 days agoModels & ProvidersApache-2.0
V

ark-cli

volcengine/ark-cli

Volcengine Ark provider plugin for DeepSeek Harness (DSH): registers Agent Plan, Coding Plan, and postpaid routes on the native pi-ai adapter so Ark models appear in the model picker.

11224 days agoModels & ProvidersApache-2.0
S

dsh-AuthInOne

stormycry-cryp/dsh-authinone

Adds account login, API and custom Provider setup, model switching, image fallback for text-only models, and token/cost attribution to DeepSeek Harness 47f.

104last monthModels & ProvidersMIT
S

dsh-design-qa

sunxin-ai/dsh-design-qa

Design-fidelity QA for text-only models: a `deepseek_vision` tool borrows an eye from any OpenAI-compatible vision route, so the model can judge whether an implementation matches its mock — shipped with the benchmark behind that judgement (four fixtures, 23 injected defects, raw transcripts) and the questioning discipline it depends on.

4414 days agoVision, Voice & MultimodalMIT
J

picturereader

jing-hy/picturereader

Image "reading" for text-only models: downscale + reduce color depth + structure/color fingerprints into text grids fed back to the conversation, letting the model zoom, sample and OCR autonomously like a multimodal model; fully local with zero external model dependency, ships an image-reading methodology skill and optional PaddleOCR.

362 days agoVision, Voice & MultimodalMIT
O

dsh-vision

oil-oil/dsh-vision

Near-native image understanding for DeepSeek Harness

24last monthVision, Voice & MultimodalMIT
S

dsh-ros2

stvli/dsh-ros2

ROS2 debugging toolset and robot-state vision analysis for DeepSeek Harness: node/topic/service/action/interface/TF enumeration, whole-graph topology JSON, rosdep checks, approval-gated builds and custom message scaffolding, GUI screenshots and multimodal vision observation, plus headless RViz2 offscreen rendering (low-poly meshes, direct pixel read, GPU passthrough - motion rendering at 30Hz) with parallel VLM realtime analysis.

214 days agoTools & CapabilitiesMIT
J

dsh-visual-plugin

jyh20030112/dsh-visual-plugin

Gives text-only models vision: forwards user images to an OpenAI-compatible vision model and shows the descriptions in a Web UI right panel.

1619 days agoVision, Voice & MultimodalMIT
F

dsh-vision-proxy

flyvhidbwo/dsh-vision-proxy

DeepSeek brain + automatic image transcription: attach images in the GUI and each one is transcribed via the official deepseek-v4-flash-vision-exp by default (a pure-text V4-Pro brain can see images), with any OpenAI-compatible VLM or local Ollama as alternatives.

1426 days agoVision, Voice & MultimodalMIT
W

dsh-mmx-bridge

welsione/dsh-mmx-bridge

MiniMax multimodal bridge: one `mmx_bridge` tool covers image understanding/generation, video, TTS, music, cover, web search and quota; optional `web_search`/`read_image` takeover; inline players/image previews right in the Web GUI (npm: `dsh-mmx-bridge`).

103 days agoTools & CapabilitiesMIT
S

dsh-deepseek-vision

siegfly/dsh-deepseek-vision

A vision-language gateway provider route: pasted images are described by a configurable VL model (Qwen-VL by default) before the DeepSeek wire.

910 days agoVision, Voice & MultimodalMIT
P

dsh-plugin-xiaomi-mimo-tts

ppy-web/dsh-plugin-xiaomi-mimo-tts

Adds Xiaomi MiMo text-to-speech controls for finalized assistant messages, with preset voices and custom voice design.

720 hours agoVision, Voice & MultimodalMIT
5

dsh-vision

54xkeee/dsh-vision

Vision for text-only DeepSeek via Doubao Web by default (zero-cost, no API key — drives your logged-in Chrome through a Windows CDP bridge), with Antigravity IDE quota (flash/pro) or Gemini fallback; auto detail escalation, vision evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.

7last monthVision, Voice & MultimodalMIT
Y

deepseek-harness-plugins (vision-bridge)

yinxe/deepseek-harness-plugins

Let text-only models see images: when a picture arrives with a placeholder like \[image omitted because this model accepts text only], the model calls the vision_describe tool and the plugin forwards the image reference plus the question to a multimodal model, retrying with a fallback model on failure. Configured in the settings page and persisted via the official settings API.

518 hours agoVision, Voice & MultimodalMIT
S

dsh-vision-bridge

sfyyy/dsh-vision-bridge

On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

5last monthVision, Voice & MultimodalMIT
A

dsh-guide-dog

atropinoltt/dsh-guide-dog

MiniMax-powered multimodal plugin: real-time voice call mode (streaming conversation, floating dock UI), voice mode and mic voice input, plus image/video/music/speech generation and vision inspection tools.

5last monthVision, Voice & MultimodalMIT
Y

dsh-opencode-free-models (dsh-opencode-free-models)

yu-wenchao/dsh-opencode-free-models

Free model provider plugin for DSH with 20+ models including Gemini, MiMo, GLM, etc.

418 days agoModels & Providers