Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

224 plugins found

Z

dsh-web-ui (dsh-tool-describe-image)

zhu1090093659/dsh-web-ui

A `describe_image` vision tool for text-only models: images (local path, URL, attachment) go to a configurable OpenAI-compatible vision endpoint and only the returned text enters the session.

8k4 days agoVision, Voice & MultimodalApache-2.0
L

modlens

liustack/modlens

Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).

4.1k2 days agoVision, Voice & MultimodalMIT
E

AI-Novel-Writer (dsh-ai-novel-writer)

ethanyoq/ai-novel-writer

Installs a dedicated AI novel-writing preset and workbench: revisioned local project assets, a compact side drawer, and native approval-gated single-file changes.

1.2k10 hours agoWorkflow & AutomationGPL-3.0
Y

dsh-vision-router

ysr666/dsh-vision-router

Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.

1.1k9 hours agoVision, Voice & MultimodalMIT
A

dsh-vision-toolkit

anionex/dsh-vision-toolkit

Vision for text-only models: paste an image and the model switches to a Vision Toolkit variant for image Q&A, multi-image comparison, long-screenshot OCR, screenshot-to-UI reproduction, element grounding, and pixel diff. No API key by default — images are processed by the author-hosted free service, 100 per machine per day; configurable to your own provider.

8844 days agoVision, Voice & MultimodalMIT
O

dsh-browser (bridge-browser)

omdsh-dev/dsh-browser

Chrome sidebar extension that lets DSH operate your browser directly, no vision capabilities required.

74616 hours agoTools & CapabilitiesMIT
L

dsh-browser (bridge-browser)

lum1104/dsh-browser

Chrome sidebar extension that lets DSH operate your browser directly, no vision capabilities required.

7188 days agoTools & CapabilitiesMIT
O

watch-skill

oxbshw/watch-skill

DeepWatch's capabilities, installable into an existing DeepSeek Harness profile

37815 days agoVision, Voice & MultimodalMIT
L

dsh-browser

lum1104/dsh-browser

Chrome sidebar extension that lets DSH operate your browser directly, no vision capabilities required.

295last monthTools & CapabilitiesMIT
Z

dsh-crew

zseven-w/dsh-crew

Dispatch work to DSH agents from Claude Code or Codex: native subagent progress, in-host worker sessions with per-tier presets, and a multimodal bridge for vision and image generation.

1535 days agoIntegrations & RemoteMIT
S

dsh-AuthInOne

stormycry-cryp/dsh-authinone

Adds account login, API and custom Provider setup, model switching, image fallback for text-only models, and token/cost attribution to DeepSeek Harness 47f.

1042 months agoModels & ProvidersMIT
L

sealos-skills

labring/sealos-skills

AI agent skills for Sealos — deploy any project, provision databases, object storage & more with one command. Works with Claude Code, Gemini CLI, Codex.

702 months agoSkills
L

gongwen-skill

linhut/gongwen-skill

Chinese government document processing toolkit: GB/T 9704 format check, auto-fix, content revision (red annotation + strikethrough), template generation, Markdown-to-docx, and document header/footer/page-number injection for 24 official document types.

694 days agoSkillsMIT
S

dsh-design-qa

sunxin-ai/dsh-design-qa

Design-fidelity QA for text-only models: a `deepseek_vision` tool borrows an eye from any OpenAI-compatible vision route, so the model can judge whether an implementation matches its mock — shipped with the benchmark behind that judgement (four fixtures, 23 injected defects, raw transcripts) and the questioning discipline it depends on.

4423 days agoVision, Voice & MultimodalMIT
J

picturereader

jing-hy/picturereader

Image "reading" for text-only models: downscale + reduce color depth + structure/color fingerprints into text grids fed back to the conversation, letting the model zoom, sample and OCR autonomously like a multimodal model; fully local with zero external model dependency, ships an image-reading methodology skill and optional PaddleOCR.

3711 days agoVision, Voice & MultimodalMIT
O

dsh-vision

oil-oil/dsh-vision

Near-native image understanding for DeepSeek Harness

242 months agoVision, Voice & MultimodalMIT
M

dsh-media-skills

mjorgin/dsh-media-skills

Free vision bridge and image generation for text-only models: paste-image reading, GLM-4V-Flash and Gemini engine failover, ModLens-style structured evidence, and a seeded free vision model route.

199 days agoVision, Voice & MultimodalMIT
D

dsh-web-ui (dsh-tool-describe-image)

damonkoy/dsh-web-ui

Gives a text-only model image understanding via a vision-language model, exposed as a `describe_image` tool.

192 months agoVision, Voice & MultimodalApache-2.0
J

dsh-visual-plugin

jyh20030112/dsh-visual-plugin

Gives text-only models vision: forwards user images to an OpenAI-compatible vision model and shows the descriptions in a Web UI right panel.

1628 days agoVision, Voice & MultimodalMIT
F

dsh-vision-proxy

flyvhidbwo/dsh-vision-proxy

DeepSeek brain + automatic image transcription: attach images in the GUI and each one is transcribed via the official deepseek-v4-flash-vision-exp by default (a pure-text V4-Pro brain can see images), with any OpenAI-compatible VLM or local Ollama as alternatives.

14last monthVision, Voice & MultimodalMIT
H

dsh-novel-forge

huangziyuan-general/dsh-novel-forge

Novel-writing guardrails: 20 novel_* tools enforce fact-ledger consistency, phase and chapter gates, machine audit with deterministic metrics, and proposal-only revisions; a right-panel workbench covers reading, TTS playback, polish, proofread, continuity checks and batch drafting, and a bypass engine runs internal steps off the main conversation.

134 days agoWorkflow & AutomationMIT
P

dsh-vision-opencode

poiuyjie/dsh-vision-opencode

Adds a configurable vision model to text-only main models: a vision_read_image tool, a composer-bar vision-model selector, and automatic image-to-text conversion for text-only routes.

1321 days agoVision, Voice & MultimodalMIT
D

dsh-auxiliary

dsh-plugins/dsh-auxiliary

Dedicated model routes, tools, and system guidance for vision, compaction, reviews, subagents, titles, and image generation.

1223 days agoTools & CapabilitiesLGPL-3.0
C

Gemini-Eyes

consolesun/gemini-eyes

MCP bridge to gemini.google.com: vision analysis of images and videos, Imagen image and Veo video generation, and conversation management using the logged-in browser session with no API key.

11last monthVision, Voice & Multimodal