Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

170 plugins found

M

dsh-windows-ocr

maxwell-feng/dsh-windows-ocr

Local OCR for attached images via the built-in Windows engine (Windows.Media.Ocr): only the recognized text is sent to the model, never the image bytes; vision passthrough is opt-in.

6yesterdayVision, Voice & MultimodalMIT
L

dsh-drop-to-path

loudmore/dsh-drop-to-path

DSH 插件:图片与文件直达纯文本模型——图片保留原生附件体验,PDF/Office/压缩包/视频/音频显示为附件栏方块,点击发送时自动转为工作区路径,配合 dsh-vision-toolkit 粘贴即看图。A DSH plugin that delivers images AND files to text-only models as workspace paths: images keep the native attachment UI, other files show as square chips in the rail, paths append on send — pairs with dsh-vision-toolkit.

64 days agoUI EnhancementsMIT
B

dsh-ocr-local

balcoz/dsh-ocr-local

Local OCR for DeepSeek Harness: paste/attach an image, get its text via PP-OCRv5 + ONNX Runtime, fully offline. TUI (cc-tui) and Web. / DeepSeek Harness 本地 OCR 插件:图片转文字,PP-OCRv5 + ONNX Runtime,完全离线,支持 TUI 与 Web。

515 hours agoUI Enhancements
S

dsh-vision-bridge

sfyyy/dsh-vision-bridge

On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

52 days agoSessions & MessagesMIT
W

dsh-mmx-bridge

welsione/dsh-mmx-bridge

MiniMax multimodal bridge: one `mmx_bridge` tool covers image understanding/generation, video, TTS, music, cover, web search and quota; optional `web_search`/`read_image` takeover; inline players/image previews right in the Web GUI (npm: `dsh-mmx-bridge`).

512 hours agoTools & CapabilitiesMIT
L

awesome-dsh-background-plugin

leavestring/awesome-dsh-background-plugin

DSH Web background settings: upload a local image or pick a preset ambiance, with live preview and persistent settings.

54 days agoThemes & AppearanceMIT
C

dsh-client-ui-skins

caoyiwei850/dsh-client-ui-skins

Skin plugin for the DSH Web UI: 4 built-in skins plus custom image skins with the photo as a full-interface background and the palette following its dominant hue.

55 hours agoThemes & AppearanceMIT
S

deepseek-harness-qqbot

sliverp/deepseek-harness-qqbot

QQ Bot text and image channel plugin for DeepSeek Harness

54 days agoTools & CapabilitiesMIT
A

dsh-desk-pet

anneheartrecord/dsh-desk-pet

Always-on-top DeepSeek Harness desk pet with a native menu and DIY skins. Shows what your agent is doing: working, waiting, finished, failed. Five skins, or make your own from one image. No dependencies.

413 hours agoThemes & AppearanceMIT
G

deepseek-vision

gou-gee/deepseek-vision

DeepSeek Harness 原生视觉 Bundle:粘贴或拖入图片,通过托管的 deepseek-vision-mcp 调用 OpenAI 兼容视觉模型。

42 days agoMCP & ConnectorsMIT
K

dsh-office-tools

kw78/dsh-office-tools

Workspace-safe Office tools for agents: create/read Word, create/read/update Excel, and create/read PowerPoint decks with PNG/JPG/GIF image placement.

43 days agoTools & CapabilitiesMIT
G

deepseek-vision (dsh-plugin-deepseek-vision)

gou-gee/deepseek-vision

Vision MCP and DSH bundle for text-only DeepSeek: analyze_image, analyze_clipboard, compare_images and vision_status tools, a visual settings page, free GLM-4.6V-Flash by default, result caching and rate-limit tolerance; keys stay out of logs.

42 days agoVision, Voice & MultimodalMIT
A

dsh-guide-dog

atropinoltt/dsh-guide-dog

MiniMax-powered multimodal plugin: real-time voice call mode (streaming conversation, floating dock UI), voice mode and mic voice input, plus image/video/music/speech generation and vision inspection tools.

4yesterdayVision, Voice & MultimodalMIT
D

dsh-auxiliary

dsh-plugins/dsh-auxiliary

Dedicated model routes, tools, and system guidance for vision, compaction, reviews, subagents, titles, and image generation.

43 days agoTools & CapabilitiesLGPL-3.0
W

dsh-plugin-codex

wss534857356/dsh-plugin-codex

Codex App Server model provider using a local Codex login, with session reuse, Harness tool bridging, native action traces, and durable generated-image projection.

43 days agoModels & ProvidersMIT
K

dsh-plugin-subhub

kinoward/dsh-plugin-subhub

Use third-party subscription accounts in DeepSeek Harness: chat, image understanding, image generation, and image editing with the models your subscription covers, with available models and reasoning levels synced from your account; currently supports OpenAI/ChatGPT subscriptions, more providers planned.

417 hours agoModels & ProvidersMIT
S

deepseek-harness-wecom

sliverp/deepseek-harness-wecom

WeCom AI Bot text and image bridge for DeepSeek Harness

44 days agoTools & CapabilitiesMIT
F

dsh-plugin-deepeye

favio8/dsh-plugin-deepeye

DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.

45 days agoUI Enhancements
Y

dsh-image-subagent

yuqingsh/dsh-image-subagent

44 days agoTools & CapabilitiesMIT
P

dsh-plugin

picgo/dsh-plugin

Upload local images and files to your image host through PicGo's existing configuration (PicGo Cloud, GitHub, S3, COS, Qiniu, or any installed uploader plugin), via a `picgo_upload` tool and a `/picgo` command.

42 days agoTools & Capabilities
M

widget-dock

morgogh/widget-dock

Draggable grid-aligned workbench of 28 mini-cards in the margins of the conversation (API balance, token usage, cost, context pressure, todos, goal, usage heatmap, GitHub repo, image relay), with S/M/L/XL size tiers.

43 days agoUI EnhancementsMIT
S

dsh-plugin-multimodal

shinjiyu/dsh-plugin-multimodal

Advertise image paste on text-only DeepSeek routes, describe attachments with a vision sidecar, and leave native vision models untouched.

32 days agoVision, Voice & MultimodalMIT
N

free-vision-skill

niyongsheng/free-vision-skill

Fully-local image understanding & OCR via macOS Vision Framework: `ocr_image` (text, table layout + coordinates) and `view_image` (scene, faces, QR) — paste multiple images into the web input box or pass path/URL/base64; images never leave your Mac.

33 days agoVision, Voice & MultimodalMIT
D

dsh-web-ui (dsh-tool-describe-image)

damonkoy/dsh-web-ui

Gives a text-only model image understanding via a vision-language model, exposed as a `describe_image` tool.

33 days agoVision, Voice & MultimodalApache-2.0