Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

83 plugins found

L

modlens

liustack/modlens

Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).

4.1k2 days agoVision, Voice & MultimodalMIT
Y

dsh-vision-router

ysr666/dsh-vision-router

Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.

1.1k5 hours agoVision, Voice & MultimodalMIT
A

dsh-vision-toolkit

anionex/dsh-vision-toolkit

Vision for text-only models: paste an image and the model switches to a Vision Toolkit variant for image Q&A, multi-image comparison, long-screenshot OCR, screenshot-to-UI reproduction, element grounding, and pixel diff. No API key by default — images are processed by the author-hosted free service, 100 per machine per day; configurable to your own provider.

8844 days agoVision, Voice & MultimodalMIT
E

Invoice-Downloader (dsh-invoice-downloader)

ethanyoq/invoice-downloader

Local IMAP invoice download, OCR, archive, and Excel reimbursement summaries for DeepSeek Harness.

4735 days agoTools & CapabilitiesApache-2.0
O

watch-skill

oxbshw/watch-skill

DeepWatch's capabilities, installable into an existing DeepSeek Harness profile

37815 days agoVision, Voice & MultimodalMIT
Z

dsh-ios

zseven-w/dsh-ios

A live iOS Simulator or USB-connected iPhone inside the conversation: 22 agent tools for booting, building, driving the UI by accessibility identity or OCR text, list-row actions and SwiftUI preview hot reload, plus a streaming sidebar panel you can tap and drag on.

3085 days agoTools & CapabilitiesMIT
Z

dsh-android

zseven-w/dsh-android

A live Android device inside the conversation — emulator or USB phone, driven entirely through adb: 20 agent tools for streaming, Gradle build and run, UI-tree or OCR interaction, logcat, processes and memory, plus a three-button navigation panel.

1645 days agoTools & CapabilitiesMIT
E

invoice-downloader

ethanyoq/invoice-downloader

Local IMAP invoice download, OCR, archive, and Excel summary bundle for DeepSeek Harness

132last monthTools & CapabilitiesApache-2.0
W

dsh-openbiliclaw

whiteguo233/dsh-openbiliclaw

OpenBiliClaw consumer panel for DSH: recommendations, saved lists, Socratic dialogue, profile views, and Agent Bridge tools.

5919 days agoUI EnhancementsBSD-3-Clause
J

picturereader

jing-hy/picturereader

Image "reading" for text-only models: downscale + reduce color depth + structure/color fingerprints into text grids fed back to the conversation, letting the model zoom, sample and OCR autonomously like a multimodal model; fully local with zero external model dependency, ships an image-reading methodology skill and optional PaddleOCR.

3711 days agoVision, Voice & MultimodalMIT
Y

dsh-computer-use-win

yu-tao-li/dsh-computer-use-win

Windows computer use for DeepSeek Harness: an MCP stdio server over a PowerShell UIA backend exposing 22 desktop tools (UIA tree, screenshots, typed input, OCR, window management, failsafe).

182 days agoTools & CapabilitiesMIT
F

dsh-vision-proxy

flyvhidbwo/dsh-vision-proxy

DeepSeek brain + automatic image transcription: attach images in the GUI and each one is transcribed via the official deepseek-v4-flash-vision-exp by default (a pure-text V4-Pro brain can see images), with any OpenAI-compatible VLM or local Ollama as alternatives.

14last monthVision, Voice & MultimodalMIT
Y

dsh-pdf-mineru

yurzi/dsh-pdf-mineru

Provider-independent DSH document parsing powered by MinerU, with native background jobs, immutable results, and safe request coalescing.

1118 hours agoTools & CapabilitiesMIT
S

dsh-docs

sqhao-o/dsh-docs

Local PDF, Office, image, and OCR document intelligence for DeepSeek Harness.

10last monthTools & CapabilitiesMIT
L

dsh-vision

linenxi-ctrl/dsh-vision

External vision plugin for DeepSeek Harness: whale-button config panel, image recognition with auto-reply, and agent screenshot/recognize tools.

9last monthVision, Voice & MultimodalMIT
F

dsh-free-vision

fuzzysoul/dsh-free-vision

Free vision bridge for text-only models: image understanding, OCR, UI and debug analysis via free-tier providers (Qwen3-VL-Flash, Doubao, DeepSeek-OCR) with a settings GUI.

8last monthVision, Voice & MultimodalMIT
M

dsh-iris

mokuyoaxis/dsh-iris

Media and vision workspace for DeepSeek Harness: image, video and speech generation, image Q&A and element locating, long-image OCR, pixel diff, HTML-screenshot verification and video summarization, with DashScope and OpenAI-compatible providers, model pools and a workbench client.

79 days agoTools & CapabilitiesMIT
5

dsh-vision

54xkeee/dsh-vision

Vision for text-only DeepSeek via Doubao Web by default (zero-cost, no API key — drives your logged-in Chrome through a Windows CDP bridge), with Antigravity IDE quota (flash/pro) or Gemini fallback; auto detail escalation, vision evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.

7last monthVision, Voice & MultimodalMIT
C

dsh-learning-mode (plugin)

chplus0/dsh-learning-mode

Learning Mode agent preset: a coding agent that teaches while coding — concrete scenario-grounded explanations, Socratic guidance, and TODO(你) practice blanks, modeled on Claude Code's Learning output style; installable via dsh plugin add (dsh-learning-mode on npm).

6last monthTools & Capabilities
G

dsh-ocr-local

grelvan/dsh-ocr-local

Local OCR fallback for text-only routes: when the session model declares it cannot accept images, the attached image is cached locally and its path injected so the model can call ocr_image — PP-OCRv5 + ONNX Runtime on CPU, no API key, images never leave the machine. Silent when the model can see images.

52 days agoVision, Voice & Multimodal
N

vision-exp-tile

nicholas023/vision-exp-tile

Large-image recognition for vision-exp models: lossless 800×800 tile recognition (smart/pipeline/full), local OCR with preprocessing & handwriting routing, optional multi-vendor GPU (DirectML/CUDA/OpenVINO) with auto CPU fallback.

5last monthVision, Voice & MultimodalMIT
B

dsh-ocr-local

balcoz/dsh-ocr-local

Local OCR for DeepSeek Harness: paste/attach an image, get its text via PP-OCRv5 + ONNX Runtime, fully offline. TUI (cc-tui) and Web. / DeepSeek Harness 本地 OCR 插件:图片转文字,PP-OCRv5 + ONNX Runtime,完全离线,支持 TUI 与 Web。

5last monthVision, Voice & Multimodal
S

dsh-vision-bridge

sfyyy/dsh-vision-bridge

On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

5last monthVision, Voice & MultimodalMIT
E

canvas-workbench

elangan1997-cmyk/canvas-workbench

本地生图工作台,零订阅费:开自己的 API 生图(任意 OpenAI 兼容接口,零订阅),画布排版+修图/擦除/去背景/OCR/转矢量,可编辑 PSD/AI 交付,Photoshop/Illustrator 图层级双向桥接

48 hours agoVision, Voice & MultimodalMIT