Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

83 plugins found

H

dsh-vision-analysis

harvey-will/dsh-vision-analysis

DeepSeek Harness vision plugin: 8 analysis modes (describe, OCR, chart data, UI review, object detection, compare, code-gen, debug), any OpenAI- or Anthropic-compatible vision API, with a built-in free vision model and automatic rate-limit failover.

14 days agoVision, Voice & MultimodalMIT
H

dsh-open-file

hyper-dsh-plugins/dsh-open-file

Workspace-bound arbitrary file upload, reading, OCR, and rendering for DeepSeek Harness.

1last monthVision, Voice & MultimodalMIT
Z

dsh-plugins

zjcdkj/dsh-plugins

DeepSeek Harness out-of-tree plugin: give a text-only coding model eyes by routing images to a Qwen-VL (DashScope) route through ctx.llm and returning text.

1last monthVision, Voice & MultimodalMIT
H

dsh-open-file

hyp6666/dsh-open-file

Workspace-bound arbitrary file upload, reading, OCR, and rendering for DeepSeek Harness.

1last monthVision, Voice & MultimodalMIT
L

dsh-eyes

leeminjing/dsh-eyes

On-demand vision for text-only DeepSeek models: upload images, and the model calls a view_image tool backed by any OpenAI-compatible vision endpoint (Qwen/DashScope by default).

1last monthVision, Voice & MultimodalMIT
G

dsh-tool-vision

gloryxpnv/dsh-tool-vision

Local-first structured vision for text-only agents: images go to a local OpenAI-compatible VLM and come back as JSON evidence (summary, verbatim OCR, layout regions, entities/relations, colors, explicit uncertainty), with anti-hallucination fallback and an optional paste/upload bridge; zero cloud cost, images never leave the machine.

12 months agoVision, Voice & MultimodalMIT
S

macos-computer-use-kit (dsh)

sur-cai/macos-computer-use-kit

AX-first macOS computer use for DeepSeek Harness: accessibility snapshots with stable refs, background input, Unicode typing, menus and apps, set-of-mark screenshots, on-device OCR and verified actions — 11 tools bridged to the macos-cu CLI.

0Tools & CapabilitiesPossibly stale
M

lookover (dsh-look)

mengxiaoxian/lookover

Scene-awareness probe for DSH on macOS: a privacy-first, pull-model `look` tool that reads the frontmost non-self window (app info, AX title/selection, gated local OCR), plus a summon-hotkey snapshot captured the instant you press.

08 days agoVision, Voice & MultimodalMIT
M

lookover (dsh-expmem)

mengxiaoxian/lookover

Personal-experience memory for DSH: turns solved tasks into sourced problem cases with condition-scoped claims, recalls them into new tasks via hidden pre-step injection, and makes corrections/deletes cascade to future recall.

08 days agoMemoryMIT
N

vision-exp-tile

nicholaskin/vision-exp-tile

Large-image recognition for vision-exp models: lossless 800×800 tile recognition (smart/pipeline/full), local OCR with preprocessing & handwriting routing, optional multi-vendor GPU (DirectML/CUDA/OpenVINO) with auto CPU fallback.

019 days agoVision, Voice & MultimodalMIT
J

dsh-markitdown

jiekesu967/dsh-markitdown

Adds a markitdown tool that converts PDF, Word, Excel, PowerPoint, HTML, CSV, EPUB and URLs to Markdown, using a real Microsoft MarkItDown installation when one is reachable and a bundled dependency-free converter otherwise.

017 days agoTools & CapabilitiesMIT
L

dsh-file-upload-local

lubenweimeiyoukaig/dsh-file-upload-local

Local file upload: paperclip button and drag-and-drop, per-session storage under .dsh-uploads, and a read_document tool that pages text and OCRs images.

025 days agoTools & CapabilitiesMIT
H

dsh-maclens

harzva/dsh-maclens

Apple on-device Vision tools for text-only dsh models: local OCR (zh-Hans + 30 langs), image classification, face detection, document layout, and a combined describe — 100% offline, no API key, tall-screenshot slicing.

0last monthVision, Voice & MultimodalMIT
J

linkdigest-mcp

jcaiagent7143-ui/linkdigest-mcp

Turns a Xiaohongshu, Douyin, TikTok, YouTube or X link into text: transcript with timecodes, on-screen text, a description and OCR of every image, caption and metadata. Mounts the hosted MCP server over streamable HTTP.

024 days agoTools & CapabilitiesMIT
L

dsh-file-attach

lucasxingg/dsh-file-attach

Drag-and-drop PDF, Office, images, and text/code files into DSH conversations. The host extracts (and OCRs) them into the prompt; attach_* tools cover notebook cells, PDF-page OCR, image describe, and save.

0last monthVision, Voice & MultimodalMIT
T

DSH-FormatForge (dsh-formatforge)

tianbuyu-wwx/dsh-formatforge

FormatForge: forge 30+ file formats (PDF, DOCX, PPTX, XLSX, EML, EPUB, TOML, archives) into AI-readable text via a drag-and-drop inbox, ff_translate/ff_formats/ff_result tools and a CLI; scanned PDFs fall back to local OCR or an enhance hint for the session model.

0Tools & CapabilitiesPossibly stale
X

dsh-vision-hub (tool-vision)

xing666173/dsh-vision-hub

Enhanced vision toolbox: 14 pixel-level vision tools (describe, ground, detect, crop, pixel-diff, OCR, long-screenshot OCR, vectorize, colors, cutout, screenshot, present, materialize, html-screenshot) driven by one OpenAI-compatible endpoint, with clean \[图片: path] bridge markers, content-safety classification and rate-limit auto-retry.

0Vision, Voice & MultimodalPossibly stale
S

dsh-vision-link

sprainjinyu/dsh-vision-link

Lightweight, route-preserving vision link for DSH: a configured vision model sees while the selected text model stays in control.

0last monthVision, Voice & Multimodal
R

dsh-omni-vision

renji004/dsh-omni-vision

Local eyes for text-only models: eyes_render draws text/shapes/Mermaid onto a canvas in the Web GUI, eyes_paste captures pasted images, eyes_ocr reads text via the built-in Windows OCR (offline), and eyes_analyze inspects pixels as structured data — no vision model required.

0last monthVision, Voice & MultimodalMIT
V

dsh-ocr-bridge

vuvanmai936-dot/dsh-ocr-bridge

Paste images into DeepSeek Harness chat and have them read by a free local backend (macOS Vision / Tesseract) before the text-only DeepSeek model answers

0last monthVision, Voice & Multimodal
H

win11-oneocr

hawkhai/win11-oneocr

Local Windows 11 OneOCR tool for DSH: `oneocr_recognize` returns OCR text and structured line/word polygons, confidence, rotation, and handwriting style.

0last monthVision, Voice & Multimodal
H

wechat-ocr

hawkhai/wechat-ocr

Local WeChat OCR tool for DSH: `wechat_ocr_recognize` returns recognized text and the engine structured result for a local image path.

0last monthVision, Voice & Multimodal
Z

dsh-kirocrew

zoahdev/dsh-kirocrew

KiroCrew bridge for DeepSeek Harness: let your dsh agent delegate to a persistent, self-evolving KiroCrew workspace over ACP (JSON-RPC 2.0 over stdio).

0last monthIntegrations & RemoteMIT
T

dsh-vision-ocr

timeflies-qyh/dsh-vision-ocr

DeepSeek Harness OCR plugin — offline image text recognition powered by PaddleOCR-json (primary) and RapidOCR-json (fallback). 让 DeepSeek Harness 直接识别图片中的文字,无需视觉模型、完全本地离线运行。

0last monthVision, Voice & MultimodalMIT