본문으로 건너뛰기

플러그인

DeepSeek-Harness 플러그인을 찾아보고, 필터링하고, 설치하세요.

플러그인 86개 찾음

1

dsh-wsl-media

173787247/dsh-wsl-media

Local media/doc pipeline: ffprobe, extract, thumbnail, PDF, ASR, pandoc, OCR, exif (allowRoots include IM inbox).

13일 전개발 및 플러그인 도구MIT
T

Mimir (mimir-skin)

tommyhedgerow/mimir

Publishes a lesson into the DSH conversation: the dependency spine, the question and the vault's drawings, drawn in the vault's own palette.

16일 전UI 확장MIT
X

computer-user-vision

xie129716/computer-user-vision

Windows computer use forked from computer-user: 13 computer_* tools that read the screen and drive the mouse and keyboard. Image-capable routes get the screenshot as a real image with an exact image-to-screen mapping, so no external OCR; elements return as UI Automation refs so a click lands on the exact control rectangle; Ctrl+Alt+Esc stops every call.

118일 전도구 및 기능MIT
J

dsh-markitdown

jiekesu967/dsh-markitdown

Adds a markitdown tool that converts PDF, Word, Excel, PowerPoint, HTML, CSV, EPUB and URLs to Markdown, using a real Microsoft MarkItDown installation when one is reachable and a bundled dependency-free converter otherwise.

13일 전도구 및 기능MIT
W

dsh-screenshot-capture

wangzhanchao883/dsh-screenshot-capture

Point-and-shoot screenshot capture: a system floating window turns a new screenshot (or copied image) into an Obsidian note with comments and key-point marking, instant Tongyi Qianwen OCR, per-day merging, and an optional evening AI organization pass.

14일 전도구 및 기능MIT
Z

dsh-file-convert

zzy-12345678/dsh-file-convert

Local-first file conversion: 26 conversions across images, PDF (with OCR and experimental PDF→DOCX), data, audio/video and office docs; 7 tools, all local, no API keys.

1지난달비전, 음성 및 멀티모달
J

dsh-mmroute

jmxsxwyzjdwl/dsh-mmroute

Transparent multimodal routing for text-only models: every image in every model call is fully transcribed (verbatim OCR, data, uncertainty zones, injection-hardened) by your own multimodal understander, with focused re-look via vision_relook and automatic retry on image-related failures. No bundled endpoints, no borrowed logins.

126일 전비전, 음성 및 멀티모달MIT
C

aura-vision

ck-epsilon/aura-vision

Free vision OCR with adaptive tile recognition for long documents and Markdown/Word/PNG/Excel export.

1지난달비전, 음성 및 멀티모달MIT
H

dsh-vision-analysis

harvey-will/dsh-vision-analysis

DeepSeek Harness vision plugin: 8 analysis modes (describe, OCR, chart data, UI review, object detection, compare, code-gen, debug), any OpenAI- or Anthropic-compatible vision API, with a built-in free vision model and automatic rate-limit failover.

13일 전비전, 음성 및 멀티모달MIT
H

dsh-open-file

hyper-dsh-plugins/dsh-open-file

DeepSeek Harness를 위한 워크스페이스에 한정된 임의 파일 업로드, 읽기, OCR 및 렌더링.

1지난달비전, 음성 및 멀티모달MIT
Z

dsh-plugins

zjcdkj/dsh-plugins

DeepSeek Harness 비트리 플러그인: 이미지를 Qwen-VL(DashScope) 경로로 ctx.llm을 통해 라우팅하고 텍스트를 반환하여 텍스트 전용 코딩 모델에 시각을 부여

1지난달비전, 음성 및 멀티모달MIT
H

dsh-open-file

hyp6666/dsh-open-file

DeepSeek Harness용 워크스페이스 바인딩 임의 파일 업로드, 읽기, OCR 및 렌더링

12개월 전비전, 음성 및 멀티모달MIT
L

dsh-eyes

leeminjing/dsh-eyes

텍스트 전용 DeepSeek 모델용 온디맨드 비전: 이미지를 업로드하면 모델이 임의의 OpenAI 호환 비전 엔드포인트(기본 Qwen/DashScope)를 사용하는 view_image 도구를 호출합니다.

12개월 전비전, 음성 및 멀티모달MIT
G

dsh-tool-vision

gloryxpnv/dsh-tool-vision

텍스트 전용 에이전트를 위한 로컬 우선 구조적 비전: 이미지를 로컬 OpenAI 호환 VLM에 보내고 JSON 증거(요약, 그대로의 OCR, 레이아웃 영역, 엔티티/관계, 색상, 명시적 불확실성)로 돌려받습니다. 환각 방지 폴백과 선택적 붙여넣기/업로드 브리지 포함. 클라우드 비용 없음, 이미지는 기기 밖으로 나가지 않습니다.

12개월 전비전, 음성 및 멀티모달MIT
M

lookover (dsh-look)

mengxiaoxian/lookover

Scene-awareness probe for DSH on macOS: a privacy-first, pull-model `look` tool that reads the frontmost non-self window (app info, AX title/selection, gated local OCR), plus a summon-hotkey snapshot captured the instant you press.

011일 전비전, 음성 및 멀티모달MIT
M

lookover (dsh-expmem)

mengxiaoxian/lookover

Personal-experience memory for DSH: turns solved tasks into sourced problem cases with condition-scoped claims, recalls them into new tasks via hidden pre-step injection, and makes corrections/deletes cascade to future recall.

011일 전메모리MIT
N

vision-exp-tile

nicholaskin/vision-exp-tile

Large-image recognition for vision-exp models: lossless 800×800 tile recognition (smart/pipeline/full), local OCR with preprocessing & handwriting routing, optional multi-vendor GPU (DirectML/CUDA/OpenVINO) with auto CPU fallback.

022일 전비전, 음성 및 멀티모달MIT
L

dsh-file-upload-local

lubenweimeiyoukaig/dsh-file-upload-local

Local file upload: paperclip button and drag-and-drop, per-session storage under .dsh-uploads, and a read_document tool that pages text and OCRs images.

028일 전도구 및 기능MIT
H

dsh-maclens

harzva/dsh-maclens

Apple on-device Vision tools for text-only dsh models: local OCR (zh-Hans + 30 langs), image classification, face detection, document layout, and a combined describe — 100% offline, no API key, tall-screenshot slicing.

0지난달비전, 음성 및 멀티모달MIT
J

linkdigest-mcp

jcaiagent7143-ui/linkdigest-mcp

Turns a Xiaohongshu, Douyin, TikTok, YouTube or X link into text: transcript with timecodes, on-screen text, a description and OCR of every image, caption and metadata. Mounts the hosted MCP server over streamable HTTP.

026일 전도구 및 기능MIT
L

dsh-file-attach

lucasxingg/dsh-file-attach

Drag-and-drop PDF, Office, images, and text/code files into DSH conversations. The host extracts (and OCRs) them into the prompt; attach_* tools cover notebook cells, PDF-page OCR, image describe, and save.

0지난달비전, 음성 및 멀티모달MIT
S

dsh-vision-link

sprainjinyu/dsh-vision-link

Lightweight, route-preserving vision link for DSH: a configured vision model sees while the selected text model stays in control.

0지난달비전, 음성 및 멀티모달
R

dsh-omni-vision

renji004/dsh-omni-vision

Local eyes for text-only models: eyes_render draws text/shapes/Mermaid onto a canvas in the Web GUI, eyes_paste captures pasted images, eyes_ocr reads text via the built-in Windows OCR (offline), and eyes_analyze inspects pixels as structured data — no vision model required.

0지난달비전, 음성 및 멀티모달MIT
V

dsh-ocr-bridge

vuvanmai936-dot/dsh-ocr-bridge

Paste images into DeepSeek Harness chat and have them read by a free local backend (macOS Vision / Tesseract) before the text-only DeepSeek model answers

0지난달비전, 음성 및 멀티모달