Plugins
Navegue, filtre e instale plugins do DeepSeek-Harness.
86 plugins encontrados
dsh-wsl-media
173787247/dsh-wsl-media
Local media/doc pipeline: ffprobe, extract, thumbnail, PDF, ASR, pandoc, OCR, exif (allowRoots include IM inbox).
Mimir (mimir-skin)
tommyhedgerow/mimir
Publishes a lesson into the DSH conversation: the dependency spine, the question and the vault's drawings, drawn in the vault's own palette.
computer-user-vision
xie129716/computer-user-vision
Windows computer use forked from computer-user: 13 computer_* tools that read the screen and drive the mouse and keyboard. Image-capable routes get the screenshot as a real image with an exact image-to-screen mapping, so no external OCR; elements return as UI Automation refs so a click lands on the exact control rectangle; Ctrl+Alt+Esc stops every call.
dsh-markitdown
jiekesu967/dsh-markitdown
Adds a markitdown tool that converts PDF, Word, Excel, PowerPoint, HTML, CSV, EPUB and URLs to Markdown, using a real Microsoft MarkItDown installation when one is reachable and a bundled dependency-free converter otherwise.
dsh-screenshot-capture
wangzhanchao883/dsh-screenshot-capture
Point-and-shoot screenshot capture: a system floating window turns a new screenshot (or copied image) into an Obsidian note with comments and key-point marking, instant Tongyi Qianwen OCR, per-day merging, and an optional evening AI organization pass.
dsh-file-convert
zzy-12345678/dsh-file-convert
Local-first file conversion: 26 conversions across images, PDF (with OCR and experimental PDF→DOCX), data, audio/video and office docs; 7 tools, all local, no API keys.
dsh-mmroute
jmxsxwyzjdwl/dsh-mmroute
Transparent multimodal routing for text-only models: every image in every model call is fully transcribed (verbatim OCR, data, uncertainty zones, injection-hardened) by your own multimodal understander, with focused re-look via vision_relook and automatic retry on image-related failures. No bundled endpoints, no borrowed logins.
aura-vision
ck-epsilon/aura-vision
Free vision OCR with adaptive tile recognition for long documents and Markdown/Word/PNG/Excel export.
dsh-vision-analysis
harvey-will/dsh-vision-analysis
DeepSeek Harness vision plugin: 8 analysis modes (describe, OCR, chart data, UI review, object detection, compare, code-gen, debug), any OpenAI- or Anthropic-compatible vision API, with a built-in free vision model and automatic rate-limit failover.
dsh-open-file
hyper-dsh-plugins/dsh-open-file
Upload, leitura, OCR e renderização de arquivos arbitrários vinculados ao workspace, para o DeepSeek Harness.
dsh-plugins
zjcdkj/dsh-plugins
Plugin out-of-tree para DeepSeek Harness: dá olhos a um modelo de codificação somente texto roteando imagens para uma rota Qwen-VL (DashScope) via ctx.llm e retornando texto.
dsh-open-file
hyp6666/dsh-open-file
Upload, leitura, OCR e renderização de arquivos arbitrários vinculados ao workspace para DeepSeek Harness.
dsh-eyes
leeminjing/dsh-eyes
Visão sob demanda para modelos DeepSeek somente texto: envie imagens e o modelo chama uma ferramenta view_image apoiada em qualquer endpoint de visão compatível com OpenAI (Qwen/DashScope por padrão).
dsh-tool-vision
gloryxpnv/dsh-tool-vision
Visão estruturada local-first para agentes somente texto: as imagens vão a um VLM local compatível com OpenAI e voltam como evidência JSON (resumo, OCR literal, regiões de layout, entidades/relações, cores, incerteza explícita), com fallback anti-alucinação e ponte opcional de colar/enviar; custo zero em nuvem, as imagens nunca saem da máquina.
lookover (dsh-look)
mengxiaoxian/lookover
Scene-awareness probe for DSH on macOS: a privacy-first, pull-model `look` tool that reads the frontmost non-self window (app info, AX title/selection, gated local OCR), plus a summon-hotkey snapshot captured the instant you press.
lookover (dsh-expmem)
mengxiaoxian/lookover
Personal-experience memory for DSH: turns solved tasks into sourced problem cases with condition-scoped claims, recalls them into new tasks via hidden pre-step injection, and makes corrections/deletes cascade to future recall.
vision-exp-tile
nicholaskin/vision-exp-tile
Large-image recognition for vision-exp models: lossless 800×800 tile recognition (smart/pipeline/full), local OCR with preprocessing & handwriting routing, optional multi-vendor GPU (DirectML/CUDA/OpenVINO) with auto CPU fallback.
dsh-file-upload-local
lubenweimeiyoukaig/dsh-file-upload-local
Local file upload: paperclip button and drag-and-drop, per-session storage under .dsh-uploads, and a read_document tool that pages text and OCRs images.
dsh-maclens
harzva/dsh-maclens
Apple on-device Vision tools for text-only dsh models: local OCR (zh-Hans + 30 langs), image classification, face detection, document layout, and a combined describe — 100% offline, no API key, tall-screenshot slicing.
linkdigest-mcp
jcaiagent7143-ui/linkdigest-mcp
Turns a Xiaohongshu, Douyin, TikTok, YouTube or X link into text: transcript with timecodes, on-screen text, a description and OCR of every image, caption and metadata. Mounts the hosted MCP server over streamable HTTP.
dsh-file-attach
lucasxingg/dsh-file-attach
Drag-and-drop PDF, Office, images, and text/code files into DSH conversations. The host extracts (and OCRs) them into the prompt; attach_* tools cover notebook cells, PDF-page OCR, image describe, and save.
dsh-vision-link
sprainjinyu/dsh-vision-link
Lightweight, route-preserving vision link for DSH: a configured vision model sees while the selected text model stays in control.
dsh-omni-vision
renji004/dsh-omni-vision
Local eyes for text-only models: eyes_render draws text/shapes/Mermaid onto a canvas in the Web GUI, eyes_paste captures pasted images, eyes_ocr reads text via the built-in Windows OCR (offline), and eyes_analyze inspects pixels as structured data — no vision model required.
dsh-ocr-bridge
vuvanmai936-dot/dsh-ocr-bridge
Paste images into DeepSeek Harness chat and have them read by a free local backend (macOS Vision / Tesseract) before the text-only DeepSeek model answers