Pular para o conteúdo principal

Plugins

Navegue, filtre e instale plugins do DeepSeek-Harness.

47 plugins encontrados

L

modlens

liustack/modlens

Ponte de visão para modelos somente-texto: cole uma imagem e receba evidências estruturadas em JSON (OCR, layout, semântica).

3.1khá 3 horasVisão, voz e multimodalMIT
Y

dsh-vision-router

ysr666/dsh-vision-router

Visão gratuita para agentes somente-texto: cadeia de visão integrada sem chave, além de ferramentas de pixel (perguntas e respostas, grounding, recorte, diff de pixels, cores, OCR, rastreamento SVG, recorte de objetos, capturas de tela); cole uma imagem para usar.

740há 3 horasVisão, voz e multimodalMIT
Z

dsh-crew

zseven-w/dsh-crew

DeepSeek Harness (DSH) plugin: dispatch work to DSH agents from Claude Code / Codex — native subagent progress, in-host worker sessions with per-tier presets, and a multimodal bridge that lends the text-only harness vision and image generation.

59ontemMCP e conectoresMIT
O

dsh-vision

oil-oil/dsh-vision

Compreensão de imagens quase nativa para o DeepSeek Harness

24há 5 diasFerramentas e funçõesMIT
S

dsh-authinone

stormycry-cryp/dsh-authinone

Self-contained DeepSeek Harness (DSH) plugin for Provider/Auth login, model switching, image fallback, token/cost analytics, and same-port Web restart. Useful? A star helps.

20há 3 diasModelos e provedoresMIT
J

picturereader

jing-hy/picturereader

Image "reading" for text-only models: downscale + reduce color depth + structure/color fingerprints into text grids fed back to the conversation, letting the model zoom, sample and OCR autonomously like a multimodal model; fully local with zero external model dependency, ships an image-reading methodology skill and optional PaddleOCR.

20há 8 horasVisão, voz e multimodalMIT
F

dsh-vision-proxy

flyvhidbwo/dsh-vision-proxy

Cérebro DeepSeek + transcrição automática de imagens: anexe imagens na GUI e cada uma é transcrita em texto por qualquer VLM compatível com OpenAI antes de chegar ao DeepSeek, que é somente texto — um caminho rápido com chave própria (padrão qwen3.7-flash; DashScope/Zhipu/OpenRouter ou qualquer endpoint compatível com OpenAI), ou Ollama local detectado automaticamente sem configuração.

11anteontemVisão, voz e multimodalMIT
J

dsh-visual-plugin

jyh20030112/dsh-visual-plugin

Dsh-visual-plugin. Dê olhos ao seu modelo somente-texto: encaminhe imagens do usuário para qualquer modelo de visão compatível com OpenAI e veja os resultados em um painel à direita da Web UI

9há 13 horasVisão, voz e multimodalMIT
5

dsh-vision

54xkeee/dsh-vision

Vision for text-only DeepSeek via Doubao Web by default (zero-cost, no API key — drives your logged-in Chrome through a Windows CDP bridge), with Antigravity IDE quota (flash/pro) or Gemini fallback; auto detail escalation, vision evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.

7anteontemVisão, voz e multimodalMIT
S

dsh-deepseek-vision

siegfly/dsh-deepseek-vision

A vision-language gateway provider route: pasted images are described by a configurable VL model (Qwen-VL by default) before the DeepSeek wire.

6ontemVisão, voz e multimodalMIT
S

dsh-vision-bridge

sfyyy/dsh-vision-bridge

On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

5anteontemSessões e mensagensMIT
W

dsh-mmx-bridge

welsione/dsh-mmx-bridge

MiniMax multimodal bridge: one `mmx_bridge` tool covers image understanding/generation, video, TTS, music, cover, web search and quota; optional `web_search`/`read_image` takeover; inline players/image previews right in the Web GUI (npm: `dsh-mmx-bridge`).

5há 12 horasFerramentas e funçõesMIT
G

deepseek-vision

gou-gee/deepseek-vision

DeepSeek Harness 原生视觉 Bundle:粘贴或拖入图片,通过托管的 deepseek-vision-mcp 调用 OpenAI 兼容视觉模型。

4anteontemMCP e conectoresMIT
G

deepseek-vision (dsh-plugin-deepseek-vision)

gou-gee/deepseek-vision

Vision MCP and DSH bundle for text-only DeepSeek: analyze_image, analyze_clipboard, compare_images and vision_status tools, a visual settings page, free GLM-4.6V-Flash by default, result caching and rate-limit tolerance; keys stay out of logs.

4anteontemVisão, voz e multimodalMIT
A

dsh-guide-dog

atropinoltt/dsh-guide-dog

MiniMax-powered multimodal plugin: real-time voice call mode (streaming conversation, floating dock UI), voice mode and mic voice input, plus image/video/music/speech generation and vision inspection tools.

4ontemVisão, voz e multimodalMIT
F

dsh-plugin-deepeye

favio8/dsh-plugin-deepeye

Plugin de visão DeepEye para o DeepSeek Harness (DSH): descrição de imagens, OCR, VQA, layout de UI e análise da área de transferência.

4há 5 diasMelhorias de UI
H

dsh-her-eyes

huashenglian/dsh-her-eyes

Um plugin dsh que permite à IA invocar automaticamente VLMs (modelos multimodais) para análise visual.

4há 5 diasVisão, voz e multimodalMIT
S

dsh-plugin-multimodal

shinjiyu/dsh-plugin-multimodal

Advertise image paste on text-only DeepSeek routes, describe attachments with a vision sidecar, and leave native vision models untouched.

3anteontemVisão, voz e multimodalMIT
W

visual-review

wang-bool/visual-review

Renders pasted/uploaded images inline in the DSH Web chat and gives text-only models vision: the model-invokable visual_review tool calls any OpenAI-compatible multimodal API first, falling back to a local Qwen3-VL worker.

2há 15 horasVisão, voz e multimodalMIT
H

dsh-open-eyes

hyp6666/dsh-open-eyes

Vision bridge for text-only DeepSeek routes that analyzes attached and local images through configurable OpenAI Responses, Chat Completions, or Anthropic Messages endpoints while leaving image-capable routes native.

2anteontemVisão, voz e multimodalMIT
H

dsh-vision-mix

haiziyao/dsh-vision-mix

Combine text, vision, and image-generation APIs into one Mix model with automatic routing: text-only requests go to the chat model, user images and agent screenshots go to the vision model, follow-ups keep using the same session image, and agents can generate or edit images with session-scoped call history.

2anteontemVisão, voz e multimodalMIT
Y

dsh-multimodal

yauntyour/dsh-multimodal

Per-file-type multimodal chains: preset processing models per wildcard convert image/video/audio files into prompt tokens before they reach the text-only session model, with per-preset fallback chains and a Multimodal settings page.

1há 4 diasSessões e mensagensMIT
I

dsh-tool-visual-primitives

inkshadewoods/dsh-tool-visual-primitives

DeepSeek Harness 视觉增强插件:将图片交给外部视觉模型分析,输出带坐标化视觉原语的纯文本证据,使不支持多模态的文本模型也能在对话中理解图片、截图与文档。

1há 3 diasFerramentas e funçõesMIT
S

dsh-design-qa

sunxin-ai/dsh-design-qa

让 DeepSeek Harness 里任何纯文本模型都能读图。识图是按需调用的 tool —— 图片不进主模型上下文,不看就不花钱;附 23 处缺陷的评测集,换模型可自测。

1há 8 horasFerramentas e funçõesMIT