Saltar al contenido principal

Plugins

Explora, filtra e instala plugins de DeepSeek-Harness.

123 plugins encontrados

W

mimo-vision

wulusai2333/mimo-vision

`describe_image` tool: a vision bridge that sends images to mimo-v2.5 through the opencode Zen API (credential `OPENCODE_GO_API_KEY`, free route first with paid fallback) and returns text descriptions for text-only models, with native passthrough and ImageMagick transcoding of SVG/TIFF/HEIC formats.

1anteayerVisión, voz y multimodalMIT
W

dsh-tool-vision

wanshichenguang/dsh-tool-vision

Adds an image_describe tool backed by the DashScope OpenAI-compatible vision API; on text-only sessions, pasted images are stored as local paths for the model and rendered inline in the chat transcript.

1anteayerVisión, voz y multimodalMIT
T

dsh-plugins (dsh-vision)

tzhr-invest/dsh-plugins

Agent-callable vision tool that describes local images via any OpenAI-compatible vision endpoint you configure, with an optional multi-model cross-check and no built-in keys.

1ayerVisión, voz y multimodalMIT
P

dsh-screenshot

paicat1/dsh-screenshot

Standalone screen capture for DeepSeek Harness (dsh). Browser hotkeys for instant capture plus an agent-facing capture and read tool that lets the agent see and analyze any screen region.

1ayerVisión, voz y multimodalMIT
N

dsh-auto-vision

normanfxxkingrockwell/dsh-auto-vision

Auto-discovery vision bridge for text-only DeepSeek Harness agents: automatically finds an image-capable model from your configured providers and returns picture descriptions as plain text via a vision tool.

1hace 14 horasVisión, voz y multimodalMIT
M

dsh-koboldcpp-hands

microherox/dsh-koboldcpp-hands

Hands repetitive text and vision labor (OCR, image analysis, comparison) to a local KoboldCpp (llama.cpp) server through koboldcpp_run and koboldcpp_vision tools, with on-demand server lifecycle management.

1hace 19 horasVisión, voz y multimodalMIT
L

dsh-eyes

leeminjing/dsh-eyes

On-demand vision for text-only DeepSeek models: upload images, and the model calls a view_image tool backed by any OpenAI-compatible vision endpoint (Qwen/DashScope by default).

1anteayerVisión, voz y multimodalMIT
K

dsh-mindseye

kanchengw/dsh-mindseye

Vision plugin for text-only DeepSeek Harness models: native image paste, layered evidence memory and cache, and intent-driven tool selection.

1hace 7 horasVisión, voz y multimodalMIT
G

dsh-tool-vision

gloryxpnv/dsh-tool-vision

Local-first structured vision for text-only agents: images go to a local OpenAI-compatible VLM and come back as JSON evidence (summary, verbatim OCR, layout regions, entities/relations, colors, explicit uncertainty), with anti-hallucination fallback and an optional paste/upload bridge; zero cloud cost, images never leave the machine.

1hace 3 díasVisión, voz y multimodalMIT
E

dsh-plugin-mm-vision

elohia/dsh-plugin-mm-vision

Synesthesia Encoder for DSH: a vision model translates images into compact structured spatial text (canvas/elements/percentage coordinates), giving text-only LLMs pixel-level image understanding via the `mm_vision` tool.

1hace 5 díasVisión, voz y multimodalMIT
C

dsh-deepseek-vision

cheng-cheng9669/dsh-deepseek-vision

Reuses DeepSeek web's built-in vision mode for text-only models: the deepseek_vision tool drives the local deepseek-vision-cli browser automation (manual login helper, deep-think enabled, auto-closes browser) and returns image descriptions as text.

1anteayerVisión, voz y multimodalMIT
5

dsh-youreyes

54xkeee/dsh-youreyes

Vision toolkit for text-only DeepSeek: model-invokable `vision` tool, wrapper adapters for deepseek/opencode-go (v4 flash/pro), Antigravity IDE quota (default, flash/pro) / any OpenAI-compatible VLM / Gemini / local Ollama channels, evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.

1hace 16 horasVisión, voz y multimodalMIT
3

dsh-vision (vision-tool)

314857493/dsh-vision

Model-facing `vision` tool for DeepSeek Harness: describe and OCR image files by calling the free Zhipu GLM vision API directly (glm-4v-flash fallback chain), no external CLI required.

1anteayerVisión, voz y multimodalMIT
3

dsh-vision (vision-route)

314857493/dsh-vision

Registers a `deepseek-vision` provider route: the Web GUI accepts pasted images and transcribes them to text via the free Zhipu GLM vision API before delegating to the DeepSeek adapter.

1anteayerVisión, voz y multimodalMIT
1

dsh-vision-fallback

1helloman1/dsh-vision-fallback

Routes chat images to a fixed OpenAI-compatible vision model, returns factual observations to the selected main model, and reuses session-scoped observations across replay, compaction, and restarts.

1hace 14 horasVisión, voz y multimodalMIT
N

dsh-vision-plugin

nexsjournal/dsh-vision-plugin

Installs an image-input checkbox into the Models catalog so a custom model can declare image input and receive images directly; includes an optional BYO vision relay.

1anteayerMejoras de UIMIT
S

multimodal-bridge

spirit4471/multimodal-bridge

Paquete de plugins de DeepSeek Harness: herramientas qwen_vision (comprensión de imágenes con Qwen-VL) y qwen_generate (texto a imagen con Qwen-Image) para modelos de solo texto

1hace 5 díasVisión, voz y multimodalMIT
Y

dsh-plugin-vision-toolkit

yytbit/dsh-plugin-vision-toolkit

Kit de herramientas de visión para DeepSeek Harness: herramientas de CLI glance, ground, detect y crop para que los agentes de solo texto comprendan imágenes

1hace 5 díasHerramientas y funcionesMIT
C

deepsee

chang416/deepsee

DeepSee: visión para DeepSeek Harness, enrutamiento multimodelo y autocomprobaciones visuales con Gemini antes de la entrega

1hace 4 díasVisión, voz y multimodalMIT
F

dsh-sight

fu3rte/dsh-sight

Visión enchufable para los modelos de solo texto de DeepSeek Harness (dsh): una herramienta `vision` con preajustes de VLM económicos/gratuitos integrados, análisis por lotes de varias imágenes, admisión de imágenes al pegarlas como pista, y una página de ajustes web con recarga en caliente.

1hace 4 díasMejoras de UIMIT
Y

dsh-workloads

yewenyell-lang/dsh-workloads

Workspace-owned durable process supervision, readiness checks, and a Runtime Center for DeepSeek Harness.

0hace 4 díasDesarrollo y herramientas de pluginsMIT
Z

dsh-vision

zoahdev/dsh-vision

vision_analyze tool: analyze a local image or URL with an OpenAI-compatible vision model.

0anteayerHerramientas y funcionesMIT
N

gewu-tools

nyantused-cpun/gewu-tools

Model-agnostic visual-inspection pipeline for text-only agents: page-by-page HTML screenshots plus a ready-made vision-subagent briefing contract (gewu_prep), then source-code truth verification of every finding (gewu_locate); validated on mimo-v2.5 & qwen3.7-plus.

0hace 14 horasFlujos y automatizaciónMIT
T

dsh-vision-api-localorweb

tipsong/dsh-vision-api-localorweb

0hace 3 díasHerramientas y funcionesMIT