Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

86 plugins found

G

deepseek-vision

gou-gee/deepseek-vision

DeepSeek Harness 原生视觉 Bundle:粘贴或拖入图片,通过托管的 deepseek-vision-mcp 调用 OpenAI 兼容视觉模型。

4last monthMCP & ConnectorsMIT
S

dsh-digipet

swaylq/dsh-digipet

Digimon-style raising game: hatch an egg that feeds on real work (turns, tool calls, errors) and evolves along four branching lines shaped by how you work; zero tokens, invisible to the model.

4last monthJust for FunMIT
G

deepseek-vision (dsh-plugin-deepseek-vision)

gou-gee/deepseek-vision

Vision MCP and DSH bundle for text-only DeepSeek: analyze_image, analyze_clipboard, compare_images and vision_status tools, a visual settings page, free GLM-4.6V-Flash by default, result caching and rate-limit tolerance; keys stay out of logs.

414 days agoVision, Voice & MultimodalMIT
F

dsh-plugin-deepeye

favio8/dsh-plugin-deepeye

DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.

4last monthVision, Voice & Multimodal
H

dsh-her-eyes

huashenglian/dsh-her-eyes

一个可以让ai自动调用VLM(多模态模型)进行视觉分析的dsh插件。A dsh plugin that allows AI to automatically invoke VLMs (multimodal models) for visual analysis.

4last monthVision, Voice & MultimodalMIT
H

dsh-omni-workstation

huashenglian/dsh-omni-workstation

Omni-modal workstation for DSH: analyze_image over an ordered multi-card VLM failover chain, a six-tool vision toolkit (zoom, colour sampling, pixel diff, OCR, element detection, inline display) sharing the same image resolver, generate_image over OpenAI/DashScope/ComfyUI protocols, multi-card async generate_video with a /build-video-tool builder, and speak/clone_voice TTS over seven providers, all driven by one auto-saving settings page.

32 days agoVision, Voice & MultimodalMIT
M

dsh-iris

mokuyoaxis/dsh-iris

Media and vision workspace for DeepSeek Harness: image, video and speech generation, image Q&A and element locating, long-image OCR, pixel diff, HTML-screenshot verification and video summarization, with DashScope and OpenAI-compatible providers, model pools and a workbench client.

312 hours agoTools & CapabilitiesMIT
W

dsh-model-capabilities

wjz-p/dsh-model-capabilities

Adds an input-modality selector to custom model settings in the DeepSeek Harness Web UI.

3last monthModels & ProvidersMIT
W

visual-review

wang-bool/visual-review

Renders pasted/uploaded images inline in the DSH Web chat and gives text-only models vision: the model-invokable visual_review tool calls any OpenAI-compatible multimodal API first, falling back to a local Qwen3-VL worker.

3last monthVision, Voice & MultimodalMIT
S

dsh-plugin-multimodal

shinjiyu/dsh-plugin-multimodal

Advertise image paste on text-only DeepSeek routes, describe attachments with a vision sidecar, and leave native vision models untouched.

3last monthVision, Voice & MultimodalMIT
H

dsh-vision-mix

haiziyao/dsh-vision-mix

Combine text, vision, and image-generation APIs into one Mix model with automatic routing: text-only requests go to the chat model, user images and agent screenshots go to the vision model, follow-ups keep using the same session image, and agents can generate or edit images with session-scoped call history.

313 days agoVision, Voice & MultimodalMIT
L

dsh-visibridge

lhbsaa/dsh-visibridge

Structured vision evidence (OCR/layout/semantics) plus a USB camera capture tool for a "shoot-look-adjust" debug loop; backends: Ollama, DeepSeek, Xiaomi.

228 days agoVision, Voice & MultimodalMIT
S

dsh-llm-qwen-local

starefinger/dsh-llm-qwen-local

LLM adapter plugin for locally deployed Qwen models behind a vLLM OpenAI-compatible endpoint, with per-model multimodal switch, fully configurable reasoning efforts, and a web settings page for editing the deployment from the frontend.

26 days agoModels & ProvidersMIT
A

dsh-multi-model-provider

alexkaiqi/dsh-multi-model-provider

Register models, assist with portraits, and select the Agent model from a secret-free catalog for DeepSeek Harness.

228 days agoModels & ProvidersMIT
S

dsh-collaboration

socialist-sister/dsh-collaboration

Multi-agent collaboration suite: user-configured specialist roster, persistent on-demand dispatch (team_call/team_message/team_status/team_close), clone instances, star-topology relay, model comparison and a multimodal vision bridge.

2last monthWorkflow & AutomationMIT
I

dsh-tool-visual-primitives

inkshadewoods/dsh-tool-visual-primitives

Analyzes conversation images through mode-specific prompts (caption, UI, document, grounding, topology, etc.) and injects structured evidence with coordinate primitives (boxes, points, refs) as text, with session-level caching for reuse across replay and compaction.

210 days agoVision, Voice & MultimodalMIT
H

dsh-open-eyes

hyp6666/dsh-open-eyes

Vision bridge for text-only DeepSeek routes that analyzes attached and local images through configurable OpenAI Responses, Chat Completions, or Anthropic Messages endpoints while leaving image-capable routes native.

215 days agoVision, Voice & MultimodalMIT
5

dsh-youreyes

54xkeee/dsh-youreyes

Vision toolkit for text-only DeepSeek: model-invokable `vision` tool, wrapper adapters for deepseek/opencode-go (v4 flash/pro), Antigravity IDE quota (default, flash/pro) / any OpenAI-compatible VLM / Gemini / local Ollama channels, evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.

2last monthVision, Voice & MultimodalMIT
W

dsh-file-attachment

wszhoho/dsh-file-attachment

Drag-and-drop / paste / upload files and images; on non-multimodal models images are auto-described by a configured VLM.

13 days agoUI EnhancementsMIT
M

dsh-bilibili

moxingovo/dsh-bilibili

DeepSeek Harness plugin: Bilibili keyword video search, video metadata, subtitle transcripts, direct play URLs, and multimodal frame viewing (bilibili_search / bilibili_video / bilibili_subtitles / bilibili_playurl / bilibili_frames). Anonymous by default

127 days agoVision, Voice & MultimodalMIT
J

dsh-mmroute

jmxsxwyzjdwl/dsh-mmroute

Transparent multimodal routing for text-only models: every image in every model call is fully transcribed (verbatim OCR, data, uncertainty zones, injection-hardened) by your own multimodal understander, with focused re-look via vision_relook and automatic retry on image-related failures. No bundled endpoints, no borrowed logins.

114 days agoVision, Voice & MultimodalMIT
H

dsh-vision-analysis

harvey-will/dsh-vision-analysis

DeepSeek Harness vision plugin: 8 analysis modes (describe, OCR, chart data, UI review, object detection, compare, code-gen, debug), any OpenAI- or Anthropic-compatible vision API, with a built-in free vision model and automatic rate-limit failover.

111 days agoVision, Voice & MultimodalMIT
E

dsh-sight

ericfetch/dsh-sight

DeepSeek Harness plugin: direct multimodal image transfer declarations + per-session image clearing (surface replace), with a settings page and composer controls.

1last monthVision, Voice & MultimodalMIT
Y

dsh-multimodal

yauntyour/dsh-multimodal

Per-file-type multimodal chains: preset processing models per wildcard convert image/video/audio files into prompt tokens before they reach the text-only session model, with per-preset fallback chains and a Multimodal settings page.

1last monthVision, Voice & MultimodalMIT