Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

35 plugins found

S

dsh-image-gen

shanliuling/dsh-image-gen

Native conversational image generation for DeepSeek Harness: ask the agent to create an image, and it handles generation and keeps the result directly in the conversation.

52517 hours agoVision, Voice & MultimodalApache-2.0
M

dsh-media-skills

mjorgin/dsh-media-skills

Free vision bridge and image generation for text-only models: paste-image reading, GLM-4V-Flash and Gemini engine failover, ModLens-style structured evidence, and a seeded free vision model route.

199 days agoVision, Voice & MultimodalMIT
Z

dsh-agnes-studio

zmm863-commits/dsh-agnes-studio

AI image and video studio as a floating DSH panel, so creation runs alongside the conversation instead of replacing it: text-to-image, image-to-image and multi-image composition at 1K-4K across eight aspect ratios; text-to-video and image-to-video with first-frame control at 4-12 seconds; short-drama mode that imports a script (.txt/.md/.json), breaks it into storyboard shots for preview and batch generation; and a prompt-expert workspace. Zero runtime dependencies.

112 days agoVision, Voice & Multimodal
W

dsh-mmx-bridge

welsione/dsh-mmx-bridge

MiniMax multimodal bridge: one `mmx_bridge` tool covers image understanding/generation, video, TTS, music, cover, web search and quota; optional `web_search`/`read_image` takeover; inline players/image previews right in the Web GUI (npm: `dsh-mmx-bridge`).

1012 days agoTools & CapabilitiesMIT
W

dsh-comfyui-canvas

wbin0001/dsh-comfyui-canvas

From chat to canvas to artwork — drive ComfyUI as a visual workflow IDE inside DSH. Embed ComfyUI (local or cloud) as a split-screen canvas in DeepSeek Harness Web: the agent sparks ideas, writes prompts and scripts right in the chat, applies them live to the canvas in front of you, and produces images, music, video, and 3D. From idea to finished output without ever leaving the conversation or switching front-ends: Canvas ops — compose and arrange pipelines, read/write workflows, edit nodes, wire links, run, tune parameters, and debug errors, all live and WYSIWYG on the exact canvas you are looking at; Production tasks — batch parameter sweeps (batch_run) and automatic output-image retrieval back into the chat (get_outputs), powering multi-modal creative and batch generation across images, music, video, and 3D; Environment upkeep — one-click launch of ComfyUI and one-click upgrade of the core plus every custom node (upgrade), keeping the stack healthy without interruption. This package is the DSH-side plugin, and it ships the ComfyUI-side bridge node too.

9yesterdayTools & CapabilitiesMIT
P

dsh-draw

perrylink/dsh-draw

Multi-engine text-to-image generation (OpenAI Images and Zhipu CogView presets) with per-session quota tracking, engine failover, credential-safe config, and a result card with regenerate.

95 days agoVision, Voice & MultimodalApache-2.0
M

dsh-iris

mokuyoaxis/dsh-iris

Media and vision workspace for DeepSeek Harness: image, video and speech generation, image Q&A and element locating, long-image OCR, pixel diff, HTML-screenshot verification and video summarization, with DashScope and OpenAI-compatible providers, model pools and a workbench client.

710 days agoTools & CapabilitiesMIT
G

dsh-image-gen

goodandready/dsh-image-gen

Adds a generate_image tool backed by the FAL queue, any OpenAI-compatible images API, or a ChatGPT/Grok subscription, and shows the result inline in the conversation.

62 days agoTools & CapabilitiesMIT
X

dsh-draw-router

xiaozhe7772222/dsh-draw-router

Unified image generation router for DeepSeek Harness (DSH): auto-discovers image models from any OpenAI-compatible endpoint, provides draw_image and draw_list_sources tools, supports SenseNova, StepFun, Agnes, Qwen, Flux, SD, Imagen and more.

4last monthVision, Voice & MultimodalMIT
H

dsh-omni-workstation

huashenglian/dsh-omni-workstation

Omni-modal workstation for DSH: analyze_image over an ordered multi-card VLM failover chain, a six-tool vision toolkit (zoom, colour sampling, pixel diff, OCR, element detection, inline display) sharing the same image resolver, generate_image over OpenAI/DashScope/ComfyUI protocols, multi-card async generate_video with a /build-video-tool builder, and speak/clone_voice TTS over seven providers, all driven by one auto-saving settings page.

311 days agoVision, Voice & MultimodalMIT
L

Deepseek-Continuity

linxuhao/deepseek-continuity

Local image, voice, music and SFX generation plus transcription, with pinned identity: characters, animals, objects and actor voices are defined once and reused on every later call, degenerate output (a near-flat image, silent audio) is rejected instead of returned as success, and a generated line can be read back as text so a clone that swallowed its ending becomes visible. The engines unload when idle, and image generation and transcription can each be pointed at an OpenAI-shaped API instead of the local Vulkan backend.

39 days agoVision, Voice & MultimodalMIT
W

wavespeed-dsh-skill

wavespeedai/wavespeed-dsh-skill

Generate and edit AI media (image, video, audio, 3D) with WaveSpeed models via the open-source wavespeed CLI: live catalog search, per-model input-schema introspection, local-file upload with @path markers, and price checks before running.

310 days agoSkillsMIT
L

dsh-image-gen

leemancheung/dsh-image-gen

GPT Image 2 `image_gen` with Codex subscription OAuth by default or explicit API-key mode: developing card, up to three live API partials, durable attachment replay/lightbox/download, text-only model output, and bounded credential-safe requests.

33 days agoVision, Voice & MultimodalMIT
H

dsh-vision-mix

haiziyao/dsh-vision-mix

Combine text, vision, and image-generation APIs into one Mix model with automatic routing: text-only requests go to the chat model, user images and agent screenshots go to the vision model, follow-ups keep using the same session image, and agents can generate or edit images with session-scoped call history.

322 days agoVision, Voice & MultimodalMIT
Q

dsh-gemini-pool

qikairo7/dsh-gemini-pool

Multi-account Google Gemini provider: picks the account with the most remaining quota, fails over on 429 with exponential cooldown, probes disabled accounts in the background, and ships a Chinese/English settings page.

213 hours agoModels & ProvidersMIT
M

dsh-ui-mockup

mackwan84/dsh-ui-mockup

ui_mockup 工具 Consumer:研讨阶段生成 UI 线框图/高保真草图,落盘设计稿并提示确认后锁定设计规格;含客户端卡片、图片路由与 i18n

219 days agoVision, Voice & MultimodalMIT
L

agnes-media

littlebeaverstudio/agnes-media

Agnes AI image and video generation plugin for DeepSeek Harness.

219 days agoTools & CapabilitiesMIT
M

dsh-ui-mockup (ui-mockup)

mackwan84/dsh-ui-mockup

Generates UI wireframes and high-fidelity mockups in DeepSeek Harness with selectable DashScope and Volcengine image providers.

213 days agoUI EnhancementsMIT
E

dsh-labnana

exoticknight/dsh-labnana

Labnana image generation for DeepSeek Harness: text-to-image / image-to-image / precise editing with credits estimation, subscription balance and web settings UI.

26 days agoVision, Voice & MultimodalApache-2.0
Z

dsh-tool-generate-image

zcldragon/dsh-tool-generate-image

Model-facing `generate_image` tool for text-only models: the model asks for a picture in natural language, Gemini draws it via the Antigravity CLI, the image is saved to a configurable output directory, and the tool returns the file paths for the model to use in its work.

2last monthWorkflow & AutomationApache-2.0
D

dsh-imagegen-skill

dingchenhui0618-arch/dsh-imagegen-skill

Registers the imagegen skill on the skill registry, so generating and editing raster images needs no hand-copied files under $DSH_HOME/skills. The agent drives a bundled image_gen.py CLI against any OpenAI-compatible GPT Image endpoint for text-to-image, reference-image editing, and transparent-background cutouts.

112 days agoSkillsApache-2.0
N

dsh-imgdraw

ninjasln-labs/dsh-imgdraw

Text-to-image for DeepSeek Harness: a `draw_image` model tool, an input-bar 生图 button with a prompt popup (async generation, 4-grid results, download / keep / delete), an /imgdraw image route, and persisted history. Backends: DashScope wan2.7-image (free

125 days agoVision, Voice & MultimodalMIT
P

dsh-tool-imagegen

pappet/dsh-tool-imagegen

Text-to-image and image-to-image generation via OpenRouter's unified Image API: configurable model aliases with parameters gated against the live model capability listing, reference-image inputs, a settings card, and inline chat display of results.

129 days agoTools & CapabilitiesMIT
R

dsh-image-gen

randomix777/dsh-image-gen

Optimized fork of dsh-image-gen with batch generation (count), gallery pagination, lazy loading, fetch retry/timeout, and 7 providers: Gemini/OpenAI/Seedream/DashScope/Agnes/GLM-Image/Stability AI.

13 days agoWorkflow & AutomationApache-2.0