Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

45 plugins found

F

dsh-memory

furongjun-1999/dsh-memory

White-box AGI architecture exploration: metacognition (self-cognition loop), continual learning (knowledge flywheel), world model (condition space, spatiotemporal memory graph), self-improvement (bootstrap discipline), zero-LLM white-box pipeline, and auditable trust guardrails.

30814 hours agoMemoryMIT
B

DSH-Plugins-Marketplace

bradegithub/dsh-plugins-marketplace

GitHub-topic-driven plugin & skill marketplace: a Settings page that browses the auto-collected registry (the whole dsh-plugin topic plus the skills index, CI-refreshed every 2 hours) with one-click install, type detection, install-script and host-shadow-dependency safety confirmations, env-key management, and the STANDARD.md recognition spec.

1693 hours agoDev & Plugin ToolsMIT
F

dsh-prompt-enhancer

fishsb/dsh-prompt-enhancer

One-click prompt enhancement (5 modes, memory chain, multi-model fallback) plus voice recognition (cloud Qwen3-ASR / offline SenseVoice, auto-stop on silence) for the composer, with one-click DSH service restart.

825 days agoTools & Capabilities
P

dsh-voice-scribe

pensivefei/dsh-voice-scribe

Voice input for the web UI: tap Alt (or Alt+Space) to start/stop dictation, browser Web Speech (zero-config) or OpenAI-compatible cloud ASR, optional polish through DSH-configured LLM, settings UI.

343 days agoVision, Voice & MultimodalMIT
W

dsh-ears

wiziscool/dsh-ears

Voice input plugin for DeepSeek Harness (dsh): a microphone button in the composer turns speech into a draft transcript, with a choice of speech-recognition backends, optional polish through dsh own LLM routes, and a native settings page.

2122 hours agoVision, Voice & MultimodalMIT
Q

dsh-voice-mode (dsh-voice-mode)

qishuilalala/dsh-voice-mode

Full-duplex voice mode for the DeepSeek Harness Web UI: toggle (2s-pause auto-send) or hold-to-talk dictation into an editable draft with zipformer2 streaming ASR, optional wake word; sentence-by-sentence Edge TTS read-aloud with live captions, and speaking interrupts playback and the running turn (true barge-in). On-device ASR, no API key.

157 hours agoVision, Voice & MultimodalMIT
L

dsh-vision

linenxi-ctrl/dsh-vision

External vision plugin for DeepSeek Harness: whale-button config panel, image recognition with auto-reply, and agent screenshot/recognize tools.

92 months agoVision, Voice & MultimodalMIT
A

dsh-highres-vision

azwosile/dsh-highres-vision

For the DeepSeek Harness native vision model deepseek-v4-flash-vision-exp, raises image admission limits to 32 MiB / 8192 px / 600 images and adds a highres_read tool that tiles large images, then returns the whole image plus 800x800 tiles through the host read_image tool.

822 days agoVision, Voice & MultimodalMIT
3

dsh-voice

3274375092/dsh-voice

Voice input for DeepSeek Harness: speak into the microphone and the recognized text is submitted as a normal chat message, via local or browser speech recognition.

82 days agoVision, Voice & MultimodalMIT
5

dsh-vision

54xkeee/dsh-vision

Vision for text-only DeepSeek via Doubao Web by default (zero-cost, no API key — drives your logged-in Chrome through a Windows CDP bridge), with Antigravity IDE quota (flash/pro) or Gemini fallback; auto detail escalation, vision evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.

72 months agoVision, Voice & MultimodalMIT
G

dsh-tool-see-image

gugu123a/dsh-tool-see-image

Provides the see_image tool for DSH: send an image file to a configurable OpenAI-compatible vision model and relay its description back to a text-only model.

529 days agoTools & CapabilitiesMIT
N

vision-exp-tile

nicholas023/vision-exp-tile

Large-image recognition for vision-exp models: lossless 800×800 tile recognition (smart/pipeline/full), local OCR with preprocessing & handwriting routing, optional multi-vendor GPU (DirectML/CUDA/OpenVINO) with auto CPU fallback.

5last monthVision, Voice & MultimodalMIT
F

dsh-chatvoice

fuzzysoul/dsh-chatvoice

Free voice closed loop for the Web UI: browser SpeechRecognition mic input with live interim results plus read-aloud speaker buttons and auto-read for assistant replies, zero configuration and no API key.

42 months agoVision, Voice & MultimodalMIT
S

dsh-linghun

syyr1987/dsh-linghun

A judgment core for DeepSeek Harness: a cognition loop that gives every decision a source and attribution and drives planning to a close, hippocampus memory that condenses running experience into reusable knowledge, and a customizable persona card (name, personality, communication style). Requires DSH ^0.1.0-rc.7.

313 hours agoMemory
A

dsh-j-space

anonyjcy/dsh-j-space

J-Space Cognition Suite SV1 native agent preset and standalone Cordis plugin: 13 modules, persistent controller and decoupled workspace.

33 days agoMemoryMIT
H

dsh-omni-workstation

huashenglian/dsh-omni-workstation

Omni-modal workstation for DSH: analyze_image over an ordered multi-card VLM failover chain, a six-tool vision toolkit (zoom, colour sampling, pixel diff, OCR, element detection, inline display) sharing the same image resolver, generate_image over OpenAI/DashScope/ComfyUI protocols, multi-card async generate_video with a /build-video-tool builder, and speak/clone_voice TTS over seven providers, all driven by one auto-saving settings page.

314 days agoVision, Voice & MultimodalMIT
W

dsh-hold-to-talk

wangzhanchao883/dsh-hold-to-talk

One-handed, keyboard-free input for the composer: press and hold the mouse on the input box, speak, release — the text lands in the draft, and sliding up cancels without leaving a half sentence behind. No mic button to aim at and no shortcut to remember; the hand never has to leave the input area, which is the point when the other hand is busy. Recognition runs fully offline (SenseVoice via sherpa-onnx, no API key, audio never leaves the machine).

321 hours agoVision, Voice & MultimodalMIT
S

dsh-voice-assistant

supersyh-sss/dsh-voice-assistant

Voice assistant for dsh web: say the wake phrase (e.g. "小鲸") to activate hands-free dictation — what you say is transcribed and typed into the chat box automatically. Supports spoken edit commands (send, clear, new line, stop reading) and reads assistant replies aloud in Chinese. Speech recognition runs locally in-browser via sherpa-onnx WASM, so it works offline without an API key.

3last monthVision, Voice & MultimodalMIT
L

dsh-voco

lgquan/dsh-voco

Continuous voice conversations for DSH with hands-free listening, push-to-talk, speech recognition, TTS replies, and background Agent delegation.

3last monthVision, Voice & MultimodalMIT
X

dsh-image-vision

xiaoyuink/dsh-image-vision

Image understanding for any DSH model: vision, OCR, grounding, and crop tools with domain presets for histopathology, cell biology, anatomy, clinical images and scientific figures.

329 days agoVision, Voice & Multimodal
N

voco-input-sh

nothree-code/voco-input-sh

Voice input for the Web UI: a mic button that drives local VocoType offline speech recognition and auto-inserts recognized text into the composer (auto-deploy, dedupe, continuous dictation).

319 days agoVision, Voice & MultimodalMIT
W

visual-review

wang-bool/visual-review

Renders pasted/uploaded images inline in the DSH Web chat and gives text-only models vision: the model-invokable visual_review tool calls any OpenAI-compatible multimodal API first, falling back to a local Qwen3-VL worker.

32 months agoVision, Voice & MultimodalMIT
M

dsh-vision-tools

moon09300731/dsh-vision-tools

Full vision-capability bundle for DeepSeek Harness: a vision_understand tool (OpenAI-compatible vision APIs, free Zhipu GLM-4V-Flash by default) plus paste/drag-and-drop/button entry points for image recognition.

32 months agoVision, Voice & MultimodalMIT
B

dsh-voice-call

biliye/dsh-voice-call

Voice call assistant for the DSH Web GUI: a draggable floating call ball, browser-side VAD that sends each utterance after a pause, FunASR HTTP or streaming speech recognition, MiniMax or OpenAI-compatible TTS replies, an optional wake-word mode with auto-sleep, and task dispatch to separate subagent sessions with progress, stop, and spoken completion reports.

24 days agoVision, Voice & MultimodalMIT