Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

76 plugins found

W

dsh-opencode-go-usage

wawo77/dsh-opencode-go-usage

Plan-usage pet for the DSH web GUI: a floating skeletal-animated character reports quota in a speech bubble, auto-matching the configured provider (OpenCode Go, Command Code, DeepSeek) to show a plan percentage or an account balance, and falling back to local token stats when the provider API is unavailable; draggable, wheel-scalable (90-260px), synthesized click sounds, loopback-only routes.

316 days agoUsage & BillingMIT
W

dsh-hold-to-talk

wangzhanchao883/dsh-hold-to-talk

One-handed, keyboard-free input for the composer: press and hold the mouse on the input box, speak, release — the text lands in the draft, and sliding up cancels without leaving a half sentence behind. No mic button to aim at and no shortcut to remember; the hand never has to leave the input area, which is the point when the other hand is busy. Recognition runs fully offline (SenseVoice via sherpa-onnx, no API key, audio never leaves the machine).

319 hours agoVision, Voice & MultimodalMIT
N

dsh-plugin-speech

nakamuraia/dsh-plugin-speech

Read assistant replies aloud in DeepSeek Harness: text-to-speech providers with streaming playback.

318 days agoVision, Voice & MultimodalMIT
S

dsh-voice-assistant

supersyh-sss/dsh-voice-assistant

Voice assistant for dsh web: say the wake phrase (e.g. "小鲸") to activate hands-free dictation — what you say is transcribed and typed into the chat box automatically. Supports spoken edit commands (send, clear, new line, stop reading) and reads assistant replies aloud in Chinese. Speech recognition runs locally in-browser via sherpa-onnx WASM, so it works offline without an API key.

3last monthVision, Voice & MultimodalMIT
D

dsh-whale-pet

dleaf6211-hash/dsh-whale-pet

Whale girl desktop pet for DeepSeek Harness: browser floating mascot plus Windows desktop pet sharing one data backend, with real-time balance board, task notices, head-pat interaction, 19 pre-synthesized voiced lines with per-line emotion, and presence-aware idle speech.

317 days agoJust for FunMIT
L

dsh-voco

lgquan/dsh-voco

Continuous voice conversations for DSH with hands-free listening, push-to-talk, speech recognition, TTS replies, and background Agent delegation.

3last monthVision, Voice & MultimodalMIT
T

dsh-gsv

taoruiliu19/dsh-gsv

Real-time local TTS for DeepSeek Harness: voice presets, auto-read, engine setup assistant, a read-aloud button, and a settings panel for the GSV-TTS-Lite engine.

3last monthVision, Voice & MultimodalMIT
L

Deepseek-Continuity

linxuhao/deepseek-continuity

Local image, voice, music and SFX generation plus transcription, with pinned identity: characters, animals, objects and actor voices are defined once and reused on every later call, degenerate output (a near-flat image, silent audio) is rejected instead of returned as success, and a generated line can be read back as text so a clone that swallowed its ending becomes visible. The engines unload when idle, and image generation and transcription can each be pointed at an OpenAI-shaped API instead of the local Vulkan backend.

312 days agoVision, Voice & MultimodalMIT
N

voco-input-sh

nothree-code/voco-input-sh

Voice input for the Web UI: a mic button that drives local VocoType offline speech recognition and auto-inserts recognized text into the composer (auto-deploy, dedupe, continuous dictation).

319 days agoVision, Voice & MultimodalMIT
B

dsh-stt-input

baisama-cloud/dsh-stt-input

Speech-to-text voice input for the web UI: a mic button in the composer transcribes speech into the draft via the browser Web Speech API (zero-config) or an OpenAI-compatible Whisper API (OpenAI / Groq), with a selectable model and language in Settings.

3last monthVision, Voice & MultimodalMIT
J

dsh-voice

jesse-njx/dsh-voice

Voice notes in, spoken answers out: dictate audio that becomes user messages (transcribe), have the agent read replies aloud (speak), local-first under ~/.dsh/voice.

32 months agoVision, Voice & MultimodalMIT
Q

dsh-mic-input

qt-chen/dsh-mic-input

Microphone voice input for the composer: browser Web Speech API live transcription, dedupe/auto-continue, smart punctuation, language and auto-send settings.

32 months agoVision, Voice & MultimodalMIT
B

dsh-voice-call

biliye/dsh-voice-call

Voice call assistant for the DSH Web GUI: a draggable floating call ball, browser-side VAD that sends each utterance after a pause, FunASR HTTP or streaming speech recognition, MiniMax or OpenAI-compatible TTS replies, an optional wake-word mode with auto-sleep, and task dispatch to separate subagent sessions with progress, stop, and spoken completion reports.

24 days agoVision, Voice & MultimodalMIT
M

dsh-voice-chat

maoyuching/dsh-voice-chat

Doubao-style voice chat for DSH: press-and-hold mic in the composer converts speech to text and auto-sends, and AI replies are read aloud. Optional LLM condensing (long replies summarized into short spoken lines, following the current conversation model), TTS-friendly text cleaning, selectable Edge TTS voices, adjustable silence auto-stop, and settings embedded in DSH settings dialog.

28 days agoVision, Voice & MultimodalMIT
G

dsh-tts

goodandready/dsh-tts

Speaks agent replies in the DeepSeek Harness web UI through a provider fallback chain (OpenAI, ElevenLabs, Google, Azure, Groq, Deepgram, OpenRouter, Edge, Piper, eSpeak), so a failing or rate-limited provider falls through to the next instead of going silent.

26 days agoVision, Voice & MultimodalMIT
T

dsh-deepseek-pet

txlznbzsdj-collab/dsh-deepseek-pet

Animated DeepSeek blue-whale desktop pet floating on the DSH Web page: draggable with speech bubbles and mood poses, a pet_say tool for the model to make it talk, and turn-start/end thinking and celebrating states.

22 months agoJust for FunMIT
P

dsh-voice-call

pandapolo/dsh-voice-call

Agent-initiated voice calls: `offer_call` rings the human (接听/拒接/稍后再说); accepted calls synthesize and play locally via CrispASR + Qwen3-TTS (9 speakers, 2 Chinese dialects), rejected calls return the decision to the agent.

23 days agoVision, Voice & MultimodalMIT
M

dsh-fish-tts

mari23333/dsh-fish-tts

Reads assistant replies aloud via Fish Audio API only (bring your own key): per-message read-aloud, auto-read toggle, and a settings page for model, voice reference_id, encrypted API key, and proxy.

22 days agoVision, Voice & MultimodalMIT
W

dsh-multi-tts

wyr-233/dsh-multi-tts

Per-reply read-aloud with a multi-provider settings page — MiniMax or any OpenAI-compatible /audio/speech endpoint, with voice, emotion, speed and model selection plus an auto-read toggle.

121 days agoVision, Voice & Multimodal
V

dsh-live-voice

victorwads/dsh-live-voice

Local-first voice conversations for DSH, with local speech recognition and synthesis and optional external providers.

117 days agoVision, Voice & MultimodalGPL-3.0
N

dsh-dictate

navneetset/dsh-dictate

Live speech-to-text dictation into the dsh web composer via OpenRouter STT

127 days agoVision, Voice & MultimodalMIT
M

dsh-minimax-asr

moluyao/dsh-minimax-asr

MiniMax 语音识别(asr-1.0)+ 语音合成(speech-2.8-hd)的 DeepSeek Harness 全局插件:转写工具、任务结束后用喇叭念一句的播报、实时免手对话、303 个音色可选

119 days agoVision, Voice & MultimodalMIT
D

dsh-voice-talk

duoduoqian708/dsh-voice-talk

Voice conversation mode for the DSH Web GUI: tap the mic to start a full-screen call, talk hands-free, and hear each reply read aloud as it streams. Bilingual (Chinese/English) UI, light and dark themes, adjustable speech rate, and switchable voices.

120 days agoVision, Voice & MultimodalMIT
J

dsh-live-voice

jstn-1g/dsh-live-voice

Consent-bound one-turn voice preview for DSH Web with a credential-free local synthetic demo, optional Qwen Audio, exact Session isolation, and explicit transcript-to-draft handoff without automatic submission.

127 days agoVision, Voice & MultimodalMIT