Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

125 plugins found

H

dsh-plugin-notify

huguangyu666/dsh-plugin-notify

Notification outbox: agent proactively notifies via toast / Chinese TTS voice / sound effects (explosion, victory, alarm), 60s confirmation window voice-calls you back, volume boost, settings panel.

315 days agoIntegrations & RemoteMIT
J

dsh-voice

jesse-njx/dsh-voice

Voice notes in, spoken answers out: dictate audio that becomes user messages (transcribe), have the agent read replies aloud (speak), local-first under ~/.dsh/voice.

32 months agoVision, Voice & MultimodalMIT
Q

dsh-mic-input

qt-chen/dsh-mic-input

Microphone voice input for the composer: browser Web Speech API live transcription, dedupe/auto-continue, smart punctuation, language and auto-send settings.

3last monthVision, Voice & MultimodalMIT
B

dsh-voice-call

biliye/dsh-voice-call

Voice call assistant for the DSH Web GUI: a draggable floating call ball, browser-side VAD that sends each utterance after a pause, FunASR HTTP or streaming speech recognition, MiniMax or OpenAI-compatible TTS replies, an optional wake-word mode with auto-sleep, and task dispatch to separate subagent sessions with progress, stop, and spoken completion reports.

22 days agoVision, Voice & MultimodalMIT
Y

dsh-tool-lipsync

yu-wenchao/dsh-tool-lipsync

Free lip-sync video generation plugin for DSH with 3000+ voices and 500+ languages.

225 days agoVision, Voice & MultimodalMIT
D

dsh-whale-pet

dleaf6211-hash/dsh-whale-pet

Whale girl desktop pet for DeepSeek Harness: browser floating mascot plus Windows desktop pet sharing one data backend, with real-time balance board, task notices, head-pat interaction, 19 pre-synthesized voiced lines with per-line emotion, and presence-aware idle speech.

214 days agoJust for FunMIT
W

dsh-voice-agent (voice-app)

wayneyu430/dsh-voice-agent

A conversational voice frontend Agent for dsh: speak naturally over ByteDance Duplex, delegate requests to background tasks, and hear their asynchronous results reported by voice.

2last monthVision, Voice & Multimodal
M

dsh-voice-chat

maoyuching/dsh-voice-chat

Doubao-style voice chat for DSH: press-and-hold mic in the composer converts speech to text and auto-sends, and AI replies are read aloud. Optional LLM condensing (long replies summarized into short spoken lines, following the current conversation model), TTS-friendly text cleaning, selectable Edge TTS voices, adjustable silence auto-stop, and settings embedded in DSH settings dialog.

25 days agoVision, Voice & MultimodalMIT
G

dsh-tts

goodandready/dsh-tts

Speaks agent replies in the DeepSeek Harness web UI through a provider fallback chain (OpenAI, ElevenLabs, Google, Azure, Groq, Deepgram, OpenRouter, Edge, Piper, eSpeak), so a failing or rate-limited provider falls through to the next instead of going silent.

23 days agoVision, Voice & MultimodalMIT
P

muxiva-dsh-voice

piyotahu/muxiva-dsh-voice

Local-first, full-duplex voice for DeepSeek Harness, orchestrated by Muxiva

2last monthVision, Voice & MultimodalApache-2.0
P

dsh-voice-call

pandapolo/dsh-voice-call

Agent-initiated voice calls: `offer_call` rings the human (接听/拒接/稍后再说); accepted calls synthesize and play locally via CrispASR + Qwen3-TTS (9 speakers, 2 Chinese dialects), rejected calls return the decision to the agent.

22 days agoVision, Voice & MultimodalMIT
M

dsh-fish-tts

mari23333/dsh-fish-tts

Reads assistant replies aloud via Fish Audio API only (bring your own key): per-message read-aloud, auto-read toggle, and a settings page for model, voice reference_id, encrypted API key, and proxy.

224 days agoVision, Voice & MultimodalMIT
J

dsh-live2d-voice

john-walks-slow/dsh-live2d-voice

Live2D session view with realtime voice: continuous voice input, streaming TTS, subtitle translation, multi-model catalog, standalone entry / Live2D 实时语音会话视图:连续语音输入、流式 TTS、字幕翻译、多模型目录、独立入口

1yesterdayVision, Voice & MultimodalMIT
Y

dsh-claude-slider

yicun0316/dsh-claude-slider

Replaces the reasoning-effort dropdown with a draggable snapping slider, and draws 13 canvas effects across five style families (thrust, fluid, fantasy, tech, texture) with three tiers that unlock as the effort level rises. Custom mode combines 10 particle forms with 8 flow trajectories; 9 accent presets and a hex picker recolor every effect, and optional key sounds cover two voice clips, a synthesized duck squeak, a mechanical click and a local audio file.

16 days agoUI EnhancementsMIT
Z

dsh-wx-bridge

zhy5/dsh-wx-bridge

Drive your local DSH from WeChat: a self-hosted iLink bridge over a persistent ACP session (context lives in DSH and is resumable), phone conversations grouped into the desktop workspace, with image recognition, file messages (Excel/Word/PDF, parsed by the agent itself) and voice transcripts. Access defaults to strict (only registered devices are served); the phone channel's permission preset is wide by design - see the README security section before sharing the bot.

13 days agoIntegrations & RemoteMIT
W

dsh-multi-tts

wyr-233/dsh-multi-tts

Per-reply read-aloud with a multi-provider settings page — MiniMax or any OpenAI-compatible /audio/speech endpoint, with voice, emotion, speed and model selection plus an auto-read toggle.

119 days agoVision, Voice & Multimodal
X

dshgo

xiazhi88/dshgo

Proxies the loopback-only dsh web onto a LAN port so phones and other machines can reach it, with an optional access password that browsers type in and the bundled Android client unlocks with biometrics. Adds a LAN-access settings page with QR codes, inlines dsh-web-mobile (MIT, attribution kept) for the mobile layout, and the Android client in the same repository adds finish/approval notifications, voice input and a home-screen widget.

112 days agoIntegrations & RemoteMIT
V

dsh-live-voice

victorwads/dsh-live-voice

Local-first voice conversations for DSH, with local speech recognition and synthesis and optional external providers.

115 days agoVision, Voice & MultimodalGPL-3.0
M

dsh-minimax-asr

moluyao/dsh-minimax-asr

MiniMax 语音识别(asr-1.0)+ 语音合成(speech-2.8-hd)的 DeepSeek Harness 全局插件:转写工具、任务结束后用喇叭念一句的播报、实时免手对话、303 个音色可选

116 days agoVision, Voice & MultimodalMIT
D

dsh-voice-talk

duoduoqian708/dsh-voice-talk

Voice conversation mode for the DSH Web GUI: tap the mic to start a full-screen call, talk hands-free, and hear each reply read aloud as it streams. Bilingual (Chinese/English) UI, light and dark themes, adjustable speech rate, and switchable voices.

117 days agoVision, Voice & MultimodalMIT
F

dsh-funasr-voice

fenglin-ai/dsh-funasr-voice

Offline voice input for the DSH Web UI: mic to local FunASR (SenseVoiceSmall), one-click install, no cloud.

1last monthVision, Voice & MultimodalMIT
I

dshtools-sensevoice-input

ilovedyou6666-hub/dshtools-sensevoice-input

基于 SenseVoiceSmall(iic/SenseVoiceSmall)多语言语音理解模型的 DSH Desktop 本地语音输入插件。

126 days agoVision, Voice & MultimodalMIT
D

dsh-voice-input-npm

difimim/dsh-voice-input-npm

语音输入插件 for Deepseek Harness

1last monthVision, Voice & MultimodalMIT
J

dsh-live-voice

jstn-1g/dsh-live-voice

Consent-bound one-turn voice preview for DSH Web with a credential-free local synthetic demo, optional Qwen Audio, exact Session isolation, and explicit transcript-to-draft handoff without automatic submission.

124 days agoVision, Voice & MultimodalMIT