Plugins
Browse, filter, and install DeepSeek-Harness plugins.
125 plugins found
dsh-plugin-notify
huguangyu666/dsh-plugin-notify
Notification outbox: agent proactively notifies via toast / Chinese TTS voice / sound effects (explosion, victory, alarm), 60s confirmation window voice-calls you back, volume boost, settings panel.
dsh-voice
jesse-njx/dsh-voice
Voice notes in, spoken answers out: dictate audio that becomes user messages (transcribe), have the agent read replies aloud (speak), local-first under ~/.dsh/voice.
dsh-mic-input
qt-chen/dsh-mic-input
Microphone voice input for the composer: browser Web Speech API live transcription, dedupe/auto-continue, smart punctuation, language and auto-send settings.
dsh-voice-call
biliye/dsh-voice-call
Voice call assistant for the DSH Web GUI: a draggable floating call ball, browser-side VAD that sends each utterance after a pause, FunASR HTTP or streaming speech recognition, MiniMax or OpenAI-compatible TTS replies, an optional wake-word mode with auto-sleep, and task dispatch to separate subagent sessions with progress, stop, and spoken completion reports.
dsh-tool-lipsync
yu-wenchao/dsh-tool-lipsync
Free lip-sync video generation plugin for DSH with 3000+ voices and 500+ languages.
dsh-whale-pet
dleaf6211-hash/dsh-whale-pet
Whale girl desktop pet for DeepSeek Harness: browser floating mascot plus Windows desktop pet sharing one data backend, with real-time balance board, task notices, head-pat interaction, 19 pre-synthesized voiced lines with per-line emotion, and presence-aware idle speech.
dsh-voice-agent (voice-app)
wayneyu430/dsh-voice-agent
A conversational voice frontend Agent for dsh: speak naturally over ByteDance Duplex, delegate requests to background tasks, and hear their asynchronous results reported by voice.
dsh-voice-chat
maoyuching/dsh-voice-chat
Doubao-style voice chat for DSH: press-and-hold mic in the composer converts speech to text and auto-sends, and AI replies are read aloud. Optional LLM condensing (long replies summarized into short spoken lines, following the current conversation model), TTS-friendly text cleaning, selectable Edge TTS voices, adjustable silence auto-stop, and settings embedded in DSH settings dialog.
dsh-tts
goodandready/dsh-tts
Speaks agent replies in the DeepSeek Harness web UI through a provider fallback chain (OpenAI, ElevenLabs, Google, Azure, Groq, Deepgram, OpenRouter, Edge, Piper, eSpeak), so a failing or rate-limited provider falls through to the next instead of going silent.
muxiva-dsh-voice
piyotahu/muxiva-dsh-voice
Local-first, full-duplex voice for DeepSeek Harness, orchestrated by Muxiva
dsh-voice-call
pandapolo/dsh-voice-call
Agent-initiated voice calls: `offer_call` rings the human (接听/拒接/稍后再说); accepted calls synthesize and play locally via CrispASR + Qwen3-TTS (9 speakers, 2 Chinese dialects), rejected calls return the decision to the agent.
dsh-fish-tts
mari23333/dsh-fish-tts
Reads assistant replies aloud via Fish Audio API only (bring your own key): per-message read-aloud, auto-read toggle, and a settings page for model, voice reference_id, encrypted API key, and proxy.
dsh-live2d-voice
john-walks-slow/dsh-live2d-voice
Live2D session view with realtime voice: continuous voice input, streaming TTS, subtitle translation, multi-model catalog, standalone entry / Live2D 实时语音会话视图:连续语音输入、流式 TTS、字幕翻译、多模型目录、独立入口
dsh-claude-slider
yicun0316/dsh-claude-slider
Replaces the reasoning-effort dropdown with a draggable snapping slider, and draws 13 canvas effects across five style families (thrust, fluid, fantasy, tech, texture) with three tiers that unlock as the effort level rises. Custom mode combines 10 particle forms with 8 flow trajectories; 9 accent presets and a hex picker recolor every effect, and optional key sounds cover two voice clips, a synthesized duck squeak, a mechanical click and a local audio file.
dsh-wx-bridge
zhy5/dsh-wx-bridge
Drive your local DSH from WeChat: a self-hosted iLink bridge over a persistent ACP session (context lives in DSH and is resumable), phone conversations grouped into the desktop workspace, with image recognition, file messages (Excel/Word/PDF, parsed by the agent itself) and voice transcripts. Access defaults to strict (only registered devices are served); the phone channel's permission preset is wide by design - see the README security section before sharing the bot.
dsh-multi-tts
wyr-233/dsh-multi-tts
Per-reply read-aloud with a multi-provider settings page — MiniMax or any OpenAI-compatible /audio/speech endpoint, with voice, emotion, speed and model selection plus an auto-read toggle.
dshgo
xiazhi88/dshgo
Proxies the loopback-only dsh web onto a LAN port so phones and other machines can reach it, with an optional access password that browsers type in and the bundled Android client unlocks with biometrics. Adds a LAN-access settings page with QR codes, inlines dsh-web-mobile (MIT, attribution kept) for the mobile layout, and the Android client in the same repository adds finish/approval notifications, voice input and a home-screen widget.
dsh-live-voice
victorwads/dsh-live-voice
Local-first voice conversations for DSH, with local speech recognition and synthesis and optional external providers.
dsh-minimax-asr
moluyao/dsh-minimax-asr
MiniMax 语音识别(asr-1.0)+ 语音合成(speech-2.8-hd)的 DeepSeek Harness 全局插件:转写工具、任务结束后用喇叭念一句的播报、实时免手对话、303 个音色可选
dsh-voice-talk
duoduoqian708/dsh-voice-talk
Voice conversation mode for the DSH Web GUI: tap the mic to start a full-screen call, talk hands-free, and hear each reply read aloud as it streams. Bilingual (Chinese/English) UI, light and dark themes, adjustable speech rate, and switchable voices.
dsh-funasr-voice
fenglin-ai/dsh-funasr-voice
Offline voice input for the DSH Web UI: mic to local FunASR (SenseVoiceSmall), one-click install, no cloud.
dshtools-sensevoice-input
ilovedyou6666-hub/dshtools-sensevoice-input
基于 SenseVoiceSmall(iic/SenseVoiceSmall)多语言语音理解模型的 DSH Desktop 本地语音输入插件。
dsh-voice-input-npm
difimim/dsh-voice-input-npm
语音输入插件 for Deepseek Harness
dsh-live-voice
jstn-1g/dsh-live-voice
Consent-bound one-turn voice preview for DSH Web with a credential-free local synthetic demo, optional Qwen Audio, exact Session isolation, and explicit transcript-to-draft handoff without automatic submission.