Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

56 plugins found

P

dsh-omi-voice

polinnizhong/dsh-omi-voice

In-chat read-aloud for DeepSeek Harness: tap to read, pause and resume AI replies with natural Doubao TTS voices (BYOK), reading only the final answer with code, tables and diagrams filtered; local engine, plugin keeps no API key.

74last monthVision, Voice & MultimodalMIT
1

dsh-plugin-tts

1624318455/dsh-plugin-tts

Reads assistant replies aloud via free Edge TTS or your own RVC voice models: read-aloud buttons + auto-read, adaptive chunked progressive playback (gapless long reads), one-click voice-pack installs from a registry, and a portable RVC runtime.

213 days agoVision, Voice & MultimodalMIT
Q

dsh-voice-mode (dsh-voice-mode)

qishuilalala/dsh-voice-mode

Full-duplex voice mode for the DeepSeek Harness Web UI: toggle (2s-pause auto-send) or hold-to-talk dictation into an editable draft with zipformer2 streaming ASR, optional wake word; sentence-by-sentence Edge TTS read-aloud with live captions, and speaking interrupts playback and the running turn (true barge-in). On-device ASR, no API key.

156 days agoVision, Voice & MultimodalMIT
P

dsh-talk

perrylink/dsh-talk

Voice I/O for DeepSeek Harness — speech-to-text and text-to-speech over the microphone and audio output.

155 days agoVision, Voice & MultimodalApache-2.0
H

dsh-novel-forge

huangziyuan-general/dsh-novel-forge

Novel-writing guardrails: 20 novel_* tools enforce fact-ledger consistency, phase and chapter gates, machine audit with deterministic metrics, and proposal-only revisions; a right-panel workbench covers reading, TTS playback, polish, proofread, continuity checks and batch drafting, and a bypass engine runs internal steps off the main conversation.

134 days agoWorkflow & AutomationMIT
B

dsh-voice-ai-girlfriend-plugin

beiyege-01/dsh-voice-ai-girlfriend-plugin

Voice AI girlfriend for the Web UI: FunASR mic input, Qwen3-TTS spoken replies, companion animation window, and two-way QQ chat (text/voice/image push) via NapCat.

1216 days agoVision, Voice & Multimodal
P

dsh-plugin-xiaomi-mimo-tts

ppy-web/dsh-plugin-xiaomi-mimo-tts

Adds Xiaomi MiMo text-to-speech to DSH Web with assistant-message read-aloud, PCM streaming, preset and custom voice design, browser speech fallback, playback controls, and optional UI sounds.

1119 hours agoVision, Voice & MultimodalMIT
W

dsh-mmx-bridge

welsione/dsh-mmx-bridge

MiniMax multimodal bridge: one `mmx_bridge` tool covers image understanding/generation, video, TTS, music, cover, web search and quota; optional `web_search`/`read_image` takeover; inline players/image previews right in the Web GUI (npm: `dsh-mmx-bridge`).

1012 days agoTools & CapabilitiesMIT
Z

dsh-voice-input-plugin

zhangbo-cn/dsh-voice-input-plugin

Composer mic for the Web UI: tap-to-monitor live transcription and hold-to-talk, with host Edge TTS reply reading that streams while the model generates, echo-pause during reading, and tap-to-stop.

6last monthVision, Voice & MultimodalMIT
1

dsh-perlica-ding

117bs/dsh-perlica-ding

Perlica (Arknights: Endfield) themed tiered sound notifications: plan-ready, task-done, needs-your-input, and error tones; silent for plain chat, system-level playback (works in background), cross-platform (Windows/macOS/Linux), custom TTS sounds.

619 days agoIntegrations & RemoteMIT
A

dsh-guide-dog

atropinoltt/dsh-guide-dog

MiniMax-powered multimodal plugin: real-time voice call mode (streaming conversation, floating dock UI), voice mode and mic voice input, plus image/video/music/speech generation and vision inspection tools.

5last monthVision, Voice & MultimodalMIT
Y

dsh-tts-bridge

yuuyuko-uu/dsh-tts-bridge

Reads DSH conversations aloud using the DeepSeek web page's built-in read-aloud, driven by a small browser extension.

412 hours agoVision, Voice & Multimodal
S

dsh-voice

stardustlc666/dsh-voice

Voice tools: free edge-tts neural speech synthesis, OpenAI-compatible ASR transcription, voice list, batch voice preview and health self-check.

42 days agoVision, Voice & MultimodalMIT
H

dsh-voice

haoku123/dsh-voice

Full-duplex voice mode for the Web UI: tap-to-toggle or hold-to-talk dictation (send key or `Ctrl`) with a live caption, host-side SenseVoice ASR via sherpa-onnx, sentence-by-sentence spoken replies, and speaking interrupts playback and the running turn (true barge-in). No API key.

4last monthVision, Voice & MultimodalMIT
F

dsh-chatvoice

fuzzysoul/dsh-chatvoice

Free voice closed loop for the Web UI: browser SpeechRecognition mic input with live interim results plus read-aloud speaker buttons and auto-read for assistant replies, zero configuration and no API key.

4last monthVision, Voice & MultimodalMIT
L

dsh-plugin-notify-sound

ldchaowin/dsh-plugin-notify-sound

Per-workspace completion ringtones plus attention sounds for approval, question, plan-review, goal-blocked, and task-failure events, with built-in synth, voice (TTS), and custom audio.

42 months agoIntegrations & RemoteMIT
H

dsh-omni-workstation

huashenglian/dsh-omni-workstation

Omni-modal workstation for DSH: analyze_image over an ordered multi-card VLM failover chain, a six-tool vision toolkit (zoom, colour sampling, pixel diff, OCR, element detection, inline display) sharing the same image resolver, generate_image over OpenAI/DashScope/ComfyUI protocols, multi-card async generate_video with a /build-video-tool builder, and speak/clone_voice TTS over seven providers, all driven by one auto-saving settings page.

311 days agoVision, Voice & MultimodalMIT
N

dsh-plugin-speech

nakamuraia/dsh-plugin-speech

Read assistant replies aloud in DeepSeek Harness: text-to-speech providers with streaming playback.

315 days agoVision, Voice & MultimodalMIT
S

dsh-audiogen

shimingming520/dsh-audiogen

AI audio generation for the DeepSeek Harness web GUI — multi-vendor TTS, music, sound effects and voice design with a sidebar panel, model comparison, resource library and Agent tools.

321 days agoVision, Voice & MultimodalApache-2.0
L

dsh-voco

lgquan/dsh-voco

Continuous voice conversations for DSH with hands-free listening, push-to-talk, speech recognition, TTS replies, and background Agent delegation.

328 days agoVision, Voice & MultimodalMIT
T

dsh-gsv

taoruiliu19/dsh-gsv

Real-time local TTS for DeepSeek Harness: voice presets, auto-read, engine setup assistant, a read-aloud button, and a settings panel for the GSV-TTS-Lite engine.

329 days agoVision, Voice & MultimodalMIT
H

dsh-plugin-notify

huguangyu666/dsh-plugin-notify

Notification outbox: agent proactively notifies via toast / Chinese TTS voice / sound effects (explosion, victory, alarm), 60s confirmation window voice-calls you back, volume boost, settings panel.

315 days agoIntegrations & RemoteMIT
J

dsh-voice

jesse-njx/dsh-voice

Voice notes in, spoken answers out: dictate audio that becomes user messages (transcribe), have the agent read replies aloud (speak), local-first under ~/.dsh/voice.

32 months agoVision, Voice & MultimodalMIT
B

dsh-voice-call

biliye/dsh-voice-call

Voice call assistant for the DSH Web GUI: a draggable floating call ball, browser-side VAD that sends each utterance after a pause, FunASR HTTP or streaming speech recognition, MiniMax or OpenAI-compatible TTS replies, an optional wake-word mode with auto-sleep, and task dispatch to separate subagent sessions with progress, stop, and spoken completion reports.

22 days agoVision, Voice & MultimodalMIT