Skip to main content

Plugins

Browse, filter, and install DeepSeek-Harness plugins.

56 plugins found

N

insta360-ai-content-studio

nicecx/insta360-ai-content-studio

Automated daily content pipeline for Insta360 GO 3S: BLE/WiFi remote capture of aquarium footage, wireless import, ffmpeg quality evaluation with decision log, Auto editing (segment selection, TTS narration, music, subtitles), and Bilibili upload via biliup (dry-run by default).

228 days agoTools & CapabilitiesMIT
M

dsh-voice-chat

maoyuching/dsh-voice-chat

Doubao-style voice chat for DSH: press-and-hold mic in the composer converts speech to text and auto-sends, and AI replies are read aloud. Optional LLM condensing (long replies summarized into short spoken lines, following the current conversation model), TTS-friendly text cleaning, selectable Edge TTS voices, adjustable silence auto-stop, and settings embedded in DSH settings dialog.

25 days agoVision, Voice & MultimodalMIT
G

dsh-tts

goodandready/dsh-tts

Speaks agent replies in the DeepSeek Harness web UI through a provider fallback chain (OpenAI, ElevenLabs, Google, Azure, Groq, Deepgram, OpenRouter, Edge, Piper, eSpeak), so a failing or rate-limited provider falls through to the next instead of going silent.

23 days agoVision, Voice & MultimodalMIT
X

dsh-showreel

xiaoyuer3921/dsh-showreel

Turns the latest completed DSH session turn into a shareable vertical recap video: editable five-scene storyboard, host-side privacy redaction, optional AI polish and TTS narration, MP4/WebM export with PNG cover.

2last monthUI EnhancementsMIT
P

muxiva-dsh-voice

piyotahu/muxiva-dsh-voice

Local-first, full-duplex voice for DeepSeek Harness, orchestrated by Muxiva

2last monthVision, Voice & MultimodalApache-2.0
P

dsh-voice-call

pandapolo/dsh-voice-call

Agent-initiated voice calls: `offer_call` rings the human (接听/拒接/稍后再说); accepted calls synthesize and play locally via CrispASR + Qwen3-TTS (9 speakers, 2 Chinese dialects), rejected calls return the decision to the agent.

22 days agoVision, Voice & MultimodalMIT
M

dsh-fish-tts

mari23333/dsh-fish-tts

Reads assistant replies aloud via Fish Audio API only (bring your own key): per-message read-aloud, auto-read toggle, and a settings page for model, voice reference_id, encrypted API key, and proxy.

224 days agoVision, Voice & MultimodalMIT
J

dsh-live2d-voice

john-walks-slow/dsh-live2d-voice

Live2D session view with realtime voice: continuous voice input, streaming TTS, subtitle translation, multi-model catalog, standalone entry / Live2D 实时语音会话视图:连续语音输入、流式 TTS、字幕翻译、多模型目录、独立入口

1yesterdayVision, Voice & MultimodalMIT
W

dsh-multi-tts

wyr-233/dsh-multi-tts

Per-reply read-aloud with a multi-provider settings page — MiniMax or any OpenAI-compatible /audio/speech endpoint, with voice, emotion, speed and model selection plus an auto-read toggle.

119 days agoVision, Voice & Multimodal
V

dsh-live-voice

victorwads/dsh-live-voice

Local-first voice conversations for DSH, with local speech recognition and synthesis and optional external providers.

115 days agoVision, Voice & MultimodalGPL-3.0
M

dsh-minimax-asr

moluyao/dsh-minimax-asr

MiniMax 语音识别(asr-1.0)+ 语音合成(speech-2.8-hd)的 DeepSeek Harness 全局插件:转写工具、任务结束后用喇叭念一句的播报、实时免手对话、303 个音色可选

116 days agoVision, Voice & MultimodalMIT
G

dsh-messenger-gateway

goodandready/dsh-messenger-gateway

Telegram bridge for DeepSeek Harness: bidirectional sessions, conversation steering, home-channel routing, and TTS voice notes sent as Telegram audio messages.

14 days agoIntegrations & RemoteMIT
A

dsh-audio-copilot

ai-yucheng/dsh-audio-copilot

Audio Copilot for DeepSeek Harness: transcribe audio (ASR) and synthesize speech (TTS) — gives text-only agents ears and a voice. Windows-local SAPI TTS out of the box; OpenAI-compatible ASR/TTS endpoints configurable. Includes an in-composer voice-input

1last monthVision, Voice & MultimodalMIT
Z

dsh-voice

zhuiyueya/dsh-voice

Voice for DeepSeek Harness(dsh) — speech-to-text input + read-aloud TTS for text-only DeepSeek, zero API key.

12 months agoVision, Voice & MultimodalMIT
J

sh-volume-knob

jianghu-lao-yao/sh-volume-knob

Speaker button beside the composer microphone — one click scrolls to the start of your newest question, marks it with a blinking caret and reads from there through the newest agent reply (dsh-tts, browser voice as fallback); press-and-drag-right picks any other reading start position on the page, press-and-drag-up opens a vertical mixer for in-page media volume and system output volume.

05 days agoVision, Voice & MultimodalMIT
L

dsh-status-chime

lijiawei255/dsh-status-chime

Speaks a short status line when DSH changes state — turn done, turn error, job done, job failed, goal complete, goal blocked, an approval waiting, or the agent needs your answer — using eight pre-recorded clips in Chinese (default) and English. The audio is played by the host process through ffplay, or through the Windows PowerShell player that ships with the OS, so it does not depend on a Web Audio tab. Windows only.

05 days agoIntegrations & Remote
F

dsh-say

fangqian616/dsh-say

Speaks your agent's reports in a voice you choose, compressing long reports before speaking. Character voices need no training - a 3-10 second reference clip clones one, and community-trained models work too - and no 6.4 GB GPT-SoVITS install: the plugin installs its own runtime and voice. An existing GPT-SoVITS can be used instead, and it is the same voice model.

09 days agoVision, Voice & Multimodal
C

dsh-read-aloud

cccc12138/dsh-read-aloud

Adds a speaker button immediately right of the Like button on every finalized assistant reply; click it to hear the reply through the browser speech engine, and hover it for a speed and voice panel.

012 days agoVision, Voice & MultimodalMIT
C

dsh-voice-mimo

ch1bug/dsh-voice-mimo

Xiaomi MiMo-powered voice for DeepSeek Harness: browser 🎤/🧠/🔊 UI, voice_transcribe/voice_understand/voice_speak tools, configurable voice map (preset/voicedesign/voiceclone). Fork of zhuiyueya/dsh-voice (MIT), Settings pattern from Anionex/dsh-vision-toolkit (MIT).

016 days agoVision, Voice & MultimodalMIT
2

dsh-tts-flash

2021heei/dsh-tts-flash

Reads AI replies aloud as they stream, with voiced waiting phrases while the model thinks. Edge TTS built in, any OpenAI-compatible cloud engine supported.

020 days agoVision, Voice & MultimodalMIT
C

dsh-live2d-avatar

clown139880/dsh-live2d-avatar

Live2D avatar stage and desktop companion for DSH: bundled Haru sample, custom Cubism 2/3+ model loading with scale and position controls, a draggable in-page pet, an optional transparent always-on-top desktop pet window, per-conversation expression prompt control, and opt-in voice with self-hosted ASR/TTS.

02 days agoUI EnhancementsMIT
A

dsh-reelsmaker

aayan-cloud/dsh-reelsmaker

DeepSeek Harness plugin: turn lines of narration into a finished vertical reel. Free neural voice-over, burned-in captions, no API keys.

025 days agoVision, Voice & Multimodal
D

dsh-xiaoshuang (plugin)

daixin315/dsh-xiaoshuang

XiaoShuang desktop pet for DSH: six-layer memory, persona injection, mood-driven video avatar, manual emotion menu, and TTS replies.

025 days agoJust for Fun
K

dsh-kitt-voice

kittcat-lab/dsh-kitt-voice

Voice for the DeepSeek Harness: speak to the agent, hear it back, and see what it is doing from a floating window that stays on top of whatever you are running.

025 days agoVision, Voice & MultimodalMIT