Plugins
Browse, filter, and install DeepSeek-Harness plugins.
76 plugins found
watch-skill
oxbshw/watch-skill
DeepWatch's capabilities, installable into an existing DeepSeek Harness profile
dsh-omi-voice
polinnizhong/dsh-omi-voice
In-chat read-aloud for DeepSeek Harness: tap to read, pause and resume AI replies with natural Doubao TTS voices (BYOK), reading only the final answer with code, tables and diagrams filtered; local engine, plugin keeps no API key.
dsh-voice-scribe
pensivefei/dsh-voice-scribe
Voice input for the web UI: tap Alt (or Alt+Space) to start/stop dictation, browser Web Speech (zero-config) or OpenAI-compatible cloud ASR, optional polish through DSH-configured LLM, settings UI.
dsh-live2d-pets
cyanfish-x/dsh-live2d-pets
Live2D desk pet overlay for the DSH web GUI: mirrors agent state (thinking, idle, error, done, awaiting approval) with motions and speech bubbles, per-part touch reactions, drag-to-dock, switchable personas, and custom models via URL or local path.
dsh-ears
wiziscool/dsh-ears
Voice input plugin for DeepSeek Harness (dsh): a microphone button in the composer turns speech into a draft transcript, with a choice of speech-recognition backends, optional polish through dsh own LLM routes, and a native settings page.
dsh-plugin-tts
1624318455/dsh-plugin-tts
Reads assistant replies aloud via free Edge TTS or your own RVC voice models: read-aloud buttons + auto-read, adaptive chunked progressive playback (gapless long reads), one-click voice-pack installs from a registry, and a portable RVC runtime.
dsh-whale-girl-live2d
andersen216/dsh-whale-girl-live2d
A Live2D whale-girl companion for the DeepSeek Harness Web UI: the model stays on screen and reacts to what the agent is actually doing (thinking, tool calls, streaming text, errors, turn completion); click her to send the agent a message and watch the reply stream into a speech bubble; right-click for a menu of expressions, dress-up items, desk scenes and one-shot animations, every action mapped from the original model author hotkey sheet so nothing overlaps and everything expires on its own. Non-commercial model assets (CC BY-NC-SA 4.0), MIT code, by Andersen216 (https://github.com/Andersen216).
dsh-talk
perrylink/dsh-talk
Voice I/O for DeepSeek Harness — speech-to-text and text-to-speech over the microphone and audio output.
dsh-voice-mode (dsh-voice-mode)
qishuilalala/dsh-voice-mode
Full-duplex voice mode for the DeepSeek Harness Web UI: toggle (2s-pause auto-send) or hold-to-talk dictation into an editable draft with zipformer2 streaming ASR, optional wake word; sentence-by-sentence Edge TTS read-aloud with live captions, and speaking interrupts playback and the running turn (true barge-in). On-device ASR, no API key.
dsh-plugin-xiaomi-mimo-tts
ppy-web/dsh-plugin-xiaomi-mimo-tts
Adds Xiaomi MiMo text-to-speech to DSH Web with assistant-message read-aloud, PCM streaming, preset and custom voice design, browser speech fallback, playback controls, and optional UI sounds.
dsh-voice
3274375092/dsh-voice
Voice input for DeepSeek Harness: speak into the microphone and the recognized text is submitted as a normal chat message, via local or browser speech recognition.
dsh-web-whale-maid
acidgr/dsh-web-whale-maid
Anime maid whale desktop pet for the DeepSeek Harness (dsh) Web UI: real-time LLM dialogues, 8 mood animations, satiety reserve system, persistent speech bubble, and mobile-adapted cupboard UI.
dsh-iris
mokuyoaxis/dsh-iris
Media and vision workspace for DeepSeek Harness: image, video and speech generation, image Q&A and element locating, long-image OCR, pixel diff, HTML-screenshot verification and video summarization, with DashScope and OpenAI-compatible providers, model pools and a workbench client.
dsh-alert-sound
machine-126/dsh-alert-sound
Audio and voice alerts for the dsh web GUI: distinct tones and optional speech for approvals, questions, completions and errors.
adhdgofly-dsh-ext
zuoguyoupan2023/adhdgofly-dsh-ext
Highlights parts of speech (nouns green, verbs red, adjectives/adverbs purple) in rendered DSH Web Markdown, with light/dark palettes, toggles, and stream-aware updates.
liang-desktop-pet (liang-harness-plugin)
flycat43/liang-desktop-pet
A draggable character companion for the DeepSeek Harness web profile with six visual and expression levels, four task modes, current-session chat, public reasoning summaries and browser speech synthesis.
dsh-qq-onebot-bridge
cheesehaqi/dsh-qq-onebot-bridge
Bidirectional QQ bridge over OneBot v11 (reverse WebSocket): per-group and per-private-chat sessions, @-quoted voice speech-to-text, private image/animated-sticker viewing, face and sticker tools.
dsh-voice-input
0nt-one/dsh-voice-input
Mic button in the composer tool row: Web Speech API speech-to-text (Chrome/Edge), language switching, and optional auto-send, zero dependencies.
dsh-guide-dog
atropinoltt/dsh-guide-dog
MiniMax-powered multimodal plugin: real-time voice call mode (streaming conversation, floating dock UI), voice mode and mic voice input, plus image/video/music/speech generation and vision inspection tools.
Herta-dsh
yunmengyuan/herta-dsh
Installs Herta as a DeepSeek Harness agent: a preset carrying her persona and memory prompt sections, six memory and voice tools, a narrative layer that separates her thinking from what she says, 80 bundled voice clips, offline and MiniMax speech synthesis, and two conversation views that render the session with her own components.
dsh-tts-bridge
yuuyuko-uu/dsh-tts-bridge
Reads DSH conversations aloud using the DeepSeek web page's built-in read-aloud, driven by a small browser extension.
dsh-voice
stardustlc666/dsh-voice
Voice tools: free edge-tts neural speech synthesis, OpenAI-compatible ASR transcription, voice list, batch voice preview and health self-check.
dsh-chatvoice
fuzzysoul/dsh-chatvoice
Free voice closed loop for the Web UI: browser SpeechRecognition mic input with live interim results plus read-aloud speaker buttons and auto-read for assistant replies, zero configuration and no API key.
dsh-omni-workstation
huashenglian/dsh-omni-workstation
Omni-modal workstation for DSH: analyze_image over an ordered multi-card VLM failover chain, a six-tool vision toolkit (zoom, colour sampling, pixel diff, OCR, element detection, inline display) sharing the same image resolver, generate_image over OpenAI/DashScope/ComfyUI protocols, multi-card async generate_video with a /build-video-tool builder, and speak/clone_voice TTS over seven providers, all driven by one auto-saving settings page.