Zum Hauptinhalt springen

Plugins

Durchsuche, filtere und installiere DeepSeek-Harness-Plugins.

56 Plugins gefunden

N

insta360-ai-content-studio

nicecx/insta360-ai-content-studio

Automatisierte tägliche Content-Pipeline für Insta360 GO 3S: BLE/WiFi-Fernaufnahme von Aquariumaufnahmen, drahtloser Import, ffmpeg-Qualitätsbewertung mit Entscheidungsprotokoll, automatische Bearbeitung (Segmentauswahl, TTS-Erzählung, Musik, Untertitel) und Bilibili-Upload über biliup (standardmäßig Trockenlauf).

2letzten MonatTools & FunktionenMIT
G

dsh-messenger-gateway

goodandready/dsh-messenger-gateway

Telegram-Brücke für DeepSeek Harness: bidirektionale Sessions, Steuerung von Konversationen, Home-Channel-Routing und TTS-Sprachnachrichten, die als Telegram-Audionachrichten gesendet werden.

2vor 3 TagenIntegrationen & FernzugriffMIT
M

dsh-voice-chat

maoyuching/dsh-voice-chat

Sprach-Chat im Doubao-Stil für DSH: Gedrückthalten des Mikrofons im Composer wandelt Sprache in Text um und sendet automatisch, und KI-Antworten werden vorgelesen. Optionale LLM-Verdichtung (lange Antworten werden zu kurzen gesprochenen Zeilen zusammengefasst, gemäß dem Modell der aktuellen Konversation), TTS-freundliche Textbereinigung, wählbare Edge-TTS-Stimmen, einstellbares automatisches Stoppen bei Stille, und Einstellungen direkt in den DSH-Einstellungsdialog eingebettet.

2vor 8 TagenVision, Sprache & MultimodalMIT
G

dsh-tts

goodandready/dsh-tts

Spricht die Antworten des Agenten in der DeepSeek Harness Web-UI über eine Anbieter-Fallback-Kette (OpenAI, ElevenLabs, Google, Azure, Groq, Deepgram, OpenRouter, Edge, Piper, eSpeak), sodass ein fehlgeschlagener oder ratenbegrenzter Anbieter auf den nächsten übergeht, statt zu schweigen.

2vor 6 TagenVision, Sprache & MultimodalMIT
X

dsh-showreel

xiaoyuer3921/dsh-showreel

Wandelt den zuletzt abgeschlossenen DSH-Sitzungszug in ein teilbares vertikales Rückblickvideo um: bearbeitbares Storyboard mit fünf Szenen, datenschutzgerechte Redigierung auf dem Host, optionales KI-Finishing und TTS-Erzählung, MP4/WebM-Export mit PNG-Cover.

2letzten MonatUI-ErweiterungenMIT
P

muxiva-dsh-voice

piyotahu/muxiva-dsh-voice

Local-first-Vollduplex-Sprache für DeepSeek Harness, orchestriert von Muxiva

2vor 2 MonatenVision, Sprache & MultimodalApache-2.0
P

dsh-voice-call

pandapolo/dsh-voice-call

Vom Agenten initiierte Sprachanrufe: `offer_call` lässt beim Menschen klingeln (接听/拒接/稍后再说); angenommene Anrufe werden lokal mit CrispASR + Qwen3-TTS (9 Sprecher, 2 chinesische Dialekte) synthetisiert und abgespielt, abgelehnte Anrufe geben die Entscheidung an den Agenten zurück.

2vor 3 TagenVision, Sprache & MultimodalMIT
M

dsh-fish-tts

mari23333/dsh-fish-tts

Liest Assistentenantworten ausschließlich über die Fish Audio API vor (eigenen Schlüssel mitbringen): Vorlesen pro Nachricht, Auto-Read-Schalter und eine Einstellungsseite für Modell, Voice reference_id, verschlüsselten API-Schlüssel und Proxy.

2vorgesternVision, Sprache & MultimodalMIT
J

dsh-live2d-voice

john-walks-slow/dsh-live2d-voice

Live2D session view with realtime voice: continuous voice input, streaming TTS, subtitle translation, multi-model catalog, standalone entry / Live2D 实时语音会话视图:连续语音输入、流式 TTS、字幕翻译、多模型目录、独立入口

1vor 4 TagenVision, Sprache & MultimodalMIT
W

dsh-multi-tts

wyr-233/dsh-multi-tts

Per-reply read-aloud with a multi-provider settings page — MiniMax or any OpenAI-compatible /audio/speech endpoint, with voice, emotion, speed and model selection plus an auto-read toggle.

1vor 21 TagenVision, Sprache & Multimodal
V

dsh-live-voice

victorwads/dsh-live-voice

Local-first voice conversations for DSH, with local speech recognition and synthesis and optional external providers.

1vor 18 TagenVision, Sprache & MultimodalGPL-3.0
M

dsh-minimax-asr

moluyao/dsh-minimax-asr

MiniMax 语音识别(asr-1.0)+ 语音合成(speech-2.8-hd)的 DeepSeek Harness 全局插件:转写工具、任务结束后用喇叭念一句的播报、实时免手对话、303 个音色可选

1vor 19 TagenVision, Sprache & MultimodalMIT
A

dsh-audio-copilot

ai-yucheng/dsh-audio-copilot

Audio Copilot for DeepSeek Harness: transcribe audio (ASR) and synthesize speech (TTS) — gives text-only agents ears and a voice. Windows-local SAPI TTS out of the box; OpenAI-compatible ASR/TTS endpoints configurable. Includes an in-composer voice-input

1letzten MonatVision, Sprache & MultimodalMIT
Z

dsh-voice

zhuiyueya/dsh-voice

Sprache für DeepSeek Harness (dsh) — Sprache-zu-Text-Eingabe + Vorlese-TTS für rein textbasiertes DeepSeek, ohne API-Schlüssel.

1vor 2 MonatenVision, Sprache & MultimodalMIT
J

sh-volume-knob

jianghu-lao-yao/sh-volume-knob

Speaker button beside the composer microphone — one click scrolls to the start of your newest question, marks it with a blinking caret and reads from there through the newest agent reply (dsh-tts, browser voice as fallback); press-and-drag-right picks any other reading start position on the page, press-and-drag-up opens a vertical mixer for in-page media volume and system output volume.

0vor 8 TagenVision, Sprache & MultimodalMIT
L

dsh-status-chime

lijiawei255/dsh-status-chime

Speaks a short status line when DSH changes state — turn done, turn error, job done, job failed, goal complete, goal blocked, an approval waiting, or the agent needs your answer — using eight pre-recorded clips in Chinese (default) and English. The audio is played by the host process through ffplay, or through the Windows PowerShell player that ships with the OS, so it does not depend on a Web Audio tab. Windows only.

0vor 8 TagenIntegrationen & Fernzugriff
F

dsh-say

fangqian616/dsh-say

Speaks your agent's reports in a voice you choose, compressing long reports before speaking. Character voices need no training - a 3-10 second reference clip clones one, and community-trained models work too - and no 6.4 GB GPT-SoVITS install: the plugin installs its own runtime and voice. An existing GPT-SoVITS can be used instead, and it is the same voice model.

0vor 11 TagenVision, Sprache & Multimodal
C

dsh-read-aloud

cccc12138/dsh-read-aloud

Adds a speaker button immediately right of the Like button on every finalized assistant reply; click it to hear the reply through the browser speech engine, and hover it for a speed and voice panel.

0vor 15 TagenVision, Sprache & MultimodalMIT
C

dsh-voice-mimo

ch1bug/dsh-voice-mimo

Xiaomi MiMo-powered voice for DeepSeek Harness: browser 🎤/🧠/🔊 UI, voice_transcribe/voice_understand/voice_speak tools, configurable voice map (preset/voicedesign/voiceclone). Fork of zhuiyueya/dsh-voice (MIT), Settings pattern from Anionex/dsh-vision-toolkit (MIT).

0vor 19 TagenVision, Sprache & MultimodalMIT
2

dsh-tts-flash

2021heei/dsh-tts-flash

Reads AI replies aloud as they stream, with voiced waiting phrases while the model thinks. Edge TTS built in, any OpenAI-compatible cloud engine supported.

0vor 23 TagenVision, Sprache & MultimodalMIT
C

dsh-live2d-avatar

clown139880/dsh-live2d-avatar

Live2D avatar stage and desktop companion for DSH: bundled Haru sample, custom Cubism 2/3+ model loading with scale and position controls, a draggable in-page pet, an optional transparent always-on-top desktop pet window, per-conversation expression prompt control, and opt-in voice with self-hosted ASR/TTS.

0vor 4 TagenUI-ErweiterungenMIT
A

dsh-reelsmaker

aayan-cloud/dsh-reelsmaker

DeepSeek Harness plugin: turn lines of narration into a finished vertical reel. Free neural voice-over, burned-in captions, no API keys.

0vor 28 TagenVision, Sprache & Multimodal
D

dsh-xiaoshuang (plugin)

daixin315/dsh-xiaoshuang

XiaoShuang desktop pet for DSH: six-layer memory, persona injection, mood-driven video avatar, manual emotion menu, and TTS replies.

0vor 28 TagenNur zum Spaß
K

dsh-kitt-voice

kittcat-lab/dsh-kitt-voice

Voice for the DeepSeek Harness: speak to the agent, hear it back, and see what it is doing from a floating window that stays on top of whatever you are running.

0vor 28 TagenVision, Sprache & MultimodalMIT