본문으로 건너뛰기

플러그인

DeepSeek-Harness 플러그인을 찾아보고, 필터링하고, 설치하세요.

플러그인 56개 찾음

N

insta360-ai-content-studio

nicecx/insta360-ai-content-studio

Insta360 GO 3S용 자동 일일 콘텐츠 파이프라인: 수족관 영상의 BLE/WiFi 원격 촬영, 무선 가져오기, 의사결정 로그가 있는 ffmpeg 품질 평가, 자동 편집(세그먼트 선택, TTS 내레이션, 음악, 자막), biliup을 통한 Bilibili 업로드(기본값은 드라이런).

2지난달도구 및 기능MIT
G

dsh-messenger-gateway

goodandready/dsh-messenger-gateway

DeepSeek Harness용 Telegram 브리지: 양방향 세션, 대화 스티어링, 홈 채널 라우팅, 그리고 Telegram 오디오 메시지로 전송되는 TTS 음성 메모.

23일 전통합 및 원격MIT
M

dsh-voice-chat

maoyuching/dsh-voice-chat

DSH용 Doubao 스타일 음성 채팅. 작성란의 마이크를 길게 누르면 음성을 텍스트로 변환해 자동 전송하고, AI 응답은 소리 내어 읽어줍니다. 선택적 LLM 축약(긴 응답을 현재 대화 모델을 따르는 짧은 음성 문장으로 요약), TTS 친화적 텍스트 정리, 선택 가능한 Edge TTS 음성, 조절 가능한 무음 자동 중지를 지원하며 설정은 DSH 설정 대화상자에 내장되어 있습니다.

28일 전비전, 음성 및 멀티모달MIT
G

dsh-tts

goodandready/dsh-tts

DeepSeek Harness 웹 UI에서 에이전트 응답을 공급자 폴백 체인(OpenAI, ElevenLabs, Google, Azure, Groq, Deepgram, OpenRouter, Edge, Piper, eSpeak)을 통해 읽어줍니다. 실패하거나 속도 제한된 공급자는 침묵하는 대신 다음 공급자로 넘어갑니다.

26일 전비전, 음성 및 멀티모달MIT
X

dsh-showreel

xiaoyuer3921/dsh-showreel

가장 최근에 완료된 DSH 세션 턴을 공유 가능한 세로형 요약 비디오로 변환합니다. 편집 가능한 5장면 스토리보드, 호스트 측 개인정보 가리기, 선택적 AI 보정 및 TTS 내레이션, PNG 커버를 포함한 MP4/WebM 내보내기를 제공합니다.

2지난달UI 확장MIT
P

muxiva-dsh-voice

piyotahu/muxiva-dsh-voice

Muxiva가 오케스트레이션하는 DeepSeek Harness용 로컬 우선 양방향 음성

22개월 전비전, 음성 및 멀티모달Apache-2.0
P

dsh-voice-call

pandapolo/dsh-voice-call

에이전트가 시작하는 음성 통화: `offer_call`이 사람에게 전화를 울리고(接听/拒接/稍后再说), 수락된 통화는 CrispASR + Qwen3-TTS(화자 9명, 중국어 방언 2개)로 로컬에서 합성·재생되며, 거절된 통화는 그 결정을 에이전트에 돌려줍니다.

23일 전비전, 음성 및 멀티모달MIT
M

dsh-fish-tts

mari23333/dsh-fish-tts

Fish Audio API만 사용해 어시스턴트 답변을 음성으로 읽어 줍니다(키는 직접 준비): 메시지별 읽어주기, 자동 읽기 토글, 모델·음성 reference_id·암호화된 API 키·프록시를 설정하는 페이지를 제공합니다.

2그저께비전, 음성 및 멀티모달MIT
J

dsh-live2d-voice

john-walks-slow/dsh-live2d-voice

Live2D session view with realtime voice: continuous voice input, streaming TTS, subtitle translation, multi-model catalog, standalone entry / Live2D 实时语音会话视图:连续语音输入、流式 TTS、字幕翻译、多模型目录、独立入口

14일 전비전, 음성 및 멀티모달MIT
W

dsh-multi-tts

wyr-233/dsh-multi-tts

Per-reply read-aloud with a multi-provider settings page — MiniMax or any OpenAI-compatible /audio/speech endpoint, with voice, emotion, speed and model selection plus an auto-read toggle.

121일 전비전, 음성 및 멀티모달
V

dsh-live-voice

victorwads/dsh-live-voice

Local-first voice conversations for DSH, with local speech recognition and synthesis and optional external providers.

118일 전비전, 음성 및 멀티모달GPL-3.0
M

dsh-minimax-asr

moluyao/dsh-minimax-asr

MiniMax 语音识别(asr-1.0)+ 语音合成(speech-2.8-hd)的 DeepSeek Harness 全局插件:转写工具、任务结束后用喇叭念一句的播报、实时免手对话、303 个音色可选

119일 전비전, 음성 및 멀티모달MIT
A

dsh-audio-copilot

ai-yucheng/dsh-audio-copilot

Audio Copilot for DeepSeek Harness: transcribe audio (ASR) and synthesize speech (TTS) — gives text-only agents ears and a voice. Windows-local SAPI TTS out of the box; OpenAI-compatible ASR/TTS endpoints configurable. Includes an in-composer voice-input

1지난달비전, 음성 및 멀티모달MIT
Z

dsh-voice

zhuiyueya/dsh-voice

DeepSeek Harness(dsh)용 음성 기능 — 텍스트 전용 DeepSeek을 위한 음성-텍스트 입력 + TTS 읽어주기, API 키 불필요

12개월 전비전, 음성 및 멀티모달MIT
J

sh-volume-knob

jianghu-lao-yao/sh-volume-knob

Speaker button beside the composer microphone — one click scrolls to the start of your newest question, marks it with a blinking caret and reads from there through the newest agent reply (dsh-tts, browser voice as fallback); press-and-drag-right picks any other reading start position on the page, press-and-drag-up opens a vertical mixer for in-page media volume and system output volume.

08일 전비전, 음성 및 멀티모달MIT
L

dsh-status-chime

lijiawei255/dsh-status-chime

Speaks a short status line when DSH changes state — turn done, turn error, job done, job failed, goal complete, goal blocked, an approval waiting, or the agent needs your answer — using eight pre-recorded clips in Chinese (default) and English. The audio is played by the host process through ffplay, or through the Windows PowerShell player that ships with the OS, so it does not depend on a Web Audio tab. Windows only.

08일 전통합 및 원격
F

dsh-say

fangqian616/dsh-say

Speaks your agent's reports in a voice you choose, compressing long reports before speaking. Character voices need no training - a 3-10 second reference clip clones one, and community-trained models work too - and no 6.4 GB GPT-SoVITS install: the plugin installs its own runtime and voice. An existing GPT-SoVITS can be used instead, and it is the same voice model.

011일 전비전, 음성 및 멀티모달
C

dsh-read-aloud

cccc12138/dsh-read-aloud

Adds a speaker button immediately right of the Like button on every finalized assistant reply; click it to hear the reply through the browser speech engine, and hover it for a speed and voice panel.

015일 전비전, 음성 및 멀티모달MIT
C

dsh-voice-mimo

ch1bug/dsh-voice-mimo

Xiaomi MiMo-powered voice for DeepSeek Harness: browser 🎤/🧠/🔊 UI, voice_transcribe/voice_understand/voice_speak tools, configurable voice map (preset/voicedesign/voiceclone). Fork of zhuiyueya/dsh-voice (MIT), Settings pattern from Anionex/dsh-vision-toolkit (MIT).

019일 전비전, 음성 및 멀티모달MIT
2

dsh-tts-flash

2021heei/dsh-tts-flash

Reads AI replies aloud as they stream, with voiced waiting phrases while the model thinks. Edge TTS built in, any OpenAI-compatible cloud engine supported.

023일 전비전, 음성 및 멀티모달MIT
C

dsh-live2d-avatar

clown139880/dsh-live2d-avatar

Live2D avatar stage and desktop companion for DSH: bundled Haru sample, custom Cubism 2/3+ model loading with scale and position controls, a draggable in-page pet, an optional transparent always-on-top desktop pet window, per-conversation expression prompt control, and opt-in voice with self-hosted ASR/TTS.

04일 전UI 확장MIT
A

dsh-reelsmaker

aayan-cloud/dsh-reelsmaker

DeepSeek Harness plugin: turn lines of narration into a finished vertical reel. Free neural voice-over, burned-in captions, no API keys.

028일 전비전, 음성 및 멀티모달
D

dsh-xiaoshuang (plugin)

daixin315/dsh-xiaoshuang

XiaoShuang desktop pet for DSH: six-layer memory, persona injection, mood-driven video avatar, manual emotion menu, and TTS replies.

028일 전재미로 즐기기
K

dsh-kitt-voice

kittcat-lab/dsh-kitt-voice

Voice for the DeepSeek Harness: speak to the agent, hear it back, and see what it is doing from a floating window that stays on top of whatever you are running.

028일 전비전, 음성 및 멀티모달MIT