Pular para o conteúdo principal

Plugins

Navegue, filtre e instale plugins do DeepSeek-Harness.

56 plugins encontrados

N

insta360-ai-content-studio

nicecx/insta360-ai-content-studio

Pipeline diário automatizado de conteúdo para Insta360 GO 3S: captura remota BLE/WiFi de filmagens de aquário, importação sem fio, avaliação de qualidade ffmpeg com registro de decisões, edição automática (seleção de segmentos, narração TTS, música, legendas) e upload no Bilibili via biliup (dry-run por padrão).

2mês passadoFerramentas e funçõesMIT
G

dsh-messenger-gateway

goodandready/dsh-messenger-gateway

Ponte Telegram para o DeepSeek Harness: sessões bidirecionais, direcionamento de conversas, roteamento por canal principal e notas de voz TTS enviadas como mensagens de áudio do Telegram.

2há 3 diasIntegrações e acesso remotoMIT
M

dsh-voice-chat

maoyuching/dsh-voice-chat

Chat de voz estilo Doubao para DSH: segurar o microfone do compositor converte a fala em texto e envia automaticamente, e as respostas da IA são lidas em voz alta. Condensação opcional por LLM (respostas longas resumidas em falas curtas, seguindo o modelo da conversa atual), limpeza de texto amigável ao TTS, vozes Edge TTS selecionáveis, parada automática por silêncio ajustável e configurações embutidas no diálogo de configurações do DSH.

2há 8 diasVisão, voz e multimodalMIT
G

dsh-tts

goodandready/dsh-tts

Fala as respostas do agente na interface web do DeepSeek Harness por meio de uma cadeia de fallback de provedores (OpenAI, ElevenLabs, Google, Azure, Groq, Deepgram, OpenRouter, Edge, Piper, eSpeak), para que um provedor com falha ou com limite de taxa passe para o próximo em vez de ficar em silêncio.

2há 6 diasVisão, voz e multimodalMIT
X

dsh-showreel

xiaoyuer3921/dsh-showreel

Transforma o último turno concluído de uma sessão DSH em um vídeo vertical de resumo compartilhável: storyboard editável de cinco cenas, redação de privacidade no host, aprimoramento opcional por IA e narração TTS, exportação MP4/WebM com capa PNG.

2mês passadoMelhorias de UIMIT
P

muxiva-dsh-voice

piyotahu/muxiva-dsh-voice

Voz full-duplex local-first para DeepSeek Harness, orquestrada por Muxiva

2há 2 mesesVisão, voz e multimodalApache-2.0
P

dsh-voice-call

pandapolo/dsh-voice-call

Chamadas de voz iniciadas pelo agente: `offer_call` faz o humano tocar (接听/拒接/稍后再说); as chamadas aceitas são sintetizadas e reproduzidas localmente com CrispASR + Qwen3-TTS (9 vozes, 2 dialetos chineses), e as rejeitadas devolvem a decisão ao agente.

2há 3 diasVisão, voz e multimodalMIT
M

dsh-fish-tts

mari23333/dsh-fish-tts

Lê em voz alta as respostas do assistente apenas via a API Fish Audio (traga sua própria chave): leitura por mensagem, alternância de leitura automática e uma página de configurações para modelo, reference_id da voz, chave de API criptografada e proxy.

2anteontemVisão, voz e multimodalMIT
J

dsh-live2d-voice

john-walks-slow/dsh-live2d-voice

Live2D session view with realtime voice: continuous voice input, streaming TTS, subtitle translation, multi-model catalog, standalone entry / Live2D 实时语音会话视图:连续语音输入、流式 TTS、字幕翻译、多模型目录、独立入口

1há 4 diasVisão, voz e multimodalMIT
W

dsh-multi-tts

wyr-233/dsh-multi-tts

Per-reply read-aloud with a multi-provider settings page — MiniMax or any OpenAI-compatible /audio/speech endpoint, with voice, emotion, speed and model selection plus an auto-read toggle.

1há 21 diasVisão, voz e multimodal
V

dsh-live-voice

victorwads/dsh-live-voice

Local-first voice conversations for DSH, with local speech recognition and synthesis and optional external providers.

1há 18 diasVisão, voz e multimodalGPL-3.0
M

dsh-minimax-asr

moluyao/dsh-minimax-asr

MiniMax 语音识别(asr-1.0)+ 语音合成(speech-2.8-hd)的 DeepSeek Harness 全局插件:转写工具、任务结束后用喇叭念一句的播报、实时免手对话、303 个音色可选

1há 19 diasVisão, voz e multimodalMIT
A

dsh-audio-copilot

ai-yucheng/dsh-audio-copilot

Audio Copilot for DeepSeek Harness: transcribe audio (ASR) and synthesize speech (TTS) — gives text-only agents ears and a voice. Windows-local SAPI TTS out of the box; OpenAI-compatible ASR/TTS endpoints configurable. Includes an in-composer voice-input

1mês passadoVisão, voz e multimodalMIT
Z

dsh-voice

zhuiyueya/dsh-voice

Voz para o DeepSeek Harness (dsh) — entrada por fala para texto + leitura em voz alta via TTS para o DeepSeek somente de texto, sem necessidade de chave de API

1há 2 mesesVisão, voz e multimodalMIT
J

sh-volume-knob

jianghu-lao-yao/sh-volume-knob

Speaker button beside the composer microphone — one click scrolls to the start of your newest question, marks it with a blinking caret and reads from there through the newest agent reply (dsh-tts, browser voice as fallback); press-and-drag-right picks any other reading start position on the page, press-and-drag-up opens a vertical mixer for in-page media volume and system output volume.

0há 8 diasVisão, voz e multimodalMIT
L

dsh-status-chime

lijiawei255/dsh-status-chime

Speaks a short status line when DSH changes state — turn done, turn error, job done, job failed, goal complete, goal blocked, an approval waiting, or the agent needs your answer — using eight pre-recorded clips in Chinese (default) and English. The audio is played by the host process through ffplay, or through the Windows PowerShell player that ships with the OS, so it does not depend on a Web Audio tab. Windows only.

0há 8 diasIntegrações e acesso remoto
F

dsh-say

fangqian616/dsh-say

Speaks your agent's reports in a voice you choose, compressing long reports before speaking. Character voices need no training - a 3-10 second reference clip clones one, and community-trained models work too - and no 6.4 GB GPT-SoVITS install: the plugin installs its own runtime and voice. An existing GPT-SoVITS can be used instead, and it is the same voice model.

0há 11 diasVisão, voz e multimodal
C

dsh-read-aloud

cccc12138/dsh-read-aloud

Adds a speaker button immediately right of the Like button on every finalized assistant reply; click it to hear the reply through the browser speech engine, and hover it for a speed and voice panel.

0há 15 diasVisão, voz e multimodalMIT
C

dsh-voice-mimo

ch1bug/dsh-voice-mimo

Xiaomi MiMo-powered voice for DeepSeek Harness: browser 🎤/🧠/🔊 UI, voice_transcribe/voice_understand/voice_speak tools, configurable voice map (preset/voicedesign/voiceclone). Fork of zhuiyueya/dsh-voice (MIT), Settings pattern from Anionex/dsh-vision-toolkit (MIT).

0há 19 diasVisão, voz e multimodalMIT
2

dsh-tts-flash

2021heei/dsh-tts-flash

Reads AI replies aloud as they stream, with voiced waiting phrases while the model thinks. Edge TTS built in, any OpenAI-compatible cloud engine supported.

0há 23 diasVisão, voz e multimodalMIT
C

dsh-live2d-avatar

clown139880/dsh-live2d-avatar

Live2D avatar stage and desktop companion for DSH: bundled Haru sample, custom Cubism 2/3+ model loading with scale and position controls, a draggable in-page pet, an optional transparent always-on-top desktop pet window, per-conversation expression prompt control, and opt-in voice with self-hosted ASR/TTS.

0há 4 diasMelhorias de UIMIT
A

dsh-reelsmaker

aayan-cloud/dsh-reelsmaker

DeepSeek Harness plugin: turn lines of narration into a finished vertical reel. Free neural voice-over, burned-in captions, no API keys.

0há 28 diasVisão, voz e multimodal
D

dsh-xiaoshuang (plugin)

daixin315/dsh-xiaoshuang

XiaoShuang desktop pet for DSH: six-layer memory, persona injection, mood-driven video avatar, manual emotion menu, and TTS replies.

0há 28 diasSó por diversão
K

dsh-kitt-voice

kittcat-lab/dsh-kitt-voice

Voice for the DeepSeek Harness: speak to the agent, hear it back, and see what it is doing from a floating window that stays on top of whatever you are running.

0há 28 diasVisão, voz e multimodalMIT