メインコンテンツへスキップ

プラグイン

DeepSeek-Harness プラグインを閲覧・絞り込み・インストールできます。

56 件のプラグインが見つかりました

N

insta360-ai-content-studio

nicecx/insta360-ai-content-studio

Insta360 GO 3S 向けの自動日次コンテンツパイプライン: 水槽映像の BLE/WiFi リモート撮影、無線取り込み、判断ログ付き ffmpeg 品質評価、自動編集(セグメント選択、TTS ナレーション、音楽、字幕)、および biliup による Bilibili アップロード(既定でドライラン)。

2先月ツールと機能MIT
G

dsh-messenger-gateway

goodandready/dsh-messenger-gateway

DeepSeek Harness向けのTelegramブリッジ:双方向セッション、会話のステアリング、ホームチャネルへのルーティング、およびTTS音声メモのTelegramオーディオメッセージとしての送信。

23 日前統合とリモートMIT
M

dsh-voice-chat

maoyuching/dsh-voice-chat

DSH 用の Doubao 風音声チャット。入力欄のマイクを長押しすると音声をテキストに変換して自動送信し、AI の返信を読み上げます。オプションの LLM 要約(長い返信を現在の会話モデルに従って短い読み上げ文に要約)、TTS に適したテキストクリーニング、選択可能な Edge TTS 音声、調整可能な無音自動停止に対応し、設定は DSH の設定ダイアログに組み込まれています。

28 日前ビジョン、音声とマルチモーダルMIT
G

dsh-tts

goodandready/dsh-tts

DeepSeek HarnessのWeb UIでエージェントの返答を、プロバイダーのフォールバックチェーン(OpenAI、ElevenLabs、Google、Azure、Groq、Deepgram、OpenRouter、Edge、Piper、eSpeak)を通じて読み上げます。これにより、失敗したりレート制限されたプロバイダーは無音になるのではなく次のプロバイダーに引き継がれます。

26 日前ビジョン、音声とマルチモーダルMIT
X

dsh-showreel

xiaoyuer3921/dsh-showreel

最新の完了済みDSHセッションターンを共有可能な縦型ダイジェスト動画に変換します。編集可能な5シーンのストーリーボード、ホスト側のプライバシー秘匿化、任意のAI仕上げとTTSナレーション、PNGカバー付きのMP4/WebMエクスポートを備えます。

2先月UI拡張MIT
P

muxiva-dsh-voice

piyotahu/muxiva-dsh-voice

Muxiva がオーケストレーションする、DeepSeek Harness 向けのローカルファースト・フルデュプレックス音声

22 か月前ビジョン、音声とマルチモーダルApache-2.0
P

dsh-voice-call

pandapolo/dsh-voice-call

エージェント起動の音声通話: `offer_call` が人間を呼び出し(接听/拒接/稍后再说)、応答された通話は CrispASR + Qwen3-TTS(9 話者、2 中国語方言)でローカル生成・再生され、拒否された通話はその判断をエージェントに返します。

23 日前ビジョン、音声とマルチモーダルMIT
M

dsh-fish-tts

mari23333/dsh-fish-tts

Fish Audio API のみを使ってアシスタントの返信を音声読み上げします(API キーは各自用意):メッセージごとの読み上げ、自動読み上げ切り替え、モデル・音声 reference_id・暗号化された API キー・プロキシを設定するページを備えます。

2一昨日ビジョン、音声とマルチモーダルMIT
J

dsh-live2d-voice

john-walks-slow/dsh-live2d-voice

Live2D session view with realtime voice: continuous voice input, streaming TTS, subtitle translation, multi-model catalog, standalone entry / Live2D 实时语音会话视图:连续语音输入、流式 TTS、字幕翻译、多模型目录、独立入口

14 日前ビジョン、音声とマルチモーダルMIT
W

dsh-multi-tts

wyr-233/dsh-multi-tts

Per-reply read-aloud with a multi-provider settings page — MiniMax or any OpenAI-compatible /audio/speech endpoint, with voice, emotion, speed and model selection plus an auto-read toggle.

121 日前ビジョン、音声とマルチモーダル
V

dsh-live-voice

victorwads/dsh-live-voice

Local-first voice conversations for DSH, with local speech recognition and synthesis and optional external providers.

118 日前ビジョン、音声とマルチモーダルGPL-3.0
M

dsh-minimax-asr

moluyao/dsh-minimax-asr

MiniMax 语音识别(asr-1.0)+ 语音合成(speech-2.8-hd)的 DeepSeek Harness 全局插件:转写工具、任务结束后用喇叭念一句的播报、实时免手对话、303 个音色可选

119 日前ビジョン、音声とマルチモーダルMIT
A

dsh-audio-copilot

ai-yucheng/dsh-audio-copilot

Audio Copilot for DeepSeek Harness: transcribe audio (ASR) and synthesize speech (TTS) — gives text-only agents ears and a voice. Windows-local SAPI TTS out of the box; OpenAI-compatible ASR/TTS endpoints configurable. Includes an in-composer voice-input

1先月ビジョン、音声とマルチモーダルMIT
Z

dsh-voice

zhuiyueya/dsh-voice

DeepSeek Harness(dsh)向けの音声機能——テキスト専用の DeepSeek に対する音声入力(STT)と読み上げ(TTS)を提供します。API キー不要。

12 か月前ビジョン、音声とマルチモーダルMIT
J

sh-volume-knob

jianghu-lao-yao/sh-volume-knob

Speaker button beside the composer microphone — one click scrolls to the start of your newest question, marks it with a blinking caret and reads from there through the newest agent reply (dsh-tts, browser voice as fallback); press-and-drag-right picks any other reading start position on the page, press-and-drag-up opens a vertical mixer for in-page media volume and system output volume.

08 日前ビジョン、音声とマルチモーダルMIT
L

dsh-status-chime

lijiawei255/dsh-status-chime

Speaks a short status line when DSH changes state — turn done, turn error, job done, job failed, goal complete, goal blocked, an approval waiting, or the agent needs your answer — using eight pre-recorded clips in Chinese (default) and English. The audio is played by the host process through ffplay, or through the Windows PowerShell player that ships with the OS, so it does not depend on a Web Audio tab. Windows only.

08 日前統合とリモート
F

dsh-say

fangqian616/dsh-say

Speaks your agent's reports in a voice you choose, compressing long reports before speaking. Character voices need no training - a 3-10 second reference clip clones one, and community-trained models work too - and no 6.4 GB GPT-SoVITS install: the plugin installs its own runtime and voice. An existing GPT-SoVITS can be used instead, and it is the same voice model.

011 日前ビジョン、音声とマルチモーダル
C

dsh-read-aloud

cccc12138/dsh-read-aloud

Adds a speaker button immediately right of the Like button on every finalized assistant reply; click it to hear the reply through the browser speech engine, and hover it for a speed and voice panel.

015 日前ビジョン、音声とマルチモーダルMIT
C

dsh-voice-mimo

ch1bug/dsh-voice-mimo

Xiaomi MiMo-powered voice for DeepSeek Harness: browser 🎤/🧠/🔊 UI, voice_transcribe/voice_understand/voice_speak tools, configurable voice map (preset/voicedesign/voiceclone). Fork of zhuiyueya/dsh-voice (MIT), Settings pattern from Anionex/dsh-vision-toolkit (MIT).

019 日前ビジョン、音声とマルチモーダルMIT
2

dsh-tts-flash

2021heei/dsh-tts-flash

Reads AI replies aloud as they stream, with voiced waiting phrases while the model thinks. Edge TTS built in, any OpenAI-compatible cloud engine supported.

023 日前ビジョン、音声とマルチモーダルMIT
C

dsh-live2d-avatar

clown139880/dsh-live2d-avatar

Live2D avatar stage and desktop companion for DSH: bundled Haru sample, custom Cubism 2/3+ model loading with scale and position controls, a draggable in-page pet, an optional transparent always-on-top desktop pet window, per-conversation expression prompt control, and opt-in voice with self-hosted ASR/TTS.

04 日前UI拡張MIT
A

dsh-reelsmaker

aayan-cloud/dsh-reelsmaker

DeepSeek Harness plugin: turn lines of narration into a finished vertical reel. Free neural voice-over, burned-in captions, no API keys.

028 日前ビジョン、音声とマルチモーダル
D

dsh-xiaoshuang (plugin)

daixin315/dsh-xiaoshuang

XiaoShuang desktop pet for DSH: six-layer memory, persona injection, mood-driven video avatar, manual emotion menu, and TTS replies.

028 日前お楽しみ機能
K

dsh-kitt-voice

kittcat-lab/dsh-kitt-voice

Voice for the DeepSeek Harness: speak to the agent, hear it back, and see what it is doing from a floating window that stays on top of whatever you are running.

028 日前ビジョン、音声とマルチモーダルMIT