メインコンテンツへスキップ

プラグイン

DeepSeek-Harness プラグインを閲覧・絞り込み・インストールできます。

125 件のプラグインが見つかりました

L

Deepseek-Continuity

linxuhao/deepseek-continuity

ローカルでの画像、音声、音楽、音響効果の生成と文字起こし。固定ID機能を搭載:キャラクター、動物、物体、俳優の音声を一度定義すれば以降のすべての呼び出しで再利用できます。退化した出力(平坦に近い画像、無音オーディオ)は成功として返さず拒否。生成された行はテキストとして読み戻せるため、末尾を飲み込んだクローンも可視化されます。エンジンはアイドル時にアンロードされ、画像生成と文字起こしはローカルのVulkanバックエンドの代わりにOpenAI互換APIを指し示すこともできます。

312 日前ビジョン、音声とマルチモーダルMIT
N

voco-input-sh

nothree-code/voco-input-sh

Web UI 用の音声入力: ローカルの VocoType オフライン音声認識を駆動するマイクボタンで、認識したテキストを自動的に composer に挿入する(自動デプロイ、重複排除、連続ディクテーション)。

319 日前ビジョン、音声とマルチモーダルMIT
B

dsh-stt-input

baisama-cloud/dsh-stt-input

Web UI 用音声入力:コンポーザーのマイクボタンがブラウザ Web Speech API(ゼロコンフィグ)または OpenAI 互換 Whisper API(OpenAI / Groq)で音声をドラフトに転写。Settings でモデルと言語を選択可能。

3先月ビジョン、音声とマルチモーダルMIT
H

dsh-plugin-notify

huguangyu666/dsh-plugin-notify

通知アウトボックス: エージェントがトースト/中国語 TTS 音声/効果音(爆発、勝利、アラーム)で能動的に通知。60 秒の確認ウィンドウ内なら音声で呼び戻し、音量ブースト、設定パネル付き。

318 日前統合とリモートMIT
J

dsh-voice

jesse-njx/dsh-voice

音声メモを入力し、音声で回答を受け取れます。話した内容がユーザーメッセージになり(文字起こし)、エージェントの返信を読み上げさせることもできます(発話)。~/.dsh/voice 配下でローカル優先に動作します。

32 か月前ビジョン、音声とマルチモーダルMIT
Q

dsh-mic-input

qt-chen/dsh-mic-input

コンポーザー向けのマイク音声入力機能。ブラウザの Web Speech API によるリアルタイム文字起こし、重複排除/自動継続、スマート句読点、言語と自動送信の設定に対応します。

32 か月前ビジョン、音声とマルチモーダルMIT
B

dsh-voice-call

biliye/dsh-voice-call

DSH Web GUI の音声通話アシスタント:ドラッグ可能なフローティング通話ボール、発話ごとにポーズ後に送信するブラウザ側 VAD、FunASR HTTP またはストリーミング音声認識、MiniMax または OpenAI 互換 TTS の返答、自動スリープ付きの任意のウェイクワードモード、進捗・停止・読み上げ完了報告を伴う別サブエージェントセッションへのタスク派遣。

25 日前ビジョン、音声とマルチモーダルMIT
Y

dsh-tool-lipsync

yu-wenchao/dsh-tool-lipsync

3000 以上の音声と 500 以上の言語に対応する、DSH 向け無料リップシンク動画生成プラグイン。

228 日前ビジョン、音声とマルチモーダルMIT
G

dsh-messenger-gateway

goodandready/dsh-messenger-gateway

DeepSeek Harness向けのTelegramブリッジ:双方向セッション、会話のステアリング、ホームチャネルへのルーティング、およびTTS音声メモのTelegramオーディオメッセージとしての送信。

23 日前統合とリモートMIT
W

dsh-voice-agent (voice-app)

wayneyu430/dsh-voice-agent

dsh 用の会話型音声フロントエンド Agent:ByteDance Duplex 上で自然に話し、リクエストをバックグラウンドタスクに委任し、その非同期の結果を音声で報告してもらえます。

2先月ビジョン、音声とマルチモーダル
M

dsh-voice-chat

maoyuching/dsh-voice-chat

DSH 用の Doubao 風音声チャット。入力欄のマイクを長押しすると音声をテキストに変換して自動送信し、AI の返信を読み上げます。オプションの LLM 要約(長い返信を現在の会話モデルに従って短い読み上げ文に要約)、TTS に適したテキストクリーニング、選択可能な Edge TTS 音声、調整可能な無音自動停止に対応し、設定は DSH の設定ダイアログに組み込まれています。

28 日前ビジョン、音声とマルチモーダルMIT
G

dsh-tts

goodandready/dsh-tts

DeepSeek HarnessのWeb UIでエージェントの返答を、プロバイダーのフォールバックチェーン(OpenAI、ElevenLabs、Google、Azure、Groq、Deepgram、OpenRouter、Edge、Piper、eSpeak)を通じて読み上げます。これにより、失敗したりレート制限されたプロバイダーは無音になるのではなく次のプロバイダーに引き継がれます。

26 日前ビジョン、音声とマルチモーダルMIT
P

muxiva-dsh-voice

piyotahu/muxiva-dsh-voice

Muxiva がオーケストレーションする、DeepSeek Harness 向けのローカルファースト・フルデュプレックス音声

22 か月前ビジョン、音声とマルチモーダルApache-2.0
P

dsh-voice-call

pandapolo/dsh-voice-call

エージェント起動の音声通話: `offer_call` が人間を呼び出し(接听/拒接/稍后再说)、応答された通話は CrispASR + Qwen3-TTS(9 話者、2 中国語方言)でローカル生成・再生され、拒否された通話はその判断をエージェントに返します。

23 日前ビジョン、音声とマルチモーダルMIT
M

dsh-fish-tts

mari23333/dsh-fish-tts

Fish Audio API のみを使ってアシスタントの返信を音声読み上げします(API キーは各自用意):メッセージごとの読み上げ、自動読み上げ切り替え、モデル・音声 reference_id・暗号化された API キー・プロキシを設定するページを備えます。

2一昨日ビジョン、音声とマルチモーダルMIT
J

dsh-live2d-voice

john-walks-slow/dsh-live2d-voice

Live2D session view with realtime voice: continuous voice input, streaming TTS, subtitle translation, multi-model catalog, standalone entry / Live2D 实时语音会话视图:连续语音输入、流式 TTS、字幕翻译、多模型目录、独立入口

14 日前ビジョン、音声とマルチモーダルMIT
M

dsh-onebot

mario841859784/dsh-onebot

QQ channel for dsh via OneBot 11 (NapCat and friends): reverse or forward WebSocket, DM and group chats, image and voice transcription, t2i text-image cards, merged forwards, and a settings page.

18 日前統合とリモートBSD-3-Clause
Y

dsh-claude-slider

yicun0316/dsh-claude-slider

Replaces the reasoning-effort dropdown with a draggable snapping slider, and draws 13 canvas effects across five style families (thrust, fluid, fantasy, tech, texture) with three tiers that unlock as the effort level rises. Custom mode combines 10 particle forms with 8 flow trajectories; 9 accent presets and a hex picker recolor every effect, and optional key sounds cover two voice clips, a synthesized duck squeak, a mechanical click and a local audio file.

19 日前UI拡張MIT
Z

dsh-wx-bridge

zhy5/dsh-wx-bridge

Drive your local DSH from WeChat: a self-hosted iLink bridge over a persistent ACP session (context lives in DSH and is resumable), phone conversations grouped into the desktop workspace, with image recognition, file messages (Excel/Word/PDF, parsed by the agent itself) and voice transcripts. Access defaults to strict (only registered devices are served); the phone channel's permission preset is wide by design - see the README security section before sharing the bot.

16 日前統合とリモートMIT
W

dsh-multi-tts

wyr-233/dsh-multi-tts

Per-reply read-aloud with a multi-provider settings page — MiniMax or any OpenAI-compatible /audio/speech endpoint, with voice, emotion, speed and model selection plus an auto-read toggle.

121 日前ビジョン、音声とマルチモーダル
X

dshgo

xiazhi88/dshgo

Proxies the loopback-only dsh web onto a LAN port so phones and other machines can reach it, with an optional access password that browsers type in and the bundled Android client unlocks with biometrics. Adds a LAN-access settings page with QR codes, inlines dsh-web-mobile (MIT, attribution kept) for the mobile layout, and the Android client in the same repository adds finish/approval notifications, voice input and a home-screen widget.

114 日前統合とリモートMIT
V

dsh-live-voice

victorwads/dsh-live-voice

Local-first voice conversations for DSH, with local speech recognition and synthesis and optional external providers.

118 日前ビジョン、音声とマルチモーダルGPL-3.0
M

dsh-minimax-asr

moluyao/dsh-minimax-asr

MiniMax 语音识别(asr-1.0)+ 语音合成(speech-2.8-hd)的 DeepSeek Harness 全局插件:转写工具、任务结束后用喇叭念一句的播报、实时免手对话、303 个音色可选

119 日前ビジョン、音声とマルチモーダルMIT
D

dsh-voice-talk

duoduoqian708/dsh-voice-talk

Voice conversation mode for the DSH Web GUI: tap the mic to start a full-screen call, talk hands-free, and hear each reply read aloud as it streams. Bilingual (Chinese/English) UI, light and dark themes, adjustable speech rate, and switchable voices.

120 日前ビジョン、音声とマルチモーダルMIT