プラグイン
DeepSeek-Harness プラグインを閲覧・絞り込み・インストールできます。
29 件のプラグインが見つかりました
watch-skill
oxbshw/watch-skill
DeepWatch の機能を、既存の DeepSeek Harness プロファイルにインストールできるようにします。
dsh-voice-scribe
pensivefei/dsh-voice-scribe
Web UI向け音声入力:Alt(またはAlt+Space)を押してディクテーションを開始/停止、ブラウザWeb Speech(ゼロ設定)またはOpenAI互換クラウドASR、DSH設定LLMによるオプションのポリッシュ、設定UI。
dsh-ears
wiziscool/dsh-ears
DeepSeek Harness(dsh)向け音声入力プラグイン:コンポーザーのマイクボタンで音声を下書き転写に変換し、複数の音声認識バックエンドから選択可能で、dsh 独自の LLM ルートによる任意の推敲とネイティブ設定ページを備えています。
dsh-talk
perrylink/dsh-talk
DeepSeek Harness の音声 I/O — マイク入力と音声出力による speech-to-text / text-to-speech。
dsh-voice-mode (dsh-voice-mode)
qishuilalala/dsh-voice-mode
DeepSeek Harness Web UI 用のフルデュプレックス音声モード:トグル(2 秒ポーズで自動送信)またはプッシュトゥトークによるディクテーションを、編集可能なドラフトに zipformer2 ストリーミング ASR で入力、オプションでウェイクワード対応。Edge TTS によるセンテンス単位の読み上げとライブキャプション、話すと再生と実行中ターンを割り込みます(真のバージイン)。オンデバイス ASR で API キー不要。
dsh-qq-onebot-bridge
cheesehaqi/dsh-qq-onebot-bridge
OneBot v11(リバース WebSocket)による双方向 QQ ブリッジ:グループ別および個別チャット別セッション、@-引用付き音声 speech-to-text、プライベートな画像/アニメスタンプ閲覧、顔およびスタンプツール。
dsh-voice-input
0nt-one/dsh-voice-input
コンポーザーツール列のマイクボタン: Web Speech API による音声テキスト変換(Chrome/Edge)、言語切替、オプションの自動送信、依存ゼロ。
dsh-hold-to-talk
wangzhanchao883/dsh-hold-to-talk
コンポーザー向けの片手・キーボード不要の入力:入力ボックス上でマウスを押し続けて話し、離すとテキストが下書きに入り、上へスライドすると中途半端な文を残さずキャンセルできます。狙って押すマイクボタンも覚えるショートカットも不要で、もう一方の手がふさがっているときに手が入力エリアから離れずに済むのが要点です。認識は完全にオフラインで動作します(sherpa-onnx 経由の SenseVoice、API キー不要、音声はマシン外に出ません)。
voco-input-sh
nothree-code/voco-input-sh
Web UI 用の音声入力: ローカルの VocoType オフライン音声認識を駆動するマイクボタンで、認識したテキストを自動的に composer に挿入する(自動デプロイ、重複排除、連続ディクテーション)。
dsh-stt-input
baisama-cloud/dsh-stt-input
Web UI 用音声入力:コンポーザーのマイクボタンがブラウザ Web Speech API(ゼロコンフィグ)または OpenAI 互換 Whisper API(OpenAI / Groq)で音声をドラフトに転写。Settings でモデルと言語を選択可能。
dsh-voice
jesse-njx/dsh-voice
音声メモを入力し、音声で回答を受け取れます。話した内容がユーザーメッセージになり(文字起こし)、エージェントの返信を読み上げさせることもできます(発話)。~/.dsh/voice 配下でローカル優先に動作します。
dsh-voice-call
pandapolo/dsh-voice-call
エージェント起動の音声通話: `offer_call` が人間を呼び出し(接听/拒接/稍后再说)、応答された通話は CrispASR + Qwen3-TTS(9 話者、2 中国語方言)でローカル生成・再生され、拒否された通話はその判断をエージェントに返します。
dsh-live-voice
victorwads/dsh-live-voice
Local-first voice conversations for DSH, with local speech recognition and synthesis and optional external providers.
dsh-dictate
navneetset/dsh-dictate
Live speech-to-text dictation into the dsh web composer via OpenRouter STT
dsh-minimax-asr
moluyao/dsh-minimax-asr
MiniMax 语音识别(asr-1.0)+ 语音合成(speech-2.8-hd)的 DeepSeek Harness 全局插件:转写工具、任务结束后用喇叭念一句的播报、实时免手对话、303 个音色可选
dsh-live-voice
jstn-1g/dsh-live-voice
Consent-bound one-turn voice preview for DSH Web with a credential-free local synthetic demo, optional Qwen Audio, exact Session isolation, and explicit transcript-to-draft handoff without automatic submission.
dsh-voice
zhuiyueya/dsh-voice
DeepSeek Harness(dsh)向けの音声機能——テキスト専用の DeepSeek に対する音声入力(STT)と読み上げ(TTS)を提供します。API キー不要。
dsh-voice-danmaku
weizhida/dsh-voice-danmaku
这是dsh的插件,语音发送弹幕。在玩游戏时通过语音输入在b站发弹幕,不切出游戏可以正常操作
dsh-stt-plugin
zemanzhang809/dsh-stt-plugin
A speech-to-text (voice input) plugin for DeepSeek Harness.
dsh-asr-voice
bittersmilezzz/dsh-asr-voice
开口即成文 · Speak-to-prompt for DeepSeek Harness:云端 ASR 语音识别 + 提示词优化 + 填入草稿/自动发送,跨平台 macOS / Windows。
dsh-voice-mimo
ch1bug/dsh-voice-mimo
Xiaomi MiMo-powered voice for DeepSeek Harness: browser 🎤/🧠/🔊 UI, voice_transcribe/voice_understand/voice_speak tools, configurable voice map (preset/voicedesign/voiceclone). Fork of zhuiyueya/dsh-voice (MIT), Settings pattern from Anionex/dsh-vision-toolkit (MIT).
dsh-voice-input-en
mohith-das/dsh-voice-input-en
Minimal English-only voice input for the DeepSeek Harness Web UI: a mic button in the composer that transcribes speech into the draft via the browser's native SpeechRecognition API. No dependencies, no subprocess, no network calls beyond whatever the brow
dsh-kitt-voice
kittcat-lab/dsh-kitt-voice
Voice for the DeepSeek Harness: speak to the agent, hear it back, and see what it is doing from a floating window that stays on top of whatever you are running.
dsh-voice-control
sucriss/dsh-voice-control
Voice control for the DSH web composer: push-to-talk speech-to-text into the input box (with optional auto-send), spoken playback of assistant replies via the Web Speech API, right-click settings popover, and a global Ctrl+M hotkey.