メインコンテンツへスキップ

プラグイン

DeepSeek-Harness プラグインを閲覧・絞り込み・インストールできます。

24 件のプラグインが見つかりました

P

dsh-voice-scribe

pensivefei/dsh-voice-scribe

Web UI向け音声入力:Alt(またはAlt+Space)を押してディクテーションを開始/停止、ブラウザWeb Speech(ゼロ設定)またはOpenAI互換クラウドASR、DSH設定LLMによるオプションのポリッシュ、設定UI。

343 日前ビジョン、音声とマルチモーダルMIT
Z

dsh-voice-input-plugin

zhangbo-cn/dsh-voice-input-plugin

Web UI向けコンポーザーマイク:タップでライブ文字起こしをモニターし、押している間話せるホールドトーク機能を提供します。モデルが生成している間にストリーミングされるホストのEdge TTSによる返信読み上げ、読み上げ中のエコー一時停止、タップで停止も備えています。

6先月ビジョン、音声とマルチモーダルMIT
0

dsh-voice-input

0nt-one/dsh-voice-input

コンポーザーツール列のマイクボタン: Web Speech API による音声テキスト変換(Chrome/Edge)、言語切替、オプションの自動送信、依存ゼロ。

52 か月前ビジョン、音声とマルチモーダルMIT
A

dsh-guide-dog

atropinoltt/dsh-guide-dog

MiniMax駆動のマルチモーダルプラグイン:リアルタイム音声通話モード(ストリーミング会話、フローティングドックUI)、音声モードとマイク音声入力、さらに画像/動画/音楽/音声生成および視覚検査ツール。

52 か月前ビジョン、音声とマルチモーダルMIT
W

dsh-hold-to-talk

wangzhanchao883/dsh-hold-to-talk

コンポーザー向けの片手・キーボード不要の入力:入力ボックス上でマウスを押し続けて話し、離すとテキストが下書きに入り、上へスライドすると中途半端な文を残さずキャンセルできます。狙って押すマイクボタンも覚えるショートカットも不要で、もう一方の手がふさがっているときに手が入力エリアから離れずに済むのが要点です。認識は完全にオフラインで動作します(sherpa-onnx 経由の SenseVoice、API キー不要、音声はマシン外に出ません)。

321 時間前ビジョン、音声とマルチモーダルMIT
N

voco-input-sh

nothree-code/voco-input-sh

Web UI 用の音声入力: ローカルの VocoType オフライン音声認識を駆動するマイクボタンで、認識したテキストを自動的に composer に挿入する(自動デプロイ、重複排除、連続ディクテーション)。

319 日前ビジョン、音声とマルチモーダルMIT
Q

dsh-mic-input

qt-chen/dsh-mic-input

コンポーザー向けのマイク音声入力機能。ブラウザの Web Speech API によるリアルタイム文字起こし、重複排除/自動継続、スマート句読点、言語と自動送信の設定に対応します。

32 か月前ビジョン、音声とマルチモーダルMIT
I

dshtools-sensevoice-input

ilovedyou6666-hub/dshtools-sensevoice-input

基于 SenseVoiceSmall(iic/SenseVoiceSmall)多语言语音理解模型的 DSH Desktop 本地语音输入插件。

129 日前ビジョン、音声とマルチモーダルMIT
D

dsh-voice-input-npm

difimim/dsh-voice-input-npm

语音输入插件 for Deepseek Harness

1先月ビジョン、音声とマルチモーダルMIT
S

dsh-voice-input-cn

schumchanvi/dsh-voice-input-cn

コンポーザー向けの中国対応音声入力です。ローカル Python ブリッジが必要です (pip install dashscope websockets、bridge/voice-bridge.py を実行)。プラグイン単体では動作しません。ブラウザマイクが 16 kHz PCM をブリッジへ送信し、Alibaba Cloud DashScope ASR (paraformer-realtime-v2) を実行します。中間結果がカーソル位置のドラフトへ流し込まれ、無音で自動停止、オプションで自動送信を行います。

115 日前ビジョン、音声とマルチモーダルMIT
A

dsh-audio-copilot

ai-yucheng/dsh-audio-copilot

Audio Copilot for DeepSeek Harness: transcribe audio (ASR) and synthesize speech (TTS) — gives text-only agents ears and a voice. Windows-local SAPI TTS out of the box; OpenAI-compatible ASR/TTS endpoints configurable. Includes an in-composer voice-input

1先月ビジョン、音声とマルチモーダルMIT
L

dsh-voice-input

lhenlihai-hub/dsh-voice-input

12 か月前ビジョン、音声とマルチモーダルMIT
B

dsh-asr-voice

bittersmilezzz/dsh-asr-voice

开口即成文 · Speak-to-prompt for DeepSeek Harness:云端 ASR 语音识别 + 提示词优化 + 填入草稿/自动发送,跨平台 macOS / Windows。

017 日前ビジョン、音声とマルチモーダルMIT
R

dsh-voice-input

rio-promax/dsh-voice-input

DSH voice input plugin: browser realtime + VAD dictation + local/cloud ASR + DeepSeek AI polish

020 日前ビジョン、音声とマルチモーダルMIT
J

dsh-voice-input-qwen-asr

jsoncode/dsh-voice-input-qwen-asr

Voice input plugin (dual-face): mic button beside the composer send action, live recording bubble streaming PCM to a local Qwen3-ASR python service managed by the host, plus an ASR environment settings page (clone runtime/model repos, create venv, run ser

027 日前ビジョン、音声とマルチモーダル
M

dsh-voice-input-en

mohith-das/dsh-voice-input-en

Minimal English-only voice input for the DeepSeek Harness Web UI: a mic button in the composer that transcribes speech into the draft via the browser's native SpeechRecognition API. No dependencies, no subprocess, no network calls beyond whatever the brow

0先月ビジョン、音声とマルチモーダルMIT
S

dsh-voice-control

sucriss/dsh-voice-control

Voice control for the DSH web composer: push-to-talk speech-to-text into the input box (with optional auto-send), spoken playback of assistant replies via the Web Speech API, right-click settings popover, and a global Ctrl+M hotkey.

0先月ビジョン、音声とマルチモーダルMIT
W

dsh-dictation

wsl043/dsh-dictation

Editable local and desktop dictation for DeepSeek Harness

0先月ビジョン、音声とマルチモーダルMIT
X

dsh-voice-input-space

xsakura666/dsh-voice-input-space

Voice input for DeepSeek Harness: hold Space to speak, release to insert. Zero dependencies, Web Speech API. / 语音输入:长按空格说话,松开上屏,零依赖。

0先月ビジョン、音声とマルチモーダルMIT
N

dsh-voice

navid-kianfar/dsh-voice

Dictate prompts into the DeepSeek Harness Web Client — a microphone in the composer, with swappable transcription: hosted Whisper API, self-hosted server, or a fully offline whisper.cpp binary.

0先月ビジョン、音声とマルチモーダルMIT
V

dsh-doubao-voice

vorpal-poem/dsh-doubao-voice

DeepSeek Harness 语音输入插件:火山引擎流式 ASR(豆包 Seed ASR),流式回填输入框。Voice input for DSH via Volcengine streaming ASR.

02 か月前ビジョン、音声とマルチモーダルMIT
M

dsh-voice

motongv/dsh-voice

给 DeepSeek Harness(DSH / DeepSeek Hermes)加语音能力的社区插件:输入框语音输入(可配快捷键)+ 回答朗读(微软 Edge 神经网络音色,可换音色、可试听),无需 API Key。

02 か月前ビジョン、音声とマルチモーダルMIT
O

dsh-voice-input

opensquad-ai/dsh-voice-input

SenseVoice 语音输入插件 for DeepSeek Harness:在对话输入框旁添加麦克风按钮,录音后调用本地 SenseVoice 服务转成文本填入输入框。首次使用自动下载模型并显示进度,后端由插件自动启动。

02 か月前ビジョン、音声とマルチモーダルMIT
N

dsh-voice-input

newdanew/dsh-voice-input

Voice input for the web UI: a mic button in the composer that transcribes speech into the draft via the Web Speech API, with an optional auto-send toggle.

02 か月前ビジョン、音声とマルチモーダルMIT