メインコンテンツへスキップ

プラグイン

DeepSeek-Harness プラグインを閲覧・絞り込み・インストールできます。

13 件のプラグインが見つかりました

P

dsh-voice-scribe

pensivefei/dsh-voice-scribe

Web UI向け音声入力:Alt(またはAlt+Space)を押してディクテーションを開始/停止、ブラウザWeb Speech(ゼロ設定)またはOpenAI互換クラウドASR、DSH設定LLMによるオプションのポリッシュ、設定UI。

33昨日ビジョン、音声とマルチモーダルMIT
Q

dsh-voice-mode (dsh-voice-mode)

qishuilalala/dsh-voice-mode

DeepSeek Harness Web UI 用のフルデュプレックス音声モード:トグル(2 秒ポーズで自動送信)またはプッシュトゥトークによるディクテーションを、編集可能なドラフトに zipformer2 ストリーミング ASR で入力、オプションでウェイクワード対応。Edge TTS によるセンテンス単位の読み上げとライブキャプション、話すと再生と実行中ターンを割り込みます(真のバージイン)。オンデバイス ASR で API キー不要。

1223 時間前ビジョン、音声とマルチモーダルMIT
G

dsh-voice

goodandready/dsh-voice

Web UI向け音声入力:無音区間で分割するディクテーションと音声メッセージを提供し、それぞれに独自のプロバイダーフォールバックチェーン(Deepgram、Groq、HuggingFace、ローカルwhisper.cpp、またはOpenAI互換エンドポイント)があります。

6昨日ビジョン、音声とマルチモーダルMIT
H

dsh-voice

haoku123/dsh-voice

Web UI向けのフルデュプレックス音声モード:タップで切り替えまたは押し続けでディクテーション(送信キーまたは`Ctrl`)、ライブキャプション付き、sherpa-onnx経由のホスト側SenseVoice ASR、文ごとの音声応答、話すことで再生と実行中のターンを中断(真の割り込み)。APIキー不要。

429 日前ビジョン、音声とマルチモーダルMIT
W

dsh-hold-to-talk

wangzhanchao883/dsh-hold-to-talk

コンポーザー向けの片手・キーボード不要の入力:入力ボックス上でマウスを押し続けて話し、離すとテキストが下書きに入り、上へスライドすると中途半端な文を残さずキャンセルできます。狙って押すマイクボタンも覚えるショートカットも不要で、もう一方の手がふさがっているときに手が入力エリアから離れずに済むのが要点です。認識は完全にオフラインで動作します(sherpa-onnx 経由の SenseVoice、API キー不要、音声はマシン外に出ません)。

310 日前ビジョン、音声とマルチモーダルMIT
S

dsh-voice-assistant

supersyh-sss/dsh-voice-assistant

dsh web 向けの音声アシスタント:ウェイクフレーズ(例:「小鲸」)を言うとハンズフリーの音声入力が有効になり、話した内容が文字起こしされて自動的にチャットボックスへ入力されます。音声による編集コマンド(送信、クリア、改行、読み上げ停止)に対応し、アシスタントの返答を中国語で読み上げます。音声認識は sherpa-onnx WASM によりブラウザ内でローカルに実行されるため、API キーなしでオフラインでも動作します。

319 日前ビジョン、音声とマルチモーダルMIT
N

voco-input-sh

nothree-code/voco-input-sh

Web UI 用の音声入力: ローカルの VocoType オフライン音声認識を駆動するマイクボタンで、認識したテキストを自動的に composer に挿入する(自動デプロイ、重複排除、連続ディクテーション)。

38 日前ビジョン、音声とマルチモーダルMIT
N

dsh-dictate

navneetset/dsh-dictate

Live speech-to-text dictation into the dsh web composer via OpenRouter STT

115 日前ビジョン、音声とマルチモーダルMIT
W

dsh-word-vault

wangzhanchao883/dsh-word-vault

English word bank for one or more children: words marked on textbook or workbook photos (highlighter, red pen) are found by their marks, then translated and stored in a local SQLite bank with an occurrence count; the clipboard picker, the photo route and dictation all land in the same bank. High-frequency words surface by count, three correct answers in a row flag a word as learned, and any selection prints as a memory card set (a word split into chunks, real IPA per chunk, a mnemonic hook, an A4 sheet of eight) in HTML, PDF or Word, with a print-and-answer exam page whose four options are set against each other for meaning.

03 日前ツールと機能MIT
R

dsh-voice-input

rio-promax/dsh-voice-input

DSH voice input plugin: browser realtime + VAD dictation + local/cloud ASR + DeepSeek AI polish

09 日前ビジョン、音声とマルチモーダルMIT
W

dsh-dictation

wsl043/dsh-dictation

Editable local and desktop dictation for DeepSeek Harness

024 日前ビジョン、音声とマルチモーダルMIT
N

dsh-voice

navid-kianfar/dsh-voice

Dictate prompts into the DeepSeek Harness Web Client — a microphone in the composer, with swappable transcription: hosted Whisper API, self-hosted server, or a fully offline whisper.cpp binary.

026 日前ビジョン、音声とマルチモーダルMIT
F

dsh-dictate

franksong2702/dsh-dictate

Browser Web Speech dictation for the Composer: recognition needs no dedicated ASR server, key, or model download; reuses Session text and a configured DSH model for contextual phrase hints and optional transcript polishing.

0先月ビジョン、音声とマルチモーダルMIT