メインコンテンツへスキップ

プラグイン

DeepSeek-Harness プラグインを閲覧・絞り込み・インストールできます。

15 件のプラグインが見つかりました

W

dsh-ears

wiziscool/dsh-ears

DeepSeek Harness(dsh)向け音声入力プラグイン:コンポーザーのマイクボタンで音声を下書き転写に変換し、複数の音声認識バックエンドから選択可能で、dsh 独自の LLM ルートによる任意の推敲とネイティブ設定ページを備えています。

21昨日ビジョン、音声とマルチモーダルMIT
C

dsh-bilibili

czx2244/dsh-bilibili

Bilibili 動画分析: メタデータ、文字起こし(Bijian/sherpa-onnx/whisper.cpp による ASR フォールバック)、コメント、弾幕、鮮明なキーフレーム。ローカルビジョンによる説明生成もオプションで対応。

92 か月前ツールと機能MIT
I

dsh-video-understand

ilps2/dsh-video-understand

低コストな動画理解ツール。video_understand ツールが Bilibili リンク / BV / ローカル動画を AVIS 情報層(ASR + シーン構造 + オブジェクト追跡 + YOLO ラベル)に変換し、要約 + Q&A を返す。質問主導のレイヤー間ダイナミックルーティング(L0 ASR / L1 オブジェクト追跡 / L2 キーフレーム VLM)、繰り返し質問向けの意味レイヤー再利用、質問ごとの予算上限を備える。Python エンジン: コア層には faster-whisper / opencv / yt-dlp が必要(約200-300MB)。任意の意味層を追加すると torch / transformers / ultralytics が約2GB 増える。付属の doctor --fix が venv をセットアップし、両方をインストールする。

8先月ツールと機能
G

dsh-voice

goodandready/dsh-voice

Web UI向け音声入力:無音区間で分割するディクテーションと音声メッセージを提供し、それぞれに独自のプロバイダーフォールバックチェーン(Deepgram、Groq、HuggingFace、ローカルwhisper.cpp、またはOpenAI互換エンドポイント)があります。

8昨日ビジョン、音声とマルチモーダルMIT
S

dsh-voice

stardustlc666/dsh-voice

音声ツール:無料のedge-ttsニューラル音声合成、OpenAI互換のASR文字起こし、音声リスト、音声の一括プレビュー、ヘルス自己チェック。

410 時間前ビジョン、音声とマルチモーダルMIT
B

dsh-stt-input

baisama-cloud/dsh-stt-input

Web UI 用音声入力:コンポーザーのマイクボタンがブラウザ Web Speech API(ゼロコンフィグ)または OpenAI 互換 Whisper API(OpenAI / Groq)で音声をドラフトに転写。Settings でモデルと言語を選択可能。

3先月ビジョン、音声とマルチモーダルMIT
Z

dsh-watch-video

zeshuochen/dsh-watch-video

字幕優先の動画文字起こし。SRT のエクスポート、キャンセル可能なジョブ制御、字幕がない場合の faster-whisper large-v3 フォールバックを備えます。

2先月ビジョン、音声とマルチモーダル
1

dsh-wsl-im

173787247/dsh-wsl-im

Bridges Feishu, WeCom, DingTalk, QQ, Slack, Discord, Telegram and Mattermost into dsh agents (outbound WS/Stream/Gateway/long-poll/webhook), with im_status, optional local Whisper ASR, plain-text outbound for QQ/DingTalk/Telegram/Mattermost, and env CSV allowlists (DSH_IM_*_ALLOWED_USER_IDS). Empty allowedUserIds means EVERYONE can drive your agent — set a whitelist before exposing a bot.

13 日前開発とプラグインツールMIT
V

dsh-live-voice

victorwads/dsh-live-voice

Local-first voice conversations for DSH, with local speech recognition and synthesis and optional external providers.

118 日前ビジョン、音声とマルチモーダルGPL-3.0
Z

dsh-voice

zhuiyueya/dsh-voice

DeepSeek Harness(dsh)向けの音声機能——テキスト専用の DeepSeek に対する音声入力(STT)と読み上げ(TTS)を提供します。API キー不要。

12 か月前ビジョン、音声とマルチモーダルMIT
Y

dsh-video-to-notes

yll-kb/dsh-video-to-notes

Opt-in DeepSeek Harness skill bundle that turns course, lecture, tutorial, documentary, meeting, and talk videos into structured study notes.

012 日前ビジョン、音声とマルチモーダルMIT
B

dsh-asr-voice

bittersmilezzz/dsh-asr-voice

开口即成文 · Speak-to-prompt for DeepSeek Harness:云端 ASR 语音识别 + 提示词优化 + 填入草稿/自动发送,跨平台 macOS / Windows。

017 日前ビジョン、音声とマルチモーダルMIT
K

dsh-kitt-voice

kittcat-lab/dsh-kitt-voice

Voice for the DeepSeek Harness: speak to the agent, hear it back, and see what it is doing from a floating window that stays on top of whatever you are running.

028 日前ビジョン、音声とマルチモーダルMIT
N

dsh-voice

navid-kianfar/dsh-voice

Dictate prompts into the DeepSeek Harness Web Client — a microphone in the composer, with swappable transcription: hosted Whisper API, self-hosted server, or a fully offline whisper.cpp binary.

0先月ビジョン、音声とマルチモーダルMIT
Z

dsh-video-understand

zeshuochen/dsh-video-understand

Subtitle-first video transcription and deterministic extractive Markdown summaries, with a faster-whisper large-v3 fallback when subtitles are unavailable.

0先月ビジョン、音声とマルチモーダル