メインコンテンツへスキップ

プラグイン

DeepSeek-Harness プラグインを閲覧・絞り込み・インストールできます。

18 件のプラグインが見つかりました

F

dsh-vision-proxy

flyvhidbwo/dsh-vision-proxy

DeepSeek の頭脳 + 自動画像文字起こし: GUI で画像を添付すると、テキスト専用の DeepSeek に届く前に任意の OpenAI 互換 VLM でテキストに変換される——自分の API キーを使うキー方式の高速パス(デフォルト qwen3.7-flash、DashScope/Zhipu/OpenRouter や任意の OpenAI 互換エンドポイントに対応)、または設定不要で自動検出されるローカル Ollama。

11一昨日ビジョン、音声とマルチモーダルMIT
S

dsh-deepseek-vision

siegfly/dsh-deepseek-vision

A vision-language gateway provider route: pasted images are described by a configurable VL model (Qwen-VL by default) before the DeepSeek wire.

6昨日ビジョン、音声とマルチモーダルMIT
O

Qwen-MM-Plugins

omdsh-dev/qwen-mm-plugins

Qwen マルチモーダルプラグインのサポート。

49 日前モデルとプロバイダー
E

dsh-llm-vision-bridge

einskyle/dsh-llm-vision-bridge

LLM プロバイダーネイティブのビジョンブリッジ。チャットに貼り付けられた画像はビジョンモデル(pi-ai/llama.cpp 経由の Qwen3-VL)によって説明文に変換され、そのテキストがテキスト専用の DeepSeek に渡されて返信が生成されます。画像の受け入れ・ルーティング・圧縮はすべて harness ネイティブの仕組みで処理し、LRU 方式の説明文キャッシュと 503 リトライを備えます。

34 日前ビジョン、音声とマルチモーダルMIT
X

dsh-draw-router

xiaozhe7772222/dsh-draw-router

Universal image generation for DeepSeek Harness: auto-discovers image models from any OpenAI-compatible endpoint (SenseNova, StepFun, Agnes, Qwen, Flux, SD, Imagen and more), with agent tools and REST API.

24 時間前ビジョン、音声とマルチモーダルMIT
X

dsh-vision

xiaoshihou514/dsh-vision

Native vision capability extension, using either Zhipu (free) or Qwen-VL (local).

2一昨日ビジョン、音声とマルチモーダルMIT
W

visual-review

wang-bool/visual-review

Renders pasted/uploaded images inline in the DSH Web chat and gives text-only models vision: the model-invokable visual_review tool calls any OpenAI-compatible multimodal API first, falling back to a local Qwen3-VL worker.

216 時間前ビジョン、音声とマルチモーダルMIT
F

dsh-free-vision

fuzzysoul/dsh-free-vision

Free vision bridge for text-only models: image understanding, OCR, UI and debug analysis via free-tier providers (Qwen3-VL-Flash, Doubao, DeepSeek-OCR) with a settings GUI.

27 時間前ビジョン、音声とマルチモーダルMIT
S

dsh-design-qa

sunxin-ai/dsh-design-qa

让 DeepSeek Harness 里任何纯文本模型都能读图。识图是按需调用的 tool —— 图片不进主模型上下文,不看就不花钱;附 23 处缺陷的评测集,换模型可自测。

19 時間前ツールと機能MIT
Z

dsh-plugins

zjcdkj/dsh-plugins

DeepSeek Harness out-of-tree plugin: give a text-only coding model eyes by routing images to a Qwen-VL (DashScope) route through ctx.llm and returning text.

15 時間前スキルMIT
M

dsh-agent-conductor

mjorgin/dsh-agent-conductor

Dispatch self-contained tasks from DSH to 11 external agent CLIs (Codex, Claude Code, TraeCode, OpenCode, Gemini, Cursor, Kimi, Qwen, Copilot, WorkBuddy, Grok): a host-only bundle registering the conductor_dispatch tool plus an auto-triggered conductor skill.

1昨日ワークフローと自動化MIT
P

dsh-voice-call

pandapolo/dsh-voice-call

Agent-initiated voice calls: `offer_call` rings the human (接听/拒接/稍后再说); accepted calls synthesize and play locally via CrispASR + Qwen3-TTS (9 speakers, 2 Chinese dialects), rejected calls return the decision to the agent.

1昨日ビジョン、音声とマルチモーダルMIT
L

dsh-eyes

leeminjing/dsh-eyes

On-demand vision for text-only DeepSeek models: upload images, and the model calls a view_image tool backed by any OpenAI-compatible vision endpoint (Qwen/DashScope by default).

1一昨日ビジョン、音声とマルチモーダルMIT
L

dsh-plugin-thinking-mode

lovedheart/dsh-plugin-thinking-mode

DeepSeek Harness プラグイン。OpenAI 互換エンドポイント上の Qwen3 系モデルに対し、リクエストごとの思考モード切り替え(enable_thinking)を提供します。

14 日前ツールと機能MIT
S

multimodal-bridge

spirit4471/multimodal-bridge

DeepSeek Harness プラグインバンドル。テキスト専用モデル向けに qwen_vision(Qwen-VL による画像理解)と qwen_generate(Qwen-Image によるテキストから画像生成)のツールを提供します。

15 日前ビジョン、音声とマルチモーダルMIT
N

gewu-tools

nyantused-cpun/gewu-tools

Model-agnostic visual-inspection pipeline for text-only agents: page-by-page HTML screenshots plus a ready-made vision-subagent briefing contract (gewu_prep), then source-code truth verification of every finding (gewu_locate); validated on mimo-v2.5 & qwen3.7-plus.

011 時間前ワークフローと自動化MIT
L

dsh-mingmu

lab-sku/dsh-mingmu

明眸 VisionBridge - 自研视觉桥:瞎子模型收图时自动调用视觉模型识别,把识别文字喂回主模型,无感、可配置、升级不丢

016 時間前ツールと機能MIT
B

dsh-voice-ai-girlfriend-plugin

beiyege-01/dsh-voice-ai-girlfriend-plugin

Voice AI girlfriend for the Web UI: FunASR mic input, Qwen3-TTS spoken replies, companion animation window, and two-way QQ chat (text/voice/image push) via NapCat.

0一昨日ビジョン、音声とマルチモーダル