メインコンテンツへスキップ

プラグイン

DeepSeek-Harness プラグインを閲覧・絞り込み・インストールできます。

47 件のプラグインが見つかりました

L

modlens

liustack/modlens

テキスト専用モデル向けのビジョンブリッジ: 画像を貼り付けると、構造化された JSON エビデンス(OCR、レイアウト、セマンティクス)を取得。

3.1k3 時間前ビジョン、音声とマルチモーダルMIT
Y

dsh-vision-router

ysr666/dsh-vision-router

テキスト専用エージェント向けの無料ビジョン機能: キー不要の内蔵ビジョンチェーンとピクセルツール群(Q&A、グラウンディング、クロップ、ピクセル差分、カラー、OCR、SVG トレース、切り抜き、スクリーンショット)。画像を貼り付けるだけで使用可能。

7403 時間前ビジョン、音声とマルチモーダルMIT
Z

dsh-crew

zseven-w/dsh-crew

DeepSeek Harness (DSH) plugin: dispatch work to DSH agents from Claude Code / Codex — native subagent progress, in-host worker sessions with per-tier presets, and a multimodal bridge that lends the text-only harness vision and image generation.

59昨日MCPとコネクタMIT
O

dsh-vision

oil-oil/dsh-vision

DeepSeek Harness にネイティブに近い画像理解機能を追加。

245 日前ツールと機能MIT
S

dsh-authinone

stormycry-cryp/dsh-authinone

Self-contained DeepSeek Harness (DSH) plugin for Provider/Auth login, model switching, image fallback, token/cost analytics, and same-port Web restart. Useful? A star helps.

203 日前モデルとプロバイダーMIT
J

picturereader

jing-hy/picturereader

Image "reading" for text-only models: downscale + reduce color depth + structure/color fingerprints into text grids fed back to the conversation, letting the model zoom, sample and OCR autonomously like a multimodal model; fully local with zero external model dependency, ships an image-reading methodology skill and optional PaddleOCR.

208 時間前ビジョン、音声とマルチモーダルMIT
F

dsh-vision-proxy

flyvhidbwo/dsh-vision-proxy

DeepSeek の頭脳 + 自動画像文字起こし: GUI で画像を添付すると、テキスト専用の DeepSeek に届く前に任意の OpenAI 互換 VLM でテキストに変換される——自分の API キーを使うキー方式の高速パス(デフォルト qwen3.7-flash、DashScope/Zhipu/OpenRouter や任意の OpenAI 互換エンドポイントに対応)、または設定不要で自動検出されるローカル Ollama。

11一昨日ビジョン、音声とマルチモーダルMIT
J

dsh-visual-plugin

jyh20030112/dsh-visual-plugin

dsh-visual-plugin。テキスト専用モデルに目を与える: ユーザーの画像を任意の OpenAI 互換ビジョンモデルに転送し、結果を Web UI の右パネルで確認。

913 時間前ビジョン、音声とマルチモーダルMIT
5

dsh-vision

54xkeee/dsh-vision

Vision for text-only DeepSeek via Doubao Web by default (zero-cost, no API key — drives your logged-in Chrome through a Windows CDP bridge), with Antigravity IDE quota (flash/pro) or Gemini fallback; auto detail escalation, vision evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.

7一昨日ビジョン、音声とマルチモーダルMIT
S

dsh-deepseek-vision

siegfly/dsh-deepseek-vision

A vision-language gateway provider route: pasted images are described by a configurable VL model (Qwen-VL by default) before the DeepSeek wire.

6昨日ビジョン、音声とマルチモーダルMIT
S

dsh-vision-bridge

sfyyy/dsh-vision-bridge

On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

5一昨日セッションとメッセージMIT
W

dsh-mmx-bridge

welsione/dsh-mmx-bridge

MiniMax multimodal bridge: one `mmx_bridge` tool covers image understanding/generation, video, TTS, music, cover, web search and quota; optional `web_search`/`read_image` takeover; inline players/image previews right in the Web GUI (npm: `dsh-mmx-bridge`).

512 時間前ツールと機能MIT
G

deepseek-vision

gou-gee/deepseek-vision

DeepSeek Harness 原生视觉 Bundle:粘贴或拖入图片,通过托管的 deepseek-vision-mcp 调用 OpenAI 兼容视觉模型。

4一昨日MCPとコネクタMIT
G

deepseek-vision (dsh-plugin-deepseek-vision)

gou-gee/deepseek-vision

Vision MCP and DSH bundle for text-only DeepSeek: analyze_image, analyze_clipboard, compare_images and vision_status tools, a visual settings page, free GLM-4.6V-Flash by default, result caching and rate-limit tolerance; keys stay out of logs.

4一昨日ビジョン、音声とマルチモーダルMIT
A

dsh-guide-dog

atropinoltt/dsh-guide-dog

MiniMax-powered multimodal plugin: real-time voice call mode (streaming conversation, floating dock UI), voice mode and mic voice input, plus image/video/music/speech generation and vision inspection tools.

4昨日ビジョン、音声とマルチモーダルMIT
F

dsh-plugin-deepeye

favio8/dsh-plugin-deepeye

DeepSeek Harness(DSH)向け DeepEye ビジョンプラグイン: 画像説明、OCR、VQA、UI レイアウト、クリップボード分析。

45 日前UI拡張
H

dsh-her-eyes

huashenglian/dsh-her-eyes

AI が VLM(マルチモーダルモデル)を自動呼び出しして視覚分析を行えるようにする dsh プラグイン。

45 日前ビジョン、音声とマルチモーダルMIT
S

dsh-plugin-multimodal

shinjiyu/dsh-plugin-multimodal

Advertise image paste on text-only DeepSeek routes, describe attachments with a vision sidecar, and leave native vision models untouched.

3一昨日ビジョン、音声とマルチモーダルMIT
W

visual-review

wang-bool/visual-review

Renders pasted/uploaded images inline in the DSH Web chat and gives text-only models vision: the model-invokable visual_review tool calls any OpenAI-compatible multimodal API first, falling back to a local Qwen3-VL worker.

215 時間前ビジョン、音声とマルチモーダルMIT
H

dsh-open-eyes

hyp6666/dsh-open-eyes

Vision bridge for text-only DeepSeek routes that analyzes attached and local images through configurable OpenAI Responses, Chat Completions, or Anthropic Messages endpoints while leaving image-capable routes native.

2一昨日ビジョン、音声とマルチモーダルMIT
H

dsh-vision-mix

haiziyao/dsh-vision-mix

Combine text, vision, and image-generation APIs into one Mix model with automatic routing: text-only requests go to the chat model, user images and agent screenshots go to the vision model, follow-ups keep using the same session image, and agents can generate or edit images with session-scoped call history.

2一昨日ビジョン、音声とマルチモーダルMIT
Y

dsh-multimodal

yauntyour/dsh-multimodal

Per-file-type multimodal chains: preset processing models per wildcard convert image/video/audio files into prompt tokens before they reach the text-only session model, with per-preset fallback chains and a Multimodal settings page.

14 日前セッションとメッセージMIT
I

dsh-tool-visual-primitives

inkshadewoods/dsh-tool-visual-primitives

DeepSeek Harness 视觉增强插件:将图片交给外部视觉模型分析,输出带坐标化视觉原语的纯文本证据,使不支持多模态的文本模型也能在对话中理解图片、截图与文档。

13 日前ツールと機能MIT
S

dsh-design-qa

sunxin-ai/dsh-design-qa

让 DeepSeek Harness 里任何纯文本模型都能读图。识图是按需调用的 tool —— 图片不进主模型上下文,不看就不花钱;附 23 处缺陷的评测集,换模型可自测。

18 時間前ツールと機能MIT