メインコンテンツへスキップ

プラグイン

DeepSeek-Harness プラグインを閲覧・絞り込み・インストールできます。

39 件のプラグインが見つかりました

L

modlens

liustack/modlens

テキスト専用モデル向けのビジョンブリッジ: 画像を貼り付けると、構造化された JSON エビデンス(OCR、レイアウト、セマンティクス)を取得。

3.1k3 時間前ビジョン、音声とマルチモーダルMIT
Y

dsh-vision-router

ysr666/dsh-vision-router

テキスト専用エージェント向けの無料ビジョン機能: キー不要の内蔵ビジョンチェーンとピクセルツール群(Q&A、グラウンディング、クロップ、ピクセル差分、カラー、OCR、SVG トレース、切り抜き、スクリーンショット)。画像を貼り付けるだけで使用可能。

7403 時間前ビジョン、音声とマルチモーダルMIT
A

dsh-vision-toolkit

anionex/dsh-vision-toolkit

テキスト専用モデル向けのビジョンタスク: 意図を汲む画像 Q&A、長尺スクリーンショット OCR、UI 再現、グラウンディング、ピクセル差分。

6946 時間前ビジョン、音声とマルチモーダルMIT
W

dsh-openbiliclaw

whiteguo233/dsh-openbiliclaw

OpenBiliClaw はローカルで動くクロスプラットフォームのパーソナライズ推薦 Agent で、あなたの興味を継続的に理解しコンテンツを能動的に探す。本リポジトリはその DeepSeek Harness プラグイン: DSH 画面に第 4 タブ(推薦/コンテンツライブラリ/対話/プロファイル/設定)を常駐させ、22 個の Agent Bridge ツールを登録。Agent も推薦を読み取り、質問に答え、クローズドループ学習を行える。

44昨日UI拡張BSD-3-Clause
J

picturereader

jing-hy/picturereader

Image "reading" for text-only models: downscale + reduce color depth + structure/color fingerprints into text grids fed back to the conversation, letting the model zoom, sample and OCR autonomously like a multimodal model; fully local with zero external model dependency, ships an image-reading methodology skill and optional PaddleOCR.

208 時間前ビジョン、音声とマルチモーダルMIT
T

dsh-openmaic

thu-maic/dsh-openmaic

OpenMAIC: 教室、スライド、インタラクティブウィジェット、ソクラテス式教授法。

165 日前ツールと機能MIT
L

dsh-vision

linenxi-ctrl/dsh-vision

DeepSeek Harness 向け外部ビジョンプラグイン: クジラボタンの設定パネル、自動返信付き画像認識、エージェント用スクリーンショット/認識ツール。

123 日前ビジョン、音声とマルチモーダルMIT
F

dsh-vision-proxy

flyvhidbwo/dsh-vision-proxy

DeepSeek の頭脳 + 自動画像文字起こし: GUI で画像を添付すると、テキスト専用の DeepSeek に届く前に任意の OpenAI 互換 VLM でテキストに変換される——自分の API キーを使うキー方式の高速パス(デフォルト qwen3.7-flash、DashScope/Zhipu/OpenRouter や任意の OpenAI 互換エンドポイントに対応)、または設定不要で自動検出されるローカル Ollama。

11一昨日ビジョン、音声とマルチモーダルMIT
S

dsh-docs

sqhao-o/dsh-docs

Local PDF, Office, image, and OCR document intelligence for DeepSeek Harness.

103 日前ツールと機能MIT
5

dsh-vision

54xkeee/dsh-vision

Vision for text-only DeepSeek via Doubao Web by default (zero-cost, no API key — drives your logged-in Chrome through a Windows CDP bridge), with Antigravity IDE quota (flash/pro) or Gemini fallback; auto detail escalation, vision evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.

7一昨日ビジョン、音声とマルチモーダルMIT
M

dsh-windows-ocr

maxwell-feng/dsh-windows-ocr

Local OCR for attached images via the built-in Windows engine (Windows.Media.Ocr): only the recognized text is sent to the model, never the image bytes; vision passthrough is opt-in.

6昨日ビジョン、音声とマルチモーダルMIT
B

dsh-ocr-local

balcoz/dsh-ocr-local

Local OCR for DeepSeek Harness: paste/attach an image, get its text via PP-OCRv5 + ONNX Runtime, fully offline. TUI (cc-tui) and Web. / DeepSeek Harness 本地 OCR 插件:图片转文字,PP-OCRv5 + ONNX Runtime,完全离线,支持 TUI 与 Web。

514 時間前UI拡張
S

dsh-vision-bridge

sfyyy/dsh-vision-bridge

On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

5一昨日セッションとメッセージMIT
G

deepseek-vision

gou-gee/deepseek-vision

DeepSeek Harness 原生视觉 Bundle:粘贴或拖入图片,通过托管的 deepseek-vision-mcp 调用 OpenAI 兼容视觉模型。

4一昨日MCPとコネクタMIT
G

deepseek-vision (dsh-plugin-deepseek-vision)

gou-gee/deepseek-vision

Vision MCP and DSH bundle for text-only DeepSeek: analyze_image, analyze_clipboard, compare_images and vision_status tools, a visual settings page, free GLM-4.6V-Flash by default, result caching and rate-limit tolerance; keys stay out of logs.

4一昨日ビジョン、音声とマルチモーダルMIT
C

dsh-learning-mode

chplus0/dsh-learning-mode

Learning Mode agent preset: a coding agent that teaches while coding — concrete scenario-grounded explanations, Socratic guidance, and TODO(你) practice blanks, modeled on Claude Code's Learning output style; installable via dsh plugin add (dsh-learning-mode on npm).

43 日前ツールと機能
F

dsh-plugin-deepeye

favio8/dsh-plugin-deepeye

DeepSeek Harness(DSH)向け DeepEye ビジョンプラグイン: 画像説明、OCR、VQA、UI レイアウト、クリップボード分析。

45 日前UI拡張
Y

dsh-ui-spec

yumimanji/dsh-ui-spec

DeepSeek Harness plugin that turns UI screenshots into implementation-grade web specs using OCR, deterministic geometry, scene graphs, assets, and render comparison.

3一昨日UI拡張MIT
N

free-vision-skill

niyongsheng/free-vision-skill

Fully-local image understanding & OCR via macOS Vision Framework: `ocr_image` (text, table layout + coordinates) and `view_image` (scene, faces, QR) — paste multiple images into the web input box or pass path/URL/base64; images never leave your Mac.

33 日前ビジョン、音声とマルチモーダルMIT
O

dsh-paddle-ocr

omdsh-dev/dsh-paddle-ocr

35 日前ツールと機能BSD-3-Clause
M

dsh-tesseract-ocr

maxwell-feng/dsh-tesseract-ocr

Local OCR for attached images via Tesseract: only the recognized text is sent to the model, never the image bytes; vision passthrough is opt-in.

2昨日ビジョン、音声とマルチモーダルMIT
F

dsh-free-vision

fuzzysoul/dsh-free-vision

Free vision bridge for text-only models: image understanding, OCR, UI and debug analysis via free-tier providers (Qwen3-VL-Flash, Doubao, DeepSeek-OCR) with a settings GUI.

26 時間前ビジョン、音声とマルチモーダルMIT
Y

dsh-computer-use-win

yu-tao-li/dsh-computer-use-win

Windows computer use for DeepSeek Harness: an MCP stdio server over a PowerShell UIA backend exposing 22 desktop tools (UIA tree, screenshots, typed input, OCR, window management, failsafe).

1昨日ツールと機能MIT
Z

dsh-plugins

zjcdkj/dsh-plugins

DeepSeek Harness out-of-tree plugin: give a text-only coding model eyes by routing images to a Qwen-VL (DashScope) route through ctx.llm and returning text.

14 時間前スキルMIT