Zum Hauptinhalt springen

Plugins

Durchsuche, filtere und installiere DeepSeek-Harness-Plugins.

15 Plugins gefunden

W

dsh-ears

wiziscool/dsh-ears

Spracheingabe-Plugin für DeepSeek Harness (dsh): Eine Mikrofon-Schaltfläche im Composer wandelt Sprache in einen Entwurfstranskript um, mit wählbaren Spracherkennungs-Backends, optionaler Nachbearbeitung über die eigenen LLM-Routen von dsh und einer nativen Einstellungsseite.

21gesternVision, Sprache & MultimodalMIT
C

dsh-bilibili

czx2244/dsh-bilibili

Bilibili-Video-Analyse: Metadaten, Transkript (ASR-Fallback über Bijian/sherpa-onnx/whisper.cpp), Kommentare, Danmaku und scharfe Keyframes mit optionalen lokalen Vision-Beschreibungen.

9vor 2 MonatenTools & FunktionenMIT
I

dsh-video-understand

ilps2/dsh-video-understand

Kostengünstiges Video-Verständnis-Tool: Ein video_understand-Tool wandelt einen Bilibili-Link / BV / ein lokales Video in eine AVIS-Informationsschicht (ASR + Szenenstruktur + Objekt-Tracking + YOLO-Labels) um und liefert Zusammenfassung + Q&A. Fragegesteuertes dynamisches Routing über Ebenen hinweg (L0 ASR / L1 Objekt-Tracking / L2 Keyframe-VLM), Wiederverwendung der semantischen Schicht für Wiederholungsfragen, Budgetobergrenze pro Frage. Python-Engine: Die Kernschicht benötigt faster-whisper / opencv / yt-dlp (~200-300MB); die optionale semantische Schicht fügt ~2GB torch / transformers / ultralytics hinzu. Ein mitgeliefertes doctor --fix richtet das venv ein und installiert beide.

8letzten MonatTools & Funktionen
G

dsh-voice

goodandready/dsh-voice

Spracheingabe für die Web-UI: nach Pausen segmentiertes Diktat und Sprachnachrichten, jeweils mit eigener Anbieter-Fallback-Kette (Deepgram, Groq, HuggingFace, lokales whisper.cpp oder ein OpenAI-kompatibler Endpunkt).

8gesternVision, Sprache & MultimodalMIT
S

dsh-voice

stardustlc666/dsh-voice

Sprachwerkzeuge: kostenlose neuronale Sprachsynthese mit edge-tts, OpenAI-kompatible ASR-Transkription, Stimmliste, Massenvorschau der Stimmen und Integritäts-Selbstprüfung.

4vor 10 StundenVision, Sprache & MultimodalMIT
B

dsh-stt-input

baisama-cloud/dsh-stt-input

Speech-to-Text-Spracheingabe für die Web-UI: ein Mikrofon-Button im Composer transkribiert Sprache per Browser Web Speech API (null Konfiguration) oder eine OpenAI-kompatible Whisper-API (OpenAI / Groq) in den Entwurf, mit auswählbarem Modell und Sprache in den Settings.

3letzten MonatVision, Sprache & MultimodalMIT
Z

dsh-watch-video

zeshuochen/dsh-watch-video

Auf Untertitel ausgerichtete Transkription von Videos mit SRT-Export, abbrechbaren Steuerelementen für den Auftrag und einem Rückgriff auf faster-whisper large-v3, wenn keine Untertitel verfügbar sind.

2letzten MonatVision, Sprache & Multimodal
1

dsh-wsl-im

173787247/dsh-wsl-im

Bridges Feishu, WeCom, DingTalk, QQ, Slack, Discord, Telegram and Mattermost into dsh agents (outbound WS/Stream/Gateway/long-poll/webhook), with im_status, optional local Whisper ASR, plain-text outbound for QQ/DingTalk/Telegram/Mattermost, and env CSV allowlists (DSH_IM_*_ALLOWED_USER_IDS). Empty allowedUserIds means EVERYONE can drive your agent — set a whitelist before exposing a bot.

1vor 3 TagenEntwicklung & Plugin-ToolsMIT
V

dsh-live-voice

victorwads/dsh-live-voice

Local-first voice conversations for DSH, with local speech recognition and synthesis and optional external providers.

1vor 18 TagenVision, Sprache & MultimodalGPL-3.0
Z

dsh-voice

zhuiyueya/dsh-voice

Sprache für DeepSeek Harness (dsh) — Sprache-zu-Text-Eingabe + Vorlese-TTS für rein textbasiertes DeepSeek, ohne API-Schlüssel.

1vor 2 MonatenVision, Sprache & MultimodalMIT
Y

dsh-video-to-notes

yll-kb/dsh-video-to-notes

Opt-in DeepSeek Harness skill bundle that turns course, lecture, tutorial, documentary, meeting, and talk videos into structured study notes.

0vor 12 TagenVision, Sprache & MultimodalMIT
B

dsh-asr-voice

bittersmilezzz/dsh-asr-voice

开口即成文 · Speak-to-prompt for DeepSeek Harness:云端 ASR 语音识别 + 提示词优化 + 填入草稿/自动发送,跨平台 macOS / Windows。

0vor 17 TagenVision, Sprache & MultimodalMIT
K

dsh-kitt-voice

kittcat-lab/dsh-kitt-voice

Voice for the DeepSeek Harness: speak to the agent, hear it back, and see what it is doing from a floating window that stays on top of whatever you are running.

0vor 28 TagenVision, Sprache & MultimodalMIT
N

dsh-voice

navid-kianfar/dsh-voice

Dictate prompts into the DeepSeek Harness Web Client — a microphone in the composer, with swappable transcription: hosted Whisper API, self-hosted server, or a fully offline whisper.cpp binary.

0letzten MonatVision, Sprache & MultimodalMIT
Z

dsh-video-understand

zeshuochen/dsh-video-understand

Subtitle-first video transcription and deterministic extractive Markdown summaries, with a faster-whisper large-v3 fallback when subtitles are unavailable.

0letzten MonatVision, Sprache & Multimodal