Перейти к основному содержимому
C

dsh-read-aloud

cccc12138/dsh-read-aloud

Adds a speaker button immediately right of the Like button on every finalized assistant reply; click it to hear the reply through the browser speech engine, and hover it for a speed and voice panel.

Установка

dsh plugin --profile web add github:cccc12138/dsh-read-aloud

README

dsh-read-aloud

English | 中文

A speaker button beside the Like button: read any DeepSeek Harness reply aloud.

Every finalized assistant message gets one extra action in the row it already has, immediately to the right of 👍👎:

copy · 👍 👎 · 🔊 · branch

Click it and the reply is spoken by your browser's own speech engine. Nothing is uploaded, no API key is involved, and no host-side code runs.

Install

dsh plugin --profile web add dsh-read-aloud

Or install it from Settings → Plugin Market. Most installs go live after a page refresh.

Use

ActionResult
Click the speakerStarts reading this reply from the top
Click again while readingPauses; the icon becomes ▶
Click again while pausedResumes from the same sentence
Click another message's speakerStops the current one and starts that one
EscStops immediately, wherever the pointer is
Hover the speakerOpens the speed and voice panel

The button shows three states: 🔈 idle, ⏸ reading, ▶ paused. Reading ends by itself, and the icon returns to 🔈.

Speed and voice

Hovering the button opens a small panel that stays with the message you are listening to:

Speed   [0.5×] [0.75×] [1×] [1.25×] [1.5×] [1.75×] [2×]
Voice   [ System default ▾ ]

Both choices take effect on the sentence being read and are remembered in the browser's local storage. Voices come from your operating system; the list is sorted with Chinese voices first, and System default picks a voice matching your interface language.

What gets read

Replies are Markdown, and reading Markdown literally sounds terrible — URLs get spelled out character by character and code blocks become noise. So the text is cleaned first:

KeptDropped
Prose and headingsFenced code blocks (silently)
List items, with their markers removedTable rows
Inline code content, without the backticksImage syntax
Link labelsLink targets and bare URLs
Emphasis text, without ** and *HTML tags, file paths, emoji

Long replies are read in full — nothing is truncated. The text is queued in sentence-sized pieces rather than handed to the engine in one lump, because several engines silently cut short or drop a single very long utterance.

Requirements

  • DeepSeek Harness 0.1.2-rc.1 or newer
  • A browser with the Web Speech API. Speech comes from your operating system's installed voices, so an OS with no voice for the reply's language will stay silent — Windows and macOS both ship usable voices, and Edge exposes additional natural voices.
  • If the engine is missing entirely, the button says so instead of failing silently.

Privacy

  • No network requests. The plugin never contacts a server.
  • No API keys, no accounts, no telemetry.
  • No host-side code: lib/index.js is an empty apply that exists only so the plugin appears in the profile's loader.
  • Speed and voice live in your browser's local storage under dsh-read-aloud/settings. Clearing site data resets them to and System default; nothing else is affected.

Compatibility

The plugin declares engines.dsh >= 0.1.2-rc.1 and registers one entry in the conversation.chat.assistant-actions slot at order: 20, so it sits beside the shipped feedback entry (order: 10) without replacing it. It reads the reply text through the slot's own useChat standard prop, so it needs no host RPC and no DOM scraping.

Development

npm test

The bundle is hand-written in the client-module format the web shell loads, so there is no build step and no bundler — lib/client.js is shipped as authored. The test loads that shipped file with a stubbed module graph and exercises the text-cleaning, chunking, message-lookup and voice-listing helpers.

License

MIT

Похожие плагины