CharlesLiuZC/deepseek-harness-voice-context--packages-client-ui-voice-context ↗★ 2

@deepseek-ai/dsh-client-ui-voice-context

网页端语音输入界面:输入框麦克风按钮与密钥设置页。 适合想在网页对话中用语音转写输入的用户,需配置云端或本地识别。

包名
@deepseek-ai/dsh-client-ui-voice-context
兼容性
待验证
Harness 依赖范围
workspace:^
Cordis 依赖范围
workspace:^
版本
0.1.0-rc.5
许可证
MIT
最近更新
2026年8月14日

安装

此插件尚未提供可验证的 bundle,或兼容性检查未通过。请先阅读仓库说明。 阅读完整 README ↗

@deepseek-ai/dsh-client-ui-voice-context

English | 中文

Voice-Context Web surface, browser half: a mic button in the conversation.input.left composer tool row (order 100) records an utterance, encodes it to base64, transcribes through ctx.remote.voiceContext.transcribe(...), and appends the text through inputActions.setDraft. A settings.section page (order 40) provides first-time routing between the cloud API and local SenseVoiceSmall/faster-whisper small, medium, or large-v3. The controlled backend/model pair is stored in browser local storage; no URL is browser-configurable. The cloud API key is written through credentials.set under SILICONFLOW_API_KEY; the page reads only configured/writable state, never the value. MediaRecorder output is decoded and re-encoded as 16 kHz mono 16-bit PCM WAV so every ASR backend accepts the container.

Model Experience

Indirectly, through the composer draft that reaches a model request only when the user submits it as an ordinary prompt.

KV Cache effect

None unless the user submits the transcribed draft; it then extends history like any other user message.

Known Limitations and Deferred Work

  • Inline status only — transcription errors surface on the mic button's tooltip, not through the composer notice channel.
  • No live transcript preview — the final transcript appears only after the Remote settles; streaming results are future work.
  • Browser-local selection — backend/model preference does not synchronize between browsers or profiles.