@biliye/dsh-voice-call
个人语音通话助手:专属工作区/会话(自动创建、跨重启保持)、悬浮球通话面板、FunASR 流式/HTTP 语音识别、云端 TTS 语音回复、唤醒词通话模式(可配置唤醒词/休眠时长,建议本地 ASR)、子代理任务分发与进度跟踪。持久化安装,重启后仍在「设置 → 插件」中可见。
57 results
个人语音通话助手:专属工作区/会话(自动创建、跨重启保持)、悬浮球通话面板、FunASR 流式/HTTP 语音识别、云端 TTS 语音回复、唤醒词通话模式(可配置唤醒词/休眠时长,建议本地 ASR)、子代理任务分发与进度跟踪。持久化安装,重启后仍在「设置 → 插件」中可见。
Aura Vision — free vision OCR plugin for DeepSeek Harness web profile: Zhipu GLM-4V-Flash (free tier), adaptive tile recognition for long documents, history with favorites and Markdown/Excel/Word/PNG export.
Voice input plugin for DeepSeek Harness web (China-ready): Alibaba Cloud DashScope ASR via a local bridge. Mic button in the composer, streaming recognition, cursor-aware insertion, silence auto-stop.
App-free mobile voice calls with existing DeepSeek Harness sessions
DeepSeek web vision bridge: give text-only DeepSeek models image recognition via a deepseek_vision tool
Session-scoped full-duplex voice assistant with tool-controlled drafting and Agent submission for DeepSeek Harness.
DSH plugin: prompt the agent to dispatch image recognition to an opencode-go mimo-v2.5 subagent
Speech plugin for DeepSeek Harness: per-message speak button, composer voice input, and auto-announce toggle, over cloud TTS/ASR with Web Speech API fallback
Adaptive image routing for DeepSeek Harness: pass images directly to native multimodal models and transcribe them only for text-only models.
OmniVision for DeepSeek Harness: an OmniParser-powered GUI agent plugin — screen capture, element recognition, click/type automation and a browser vision dock with recognition history, diffing and summary
Microphone speech-to-text input for the DeepSeek Harness Web UI
SenseVoiceSmall-powered local speech-to-text for the DSH Desktop input box: click the mic beside the composer, speak, and the recognized draft (with language/emotion tags) is written into the input.
A microphone button for DeepSeek Harness that writes browser speech recognition into the composer draft.
Optional Web UI enhancements for DeepSeek Harness
DSH 插件:通过硅基流动(SiliconFlow)视觉大模型识别/分析图片,支持本地文件路径、http(s) 图片 URL 与 base64 data URL。含持久化的粘贴识别面板(web 端)。
Vision recognition plugin for DeepSeek Harness: paste images into the composer, recognize them via GLM-4V on the host side, and inject the result into the conversation.
Voice input plugin for DeepSeek Harness web: a mic button in the composer that recognizes speech (Web Speech API) and fills the draft, with optional auto-send.
明眸 VisionBridge - 自研视觉桥:瞎子模型收图时自动调用视觉模型识别,把识别文字喂回主模型,无感、可配置、升级不丢
Offline Parakeet voice input for DeepSeek Harness
SimpleTex 专用 OCR/公式识别工具 simpletex_recognize:task=formula 返回 LaTeX(含置信度),task=text 返回 Markdown,task=auto 按问题关键词路由。GUI 设置卡热配置 token/端点/上限。
识别昆虫或其他节肢动物名称(或所属目, 科, 属, 种)。
对包含主体物体的图像进行标签识别,输出主体物体的类别标签,目前已经覆盖了5万多类的物体类别。
对含有动物的图像进行标签识别,无需任何额外输入,输出动物的类别标签。
Editable local and desktop dictation for DeepSeek Harness