dsh-mic-dictation
DeepSeek Harness client plugin: a microphone dictation button to the left of the Full access control in the web composer.
57 results
DeepSeek Harness client plugin: a microphone dictation button to the left of the Full access control in the web composer.
Voice input plugin for DeepSeek Harness web (China-ready): Alibaba Cloud DashScope ASR via a local bridge. Mic button in the composer, streaming recognition, cursor-aware insertion, silence auto-stop.
dsh bundle plugin: grok2api image & video generation tools plus image recognition (generate_image / generate_video / recognize_image) for the DeepSeek Harness
Speech plugin for DeepSeek Harness: per-message speak button, composer voice input, and auto-announce toggle, over cloud TTS/ASR with Web Speech API fallback
Create production-ready AI image prompts and image descriptions with the Find Image Prompt image-to-prompt API.
App-free mobile voice calls with existing DeepSeek Harness sessions
Microphone speech-to-text input for the DeepSeek Harness Web UI
SenseVoiceSmall-powered local speech-to-text for the DSH Desktop input box: click the mic beside the composer, speak, and the recognized draft (with language/emotion tags) is written into the input.
DSH plugin: prompt the agent to dispatch image recognition to an opencode-go mimo-v2.5 subagent
Voice input plugin for DeepSeek Harness web: a mic button in the composer that recognizes speech (Web Speech API) and fills the draft, with optional auto-send.
Vision recognition plugin for DeepSeek Harness: paste images into the composer, recognize them via GLM-4V on the host side, and inject the result into the conversation.
Session-scoped full-duplex voice assistant with tool-controlled drafting and Agent submission for DeepSeek Harness.
Aura Vision — free vision OCR plugin for DeepSeek Harness web profile: Zhipu GLM-4V-Flash (free tier), adaptive tile recognition for long documents, history with favorites and Markdown/Excel/Word/PNG export.
DSH vision/image-recognition plugin: enables AI to understand images with a configurable vision model
A microphone button for DeepSeek Harness that writes browser speech recognition into the composer draft.
Persistent voice conversations for DSH with cloud speech recognition, Edge TTS, and background Agent delegation
Offline Parakeet voice input for DeepSeek Harness
明眸 VisionBridge - 自研视觉桥:瞎子模型收图时自动调用视觉模型识别,把识别文字喂回主模型,无感、可配置、升级不丢
SimpleTex 专用 OCR/公式识别工具 simpletex_recognize:task=formula 返回 LaTeX(含置信度),task=text 返回 Markdown,task=auto 按问题关键词路由。GUI 设置卡热配置 token/端点/上限。
识别昆虫或其他节肢动物名称(或所属目, 科, 属, 种)。
对含有动物的图像进行标签识别,无需任何额外输入,输出动物的类别标签。
包括通用文本识别、手写识别、车牌识别、身份证识别、护照识别、港澳台通行证识别、银行卡识别、营业执照识别、驾驶证识别、行驶证识别。
识别植物名称(或所属科, 属, 种或亚种)。
对包含主体物体的图像进行标签识别,输出主体物体的类别标签,目前已经覆盖了5万多类的物体类别。