@difimim/dsh-voice-input
DeepSeek Harness 语音输入插件:在输入框加入麦克风按钮,用浏览器 Web Speech API 把语音实时转成文字填入输入框。
129 results
DeepSeek Harness 语音输入插件:在输入框加入麦克风按钮,用浏览器 Web Speech API 把语音实时转成文字填入输入框。
Speech capability plugin for the DeepSeek Harness (dsh) web host: a token-gated /s/api route family serving audio transcription (ASR) and synthesis (TTS) over configurable providers
Telegram messenger bridge for DeepSeek Harness: sessions, steer, homes, inline asks, notify bridge, and optional TTS voice notes.
语音朗读(MiniMax / OpenAI 兼容 TTS,服务商与音色自选)——每条回复旁的朗读按钮、输入框自动朗读开关、设置页自选服务商与音色。零构建,纯 JS。
MiniMax speech-to-text (asr-1.0) and text-to-speech (speech-2.8-hd) as a global DeepSeek Harness plugin: a transcribe_audio tool, an announce_speech tool, a settings card, composer voice input, spoken turn announcements, and handsfree conversation.
Voice conversation mode for the DeepSeek Harness Web GUI: mic input with auto-submit, streaming spoken readout of replies, and a hands-free multi-turn loop with barge-in.
dsh-gal —— DeepSeek Harness 的立绘挂件:实时显示余额与今日消耗、每轮对话结束结算 token 与花费、点击立绘随机播语音并逐字显示台词,立绘包与语音包都能换
Speech-to-text (voice input) plugin for DeepSeek Harness — a microphone toggle in the composer tool row and a Voice Input settings page.
Voice for DeepSeek Harness: an AI voice-reply tool, a persistent voice bar, a Voice settings page, and a boot sound module — all backed by Alibaba Cloud Bailian TTS.
Voice chat for DSH Web: speech-to-text input (Web Speech) + auto/manual read-aloud of assistant replies (TTS) with voice/rate/pitch settings. Pure client UI, minimal Node entry.
Voice alert plugin for DSH: plays voice, shows popup and system notification when a task completes or the agent asks the user
Voice input plugin for DeepSeek Harness web: a mic button in the composer that recognizes speech (Web Speech API) and fills the draft, with optional auto-send.
Multi-platform IM gateway for DeepSeek Harness (fork of dsh-im-hub with media): Telegram voice/photo/document/video + reply handling, STT (Deepgram primary, HF Whisper fallback), outbound MEDIA: markers. Feishu/WeCom kept as-is.
Speech suite for DeepSeek Harness: free edge-tts page announce, speech-to-text voice input (Bailian paraformer-realtime-v2) with Alt+Q hotkey, tap/hold modes, auto-send, and stop-playback-on-record
Offline Parakeet voice input for DeepSeek Harness
SenseVoice 语音输入插件 for DeepSeek Harness:在对话输入框旁添加麦克风按钮,录音后调用本地 SenseVoice 服务转成文本填入输入框。首次使用自动下载模型并显示进度,后端由插件自动启动。
Voice input for DeepSeek Harness: a microphone seat inside the composer plus a Web settings page, transcribing through the browser engine or any OpenAI-compatible /audio/transcriptions endpoint.
Real-time duplex voice for DeepSeek Harness: Volcengine streaming ASR/TTS, agent reply narration, barge-in, wake word, live captions. | 实时双工语音插件:火山流式 ASR/TTS、回复朗读、打断、唤醒词、实时字幕。
Voice input for DeepSeek Harness: dictation chunked by pauses and voice messages, each with its own provider fallback chain (Deepgram, Groq, HuggingFace, local whisper.cpp, plus any OpenAI-compatible API of your own).
DSH plugin: compress verbose voice-dictation text into token-efficient prompts, fully local.
DSH skill bundle: the Agora skill (RTC, RTM, ConvoAI, CLI, Cloud Recording, tokens) synced verbatim from AgoraIO/skills at a pinned release tag. Install to add the `agora` skill to DeepSeek Harness.
DSH WebUI 语音输入插件(火山引擎流式 ASR / 豆包 Seed ASR):输入框麦克风按钮(Alt+V)→ 浏览器采集 16kHz PCM → 经宿主 WebSocket 中继到豆包流式识别 → 实时回填(跟随光标/Proma 式输出,失败兜底剪贴板)。协议实现源自 Proma 桌面端调研复刻。
TTS 语音播放插件:为 DSH 聊天界面添加语音朗读功能,支持配置本地 TTS 服务,每条助手消息可朗读,AI 可调用 tts-speak 工具发送语音。
dsh-voice — turn-based voice loop for DeepSeek Harness: pluggable Qwen / MiMo / local ASR+TTS engines, agent-driven speak/listen tools and browser PTT UI, built for interviewer presets