@allmodels/dsh-speech
Speech-to-text and spoken answer summaries for DeepSeek Harness using AllModels.io.
79 results
Speech-to-text and spoken answer summaries for DeepSeek Harness using AllModels.io.
DeepSeek Harness 本地离线语音插件:STT 语音识别 + TTS 语音合成 + WebUI 按住说话(Sherpa-ONNX)
DSH Web 语音输入插件:默认使用浏览器内置 Web Speech API,也可按需下载本地模型离线识别。
DSH WebUI 语音输入插件:输入框麦克风按钮(Alt+V)→ 录音 → 浏览器 Web Speech API 或本地 FunASR 后端转写 → 文本回填输入框(不自动发送)
Context-aware voice input for DeepSeek Harness with Web Speech, local SenseVoice transcription, model polish, editable Composer drafts, and user-controlled sending
Local IndexTTS 2.5 and GPT-SoVITS sentence-level speech synthesis and playback for DeepSeek Harness
DeepSeek Harness client plugin: a microphone dictation button to the left of the Full access control in the web composer.
Voice input plugin for DeepSeek Harness web (China-ready): Alibaba Cloud DashScope ASR via a local bridge. Mic button in the composer, streaming recognition, cursor-aware insertion, silence auto-stop.
Local-first, full-duplex voice for DeepSeek Harness, orchestrated by Muxiva
Voice input plugin for DeepSeek Harness
Voice practice mode for the DSH web app: bilingual (中文 / English) conversation output plus read-aloud (TTS) and speech input (STT). Adds a 语音交流 dialog-mode toggle to the composer.
Speech plugin for DeepSeek Harness: per-message speak button, composer voice input, and auto-announce toggle, over cloud TTS/ASR with Web Speech API fallback
Voice input (Web Speech API) and read-aloud (speechSynthesis) for the DeepSeek Harness web GUI — DSH 语音输入 + 回复朗读套件
OpenRouter image, video, and speech generation as dsh tools, shipped as an out-of-tree profile bundle
DSH Web 本地离线语音输入插件:浏览器采集麦克风 → host 拉起本地 FunASR (SenseVoiceSmall) 识别 → 文字填入输入框。
DeepSeek Harness plugin: a desktop pet pinned to the bottom-right of the screen that shows your DeepSeek account balance in its speech bubble (Windows)
App-free mobile voice calls with existing DeepSeek Harness sessions
Microphone speech-to-text input for the DeepSeek Harness Web UI
聊天框语音输入按钮 for DeepSeek Harness: 点击麦克风说话,多引擎转写(智谱 GLM-ASR-2512 / 本地 faster-whisper / Gemini / OpenAI)自动填入输入框。一个按钮,所见即所得。
Text-to-speech for DeepSeek Harness: speak agent replies in the Web UI with a provider fallback chain (OpenAI, ElevenLabs, Google, Azure, Groq, Deepgram, OpenRouter, Edge, Piper, eSpeak).
DSH web plugin: speak into the composer — macOS native speech-to-text (Apple Speech framework) via a bundled Swift helper
SenseVoiceSmall-powered local speech-to-text for the DSH Desktop input box: click the mic beside the composer, speak, and the recognized draft (with language/emotion tags) is written into the input.
Voice (speech-to-text) input for the DSH Web composer via the browser Web Speech API — tap the mic, speak, the transcript fills the input box. Zero server, zero API keys, nothing leaves the machine. · DSH Web 语音输入:点麦克风说话,识别文字回填输入框(Chrome/Edge,无需 API key)。
Whale girl desktop pet for DeepSeek Harness: browser floating mascot + Windows desktop pet, balance watching, task notices, 19 voiced lines with per-line emotion, head-pat interaction and presence-aware idle speech.