dsh-omi-voice
DeepSeek Harness 对话内朗读:豆包音质 · 点读/暂停/继续 · 豆包 Key 只留在 Omi 引擎(BYOK)
110 results
DeepSeek Harness 对话内朗读:豆包音质 · 点读/暂停/继续 · 豆包 Key 只留在 Omi 引擎(BYOK)
DeepSeek Harness (DSH) plugin: one-click prompt enhancement (✨) and voice recognition (💬, cloud/local dual engines) for the composer, plus one-click DSH service restart.
Agent skill for DeepSeek Harness: rewrites Russian text to remove 64 markers of AI generation (bureaucratese, calques, ChatGPT fingerprints), with a corpus-calibrated scanner, audit mode and author-voice calibration.
DSH 专属语音输入插件:点按或按住 Alt 说话、松开/再点按转文字,支持热词替换表(hot.txt)、自定义润色提示词、录音电平指示。默认本地离线识别(SenseVoice,零配置零 key、音频不出本机),自动回退浏览器 Web Speech,可选云端 ASR 与润色(复用 DSH 模型)。Voice input for DeepSeek Harness: tap or hold Alt to talk, get text in the composer — local SenseVoice by default, zero config, zero API key.
Voice input plugin for DeepSeek Harness
Make your AI harness speak — voice announcements for DSH and other AI coding harnesses (Windows SAPI5 + macOS system voices)
Voice-first session loop for DeepSeek Harness: a composer microphone button with browser/local speech-to-text (Web Speech, FunASR, whisper.cpp), a speak tool for text-to-speech replies (browser, edge-tts, piper), event announcements with mute, and speak-to-interrupt.
DeepSeek Harness plugin: per-event task-completion and attention sounds for turn-end / approval / question / plan-review / goal-blocked / task-failure, each with its own sound (built-in synth, local audio file) and volume (Web UI)
Voice input plugin for DeepSeek Harness: microphone → local/browser speech recognition → text submitted as a normal chat message (input-only, preset-agnostic)
Voice AI girlfriend for DeepSeek Harness: FunASR mic input, Qwen3-TTS voice replies, companion animation window, QQ two-way chat. Needs the repo's voice bridge + NapCat.
DeepSeek Harness (dsh) plugin: run work as a small crew of role agents (product manager, researcher, architect, engineer, test engineer, code engineer, QA, code reviewer, security reviewer, doc reviewer) that talk through files on disk, with the PM as the only voice to the user.
夸夸、运势、战报、番茄钟、摸鱼、沉浸氛围、桌宠语音、Live2D、Boss 隐身与代码花园一体化的 DeepSeek Harness 插件
Composer mic for DeepSeek Harness Web: tap-to-monitor live transcription and hold-to-talk, with host Edge TTS reply reading that streams while the model generates, echo-pause during reading, and tap-to-stop.
Voice input for DeepSeek Harness Web UI: a mic button in the composer tool row that uses the browser's Web Speech API (Chrome/Edge) to transcribe speech directly into the message draft. Zero dependencies, no API keys.
DSH web GUI notification alerts: distinct synthesized tones or a spoken voice (zh/en) for 'needs approval', 'needs answer', 'output complete' and 'error', with per-type sound, enable, volume, repeat, browser notifications and an i18n interface-language setting.
Your DeepSeek Harness agent rings your actual phone. It asks out loud, you answer out loud, and what you said steers the run.
Guide Dog for DSH, powered by MiniMax — multimodal plugin: image/video/music/speech generation, vision inspection tools, voice mode, microphone voice input and real-time voice call mode.
Yukino (Yukinoshita Yukino) standalone route plugin for DSH: session tree by workspace, task-done voice alerts, no context injection into DSH.
DeepSeek Harness 插件:通知出口——agent 通过桌面通知 / 中文语音播报 / 提示音主动联系用户(长任务完成、出错、呼叫用户回来)。Windows 本机零依赖。
Full-duplex voice mode for DeepSeek Harness: streamed ASR -> LLM -> TTS with barge-in
DSH Web 麦克风语音输入插件:浏览器内置 Web Speech API 实时转写进输入框,自动去重/续听、智能标点、语言与自动发送设置(Edge=微软语音、Chrome=谷歌语音)。Microphone voice input for the DSH Web UI using the browser's Web Speech API.
Text-to-speech for dsh web: speaks each assistant reply out loud using Edge TTS neural voices (host bridge + browser player)
ChatVoice — free voice input + AI reply read-aloud for DeepSeek Harness (dsh). Zero config, zero cost, no API key, built on the browser's native Web Speech API. 给 DSH 装上「免费、免 API key、开箱即用」的语音输入 + 回复朗读闭环(中文优先)。
VocoType voice input bridge for DSH Web: mic button + recording panel in the composer, auto-insert recognized text (dedupe, auto-launch/deploy, mtime-optimized polling)