dsh-multimodal-bridge
DeepSeek Harness plugin bundle: qwen_vision (Qwen-VL image understanding) and qwen_generate (Qwen-Image text-to-image and image editing) tools for text-only models
28 results
DeepSeek Harness plugin bundle: qwen_vision (Qwen-VL image understanding) and qwen_generate (Qwen-Image text-to-image and image editing) tools for text-only models
Share DSH Q&As or selected conversation groups as PNG or Markdown.
Session-aware Pencil integration for DeepSeek Harness with official MCP tools and an on-demand browser canvas.
Persistent DeepSeek-inspired whale-dive and reactive-water animation for the Harness Web turn status.
Eyes for text-only DeepSeek on DeepSeek Harness: image understanding via Doubao Web by default (zero cost, no API key), Antigravity IDE quota (flash/pro), Gemini API, or Cockpit proxy — model-invokable vision tool, wrapper adapters, evidence memory, and a bilingual client panel.
DSKIN - cartoon pixel kittens for DeepSeek Harness (DSH): 1-4 random kittens spawn each session, they play together, you can pet them (hearts) and drag them around. Original UI untouched.
Native Cordis WeShop canvas, tools, skills, and Web UI for DeepSeek Harness
DSH 视频创作技能插件:注册 Remotion 官方移植技能(React 编程式视频:动画/音频/字幕/3D/图表/字体,38 个规则文件),安装即用。
Eight DeepSeek Harness plugins: persona, language guard, vision fallback, python workdir guard, windows encoding guard, cross-agent memory, image generation, and skill shell injection.
DSH-native video understanding with configurable multimodal providers
DeepSeek Harness 原生视觉 Bundle:粘贴或拖入图片,通过托管的 deepseek-vision-mcp 调用 OpenAI 兼容视觉模型。
DeepSeek Harness Web UI plugin that renders settled Mermaid fenced code blocks as diagrams
SeekWhale — the DeepSeek pixel whale as a DSH web overlay pet.
MiniMax 文生图插件:把一句话画面描述生成图片并保存到工作区,作为 `image-gen` 工具提供给模型调用。
Multi-provider AI video/image generation plugin for DeepSeek Harness (DSH). ComfyUI-style node canvas studio + multi-provider settings, with agent tools list_video_models / generate_video / poll_video_task. | DeepSeek Harness AI 视频/图片生成插件:多提供者、节点画布工作室、Agent 工具与 REST RPC。
低成本视频理解工具:B站链接/BV/本地视频 → 信息层(ASR+场景+对象轨迹+YOLO)→ 摘要+问答。问题驱动动态路由分层(L0/L1/L2)、语义层复用、预算上限。引擎自包含,无需外部依赖。
Configure model input capabilities and reasoning efforts for DeepSeek Harness
Visual plan mode for DeepSeek Harness: structured plan.json + plan.md, an editable React Flow canvas, comments, plan diff, versioned revisions, and reliable write-back to the agent.
视觉层级构建参考
Zhipu BigModel capabilities for the DeepSeek Harness in one plugin: web_search_prime search provider, webReader fetch provider (server-side rendered), and GLM-4.6V vision (vision_analyze tool + pasted-image pre-step hook), all behind one ZAI_API_KEY
常驻视觉服务:直连视觉模型(默认 opencode-go/minimax-m3,回退 zai-coding-cn/glm-4.6v)。describe_image / subagent_vision 工具 + 粘贴图片自动转译(llm/stream 钩子)+ 输入框视觉状态小胶囊与详情页(活动日志:指令/思考过程/输出)。零子代理、零 agent 上下文开销,按会话记忆窗支持视觉追问与验收。
Configurable anchored popover and settings animations for DeepSeek Harness.
Image translation for non-multimodal models via GLM-4V-Flash: intercepts images, generates descriptions, injects as text.
Shots tab for DeepSeek Harness: a video-player view (live + YouTube-style scrubber) over a browser daemon's screenshot feed in <workspace>/shots/