Skip to content
dsh.fish

Browse

36 results

Bundle

oh-my-knowledge

OMK — Observe. Measure. Know. Evidence-backed knowledge changes for AI applications.

lizhiyao17
Bundle

dsh-eval-harness

DSH 插件回归评测门禁:yaml 用例 + headless 驱动 + trace 断言 + baseline 门禁(eval_run / eval_gate)

BiBoyang12
Bundle

tuningengines-cli

Tuning Engines CLI, MCP server, and Python agent runtime adapters for governed model, agent, skill, and MCP workflows. Fine-tune open-source LLMs, run inference, manage datasets/evaluations, and connect LangGraph or Temporal while Tuning Engines handles policy, audit, usage, and token economics.

cerebrixos-org5
Bundle

dsh-ecc-skills

ECC (227k-star operator system) skills for DeepSeek Harness — progressive port of 274 curated single-file skills (agentic engineering, evaluation, testing, patterns, vertical domains, docs). Adapted from affaan-m/ECC (MIT)

gongyijie855
Bundle

dsh-arena

Local-first experiment and evaluation workbench for DeepSeek Harness.

Apageoflove4
Bundle

dsh-llm-verifier

Best-of-3/5 orchestration and LLM-as-a-Verifier selection for DeepSeek Harness

Web09264
Bundle

dsh-profile-lab

Reproducible local experiment matrices for DSH profiles

young-tim3
Bundle

dsh-self-evolution

Benchmark-driven self-evolution plugin for DeepSeek Harness: evaluate, optimize, snapshot, accept or roll back.

Lhy7233
Bundle

dsh-eval-regression

Deterministic, CI-safe golden-output evaluation for DeepSeek Harness

aryswisnu2
Bundle

dsh-blind-arena

A blind, fair, local Agent arena inside DSH Web: same task, same commit, isolated worktrees, shared verification, judge before you reveal.

changer-changer2
Bundle

auto-pwa

AI-driven partial wave analysis for DeepSeek Harness: physics-gated config editing, ctpwa fit execution, numeric evaluation, and goal-driven iterative convergence. Physics knowledge (PDG-2026) as pure functions; DSH integration as a thin pwa_* tool plugin.

BHXiang2
Bundle

@webwalkerhq/dsh-replay-lab

Replay real DeepSeek Harness turns against Standard, Minimal, Anchored, or plugin candidates with frozen request-surface evidence

tbxy092
Bundle

dsh-skill-eval

Skill-trigger evaluation: an LLM judge recreates the DSH skill catalog and measures how reliably a skill description routes matching queries.

renjianguojinqianfan2
Bundle

dsh-harness-audit

Audit an agent harness against the harness-evaluation criteria, with machine-enforced evidence validation.

Leeaoyin1
Bundle

dsh-agent-observe

Agent observability plugin for DSH — behavior audit, cost tracking, anomaly detection

dsh-plugin-evaluation1
Bundle

dsh-northstar

A local DeepSeek Harness north-star guard with explicit AI indicator evaluation and task alignment context.

xDylanLong1
Bundle

dsh-dual-model-eval

Compare multiple coding models side by side in DeepSeek Harness with isolated Git worktrees

huangdaxianer1
Bundle

deepseek-harness-flow

Community visual workflow and multi-model evaluation plugin for DeepSeek Harness

alison-xx1
Bundle

@trapstreet/dsh-trapstreet

Check which DeepSeek Harness plugins actually loaded, and look up public evaluation boards on trapstreet.run

trapstreet1
Bundle

dsh-regression

Turn explicit coding-agent corrections into executable DeepSeek Harness regression tests.

chenghaoYang
Bundle

dsh-plugin-abtest

Paired experiments and promotion gates for DSH plugins.

Morriaty-The-Murderer