fable-method
Sahir619
The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.
GITHUB TOPIC
9projects include this topic
The "evaluation" topic on GitHub groups 9 open-source projects in the DeepSeek Harness (DSH) ecosystem, led by fable-method with 2.3k GitHub stars. fable-method — The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Every project here is indexed by DSH Universe with live GitHub data — stars, activity and install status — so you can compare and install directly.
{count} projects
Exact GitHub Topic match
Sahir619
The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.
BiBoyang
DSH plugin eval tool: YAML case-driven real agent regression eval + baseline comparison PASS/WARN/FAIL gates|Regression eval harness for DeepSeek Harness plugins
Apageoflove
Local-first experiment and evaluation workbench plugin for DeepSeek Harness (DSH).
aryswisnu
No project description provided for this repository yet.
young-tim
Reproducible DSH profile and patch experiment matrices with reports and policy gates
CZ-ZL
Composable DSH-native evaluation and optimization with explicit evidence modes, bounded budgets and traceable results
hj01857655
Measure whether a change to your dsh setup actually helped: register repeatable cases, run them, diff before/after. DeepSeek Harness plugin.
ruby1304
Public, reusable DeepSeek Harness plugins and skills: workflow canvas toolkit, blind eval harness, LLM cost lab, incident ledger.
ubik-dsh
Clean, portable agent skills for DeepSeek Harness, Claude Code and any harness reading the Agent Skills standard. Plain text, nothing to build. Every skill ships a checker and a measured evaluation protocol — and says what was not verified.