sunxin-ai/dsh-design-qa
C3Design-fidelity QA for DeepSeek Harness: lend any text-only model an eye, then judge whether the implementation matches the mock. Ships the benchmark behind that judgement — four fixtures, 23 injected
★ 10+ · sunxin-ai/dsh-design-qa source on GitHub · this plugin in the registry
sunxin-ai/dsh-design-qa is a DeepSeek Harness plugin rated C3 — powerful capability combined with sensitive behavior. It patches the dsh runtime, executes system commands, reads credential-class env vars.
Installable plugin — declares a dsh.bundle manifest
What it can do
| Capability | Flag | Evidence |
|---|---|---|
| patches the dsh runtime | runtime_patch | ./cordis.patch.yml |
| executes system commands | exec | ×7 in authored code, e.g. install.mjs:29, install.mjs:361 |
| reads credential-class env vars | token_env | BAILIAN_API_KEY |
Services it injects
attachments llm tools
Hooks it attaches
agent/pre-step
Outbound domains
dashscope.aliyuncs.com
Environment variables it reads
DSH_CWD DSH_RESTART_CMD DSH_HOME DSH_PROCESS_PATTERN DSH_RELAUNCH_PID DSH_RELAUNCH_CMD DSH_RELAUNCH_CWD DSH_RELAUNCH_ARGV DSH_REPO BAILIAN_API_KEY
How to read this
Levels measure capability surface and transparency, not maliciousness. A C3 plugin can be entirely legitimate — a desktop shell genuinely needs subprocesses. The point is that you can see this before installing. See the levels explained and how dsh plugins work.
Findings come from static analysis of shipped code; nothing is executed. Think a flag is wrong? Open an issue — every flag cites the file and line it came from.