dsh-aris-panel
Đã xác minhdsh-aris-panel · v0.2.8 · MIT · Giao diện web
ARIS research-workflow skills + interactive workbench panel for DeepSeek Harness (user fork of dsh-aris: 83 skills, bundled upstream dsh-aris 0.1.1 payload, ppt-master PPTX backend for the talk pipeline, cross-model Codex review, idea/workflow launcher, r
Cài đặt
dsh plugin add dsh-aris-panel Xác nhận layer đã áp bằng dsh --profile default --dump-config — xem hướng dẫn cài plugin.
Mã nguồn
Phát hành lên npm mà không có repository công khai. Hãy kiểm tra nội dung package trước khi cài.
Thẻ
Readme
ARIS on DeepSeek Harness
English | 中文
The
dsh-arisdistribution branch. The full ARIS project — every workflow, the docs, and the other host adaptations — lives onmain.
Runs the ARIS research workflow inside DeepSeek Harness: all 82 skills in the native skill catalog, with cross-model adversarial review through Codex.
The skills are unmodified. This bundle is one configuration layer plus a short adapter — it patches no Harness code and forks nothing.
Install
dsh plugin --profile web add dsh-aris
Restart the profile afterwards; plugin code loads at startup, so reloading the page is not enough.
dsh plugin shells out to pnpm, which the Harness does not bundle. Without it on PATH the install exits before doing anything.
Prerequisites
Codex CLI, installed and authenticated. It is the independent reviewer. The bundle spawns codex mcp-server and never overrides its model or reasoning effort — ~/.codex/config.toml is the reviewer posture contract. ARIS expects a non-DeepSeek family at xhigh:
model = "gpt-5.6-sol"
model_reasoning_effort = "xhigh"
If Codex cannot start, the Harness fails to boot rather than running a composition with no reviewer. That is deliberate: ARIS without an independent reviewer is not ARIS.
A DeepSeek API key, through the Harness Models page or DEEPSEEK_API_KEY.
Behind an HTTP proxy, start the Harness with NODE_USE_ENV_PROXY=1. Node's fetch ignores http_proxy otherwise, and model requests fail with TRANSPORT. It must be set when Node starts; configuration cannot repair an already-started process.
Optional — bound the executor too. ARIS's reviewers already carry scope limits: the block that forbids proposing hashes, defensive scaffolding, corner-case hardening, and over-mechanized judgement is embedded in the skills that produce review prompts, and applies with no setup. Nothing bounds the executor the same way. Two ways to close that, and they are different things:
- ARIS's own limits, one flag. The
aris-scope-limitsrow ships disabled; setdisabled: falsein your profile's patch to apply the same block to the executor. It reads the packagedskills/shared-references/review-scope-limits.mdat load, so the reviewer path and the executor path cannot drift apart. Off by default because it spends tokens on every request. - The full HERO contract. Paste its canonical block into the project's
AGENTS.mdorCLAUDE.md, or into$DSH_HOME/AGENTS.mdfor every dsh project — the Harness loads those files itself. This bundle does not vendor HERO's text; its canonical home is HERO's ownRULES.md.
Verify
dsh --profile web --dump-config | grep -A2 aris-
Then, in a session, type / — the skill menu lists the ARIS skills. To check the part that matters, ask the model to call mcp__codex__codex with a trivial prompt, report the threadId it can see, then continue that thread once with mcp__codex__codex-reply. A visible threadId is what makes multi-round review work.
The ARIS tab
In the Web UI a session gains an ARIS tab after Chat and Trajectory
(order: 15, between Trajectory and LiveBench). The tab has two layers:
- A read-only run-status block on top: it shows what the reviewer and the
loop's own artifacts say, and never writes, advances a round, or turns a score
into a completion decision. It reads
review-stage/REVIEW_STATE.jsonfrom the session's workspace, so it stays empty until anauto-review-looprun finishes a round there. - The interactive ARIS workbench below it: skill launcher, workflow guide,
idea board, experiments, wiki graph, and audit chain. It reads workspace
artifacts through two JSON endpoints —
/api/aris-run(run status) and/api/aris-wb(workbench) — and is always bound to the current session. The old bottom-left full-screen workbench overlay is gone; the composer stays available while the tab is open.
The top run-status summary is the reason the tab exists: who reviewed this
round, and
whether ARIS could verify that the reviewer belongs to a different model family
than the executor. On the Codex backend that verification is unverified by
design — Codex reports its own model, but nothing independently attests the
executor's, so ARIS records identity_assurance: caller_declared. Read it as
"route-consistent, not attested".
Two honest limits are shown in the tab and worth repeating: state is written
when a round finishes, not continuously, so a long round displays the previous
one; and completed means the loop ended — a positive assessment or the round
cap — never that the work was acquitted.
What the layer changes
| Row | Effect |
|---|---|
agent-default-model |
executor becomes deepseek-v4-pro |
aris-skills |
mounts the 82-skill corpus, publishes ARIS_REPO, restores Codex's threadId, serves the ARIS tab |
aris-codex |
codex mcp-server over MCP, 20-minute call budget, pinned to a stable working directory |
The corpus mounts at the bundled rank, so a project or user skill of the same name wins. The executor default is a deployment default, not a lock: a saved model setting or a per-session choice overrides it.
Developing against a checkout
ARIS_REPO=/absolute/path/to/aris NODE_USE_ENV_PROXY=1 \
dsh --profile web --patch /absolute/path/to/aris/dsh/checkout.patch.yml
Skill edits take effect without a restart. ARIS_REPO is required; without it the load fails naming the variable.
This overlay is not equivalent to the installed bundle: it cannot restore Codex's threadId, so multi-round codex-reply — and therefore the hard-tier Debate Protocol — works only through the installed bundle. The two are mutually exclusive: applying the overlay to a profile that already has the bundle fails on a duplicate row id. Remove one.
Known limits
- Tracks a Harness window, declared as a peer dependency. DeepSeek Harness is a developer preview; this bundle declares
@deepseek-ai/dsh >=0.1.7-rc.2 <0.3.0-0as an optional peer. The transport test passes against0.2.0-rc.2; theHostConnectionFetchregistry it uses is present and identical in0.1.7-rc.2, so the range covers both. Thedshpeer-dependency gate reads that range: an incompatible runtime skips the whole bundle with a line in the boot log instead of loading a panel whose transport may have moved. On a newer runtime, re-runnode node_modules/dsh-aris-panel/test/transport-smoke.mjs, then widen the range, or grant the exact-version exemption withdsh plugin allow-version. - The panel's browser transport is HTTP, not an RPC channel. dsh's generic channel registry (
connection.rpc.handle) resolveswebServerfrom Connection's own context, which dsh0.1.7-rc.2and0.2.0-rc.2do not inject there, so an out-of-tree channel registration throws and every call dies as HTTP 405 from the SPA fallback. ARIS mounts exact Fetch routes on the/apiprefix instead (documented asHostConnectionFetch.register), which dsh already serves behind its Host/Origin fence and browser-session cookie.test/transport-smoke.mjspins both halves of that contract against the installed dsh; it cannot cover the carrier's own auth path, so verify a live host withPOST /api/aris-runand/api/aris-wb— 400 means mounted, 404 means not. - No dsh packages are declared as npm dependencies. In-box packages are host-provided and resolve from the Harness installation through the profile module fallback. The Harness's own rule keeps
@deepseek-ai/dsh-*out ofdependencies, and the packages version in lockstep with the CLI, so any range this bundle pinned would fight the version the user already has installed. The peer range above is a compatibility declaration, marked optional so it never pulls a second copy of the runtime. web_fetchis off. Stock dsh ships it disabled and this bundle does not enable it, which would mean depending on a provider package. Skills that reach the web useweb_search, orbashwithcurl.- Reviewer thread continuity is process-local. A Harness restart, or any MCP reconnect that replaces the Codex child, loses saved
threadIds. Rounds after that start fresh;review-stage/REVIEWER_MEMORY.mdremains the durable record either way. - Codex's own reasoning is not in the Harness log. Only the verdict returns. The call arguments and the verdict are logged; the reviewer's intermediate work stays on the Codex side.
- A verdict over 50 KB is spilled to a file with a preview left in context. Raise
maxInlineByteson thespill-policyrow if your reviews run longer.