Back to home

Jungod1121

dsh-anchored-standard

Two-phase DeepSeek Harness preset: Minimal-aligned bootstrap (bash+read), then full Standard tools after the first tool call or reply

Stars
1
Language
JavaScript
Created
Aug 15, 2026
Updated
Aug 15, 2026

Introduction

dsh-anchored-standard

Anchor the first request. Unlock the full stack.

A two-phase agent preset for DeepSeek Harness: bootstrap DeepSeek V4 Pro with the Minimal-aligned prompt and two tools, then expose the complete Standard catalog after the first durable tool call or reply.

harness license stars dsh-plugin

中文说明 · Landing page · Design origin


The one-line idea

V4 Pro's ability ceiling is high — but which reasoning strategy it uses is decided by what the first API request shows it. The community evaluation at xiaobright/modeltest measured:

PresetFirst request seesFull task seesProject2 V4.1b (max)
Standard25 tools + long prompt25 tools91 / 92
Minimal1-line prompt + 2 tools2 tools only99 / 96 — but too narrow for real work
Anchored Standard1-line prompt + 2 toolsall 25 tools98 / 99

The gain comes from anchoring the first-turn trajectory — not from keeping the tool surface small forever. This preset keeps Minimal's brain and gives it Standard's hands.

How it works

sequenceDiagram
    participant U as User
    participant A as Agent loop
    participant M as DeepSeek V4 Pro

    U->>A: first message (new session)
    A->>M: request #1<br/>task-matched anchor persona<br/>tools: bash + read (± edit/write)
    Note over M: reasoning chain anchors on the task-matched scaffold
    M->>A: first tool/call
    Note over A: promotion signal (durable session event)
    A->>M: request #2<br/>same persona · all 25 Standard tools
    Note over M: full-capability work continues on the anchored trajectory
  1. The session's first user message picks the anchor:

    TaskPersonaFirst-request tools
    spec — fix / maintain / debugMinimal's exact promptbash + read + edit
    react — build / create from zerohands-on doer promptbash + read + write
    weak — ambiguousmodel self-picks (Pro: classify instruction)bash + read

    glob/grep stay out of every bootstrap catalog — measured trajectory boundary for V4 Pro; edit/write are anchor-safe.

  2. On request #1 the persona is the ONLY prompt section and runtime contexts are cleared — the cleanest possible opening.

  3. After the session records its first durable tool/call, every later request sees the full Standard catalog; the chosen persona stays constant and the remaining sections (plan-mode etc.) return. Nothing is injected after turn one (measured: post-anchor guidance hurts Pro).

Mode and promotion state live in durable session events, so refresh and resume preserve them.

Install

Option A — installer bundle (recommended)

dsh plugin --profile web add github:Jungod1121/dsh-anchored-standard

Restart DeepSeek Harness. The bundle copies the preset into $DSH_HOME/.agent-presets/anchored-standard/ (idempotent — local edits are never overwritten), then create a blank session and select Anchored Standard (experimental). To make it the default preset for new sessions:

# $DSH_HOME/settings.yaml
agent-presets:
  default: anchored-standard

Option B — manual preset directory

dsh_home="${DSH_HOME:-$HOME/.dsh}"
mkdir -p "$dsh_home/.agent-presets"
cp -R preset "$dsh_home/.agent-presets/anchored-standard"

Verify

Export the session JSONL and inspect request/header events — the first header contains only bash/read, every later header the full catalog:

seq 18  -> 2 tools  ['bash', 'read']                       # request #1
seq 137 -> 25 tools ['ask_user_question', 'bash', ...]     # request #2 (promoted)

What you get vs what you give up

You getTrade-off
vs Minimal23 more tools: glob/grep/edit/write/web_search/subagents/workflows/skills/goals/todo…one extra reasoning round for the bootstrap
vs Standardtask-matched trajectory anchors (98/99 vs 91/92 in the reference eval) + a first turn that is cleaner than Standard'sthe first-request catalog is decided by keyword classification — ambiguous tasks fall back to the weak anchor and let the model pick

Evidence & honest boundaries

  • Mechanism (task-matched bootstrap catalog → full catalog) is verified on harness 0.1.0-rc.6 at the wire level; the classifier and anchor tables are unit-tested.
  • The 98/99 ability scores are xiaobright/modeltest Project2 V4.1b — n=2 on one frozen task. They are reproducible evidence for that task, not a universal guarantee across models or workloads. The spec/react/weak anchor design incorporates measured findings from yjh051108/dsh-router-standard (P1-P24) and xiaobright's trajectory boundary probes.
  • V4 Flash does not need this: it generalizes across harnesses (style changes, scores don't). This preset targets Pro with reasoningEffort: max.

Uninstall

dsh plugin --profile web remove dsh-anchored-standard
rm -rf "${DSH_HOME:-$HOME/.dsh}/.agent-presets/anchored-standard"

Compatibility

Developed and verified against DeepSeek Harness 0.1.0-rc.6. The harness is a developer preview with breaking changes — review upstream changes before upgrading.

Community & license

This is a community project: not an official DeepSeek preset, not affiliated with or endorsed by DeepSeek. Design inspired by xiaobright/dsh-anchored-standard (see NOTICE). MIT — LICENSE.