dsh-model-garden
August 27, 2026 ยท View on GitHub
A searchable, sortable model picker for the DeepSeek Harness Web UI (dsh web). It replaces the native composer model seat with a table-style picker that adds everything the stock selector is missing:
![]() |
![]() |
Features
- ๐ Instant search across model names and descriptions
- ๐ Sortable table columns โ click
Name,CtxorPriceto sort asc/desc; a third click returns to the provider-grouped view - โญ Favorites โ star models, toggle favorites-only from the table header; persisted in
localStorage - ๐ Local tag โ providers are flagged local by their real endpoint (baseURL from settings: loopback / RFC1918 / LAN hostnames), never by price guesswork; the Local box next to the search input filters to them
- โพ Collapsible provider groups โ collapse state is persisted per provider
- ๐ Auto model-list update โ while the picker is mounted, the provider/model directory is re-loaded automatically every 5 minutes (configurable), so locally added models show up without reopening the panel
- ๐ฐ Model prices from models.dev (the same source OpenCode uses), shown as
$input/$outputper 1M tokens, cached for 24 h. Subscription routes (all-zero cost in the catalog, e.g. coding-plan providers) resolve a reference price from their pay-as-you-go provider viaPROVIDER_ALIASES, so plan models still show what their tokens would cost at API rates; only true local models stay unpriced - ๐ง Context windows โ read live from the host
llmservice (adapter-owned data, works for local providers like llama.cpp / Ollama-style gateways too), with models.dev as fallback - ๐๏ธ Reasoning effort picker โ models that support reasoning levels get a compact dropdown right next to the model name in the chat composer, styled and opening exactly like the model picker (same trigger pill, same floating menu surface, โ marks the active level, click outside /
Esc/ selecting closes it). Picking a model always starts it at its strongest reasoning level; the dropdown re-selects the current model with the chosen effort โ no clutter inside the picker panel - ๐ธ Live per-task usage & cost โ real provider-reported token usage (from the session log) is always shown while the panel is open (
in / out / cache). Each model's usage is multiplied by its own (reference) price โ properly attributed even when a session switched models mid-way โ the same math OpenCode uses (usage ร price, not a heuristic) - ๐งพ Session cost breakdown โ hover the
approx costfigure for a popup sized like the picker and parked parallel on its left (1 px gap): a table-style breakdown with per-model totals (steps, In/Out/Cache in their own columns, โ cost) and a timestamped step table. Clickable column headers work like Excel / the main list (asc โ desc โ off, โฒ/โผ indicator) on both tables; copy the summary or export the (sorted) step list as CSV (Excel-ready). The invisible hover target spans the cost row all the way to its left edge. Attribution of each step to its model comes straight from the session log (request/contextevents); nothing extra is stored - ๐ฑ๏ธ Detail tooltip that opens beside the panel (never covers the list): description, price, context window, max output, reasoning efforts
- ๐จ Native look โ built entirely on the harness design tokens (
--dsw-alias-*); the picker panel, the effort menu, the detail tooltip and the cost popup all share the same surface color and geometry, matching light & dark theme automatically
How it works
The package is a static profile plugin with two halves:
| Half | File | Role |
|---|---|---|
| Client | client.js | Registers the conversation.input.model slot (priority -1, shadowing the native seat) and renders the picker |
| Host | index.js | Serves three same-origin JSON routes on the harness webServer service |
Host endpoints
GET /model-garden/cost?session=<sessionId>
โ { inputTokens, outputTokens, cacheReadTokens, cacheWriteTokens, reasoningTokens, steps }
GET /model-garden/cost-history?session=<sessionId>&limit=<n>
โ { steps: [ { time, provider, model, turn, step, inputTokens, outputTokens, cacheReadTokens, cacheWriteTokens, reasoningTokens } ] (newest first, capped),
models: [ { provider, model, steps, inputTokens, outputTokens, cacheReadTokens, cacheWriteTokens } ],
totalSteps }
GET /model-garden/catalog
โ { "provider::model": { local, context?, maxOutput? } } (cached 10 min)
The cost endpoint aggregates the real usage payloads of assistant/message events from the durable session log โ no estimation. The cost-history endpoint additionally attributes each usage step to the model in effect: assistant/message events carry usage but not the model, so it tracks request/context (and request/header) events, which precede the request they describe with { provider, model } โ a single pass over the same in-memory events, no extra persistence. The catalog endpoint resolves contextWindow / defaultMaxTokens per model through the host llm service (resolveModelInfo), so local/self-hosted providers report their real limits.
Installation
One command โ the official plugin CLI installs the package and mounts it (the package carries a dsh.bundle.patch layer, so the CLI automatically appends it to the profile's bundle stack):
dsh plugin --profile <profile> add dsh-model-garden
Then restart the DSH server and hard-refresh the browser (Cmd/Ctrl+Shift+R).
The host half needs the web stack (
webServerservice). In minimal/TUI profiles without it the plugin stays inert by design โ boot is never blocked.
Upgrading from a manual install? Remove the old
model-gardendependency and any manual- insert:row for it from your profile'scordis.patch.ymlfirst โ otherwise the plugin mounts twice.
Verify
curl -s http://127.0.0.1:3080/model-garden/catalog | head -c 200
# โ {"deepseek::deepseek-chat":{"local":false,"context":...}, ...} (JSON, not HTML)
Manual install (without the CLI)
If you manage the profile with plain npm: add the dependency, list dsh-model-garden in dsh.profile.bundles in the profile package.json, reinstall, restart. The bundle patch inside the package inserts the loader row for you โ no cordis.patch.yml edit needed.
Configuration
No configuration is required. Several tweakable constants live at the top of the respective file:
- Auto model-list update interval โ
MODEL_LIST_REFRESH_MSinclient.js(default 5 min) controls how often the provider/model directory is reloaded while the picker is mounted. - Hidden provider routes โ
HIDDEN_PROVIDER_PREFIXESinclient.js(andSKIP_PREFIXESinindex.js). Some plugins mirror providers as internal routes (e.g. a vision toolkit duplicating every provider asvision-toolkit-<id>); such prefixes are excluded from the picker and the catalog. - Price aliases โ
PROVIDER_ALIASES/MODEL_ALIASESinclient.jsmap DSH route ids to models.dev catalog ids. They serve two cases: renamed routes (deepseek-officialโdeepseek) and subscription routes whose catalog entry is all-zero (kimi-for-codingโmoonshotai,alibaba-tpโalibaba-cn,oneproviderโanthropic), giving plan models their pay-as-you-go reference price. - Price cache TTL โ
PRICE_TTL(default 24 h) and catalog TTL โCATALOG_TTL(default 10 min).
Favorites, collapsed providers and the price cache live in the browser's localStorage under dsh.modelgarden.*.
Compatibility
Developed and tested against DeepSeek Harness 0.1.0-rc.8 (@deepseek-ai/dsh-host-webserver, dsh-session, dsh-llm, dsh-client-ui-model-selection); first released against 0.1.0-rc.6. The client half is plain React via window.__ModuleLoader__ โ no build step, no dependencies.
Credits
- Pricing data: models.dev API (also used by OpenCode)
- Design tokens & slot API: DeepSeek Harness

