feat(model-router): wave 3-B — first-use dialog with context, model then effort on change

The dialog now explains what it asks about: the skill's description (from
its SKILL.md frontmatter, five YAML forms, first sentence, cleaned), the
agent's description (from agent.offer), the phase's new 'about' line in
routing.json, and the real model id and effort the next step runs on.
Options Later / Keep / Change; Change asks the model (fable, opus, sonnet,
haiku with their tier role), then the effort among those the phases of
that model offer, then the scope; the pair maps to an existing phase (rows
stay phase names; the row's current phase wins a tie; same phase = Keep).
A main-loop phase (T3) is Later / Keep only. A main-row change toasts the
real decision and the /route switch hint when the pick is a downgrade.
Descriptions pass a hardened read (name allowlist, stat kind, size cap)
and clean(). Kit suite 190 → 232; W3-A dialog tests migrated.

Contract .claude/tasks/contracts/2026-10-11-model-router-w3b-dialog-1240.md,
plan r3: 3 lenses (2 BLOCKERs: inline rows, phase edits) + 1 confirmation,
feater + 2 rounds, GATE 0 MET, verifier CONFORME (3rd pass), security PASS.
This commit is contained in:
bchanot
2026-10-11 14:10:24 +02:00
parent f2f404a002
commit 604a6c4411
4 changed files with 769 additions and 102 deletions
+27 -14
View File
@@ -2,47 +2,58 @@
"phases": {
"plan": {
"tier": "best",
"effort": "xhigh"
"effort": "xhigh",
"about": "design and architecture: the deepest thinking, before any code exists"
},
"reflect": {
"tier": "best",
"effort": "high"
"effort": "high",
"about": "analysis, review and synthesis: reasoning about what already exists"
},
"orchestrate": {
"tier": "best",
"effort": "medium"
"effort": "medium",
"about": "dispatching and coordinating sub-agents: mostly handing out work"
},
"escalate": {
"tier": "best",
"effort": "max"
"effort": "max",
"about": "a stuck problem or a judged need: maximum effort, diagnosis only"
},
"judge": {
"tier": "big",
"effort": "xhigh"
"effort": "xhigh",
"about": "independent judgment of work: audits, challenges and verdicts"
},
"implement": {
"tier": "work",
"effort": "medium"
"effort": "medium",
"about": "writing the code of an already planned change"
},
"write": {
"tier": "work",
"effort": "high"
"effort": "high",
"about": "writing prose and docs, or a careful edit that needs polish"
},
"verify": {
"tier": "work",
"effort": "xhigh"
"effort": "xhigh",
"about": "checking a result against its spec: tests, security review, gates"
},
"explore": {
"tier": "work",
"effort": "medium"
"effort": "medium",
"about": "reading and searching the codebase to answer a question"
},
"apply": {
"tier": "work",
"effort": "low"
"effort": "low",
"about": "bookkeeping: commits, memory entries and other small edits"
},
"mechanical": {
"tier": "cheap",
"effort": "low"
"effort": "low",
"about": "scripted or trivial work needing no judgment, on the cheapest model"
}
},
"skills": {
@@ -125,7 +136,7 @@
"plugin-probe": "apply",
"validator-analyzer": "apply",
"verifier": "verify",
"security-auditor": "verify",
"security-auditor": "judge",
"status-reporter": "mechanical"
},
"projects": {},
@@ -133,12 +144,14 @@
"agents": {
"verifier": "verify",
"feater": "implement",
"doc-syncer": "write"
"doc-syncer": "write",
"security-auditor": "judge"
},
"phases": {
"verify": "verify",
"implement": "implement",
"write": "write"
"write": "write",
"orchestrate": "orchestrate"
}
},
"changed": {},