diff --git a/design/tasks.md b/design/tasks.md index b04de48..05c2801 100644 --- a/design/tasks.md +++ b/design/tasks.md @@ -89,13 +89,13 @@ These block design, not just code. Answers go into `Methodology/Calibration.md`. ## Phase 4 — Rubric and classification -- [ ] **4.1** `references/boundary-rubric.md` = `rubric.md` verbatim. -- [ ] **4.2** `skills/boundary-rubric/SKILL.md` — usable standalone on a single +- [x] **4.1** `references/boundary-rubric.md` = `rubric.md` verbatim. +- [x] **4.2** `skills/boundary-rubric/SKILL.md` — usable standalone on a single step, before any code exists. -- [ ] **4.3** `agents/offload-analyst.md` with the PLAN §9 behavioral rules, +- [x] **4.3** `agents/offload-analyst.md` with the PLAN §9 behavioral rules, including the hard gate: **no overrule case → boundary question, never filed.** -- [ ] **4.4** **Calibration against known answers — the load-bearing test.** +- [x] **4.4** **Calibration against known answers — the load-bearing test.** - `token-budget` → **must** yield zero `high`-confidence candidates (S2). - `worktree-discipline` with `bin/worktree-audit` masked out of the inventory → **must** flag worktree classification as `high` (S3).