plan(seed): Fable 5 long-range remediation roadmap (planning-only, drafts) - #1347
Conversation
Run dir + supervisor identity, owner eval waivers (PLAN-EVAL/IMPL-EVAL), lane overrides (Fable 5 high orchestrator, Opus 5 workflow contributors), drafts-only mutation boundary. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
|
[PHASE: RESEARCH] — Stage A charter read-back (seed run Charter. Planning-only seed run: long-range NetScript remediation roadmap as drafts under Owner waivers (recorded in supervisor.md / worklog.md / drift.md): PLAN-EVAL waived; IMPL-EVAL waived. No formal evaluator, OpenHands, or OpenRouter session will be launched. Owner personally reviews the plan and decides on later adversarial passes / board filing. Mutation boundary. Zero GitHub board mutation (issues/epics/milestones/labels/comments outside this PR). No framework/product source edits. Writable surfaces: this branch, its commits, this draft PR (body/comments/PR labels). Filing (stage H) is explicitly out of scope. CI lane. Evidence inputs. (1) Codex pre-plan package (README / ISSUE-DEDUP-MATRIX / EVIDENCE-REGISTER); (2) all agent-posts waves (archive + waves 3–6); (3) live GitHub board (wins over carried-in reports); (4) deep repo + published-docs audit; (5) eis-chat + competing meta-frameworks + oRPC; (6) 0.0.5 milestone/orchestrator history. Baseline. Next: Stage B discovery corpus (S2). |
Two read-only Opus-subagent research workflows (prior waves + live board; repo/docs domain audit + external comparisons), per seed-run Tier-C rule. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
… Stage-B fan-out Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
…ound truth Seven Opus-5 research artifacts (3.8k lines): waves 1-6 evidence, exhaustive open-board snapshot (259 issues/13 milestones), 0.0.5 orchestration history, live conventions incl. label/milestone/RFC practice. Supervisor-reviewed. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
|
[PHASE: RESEARCH] — S2a landed: prior-wave evidence + live board ground truth ( Workflow Load-bearing corrections vs the carried-in pre-plan (current GitHub wins):
Note: this commit also carried the first 8 domain files under Next: S2b — repo/docs domain audit + external comparisons (eis-chat, meta-frameworks, oRPC). |
…ternal bar Eleven Opus-5 artifacts (5.2k lines): docs/MCP/CLI/web/SDK/auth/runtime/ observability/scaffold-doctrine audits with execution-verified defect verdicts, plus eis-chat teardown, meta-framework competitive bar, oRPC extension-model deep dive. Supervisor-reviewed (see worklog S2b note). Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
|
[PHASE: RESEARCH] — S2b landed: repo/docs domain audit + external bar ( Workflow Highest-leverage findings:
Next: Stage C synthesis ( |
…cks, train direction Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
…afts) Two-milestone insertion via house rename pattern; complete per-issue move ledger; entry/exit criteria; thin Wave-7 A/B/C measured smoke matrix. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
… amendments Three milestone directories (0.0.6 verification/docs/RFCs; 0.0.7 typed seams+generation; 0.0.8 runtime truth+service slice), 929-line amendments file, dedup table + Stage-E ledger. Supervisor-reviewed; six corpus corrections recorded. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
|
[PHASE: PLAN] — S4 landed: design packs ( Workflow
Quality note: every pack re-verified corpus claims at source before drafting, producing six corrections now recorded in the dedup §3.1 — most notably the wave-6 All 44 files carry the Next: Stage E plan lock — |
…run close 12-fork owner sweep, program DAG, filing gate + first executable PR groups; plan.md locked, context-pack finalized, session record written. No board mutation occurred in this run. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
|
[PHASE: PLAN] [VERDICT: LOCKED — awaiting owner review] — Stage E complete ( The complete long-range remediation plan is committed under Program in one paragraph: the gap is not missing capability but a product surface that contradicts it — one central seam (contract → typed client → query → page) broken across docs, generation, API sealing, and runtime proof. The plan lands five contracts (typed extension chain RFC-A; generated canonical vertical slice; service slice + command kit RFC-B; runtime truth; agent-native surface) across two inserted milestones (0.0.7 typed seams + generation, 0.0.8 runtime truth + service slice) with 0.0.5 closing as scoped, RFCs ratifying in 0.0.6, existing 0.0.7–0.0.13 shifting to 0.0.9–0.0.15 by title-only rename (every issue retained), and a Wave-7 measured A/B/C adoption smoke as the exit gate before #922's train resumes. Owner decisions requested: forks F1–F12 in Closing statement: No GitHub issue, epic, milestone, project, label-set, or comment mutation occurred in this run. The only GitHub writes were this draft PR's own branch/commits/body/comments/PR-labels. All 44 draft files carry the no-mutation H1 marker; filing happens once, from the committed manifest, only after owner ratification, in a separately authorized session. Evaluator waivers (PLAN-EVAL, IMPL-EVAL) are recorded in |
POST /api/v1/auth/signout revokes any session id an unauthenticated caller supplies
#1384
|
[PHASE: FILING] [VERDICT: COMPLETE] Owner-ratified Stage H is complete. The milestone train was shifted highest-first, fresh 0.0.6 and 0.0.7 milestones were created, the full former-0.0.6 membership was preserved, and all 41 roadmap issues (#1348–#1388) were filed and audited. Live dependencies use live issue numbers; 20 additive existing-owner amendments were posted; no issue was closed. Durable receipt: |
…luator delegation record) Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
|
[PHASE: REVIEW] — Cross-RFC PLAN-EVAL advisory (RFC-A #1390 × RFC-B #1389) Both RFC PLAN-EVALs are complete (owner-designated cross-family evaluator: this Fable 5 session; generators: Codex Sol xhigh; deep dives: 3× Opus 5 xhigh, recorded). Both verdicts: CHANGES_REQUESTED, FAIL_PLAN cycle 1 — full findings on each PR and in each run's Cross-RFC verdict: they compose cleanly. No circular dependency (#1350 is a shared one-way stage-0/1 prerequisite); error policies are disjoint by construction (local preparation failures vs route-opt-in contract errors, both deferring client typing to #1350's literal-preserving spelling — whichever lands first establishes it, the other reuses it); context/telemetry ownership does not overlap (transport-owned trace headers vs persisted W3C row context). Sequencing matches the filed train (ratify 0.0.6 → RFC-A impl 0.0.7 → RFC-B impl 0.0.8). Headline findings: RFC-A — the zero-oRPC-symbol gate fails on unchanged code ( Board: existing children #1348–#1364 are sufficient — no duplicates needed; the named amendments (#1351 OTel-rename/pin-policy/dedupe-trap, #1349 channel/port-location/key-algebra, #1350 stage-0 spelling, #1362/#1363 relay-ownership + Handoff: root orchestrator resumes both Codex generators with their finding lists (branch HEADs |
…ff brief Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
|
@openhands-agent model=openrouter/minimax/minimax-m3 output=pr-comment iterations=400 use harness SKILL
Act as a cheap-and-quick documentation accuracy evaluator. Do not edit source, documentation,
Keep the iteration budget small. Prefer one to three decisive manual checks over broad exploration, |
OpenHands Agent — Agent failedOPENHANDS_VERDICT: NONE Model: OpenHands Agent Summary — INCOMPLETE (iteration limit)The agent hit the maximum iterations limit (400) before Re-trigger with a narrower task or a higher Run: https://github.com/rickylabs/netscript/actions/runs/31279410211 |
#1395 merged as aa8e151, closing #1329. Wave 2 landed four issues across three PRs. Its IMPL-EVAL required two verdicts: the first passed the implementation, then 75832db landed generic suite-runner and deferred-gate semantics after it, so the mandatory evaluation was re-opened against the merging head. The correction review passed, establishing that the deferral machinery cannot hide a failing gate by construction. C17's payload is frozen from first-parent history: seven merges since canary.16's source. Four of them — #1391, #1337, #1347, #1215 — were never dispatched as part of the wave and are in the payload regardless, which is the membership rule working as canary-cadence.md intends: the wave is a dispatch unit, the canary is a content unit. Refs #1329 Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01LCtRhvYeuhQpvxwQAibz5a
…iled in one milestone Three 0.0.6 issues prescribed a fix that, implemented literally, would have produced a guard that could never fire. Two were caught in this lane (#1377's group-level command gate is green on arrival; #1374's locked deno check line exits 1 on every input). The third was caught independently by the internals lane (#1436 prescribed adding a word boundary that was already present and is the cause, not the cure), which is what makes it a pattern rather than a coincidence. Common cause: all three came from the same planning seed (PR #1347). Its evidence was excellent and still holds; its prescriptions were reasoned from that evidence rather than executed against it. Each lane caught its instance by executing something -- re-measuring before dispatch, building the mechanism and running it, or committing a baseline probe recording RED before the fix. Proposes a predecessor clause for milestone-run.md Gate integrity: an issue that prescribes a fix has its prescription re-verified before implementation, executed rather than read. 0.0.4 taught that shipped guards can be inert; 0.0.6 shows filed prescriptions can be inert before anyone writes a line. Refs #1374 #1377 Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01QGrdeXR3yuCZt78FxtpMy5
Summary
Planning-only seed run that produced and then, after explicit owner ratification, filed the complete long-range NetScript remediation roadmap: master plan with a 12-fork owner sweep, milestone train, 41 issues, two RFC entry points, existing-owner amendments, Wave-7 measured-adoption design, implementation handoff, and a cited research corpus.
GitHub is now authoritative. This draft PR preserves the planning rationale and filing receipts; it must not be merged.
Scope
.llm/runs/plan-fable5-remediation-roadmap--seed/).llm/harness/workflow/seed-run.md; profileSCOPE-docsSlices
Stage H — owner-ratified filing
status:triage, and type/area/priority labels present. No issue was closed.Full durable receipt:
fable-5-remediation-plan/FILING-LOG.mdat commit1ef78e29f.Validation and harness
ci:skip-e2e+ci:skip-scaffoldintentionally applied.origin/main@fac9e339042c.wf_e2194004-808,wf_03b88126-e7e,wf_ebfe8327-306; 25 Opus 5 contributors, 0 workflow errors.🤖 Generated with Claude Code