Conversation
bd6be47 to
1cbdf16
Compare
SummaryThis adds detected-catalog reasoning-budget rewrites and ports the requested pstack prompt reductions and shipping verification-reuse exception. It depends on the priority workflow PR; its base is that PR's branch so the review diff contains only this second set of changes.
Scope
Verification
Review risksThe prompt reductions apply to the shared skill tree. Upstream motivated them with Opus 5.5, but this PR does not establish behavioral equivalence for other models. Build-output reuse is an agent-facing procedure, not an automatic classifier. Dev-server lanes and lanes without saved build output must rerun. Agent: gpt-6.1-sol via pi |
Summary
This adds detected-catalog reasoning-budget rewrites and ports the requested pstack prompt reductions and shipping verification-reuse exception. It depends on the priority workflow PR; its base is that PR's branch so the review diff contains only this second set of changes.
model-budget.mjspreviews changes, preserves aliases and other Harness choices, and refuses writes for unresolved models or duplicate panels. Shipping reuses individual lane results only after documented build-output comparisons, while current-head CI and review remain mandatory.Scope
Verification
npm test: 211 passed, 1 Unix-only test skipped on Windows.git diff --checkpassed.Review risks
The prompt reductions apply to the shared skill tree. Upstream motivated them with Opus 5.5, but this PR does not establish behavioral equivalence for other models. Build-output reuse is an agent-facing procedure, not an automatic classifier. Dev-server lanes and lanes without saved build output must rerun.
Agent: gpt-6.1-sol via pi