Player select · 04 files
Claude · Anthropic

Claude

Opus 4.8 · Fable 5 · Sonnet 5 · effort varies by seat · Planning + co-planning

← →  Switch player
Blue owl-like archivist mascot representing the Claude family
PLANNING + CO-PLANNING
Policy · Planning
Claude capability reading: Discovery 85, Planning 85, Build 75, Review judgment 45, Grounding 75, Cost efficiency 50DiscoveryPlanningBuildReview judgmentGroundingCost efficiency
Discovery85
Planning85
Build75
Review judgment45
Grounding75
Cost efficiency50

Editorial ratings · recomputed 2026-09-10. 70% evidence reading + 30% harness fit, rounded to five points.

Scoring breakdown & sources

Claude-family reading. Opus supplies the discovery/build baseline; Fable 5.1 supplies the newer review evidence. Models are named in each explanation.

Each input uses a 0–4 rubric. Evidence: negative (0), weak or indirect (1), mixed (2), strong within bounds (3), strong across tested cases (4). Fit: excluded (0), trial/reserve (1), bounded/additive (2), regular role (3), core role (4). Policy estimates are labeled below; these ratings are not measured success percentages.

Discovery 85

Evidence 3/4 · fit 4/4 · study + policy

Opus averaged 73.35 in the two discovery comparisons, with strong anchoring and retained deep-synthesis duties. GLM won pooled breadth, so this is below the top evidence band.

Planning 85

Evidence 3/4 · fit 4/4 · policy estimate

Policy-based estimate: Opus owns plan and planloop leadership; Fable is the fixed co-plan drafter. This records established responsibility, not a measured planning win rate.

Build 75

Evidence 3/4 · fit 3/4 · study + policy

Opus’s shipped W3-EMP build scored 85/100. Opus remains a build-lead candidate, but shares the route with Terra.

Review judgment 45

Evidence 2/4 · fit 1/4 · study + policy

Fable 5.1 was strong at FIND and bounded DECLARER work, but recalled only one of three ITERATE musts and lacks a dedicated confirm-seat test. Current review-lead routes list Terra.

Grounding 75

Evidence 3/4 · fit 3/4 · study + policy

Opus’s discovery anchoring was strong; Fable’s eight cells had no fabricated citations, with one immaterial location error. These support a strong scoped reading, not a family-wide zero-error claim.

Cost efficiency 50

Evidence 2/4 · fit 2/4 · study reading

Relative resource-use estimate: the lead study shifted lead-token load off Claude, while Fable xhigh used about 1.3–1.5× high’s tokens. Comparable billed cost per successful task is unavailable, so neither gets a top efficiency band.

Assignment · checked 2026-09-10
Planning + co-planning

Opus leads planning and the planning loop, with deep discovery work. Fable is the fixed co-plan drafter. Opus is also eligible for the build lead; review and cleanup lead routes currently list Terra.

Claude is a family: Opus, Fable and Sonnet have different assignments.

Configured seatsPolicy snapshot
Plan + planning-loop leadconfiguredOpus 4.8; both lead routes are enabled. Source ↗
Deep discoveryconfiguredOpus 4.8 retains the deep codebase and web seats. Source ↗
Co-plan draftconfiguredFable 5 at max; this draft cannot be substituted or skipped. Source ↗
Build leadeligibleOpus is a candidate alongside Terra; the router selects by pool availability and burn. Source ↗
Rote build workerconfiguredSonnet 5 is the static worker default; eligible bounded steps can route to GLM. Source ↗
Review + cleanup leadexcludedThose lead routes currently list Terra only. This does not rule out future Claude routing. Source ↗
Published evidence
Select player01 / 04
Fact-checked · Configured roles checked across Spindle, Flowmaster and ScriptGen. Router eligibility and per-unit overrides determine each dispatch. This is a dated policy snapshot, not live usage or a lifetime run count. Policy source ↗

I. Nocturne in E♭ major Op. 9 № 2 · Chopin

0:00 / 4:16