Claude Fable 5.1
Coding / long-horizon agent work; Max effort; strong cache economics.
Deck 1 · Frontier model comparison
Pick the tool by the job — not the brand. A CEO-ready lens on the September 2026 frontier wave.
Thesis
The September wave put three strong options on the table within three days. The winning move is routing — not picking a favorite.
Release wave · September 2026
Coding / long-horizon agent work; Max effort; strong cache economics.
Cost-efficient volume · ~1M context · high throughput class.
Ceiling work: computer-use, hard math/research, xhigh to max effort.
Dates and product names from public launch coverage (Sept 2026). Verify current vendor cards before procurement.
Fair comparison
| Signal | Astra | Fable 5.1 | Gemini 3.8 Flash |
|---|---|---|---|
| List input / 1M tok | $10 | $10 | $0.75 (promo) |
| List output / 1M tok | $50 | $50 | $3.75 (promo) |
| Cache reads (where reported) | ~$1.00 | ~$0.25 | ~$0.075 |
| Context (approx.) | ~1.05M · 128K max out | ~200K–1M (check card) | ~1M |
| AA Intelligence Index | ~61.2 | ~65.7 | ~59.0 |
| Coding Agent Index | ~67 | ~70 | — |
| Effort / modes | low → medium → high → xhigh → max | Max effort mode | High-throughput class |
Pricing: vendor / launch coverage, standard list. Flash promo through ~Dec 31 2026; often listed ~$1.50 / ~$7.50 after. Fable context reported inconsistently — prefer current Anthropic card. Indexes: Artificial Analysis / Sept coverage — third-party; not vendor-owned.
Pricing reality
Input / output per 1M tokens (standard list). Same sticker — different strengths and cache profiles.
Cache reads: Astra ~$1.00 · Fable ~$0.25 — material for agent loops.
~13× cheaper input vs $10 frontier list (promo)
Promo through ~Dec 31 2026; then often listed ~$1.50 / ~$7.50. Still the volume lane.
Cite as vendor / launch coverage. Confirm current price cards before budgeting.
Effort modes
Use xhigh/max for frontier agent, computer-use, and hard math/research. Lower effort for ordinary drafting.
Strong for long-horizon coding and software-agent loops — especially where cheaper cache reads compound.
Astra showcase · 01
When the job is operating a machine — not just answering a prompt — Astra’s ceiling lane is the bet.
Cited in OpenAI / launch coverage. Treat as self-reported / harness-sensitive.
Governance matter — not a feature to celebrate casually in the boardroom.
Astra showcase · 02
Signature claim for research-grade math. Confirm methodology before using in diligence.
Reserve top effort for problems where “almost right” is expensive.
Docs, CRM copy, and bulk summarization do not need Astra max.
Astra showcase · 03
Near-ceiling results attract headlines. Your job is to keep the harness footnote visible.
Cited as near-ceiling with harness / scaffold note. Treat as OpenAI / launch coverage — not a blank check for every reasoning task.
Ask: What harness? What holdout? What failure cost if the scaffold is not there in production?
AA Intelligence Index puts Astra ~61.2 vs Fable ~65.7 vs Flash ~59.0 — broad capability is not identical to signature demos.
Honest strengths
Indexes: third-party / Sept coverage
Throughput and Terminal-Bench: AA / coverage — third-party
CEO decision framework
Computer-use · hard math/research · agentic OS work
Long-horizon software · repo work · tool loops
Docs · extraction · high-throughput assistants
Routing
Default posture: route by task class, not brand loyalty.
Next session
Models are the engines. Next we show the agent stack that persists, uses tools, and works while you are away — without turning this block into an agents seminar.
Persistence · tools · routines · multi-app orchestration · security gatekeeping — live with GrokBot.
Takeaways
Hive Research Institute · AI Practicum · Deck 1 · Sept 2026 figures as cited in sources.md