What happened?
On 1 September 2026, Anthropic shipped Claude Fable 5.1 for general use and Claude Mythos 5.1 for trusted cyber and life-science access. Same weights as each other. The fork is still policy, not a second training run: Fable is the gated SKU; Mythos is the same model with looser stop conditions, US-org only for now.
List price did not move: $10 / MTok in, $50 / MTok out. The cut is cache reads — 75% cheaper, $0.25 / MTok. Anthropic’s own August usage mix: about 25% less than Fable 5 on typical Claude Code / Enterprise / API work, up to about 45% on long, tool-heavy loops where cache is most of the bill. Defaults: High effort in Claude Code, Medium in Cowork and on Claude.ai.
The 74-second demo is the product, not the bench. One analyst’s mixed contract-and-consumption forecast runs unattended on the API overnight. In the morning they ask how the last forecasts did; the model backtests on the spot. Nothing ships until a human signs.
Capability claims sit on that loop. Terminal-Bench-Science 0.1: 52.6% vs 24.7% for Fable 5 and 29.0% for Opus 5 on Anthropic’s harness (public leaderboard is noisier; they flag ±3.5–4.5 pts). Terminal-Bench 4.0: 55.8% Fable, 60.9% Mythos, vs 42.0% Fable 5. AutomationBench 31.4% vs 17.1%. CursorBench 3.2.0 is a smaller step, 73.4% vs 70.5%. Millennium’s quote is the qualitative one: a one-in-a-million crash, years unexplained, traced to a vendor library after the model disassembled it against a core dump.
Science is the other half of the launch. Mythos-designed binders, lab-checked, hit ~50% on 12 targets against a 10–15% typical rate, and 10× the Adaptyv competition affinities on three named proteins. A Magellan-era Venus elevation map is on Zenodo. GPU-kernel rewrites on seven open genomics models, up to 2.5× on an H100. Treat those as Anthropic’s demos with external wet-lab checks, not a Nobel press release.
Safeguards got cheaper false positives, not a new philosophy. Cyber classifiers: about 60% fewer interventions per Claude Code session vs Fable 5; Fable may flag vulnerabilities, not write exploits. Dual-use cyber still falls back to Opus. Biology R&D still sits on Mythos behind a US-government access program. Enterprise Frontier Safeguards (customer-owned cloud, Anthropic-grade misuse detection, human review by the customer) starts this fall; eligible buyers get zero data retention until then. New API accounts lose a documented distillation trick: you can no longer edit prior turns and keep the thinking transcript.
Why this is interesting
- The SKU split survived contact with cost: Fable vs Mythos was “same brain, different stop conditions.” 5.1 keeps that, then makes the gated SKU cheap enough that Cognition moves Opus 5 Devin traffic onto Fable on day one. Cache-read pricing is how a $10/$50 model eats the daily-driver slot.
- Unattended hours are the unit: Ramp’s 38-hour ML run, MongoDB’s three-day prototype, the overnight forecast. The pitch is not a better autocomplete. It is a worker that keeps its own notes, checks its numbers, and waits for a signature. That is also where alignment coverage is thinnest — Anthropic says so in the system card.
- Science jump is real, and still a lab demo: Doubling Terminal-Bench-Science in one point release is the number that will get copied. Protein hit-rates and a Venus DEM are the receipts they chose. Neither is “AI discovered a drug.” Both are “the model can run the loop that used to be a specialist team, then a human still has to believe it.”
- Precision, not permission: 60% fewer cyber false positives, 85% fewer elementary-bio misfires, vuln-finding allowed, exploit-writing not. The containment level is still a commercial product. Mythos remains a US trusted-access object. Operators who need ugly tokens still do not get them on the GA SKU.
- Judgment stays on the morning desk: The demo’s last beat is review-and-approve. EFS puts the logs in the customer’s cloud. Distillation locks close a public extraction path for new accounts. None of that is the model choosing its own leash.
Bottom line
Fable 5.1 is not a new fork. It is the June SKU with a cheaper night shift: same $10/$50 sticker, cache at a quarter, longer unattended jobs, a science bench that finally moved, and a human still on the approve button. If you kept Opus on because Fable was too expensive to leave running, the rate card just argued with you. If you needed Mythos-class cyber or wet-lab, you still fill a form.
