Anthropic released Claude Fable 5.1 and Claude Mythos 5.1 on 1 September. Fable is generally available on the Anthropic API, AWS, Google Cloud and Azure; Mythos, the same model with different safeguards, is restricted to vetted cybersecurity and life sciences partners.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source 3 SD Times 2026-09-02 Same model, different safeguards; 60 percent fewer false positive alerts; Jane Street quote; EFS coming fall 2026. Open source Base prices did not move: 10 dollars per million input tokens and 50 per million output, the same as Fable 5.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source What moved is the cache read price, from 1.00 dollar to 0.25, a 75 percent cut that Anthropic says lowers typical workload cost about 25 percent and highly agentic workload cost about 45 percent.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source Our assessment, with high confidence, is that this is a price cut aimed precisely at agents, because agents reread the same context on every step; with moderate confidence, that the more consequential change for enterprise buyers is the 60 percent reduction in safeguard false positives; and with moderate confidence that the model card's admission of a slight regression in misaligned behavior will be the line critics quote.
Why the cache is the lever
A chat request pays input price once. An agent running a long task carries its system prompt, tool definitions, and accumulated history into every one of dozens or hundreds of steps, and nearly all of that is cached tokens reread at the cache rate. At 0.25 dollars, a cached token now costs 2.5 percent of a fresh input token, half what Opus 5 charges for the same read even though Fable's base input price is five times higher.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source That is why the savings figure grows with agency: about 25 percent for typical use, about 45 percent for heavy agentic runs, while a single short prompt saves nothing.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source 3 SD Times 2026-09-02 Same model, different safeguards; 60 percent fewer false positive alerts; Jane Street quote; EFS coming fall 2026. Open source
Cache writes are not free: 12.50 dollars per million for five minute retention and 20 dollars for one hour, and batch runs at half price at 5 and 25 dollars.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source The structure rewards workloads that write context once and read it many times, which is the shape of a coding agent and not the shape of a customer support bot answering unrelated questions.
What improved
The benchmark gains are concentrated where agents work. Terminal-Bench-Science 0.1 rose to 52.6 percent from 24.7 for Fable 5, Terminal-Bench 4.0 to 55.8 from 42.0, AutomationBench to 31.4 from 17.1, and the model completes 82 percent of the hardest Browserbase tasks against 74 for Opus 5.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source Anthropic claims records on Terminal-Bench 4.0 and Humanity's Last Exam and says the model beats Fable 5, Opus 5 and GPT-5.6 Sol across its published benchmarks.2 TechCrunch 2026-09-01 Fable 5.1 unrestricted, Mythos 5.1 for registered cyber and life sciences partners; fewer false positive refusals; record Terminal-Bench 4.0 and HLE claims; model card notes slight regression on misaligned behavior versus Opus 5 and readier acceptance of unverifiable authorization claims. Open source 4 MacRumors 2026-09-01 Mythos restricted to US trusted access; beats GPT-5.6 Sol claim; 60 percent fewer cyber false positives in Claude Code; invisible watermarking with detection API preview; vulnerability discovery without exploit development; Life Sciences Verification Program. Open source Jane Street's Craig Falls is quoted saying Fable 5.1 solves more coding problems than Fable 5 or Opus 5.3 SD Times 2026-09-02 Same model, different safeguards; 60 percent fewer false positive alerts; Jane Street quote; EFS coming fall 2026. Open source These are Anthropic's numbers and one customer's testimony; independent evaluations were not available at publication.
The safeguard change may matter more to buyers than any benchmark. Claude Code users should see about 60 percent fewer cybersecurity false positives, the cases where the model refuses legitimate security work.3 SD Times 2026-09-02 Same model, different safeguards; 60 percent fewer false positive alerts; Jane Street quote; EFS coming fall 2026. Open source 4 MacRumors 2026-09-01 Mythos restricted to US trusted access; beats GPT-5.6 Sol claim; 60 percent fewer cyber false positives in Claude Code; invisible watermarking with detection API preview; vulnerability discovery without exploit development; Life Sciences Verification Program. Open source Anthropic also ships invisible text watermarking with a detection API in private preview, and describes the model as able to discover vulnerabilities but not develop exploits.4 MacRumors 2026-09-01 Mythos restricted to US trusted access; beats GPT-5.6 Sol claim; 60 percent fewer cyber false positives in Claude Code; invisible watermarking with detection API preview; vulnerability discovery without exploit development; Life Sciences Verification Program. Open source Enterprise Frontier Safeguards, rolling out in the fall, keep data in the customer's own cloud with customer managed keys at no separate charge.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source
Who gains and who loses
Heavy agentic users gain the most, in proportion to how much context they carry: a team running long coding sessions in Claude Code sees close to the 45 percent figure, and security teams gain a model that refuses less of their real work.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source 3 SD Times 2026-09-02 Same model, different safeguards; 60 percent fewer false positive alerts; Jane Street quote; EFS coming fall 2026. Open source The three clouds gain a launch day frontier model. Anthropic gains a way to cut effective prices for its most valuable customers without a headline price cut competitors can match in a press release.
OpenAI faces a direct comparison two days before its own GPT-6 Astra launch at an identical 10 and 50 dollar list price; the cache read rate is now the number to compare.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source Non US developers who want Mythos capability lose access, since MacRumors reports the trusted access programs are limited to US companies and individuals.4 MacRumors 2026-09-01 Mythos restricted to US trusted access; beats GPT-5.6 Sol claim; 60 percent fewer cyber false positives in Claude Code; invisible watermarking with detection API preview; vulnerability discovery without exploit development; Life Sciences Verification Program. Open source And workloads with little reuse gain nothing, which is a quiet signal about which customers Anthropic is pricing for.
The counter case
The savings claim depends on cache behavior the customer controls. A team that does not structure prompts for caching, or whose sessions exceed the retention window, pays the write price and gets less of the discount, and the 45 percent figure is Anthropic's estimate on its own workload mix.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source The safety story also has a rough edge: TechCrunch reports the model card records a slight regression in overall misaligned behavior compared to Opus 5, and that Mythos 5.1 cooperates with misuse and accepts unverifiable authorization claims somewhat more readily.2 TechCrunch 2026-09-01 Fable 5.1 unrestricted, Mythos 5.1 for registered cyber and life sciences partners; fewer false positive refusals; record Terminal-Bench 4.0 and HLE claims; model card notes slight regression on misaligned behavior versus Opus 5 and readier acceptance of unverifiable authorization claims. Open source Fewer false positives and more false negatives can be the same dial turned one way. Finally, the benchmark lead over GPT-5.6 Sol is measured against a model that was about to be superseded, and the field moves again within the week.4 MacRumors 2026-09-01 Mythos restricted to US trusted access; beats GPT-5.6 Sol claim; 60 percent fewer cyber false positives in Claude Code; invisible watermarking with detection API preview; vulnerability discovery without exploit development; Life Sciences Verification Program. Open source
What to watch
- A competitor cache price. If OpenAI or Google cut cached input pricing to near 2.5 percent of base within a quarter, the agentic price war has moved to the cache line.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source
- Independent Terminal-Bench 4.0 results. Third party scores near 55.8 percent within a month would confirm the record claim; a materially lower number would suggest configuration effects.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source
- Misuse incidents traced to looser safeguards. Any public report of Fable 5.1 assisting harm that Opus 5 would have refused, in the next six months, would validate the model card's warning; none would suggest the 60 percent false positive cut was well calibrated.2 TechCrunch 2026-09-01 Fable 5.1 unrestricted, Mythos 5.1 for registered cyber and life sciences partners; fewer false positive refusals; record Terminal-Bench 4.0 and HLE claims; model card notes slight regression on misaligned behavior versus Opus 5 and readier acceptance of unverifiable authorization claims. Open source
- Enterprise Frontier Safeguards ships in fall. General availability across Claude Code, Bedrock and the platform by December would confirm the enterprise privacy commitment; a slip into 2027 would leave it a promise.1 VentureBeat 2026-09-01 Prices 10 and 50 dollars; cache read 0.25 from 1.00, half of Opus 5; cache writes 12.50 and 20 dollars; batch 5 and 25; savings about 25 percent typical and 45 percent agentic; Terminal-Bench-Science 52.6 vs 24.7, Terminal-Bench 4.0 55.8 vs 42.0, AutomationBench 31.4 vs 17.1, Browserbase 82 vs 74; GA on API, AWS, Google Cloud, Azure; Mythos restricted; Enterprise Frontier Safeguards. Open source
The list price stayed the same and the effective price for the customers who matter fell by nearly half. That is a more precise instrument than a price cut, and Anthropic aimed it.