
← YPO Technology Network AI Brief3 sep · 10 min
The Sticker Price Did Not Move
Anthropic shipped two new frontier models this week and left the headline price exactly where it was: $10 per million input tokens, $50 output, unchanged. The number that moved is one almost nobody looks at. Cached input reads fell 75%, from $1.00 per million tokens to $0.25.
The price you get quoted is the price of answering once. Your bill is set by re-reading.
In this episode, Stephen Forte covers:
What a cached read actually is, and why it decides agent economics: an agent is not answering one question. Every step, it is handed the whole situation again — your instructions, every tool definition, the document or codebase, and a conversation that keeps getting longer. On the new models a cache hit costs 2.5% of the standard input rate, against 10% on Anthropic's other models.
Anthropic's own estimate that the change makes ordinary workloads ~25% cheaper and heavily agentic ones up to ~45% cheaper — aired as the company's figure, not an independent measurement. The gap between those two numbers is the lesson: the more autonomously software operates, the more of the bill was sitting in that one line.
Why a quoted per-token price is very nearly useless for budgeting anything that works on your behalf over time.
The demand side: Cisco said last week it is rolling an agent out to all 90,000 employees — not a pilot, not a department — working across email, chat, project tracking and documents. And agentic interactions on its internal AI platform grew nearly 350% in a single quarter. Cost per unit of agent work is falling sharply while volume grows at that rate; those do not cancel out.
The quieter item in the same announcement: Anthropic shipped two models with identical architecture that differ only in the strength of their safety limits. The more constrained one is generally available; the less constrained one goes only to vetted cybersecurity and life-sciences organisations, through verification built in coordination with the US government. Not a better model for more money — the same model twice, with access to the looser one decided by who you are rather than what you pay.
The close: a company that wanted you to believe its product had gotten cheaper would have cut the headline number. Anthropic left it alone and cut a line most buyers have never looked at. That is information about where the money actually is.
Also mentioned: In November, alongside the YPO Global Business Summit in Istanbul, the YPO Technology Network is running a full-day AI Global Summit on 6 November. Stephen is speaking, along with people from Microsoft and other leading AI companies. Registration is open.
Sources:
Anthropic, Claude Fable 5.1 and Mythos 5.1, announced 1 September 2026. Pricing cross-verified across VentureBeat, TechSpot, implicator.ai and CybersecurityNews, plus the Claude Platform pricing documentation. The 25% / up-to-45% effective-cost figures are Anthropic's own estimate.