YPO Technology Network AI Brief

← YPO Technology Network AI Brief3 sep · 10 min

The Sticker Price Did Not Move

The Sticker Price Did Not Move3 sep10 min

Anthropic shipped two new frontier models this week and left the headline price exactly where it was: $10 per million input tokens, $50 output, unchanged. The number that moved is one almost nobody looks at. Cached input reads fell 75%, from $1.00 per million tokens to $0.25.

The price you get quoted is the price of answering once. Your bill is set by re-reading.

In this episode, Stephen Forte covers:

What a cached read actually is, and why it decides agent economics: an agent is not answering one question. Every step, it is handed the whole situation again — your instructions, every tool definition, the document or codebase, and a conversation that keeps getting longer. On the new models a cache hit costs 2.5% of the standard input rate, against 10% on Anthropic's other models.

Anthropic's own estimate that the change makes ordinary workloads ~25% cheaper and heavily agentic ones up to ~45% cheaper — aired as the company's figure, not an independent measurement. The gap between those two numbers is the lesson: the more autonomously software operates, the more of the bill was sitting in that one line.

Why a quoted per-token price is very nearly useless for budgeting anything that works on your behalf over time.

The demand side: Cisco said last week it is rolling an agent out to all 90,000 employees — not a pilot, not a department — working across email, chat, project tracking and documents. And agentic interactions on its internal AI platform grew nearly 350% in a single quarter. Cost per unit of agent work is falling sharply while volume grows at that rate; those do not cancel out.

The quieter item in the same announcement: Anthropic shipped two models with identical architecture that differ only in the strength of their safety limits. The more constrained one is generally available; the less constrained one goes only to vetted cybersecurity and life-sciences organisations, through verification built in coordination with the US government. Not a better model for more money — the same model twice, with access to the looser one decided by who you are rather than what you pay.

The close: a company that wanted you to believe its product had gotten cheaper would have cut the headline number. Anthropic left it alone and cut a line most buyers have never looked at. That is information about where the money actually is.

Also mentioned: In November, alongside the YPO Global Business Summit in Istanbul, the YPO Technology Network is running a full-day AI Global Summit on 6 November. Stephen is speaking, along with people from Microsoft and other leading AI companies. Registration is open.

Sources:

Anthropic, Claude Fable 5.1 and Mythos 5.1, announced 1 September 2026. Pricing cross-verified across VentureBeat, TechSpot, implicator.ai and CybersecurityNews, plus the Claude Platform pricing documentation. The 25% / up-to-45% effective-cost figures are Anthropic's own estimate.