AI & Machine Learning
August 31, 2026Anthropic prompt caching pricing: write, read and TTL math
Anthropic prices prompt caching with three numbers: a 1.25x or 2x premium on cache writes depending on TTL, a 0.1x rate on cache reads, and the base input rate for everything after the last breakpoint. This deep-dive verifies every figure against the current official docs and covers per-model minimums, break-even math, batch stacking and how Claude on Azure converts it all into CCUs.
Prompt Caching
Anthropic
Claude
Claude API
LLM Cost Optimization
Microsoft Foundry
Azure
Batch API
By Falak Mahmood