Free tool
Workload Cost Estimator
Estimate your monthly Claude API spend before you build — with prompt caching and Batch API discounts.
Get started-
Breaks down cost into fresh input, cache reads, and output — not one collapsed number.
-
Compares every current Claude model side by side.
-
Shows exactly what prompt caching and the Batch API save you.
Frequently asked questions
Prompt caching lets Anthropic reuse a previously-sent system prompt or context instead of reprocessing it, billing repeated (cached) tokens at a lower rate than fresh input tokens.
The Batch API processes requests asynchronously (results within 24 hours) at a discount off standard pricing — a good fit for non-realtime workloads like bulk summarization or classification.
Pricing comes from Costiva's admin-editable price list, shown on this page as "pricing last verified". Always confirm against Anthropic's official pricing page before making a purchasing decision.