Free tool

Workload Cost Estimator

Estimate your monthly Claude API spend before you build — with prompt caching and Batch API discounts.

Get started
  • Breaks down cost into fresh input, cache reads, and output — not one collapsed number.
  • Compares every current Claude model side by side.
  • Shows exactly what prompt caching and the Batch API save you.
Get started

Frequently asked questions

Prompt caching lets Anthropic reuse a previously-sent system prompt or context instead of reprocessing it, billing repeated (cached) tokens at a lower rate than fresh input tokens.

The Batch API processes requests asynchronously (results within 24 hours) at a discount off standard pricing — a good fit for non-realtime workloads like bulk summarization or classification.

Pricing comes from Costiva's admin-editable price list, shown on this page as "pricing last verified". Always confirm against Anthropic's official pricing page before making a purchasing decision.