Claude Fable 5 API pricing

Anthropic's flagship Claude 5 model — the Mythos-class tier above Opus.

VisionTool callingReasoningPrompt cachingPDF input

Input

$10.00

per 1M tokens

Cached input

$1.00

per 1M tokens

Cache write

$12.50

per 1M tokens

Output

$50.00

per 1M tokens

Context window

1M

max output 125K

Prices last synced Aug 20, 2026 from the LiteLLM community price catalog — the same catalog Marginal prices real usage events against, refreshed daily. Hand-checked against the provider's official pricing page on Aug 20, 2026. Spot an error? Tell us.

Pricing notes

  • Cache writes bill $12.50/1M (5-minute TTL) or $20.00/1M (1-hour TTL); cache reads $1.00/1M. Batch API halves input and output ($5.00/$25.00).
  • 1M-token context window at standard rates — no long-context surcharge. US-pinned inference (inference_geo: "us") bills a 1.1x multiplier.

What would Claude Fable 5 cost you?

Enter your workload — add other models to compare the same traffic across providers.

Your monthly workload

Cache hit rate = share of input tokens served from the provider's prompt cache, billed at the model's cached-input price. Estimates use standard per-token rates — batch discounts, service tiers, and long-context surcharges are not applied.

Claude Fable 5
ModelInputCached inputOutputPer requestMonthly
Claude Fable 5$6,000$400$12,500$0.038$18,900
Rates used: Claude Fable 5 $10.00 in / $50.00 out

Compare Claude Fable 5

More from Anthropic

Tracking Claude Fable 5 in production?

Marginal prices every call your app actually makes — cached tokens, dated snapshots, model switches included — and slices spend by customer or feature. One track() call to integrate.