Skip to main content
BenchLM
Data
Provider pricing hub

Claude API Pricing (October 2026)

Current Claude API pricing is $4/$20 per million input/output tokens for Claude Opus 5.5, $5/$25 for Claude Opus 5, $2/$10 for Sonnet 5, $1/$5 for Haiku 4.5, and $10/$50 for Fable 5.1. Fable 5.1 cache reads cost $0.25/M and Opus 5.5 cache reads cost $0.20/M; other current Claude cache hits cost 10% of standard input. The Batch API cuts input and output prices by 50%.

Last synced
35 confirmed releases in the last 30 daysFollow Anthropic price changes with RadarSources attached

The Claude API is Anthropic's developer interface for sending messages, images, tools, and documents to Claude from an application. The current self-serve model IDs are claude-opus-5-5, claude-fable-5-1, claude-opus-5, claude-sonnet-5, and claude-haiku-4-5-20251001. The table keeps those current rows ahead of previous models and links every verified rate to Anthropic's first-party price sheet.

For capability per dollar, use the Claude model rankings or the price-vs-performance view.

Citable stat5 current Claude API tiers span $1–$10 per million input tokens and $5–$50 per million output tokens as of October 5, 2026; Opus 5.5 is $4/$20.

Operator receipt

We checked Anthropic's September 22 Opus 5.5 model documentation, September 1 Fable 5.1 launch, and the current price sheet before publishing these rows. Opus 5.5 lists at $4/$20 with $0.20/M cache reads; Fable 5.1 keeps the $10/$50 base rate but cuts cache reads from $1/M to $0.25/M. Anthropic estimates that Fable change reduces typical usage-based Fable workloads by about 25%, with larger savings on cache-heavy work. The separate Enterprise Frontier Safeguards announcement says eligible customers can use zero data retention until its phased rollout reaches them.

Claude model pricing per 1M tokens

Claude API prices per one million tokens, by model.
ModelAPI model IDInput $/MCached input $/MOutput $/MBatch input $/MBatch output $/MContextEffectivePrice source
Claude Opus 5.5claude-opus-5-5Current$4$0.2$20$2$101MSeptember 22, 2026Claude Opus 5.5 model documentation
Claude Fable 5.1claude-fable-5-1Current$10$0.25$50——1MSeptember 1, 2026Anthropic Fable 5.1 launch
Claude Opus 5claude-opus-5Current$5$0.5$25$2.5$12.51MJuly 24, 2026Claude API pricing
Claude Sonnet 5claude-sonnet-5Current$2$0.2$10$1$51MJune 30, 2026Claude API pricing
Claude Haiku 4.5claude-haiku-4-5-20251001Current$1$0.1$5$0.5$2.5200K—Claude API pricing
Claude Mythos 5claude-mythos-5Limited access$10$1$50$5$251MJune 9, 2026Claude API pricing
Claude Opus 4.8—Previous$5—$25$2.5$12.51M——
Claude Opus 4.7—Previous$5—$25$2.5$12.51M——
Claude Opus 4.6—Previous$5—$25$2.5$12.51M——
Claude Opus 4.5—Previous$5—$25$2.5$12.5200K——
Claude Sonnet 4.6—Previous$3—$15$1.5$7.5200K——
Claude Sonnet 4.5—Previous$3—$15$1.5$7.5200K——
Claude Mythos Previewclaude-mythos-previewDeprecated limited access$25—$125$12.5$62.5—April 7, 2026Anthropic Project Glasswing
Claude 3 Opus—Legacy$15—$75$7.5$37.5200K——
Claude 4.1 Opus—Legacy$15—$75$7.5$37.5200K——
Claude 3.5 Sonnet—Legacy$3—$15$1.5$7.5200K——
Claude 4 Sonnet—Legacy$3—$15$1.5$7.5200K——
Claude Sonnet 5.5claude-sonnet-5-5Current$2$0.2$10$1$51MSeptember 28, 2026Claude API pricing
Claude 3 Haiku—Legacy$0.25—$1.25$0.125$0.625200K——

Current Claude API base rates are shown in USD per million tokens. Fable 5.1 cache reads are $0.25/M and Opus 5.5 cache reads are $0.20/M; other current Claude cache hits cost 0.1x base input. Five-minute and one-hour cache writes cost 1.25x and 2x. Batch columns apply Anthropic's 50% discount. US-only inference on Claude 4.6 and later adds 10%; fast-mode rates and tool charges are not included.

Claude API pricing calculator

Estimated monthly Claude API cost for the workload above, by model.
ModelEst. monthly cost
Claude Opus 5.5$80
Claude Fable 5.1$200
Claude Opus 5$100
Claude Sonnet 5$40
Claude Haiku 4.5$20
Claude Sonnet 5.5$40

Estimates use Claude's standard per-token rates from the table above. For token-level presets and cross-provider comparison, use the full LLM API pricing calculator.

Claude API quickstart: keys, model IDs, and docs

Start in the Claude Console, create an API key, then send a request to https://api.anthropic.com/v1/messages. Anthropic's official quickstart includes cURL, Python, and TypeScript examples. The API key has no subscription fee; the selected model, token counts, tools, and pricing modifiers determine the bill.

Use claude-opus-5-5 for the newest Opus tier, claude-opus-5 for the separately pinned prior Opus release, claude-sonnet-5 for Sonnet, and claude-fable-5-1 for Fable. Haiku's pinned ID is claude-haiku-4-5-20251001; claude-haiku-4-5 is its convenience alias. Anthropic says dateless IDs from the 4.6 generation onward are pinned snapshots, not pointers that silently move to a later model. The Models API returns the models available to your account.

Claude prompt caching and Batch API pricing

Prompt caching separates writes from reads. A 5-minute cache write costs 1.25x the model's base input rate, a 1-hour write costs 2x, and each cache hit costs 0.1x. Opus 5.5 is the exception on reads: its $0.20/M rate is 0.05x its base input price; a 5-minute write costs $5/M and a 1-hour write costs $8/M. Cache the stable system prompt, tool definitions, or document prefix when the same content will be read again.

Batch processing halves input and output for asynchronous jobs. Anthropic states that the Batch discount stacks with prompt-cache pricing, so a batched Opus 5.5 cache hit costs $0.10/M and batched output costs $10/M. Fast mode does not work with Batch. The calculator above supports cache and Batch together; compare providers with the LLM API pricing calculator.

Claude Code API pricing is separate from Claude plans

Claude Free, Pro, Max, Team, and Enterprise are subscription products. They do not create a shared pool of standard Claude API tokens. API usage is metered through a developer account and billed from input, cached-input, output, and any paid server-side tool usage.

Claude Code can run under an eligible subscription for interactive use or with API credentials for usage-based automation. When Claude Code uses an API key, the selected model's token rates apply. Use the table and calculator for API-key workloads; use the consumer plan page for a seat or subscription decision.

Subscription or API?

The question underneath most Claude pricing searches is not the token rate but whether to pay $20 or $100 a month for a plan or pay per token. The honest answer has a missing number in it. Anthropic publishes the API rates on this page to the cent, and publishes Pro and Max usage as multiples of a base allowance that is not itself stated in tokens. So the comparison people make on Reddit, "36x cheaper on a subscription", cannot be checked either way from Anthropic's own pages.

What can be checked is the API side. At Sonnet 5's $2/$10, one million input tokens and 250,000 output tokens cost $4.50; at Fable 5.1's $10/$50, the same tokens cost $22.50. If your month looks like that and stays inside a plan's cap, the plan is the cheaper product. If you run agents that write far more than they read, the API bill grows with output and the plan's cap arrives first. Run your own month through the calculator above, and compare the plan side on Claude's plans and their limits.

What you pay forClaude APIClaude Pro / Max
UnitTokens, at the rates on this pageFlat monthly fee: $20, $100 or $200
CapRate limits and a monthly spend cap by usage tier ($500 on Start)Usage caps published as multiples, not tokens
Past the capBilled at list rates up to the spend capExtra usage billed at API rates, or wait for the reset
Cache and Batch discountsCache reads at 2.5%–10% of input, by model; Batch 50% offNot applicable
Best fitProduction, agents, anything meteredIndividual interactive use inside the cap

Which Claude API model should you pay for?

Haiku 4.5 ($1/$5) is the low-cost tier for classification, extraction, routing, and short summaries. Sonnet 5 is the faster general-purpose tier at $2/$10, the launch rate Anthropic made permanent in August 2026. Opus 5.5 ($4/$20) is Anthropic's newest Opus option for long-running agentic coding and knowledge work. Fable 5.1 ($10/$50) is the highest-cost self-serve tier, with $0.25/M cache reads for context-heavy and long-running work.

The limit matters: the newest expensive model is not automatically the cheapest system. A routed workload can keep routine traffic on Haiku or Sonnet and escalate only failures to Opus or Fable. Fable loses when its quality gain does not offset a 2.5x input and output premium over Opus 5.5.

Claude vs OpenAI and Gemini API pricing

At the top tier, Fable 5.1 and GPT-6 Astra both list $10/$50 per million input/output tokens. Opus 5.5 at $4/$20 matches GPT-5.6 Sol's promotional rate, which OpenAI says runs through at least November 21, 2026. Sonnet 5 at $2/$10 matches GPT-6 Sol and costs less than GPT-5.6 Terra on output ($12), while Gemini 3.1 Pro lists at $2/$12 on prompts up to 200K tokens.

List price does not settle model quality, context behavior, or tokenization. Claude 4.7 and later use a tokenizer that Anthropic says produces about 30% more tokens for the same text, depending on the workload. Compare the full OpenAI and Gemini price tables, then rerun cost math with tokens measured from the model you plan to deploy.

Claude API pricing FAQ

How much does the Claude API cost?

Claude API pricing is $1/$5 per million input/output tokens for Haiku 4.5, $2/$10 for Sonnet 5, $4/$20 for Opus 5.5, $5/$25 for Opus 5, and $10/$50 for Fable 5.1. Opus 5.5 cache reads cost $0.20/M, Fable 5.1 cache reads cost $0.25/M, and Batch requests receive a 50% discount.

What is the Claude API?

The Claude API is Anthropic's developer interface for adding Claude to applications and agents. Developers create a key in the Claude Console and call the Messages API with a model ID such as claude-opus-5-5 or claude-sonnet-5. Usage is billed by tokens and paid tools, separately from Claude subscriptions.

What is Claude token pricing per million tokens?

Claude token pricing separates input, cache reads and writes, and output. Standard input/output rates range from Haiku 4.5 at $1/$5 per million tokens to Fable 5.1 at $10/$50. Opus 5.5 cache reads cost $0.20/M, Fable 5.1 cache reads cost $0.25/M, and other current Claude cache hits cost one-tenth of base input. Cache-write rates remain separate.

Why is Claude API pricing so expensive?

It is priced like its rivals at each tier. Fable 5.1 lists $10/$50 per million tokens, the same as GPT-6 Astra, and Sonnet 5 at $2/$10 matches GPT-6 Sol and costs less than GPT-5.6 Terra on output. The bill grows with the tier you pick, output length and missed cache hits, not with a Claude premium.

Is Claude API pricing included with Claude Pro or Max?

No. Claude Pro and Max are subscriptions for Claude's consumer and interactive coding surfaces; they do not supply a general pool of Claude API tokens. Requests made with an API key are metered separately at the model rates above. Claude Code can use a subscription or API credentials, depending on how it is configured.

Is a Claude subscription cheaper than the API?

They buy different things. Pro and Max are flat monthly fees with usage caps published as multiples, not token counts. The API bills every token at the rates above, up to your tier's spend cap. A subscription is cheaper only while you stay inside its cap, which has no published number. See Claude's plans.

Is there a free Claude API tier?

Anthropic's official API price sheet does not list a standing free production tier. The key itself has no monthly fee, but requests consume paid usage credits. Anthropic's pricing page says new users receive a small amount of free credits to test the API and does not state the amount. Claude's free consumer plan is a separate product and does not provide general API access.

Keep comparing