← Back to all pricing

Claude Sonnet 5 API Pricing: Industry Standard for Autonomous Coding

Comprehensive Claude Sonnet 5 API pricing analysis ($2.00/M input, $10.00/M output), 1M context caching, SWE-bench coding dominance, and enterprise software ROI.

Full specs, context window and API limits →

How much does Claude Sonnet 5 cost per million tokens?

Claude Sonnet 5 costs $2.00 per million input tokens and $10.00 per million output tokens ($4.00/M blended at 3:1). Anthropic premier model for autonomous software engineering, complex system architecture, and deep code refactoring. Verified 2026-09-08.

Verified 2026-09-07 source
Input
$2.00/M
Output
$10.00/M
Blended
$4.00/M
Provider
Verified 2026-08-14source

How much does Claude Sonnet 5 cost per 1,000 requests?

Computed from generated token pricing. Each row assumes the listed input and output tokens per request; this model has no measured verbosity factor, so the unadjusted output estimate is shown.

Request shapeInput tokensOutput tokensCost / 1,000 requests
Short10050$0.7000
Medium1,000500$7.0000
Long4,0002,000$28.0000

Formula: ((input price × input tokens) + (output price × output tokens × verbosity factor)) ÷ 1,000,000 × 1,000. Assumptions: short 100/50, medium 1,000/500, long 4,000/2,000 input/output tokens per request. Unadjusted — no measured verbosity factor is available.

Batch 61 · exact-model pricing decision contributions · verified 2026-09-07

Exact model boundary: Anthropic Claude Sonnet 5 (claude-sonnet-5). Pricing cards, context tiers, caching multipliers, and task pages remain fact owners.

5-minute ephemeral cache read/write multiplier table

Frozen Batch 61 scenario board. Formula / deterministic rule: net_input_cost = write_tokens * 1.25 * rate_in + read_tokens * 0.10 * rate_in; cache TTL = 300s Boundary: Owns Anthropic prompt caching economics for Sonnet 5.

Frozen scenario / field IDExact identity and evidence fieldsResultState
batch61-claude-sonnet-5-m1-r1
uncached single turn
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=uncached single turn; prompt tokens; cache write units; cache read units; base rate; effective token price; net session cost; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — uncached single turn is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m1-r2
initial cache write turn
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=initial cache write turn; prompt tokens; cache write units; cache read units; base rate; effective token price; net session cost; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — initial cache write turn is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m1-r3
subsequent cache read turn
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=subsequent cache read turn; prompt tokens; cache write units; cache read units; base rate; effective token price; net session cost; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — subsequent cache read turn is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m1-r4
5-turn agent loop session
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=5-turn agent loop session; prompt tokens; cache write units; cache read units; base rate; effective token price; net session cost; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 5-turn agent loop session is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m1-r5
cache TTL expiration
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=cache TTL expiration; prompt tokens; cache write units; cache read units; base rate; effective token price; net session cost; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — cache TTL expiration is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m1-r6
unsupported cache boundary
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=unsupported cache boundary; prompt tokens; cache write units; cache read units; base rate; effective token price; net session cost; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — unsupported cache boundary has no matched, dated bilateral observation.FAIL CLOSED — manual, probe, or source evidence required

First-party provenance: Anthropic official API pricing; verification date 2026-09-07. Missing or conflicting joins fail closed.

Interactive vs asynchronous Batch API cost frontier

Frozen Batch 61 scenario board. Formula / deterministic rule: savings = standard_cost * 0.50; SLA = 24h turnaround window Boundary: Owns Batch API workload trade-off modeling for Sonnet 5.

Frozen scenario / field IDExact identity and evidence fieldsResultState
batch61-claude-sonnet-5-m2-r1
interactive immediate call
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=interactive immediate call; job type; call volume; interactive cost; batch cost; net savings; latency trade-off; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — interactive immediate call is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m2-r2
10K batch evaluation calls
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=10K batch evaluation calls; job type; call volume; interactive cost; batch cost; net savings; latency trade-off; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 10K batch evaluation calls is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m2-r3
50K batch code review jobs
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=50K batch code review jobs; job type; call volume; interactive cost; batch cost; net savings; latency trade-off; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 50K batch code review jobs is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m2-r4
mixed interactive/batch traffic
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=mixed interactive/batch traffic; job type; call volume; interactive cost; batch cost; net savings; latency trade-off; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — mixed interactive/batch traffic is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m2-r5
urgent SLA override
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=urgent SLA override; job type; call volume; interactive cost; batch cost; net savings; latency trade-off; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — urgent SLA override is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m2-r6
unresolved batch request
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=unresolved batch request; job type; call volume; interactive cost; batch cost; net savings; latency trade-off; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — unresolved batch request has no matched, dated bilateral observation.FAIL CLOSED — manual, probe, or source evidence required

First-party provenance: Anthropic official API pricing; verification date 2026-09-07. Missing or conflicting joins fail closed.

Multi-step coding agent token envelope and loop budget

Frozen Batch 61 scenario board. Formula / deterministic rule: loop_cost = sum(turn_tokens_in * rate_in + turn_tokens_out * rate_out); context accumulation tracked Boundary: Owns multi-turn software engineering token economics.

Frozen scenario / field IDExact identity and evidence fieldsResultState
batch61-claude-sonnet-5-m3-r1
3-step quick bug fix
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=3-step quick bug fix; step count; accumulated context; generated tokens; turn cost; cumulative spend; budget status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 3-step quick bug fix is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m3-r2
5-step feature implementation
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=5-step feature implementation; step count; accumulated context; generated tokens; turn cost; cumulative spend; budget status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 5-step feature implementation is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m3-r3
10-step full-file refactoring
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=10-step full-file refactoring; step count; accumulated context; generated tokens; turn cost; cumulative spend; budget status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 10-step full-file refactoring is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m3-r4
tool schema overhead (30K)
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=tool schema overhead (30K); step count; accumulated context; generated tokens; turn cost; cumulative spend; budget status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — tool schema overhead (30K) is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m3-r5
context window limit reach (500K)
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=context window limit reach (500K); step count; accumulated context; generated tokens; turn cost; cumulative spend; budget status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — context window limit reach (500K) is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch61-claude-sonnet-5-m3-r6
exhausted loop budget
model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=exhausted loop budget; step count; accumulated context; generated tokens; turn cost; cumulative spend; budget status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — exhausted loop budget is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict

First-party provenance: Anthropic Claude documentation; verification date 2026-09-07. Missing or conflicting joins fail closed.

Method and limitations: formulas are deterministic; observed and assumed inputs are labeled; no missing provider, host, account, region, realm, alias, snapshot, revision, weight, artifact, control, tool, modality, workload, rate-period, timestamp, or result is transferred. Run the Claude Sonnet 5 Batch 61 scenario →

Continuous SEO Builder · Batch 73 Audit · 2026-09-08Owner: claude-sonnet-5

Claude Sonnet 5 API Pricing: Industry Standard for Autonomous Coding

Claude Sonnet 5 costs $2.00 per million input tokens and $10.00 per million output tokens ($4.00/M blended at 3:1). Anthropic premier model for autonomous software engineering, complex system architecture, and deep code refactoring. Verified 2026-09-08.

Module 1 · Claude Sonnet 5 Autonomous Coding Unit Economics
Blended Cost = (Input Tokens × $2.00 + Output Tokens × $10.00) / 1,000,000

Claude Sonnet 5 provides industry-leading software engineering capabilities at $4.00/M blended.

Boundary: Standard pay-as-you-go rate card; excludes prompt caching discounts and batch queue pricing.
ScenarioRendered Evidence & Bounds
Scenario 1Multi-file bug isolation and unit test fix (20K in, 3K out): $0.07000 per engineering task
Scenario 2Full-stack feature implementation PR (40K in, 6K out): $0.14000 per PR generated
Scenario 3Large monorepo architectural refactor (80K in, 8K out): $0.24000 per refactor pass
Scenario 4Automated code review across 10 pull requests (30K in, 2K out): $0.08000 per review set
Scenario 5Security vulnerability audit of third-party dependencies (15K in, 2K out): $0.05000 per audit
Scenario 6Monthly developer seat allocation (50M blended tokens): $200.00 predictable cost ceiling
Module 2 · Claude Sonnet 5 Repository Caching & 90% Amortization
Cached Cost = (Cached Repository × $0.20 + Delta Input × $2.00 + Output × $10.00) / 1,000,000

Prompt caching transforms entire codebase context from a luxury into routine developer infrastructure.

Boundary: 90% discount on prompt prefixes >1,024 tokens held in Anthropic 5-minute rolling cache.
ScenarioRendered Evidence & Bounds
Scenario 1Monorepo codebase cache (120K tokens prefix, 5K delta): 83% input cost savings
Scenario 2Shared developer system prompt and coding standards (10K prefix): $0.00200 vs $0.02000 per call
Scenario 3Interactive IDE developer chat session (12 turns cached): 82% cumulative input savings
Scenario 4Cache write fee ($2.50/M) fully amortized after only 1.25 repeat requests within TTL
Scenario 5Zero latency prompt retrieval: cuts time-to-first-token by 55% during active coding sessions
Scenario 6Net engineering infrastructure spend reduced by over 68% for active developer teams
Module 3 · Claude Sonnet 5 Software Defect Elimination & Developer ROI
Net Developer ROI = (Engineering Hours Saved × $150) - Sonnet 5 API Spend

Frontier coding precision yields massive developer productivity and software defect reduction.

Boundary: Quantifies measurable developer productivity gains and production outage prevention value.
ScenarioRendered Evidence & Bounds
Scenario 1Saves an average of 4.5 developer hours per week on boilerplate and test generation
Scenario 2Catches edge-case race conditions and memory leaks prior to production staging deployment
Scenario 3Single prevented production incident ($50,000 outage cost) offsets annual team API spend
Scenario 4First-attempt PR acceptance rate exceeds 88% on standardized ticket specifications
Scenario 5Achieves top rankings on SWE-bench verified benchmarks with minimal human intervention
Scenario 6Estimated team net ROI exceeds 25x total annual Anthropic API expenditures
Explore Related Analyses:Anthropic provider profileCompare vs Claude Sonnet 4.6Compare vs Claude Opus 5Best LLM for coding

How fast is Claude Sonnet 5?

Not yet measured — see the speed benchmark leaderboard for models we do track.

How much does Claude Sonnet 5 cost at scale?

Tokens / monthEst. cost (blended 3:1)
100,000$0.40
1,000,000$4.00
10,000,000$40.00
100,000,000$400.00

How does Claude Sonnet 5 compare with other models?

Claude Haiku 4.5$2.00/MClaude Sonnet 4.6$6.00/MClaude Sonnet 4.5$6.00/MClaude Sonnet 4$6.00/MClaude Opus 4.8$10.00/MGPT-4o$4.38/MGPT-4.1$3.50/MGemini 3.1 Pro$4.50/M
See all Anthropic models →

What is Claude Sonnet 5 best for?

#12 for Agents & Tool Use#15 for Math & Reasoning#19 for Image Understanding
Looking for a cheaper option?
Gemini 3.5 Flash Lite is 78.8% cheaper — a code-change migration. See all 8 alternatives to Claude Sonnet 5

Which Claude Sonnet 5 head-to-head comparisons are available?

Claude Sonnet 5 vs DeepSeek V4 ProClaude Sonnet 5 vs Gemini 3.1 ProClaude Sonnet 5 vs Grok 4.3Claude Sonnet 5 vs Claude Opus 4.8

What are common questions about Claude Sonnet 5?

Is Claude Sonnet 5 cheaper than GPT-4o?

Claude Sonnet 5 costs $4.00/M blended tokens, GPT-4o costs $4.38/M — Claude Sonnet 5 is cheaper.

How much does 1 million tokens cost with Claude Sonnet 5?

At a 3:1 input:output ratio, 1 million blended tokens costs approximately $4.00. Pure input costs $2.00/M; pure output costs $10.00/M.

What does Claude Sonnet 5 cost at high volume?

At 100 million blended tokens a month, Claude Sonnet 5 costs approximately $400.00. See the cost-at-scale table below for other volumes.

Try Claude Sonnet 5 for free

Run real prompts against Claude Sonnet 5 and every other model on this page in one workspace.

Try Claude Sonnet 5 Free